Skip to content

Monitoring

Overview · Updated May 2026

Coming from another cloud?

▸AWS·Cloudwatch

This Quake AI feature maps to AWS’s Cloudwatch.

▸DigitalOcean·Monitoring

This Quake AI feature maps to DigitalOcean’s Monitoring.

Monitoring

Monitoring is the practice of collecting, analyzing, and acting on data about your infrastructure and applications. On Quake AI, monitoring is a customer responsibility: the platform provides the compute, network, and storage resources, and you instrument and observe your workloads.

This section covers what to monitor, how to set up collection pipelines, and when to alert.

What to monitor#

Effective monitoring covers four layers:

  • Infrastructure: CPU, memory, disk I/O, and network throughput on your instances. These metrics tell you whether your resources are sized correctly and when to scale.
  • Application: request rates, error rates, latency, and business-specific metrics. These tell you whether your software is working correctly from your users' perspective.
  • Logs: system logs (syslog, auth.log), application logs, and audit trails. Logs provide the diagnostic detail that metrics alone cannot capture.
  • Quotas and billing: resource usage against your project quotas. Unexpected quota consumption may indicate runaway automation or compromised credentials.

Templates and tutorials#

See also#

Was this page helpful?