Skip to main content
Metrics show how much traffic a gets, how fast it responds, and how much CPU and memory it uses. You don’t need to add anything to your code. Open the , click Deployments, and open a deployment. Its overview has the charts, and its Network view shows where requests are going.

Traffic

Traffic charts cover the last 6 hours in 15 minute buckets.
  • Requests per second: the average over the last 6 hours.
  • Latency: the total time of each request, as in the request log. Pick p50, p75, p90, p95, or p99.
Like the request log, these don’t count requests a policy rejected.

Resources

Resource charts add up all the deployment’s instances. You can pick a time window from 15 minutes to a year. Compare CPU and memory with the limits in runtime settings to decide whether to raise them or add instances. See Instances and autoscaling. Longer windows show coarser points: 15 seconds up to an hour, 1 minute up to a day, 1 hour up to 30 days, and 1 day beyond that. Fine-grained points are deleted first. See retention.

Why an instance restarted

The deployment lists instance events: an instance becoming running, waiting, or terminated, with its restart count, exit code, and a reason. Two reasons to know:
  • OOMKilled (often with exit code 137): the instance ran out of memory. Raise the memory limit.
  • CrashLoopBackOff: the process keeps exiting before the health check passes. The runtime logs for that instance say why.
Events are kept for 90 days.

Network view

The Network view shows each region and running instance with its request rate over the last 15 minutes. Use it to check that every region gets traffic and that load is spread across instances. See Regions.
Deployment page Network view showing a region and its running instance
Last modified on September 29, 2026