Skip to main content

19 docs tagged with "Metrics"

View all tags

Alerting on Worker metrics

A recommended alert set for Temporal Workers, with tag filters, thresholds, and links to triage guidance

Monitor Temporal Cloud

Detect Task Queue backlogs, Worker capacity issues, and Temporal Cloud service errors, then route alerts to your monitoring tools.

Monitor Worker health

Detect and configure for Task backlogs, greedy Worker resources, misconfigured Workers, and Sticky cache settings. Optimize alert systems and get actionable insights on metrics like Schedule-To-Start latency, Sync Match Rate, and Poll Success Rate for improved application health.

Observability

Query live and closed Workflow Executions by your own business identifiers, export Prometheus-compatible metrics, and trace Executions across Worker processes.

OpenMetrics FAQ

Answers to common questions about querying, scraping limits, and missing Temporal Cloud OpenMetrics data.

Retry Alerting via Metrics

Emit a metric from the Activity when attempts cross a threshold, so on-call teams see persistent failures before an SLA breach.

Troubleshoot missed Schedule Actions

Diagnose missed or delayed Schedule Actions caused by overlap policies, buffering, cancellation, and Catchup Window expiry, with metrics and examples.