Skip to main content

20 docs tagged with "Metrics"

View all tags

Alerting on Worker metrics

A recommended alert set for Temporal Workers, with tag filters, thresholds, and links to triage guidance

Monitor Temporal Cloud

Detect Task Queue backlogs, Worker capacity issues, and Temporal Cloud service errors, then route alerts to your monitoring tools.

Monitor Worker health

Detect and configure for Task backlogs, greedy Worker resources, misconfigured Workers, and Sticky cache settings. Optimize alert systems and get actionable insights on metrics like Schedule-To-Start latency, Sync Match Rate, and Poll Success Rate for improved application health.

Observability

Query live and closed Workflow Executions by your own business identifiers, export Prometheus-compatible metrics, and trace Executions across Worker processes.

OpenMetrics FAQ

Answers to common questions about querying, scraping limits, and missing Temporal Cloud OpenMetrics data.

Retry Alerting via Metrics

Emit a metric from the Activity when attempts cross a threshold, so on-call teams see persistent failures before an SLA breach.

Temporal SDK metrics reference

Reference for the SDK metrics with guaranteed, defined behavior across Temporal Clients and Workers, including each metric's type and availability.

Troubleshoot missed Schedule Actions

Diagnose why a Schedule did not fire by alerting on the missed catchup window metric, then narrow down to the affected Schedule with ListSchedules and DescribeSchedule.