What if your agent's hallucinations had a budget? How to start using SLOs for agent behavior

From the source: Grafana Labs Blog

At Grafana Labs, observability is what we do. So as we started building AI agents, we naturally reached for the same instincts we bring to every system: measure it, set targets, and make reliability something you can reason about instead of hope for. That instinct led us somewhere unexpectedly useful. It turns out one of the oldest ideas in reliability engineering, the error budget, maps…

Read the full story on Grafana Labs Blog
Originally published by Grafana Labs Blog on 25 Sept 2026. Techarda links to the original rather than republishing it. Read the full article →

Have you worked with this?

The story is what was announced. Nobody has discussed it yet, so if it touches your team, a short post about what you’ve seen helps the next reader.

Start the discussion

More from Grafana Labs

Recent updates from Grafana Labs, so you can tell whether this is a one-off or part of a pattern.

All Grafana Labs news →
Observability

Grafana Alerting: Scale alert routing without scaling complexity using multiple notification policies

Alert routing often starts simple. A team creates a few contact points, adds some label matchers, and builds a notification policy tree that sends each alert to the right destination. But alerting configurations rarely stay simple. As an organization grows, its notification policy tree must accommodate more teams, services, and routing requirements. Changes for one team still require editing a…

Grafana Labs·via Grafana Labs Blog
Observability

Digital Experience Monitoring with Grafana Cloud: Session Replay, synthetic checks, and faster investigations

When something breaks in production, the questions that matter most are also the toughest to answer from metrics alone: who was affected, what did they actually see, and is this worth waking someone up for? Answering those questions requires a fuller picture of the issue and its impact on your users. That’s where Digital Experience Monitoring (DEM) in Grafana Cloud comes in. By combining Frontend…

Grafana Labs·via Grafana Labs Blog
Observability

Custom labels in Grafana Cloud Synthetic Monitoring: New updates for consistency and ease-of-use

Labels are a powerful way to organize telemetry and define policies across Grafana Cloud, helping to streamline alerting, attribution, access control, and more. But traditionally, custom labels in Synthetic Monitoring have worked a little differently: they only lived on a single sm_check_info metric, and Grafana Cloud prefixed each one with label_ . To make custom labels in Synthetic Monitoring…

Grafana Labs·via Grafana Labs Blog

More in Observability

What other companies in Observability are doing. The category page shows who’s active, side by side.

Compare companies in Observability →
Observability

Extend Datadog RUM and Product Analytics to Shopify and Salesforce

Use Datadog RUM and Product Analytics to monitor checkout journeys on Shopify and customer experiences in Salesforce Experience Cloud.

Datadog·via Datadog Blog
Observability

DevRel newsletter — September 2026

Hello from the Elastic DevRel team! In this newsletter, we cover jina-ocr-v1, the latest blogs and videos, and upcoming events like Elastic{ON}.

Elastic·via Elastic Blog
Observability

Cribl Stream and Databricks: A faster path to analytics-ready telemetry

The new Databricks Zerobus Destination for Cribl Stream helps organizations move analytics-ready telemetry into Delta Tables with fewer moving parts, less latency, and more control.

Cribl·via Cribl Blog
Observability

Using TypeSafe’s Jev for evals in Datadog Agent Observability

See how we wired Jev into online evals on live spans and offline evals inside Datadog experiments, using a single rubric for both.

Datadog·via Datadog Blog