LIVE

Rheostat Summit 2026 — twelve talks on shipping agents safely. Oct 14, San Francisco.

Save your seat

Rheostat

0 Book a demo
Release observability

See what every
release actually did.

Every chart knows which flag changed and when. Spot the regression, click the marker, dial it back.

Error rate0.02%↓ 0.11 since 14:02
p95 latency212 ms↓ 700 ms after rollback
Releases today386 in progress
Auto-rollbacks3all within 90 s

p95 latency · checkout

Last 6 h
v4.2 → 25%Auto rollback

Live releases

  • new-checkoutrolled back25%
  • search-ranker-v3healthy50%
  • agent-prompt-v15healthy5%
  • pricing-page-bcomplete100%
  • onboarding-tourwatching10%

Watch a regression
dial itself back.

Two minutes, one bad deploy: the latency spike, the marker on the chart, the automatic rollback and the all-clear.

Three questions every
release review asks.

Which flag caused it?

Every metric is sliced by flag and variant. The answer is one click away, not one meeting away.

errors / 1k requests · by variant

new-checkout · on4.1
new-checkout · off0.6
search-ranker · on0.9

exposure at detection

1%
5%
25%
50%
100%

caught at 5% · 412 users affected · rolled back in 38 s

Did we catch it early?

Regressions show up at 5% of traffic, when a few hundred people see them — not a few million.

What should we do now?

Alerts arrive with the fix attached. Dial back, pause the experiment or page the owner — from Slack or the dashboard.

p95 latency > 800 msSev 2

Started 14:02 · new-checkout moved to 25% at 14:01

Rheostat vs. the flags
you built in-house.

CapabilityRheostatIn-house
Metrics sliced by flag and variantYesNo
Automatic rollback on a ruleYesNo
Percentage rollouts with schedulesYesYes
Agent traces linked to releasesYesNo
Audit log of every changeYesYes
Evaluation in under 20 ms at the edgeYesNo

On-call got quieter.

“The release marker on the latency chart ended every “was it the deploy?” thread before it started.”

Priya RamanHead of Platform, Northwind

“We cut our mean time to recover from 47 minutes to 11. Most of that is the rollback nobody had to press.”

Dana WhitfieldStaff Engineer, Kilnworks

“One dashboard for flags, errors and latency. Our incident channel went from daily to monthly.”

Aiko MoriPlatform Lead, Meridian

Questions about metrics.

Which metrics sources work?
Prometheus, Datadog, New Relic, Honeycomb and anything that speaks OpenTelemetry. Custom metrics arrive over a webhook.
Do you store my metrics?
We keep 30 days of release-scoped aggregates. Raw data stays in your own observability stack.
How fast does a rollback fire?
Rules evaluate every ten seconds at the edge. Most rollbacks complete in under ninety seconds.
Can I tune the thresholds?
Every rule has its own threshold, window and minimum traffic, per flag and per environment.

Stop reading dashboards
after the incident.

−61%
median time to recover across 1,900 teams in 2026