Risk analysis
120sRollback touches payment-service and its retry queue.
An AI runbook executor that gathers evidence, runs safe diagnostics, and asks before production changes.
Risk score
Checkout incident
Based on logs, metrics, and latest deploy.
Evidence trail
Runbook: checkout-failure
Alert received
Checkout error rate increased on payment-service.
Evidence gathered
Logs, deploy history, and metrics agree on one likely cause.
Sandbox check
Diagnostic script reproduced timeout failure in isolation.
Approval required
Rollback stays locked until an engineer approves it.
Approval gate
Action: rollback payment-service
timeout_ms=3000 failed_requests=47 likely_commit=8f31c2b recommendation=rollback
RunProof turns a page into a staged workflow: gather the signal, replay the failure, and approve the action with proof attached.
Current gate
Rollback payment-service stays locked until an engineer approves the evidence packet.
Use agents for investigation and diagnosis while keeping the production boundary explicit.
Policy-aware by default
RunProof connects the parts of an incident that usually stay scattered: issue context, logs, runbooks, sandbox output, and the final approval gate.
Every alert becomes a structured workspace holding the deploys, logs, metrics, and runbook rules needed to reason about the fix.
Rollback touches payment-service and its retry queue.
Checkout failure, likely cause, and operator decision in one place.
Deploy diffs, logs, and metrics collected before any action.
Connect the tools that create incident context, then move the work through one visible proof and approval flow.
Source systems
RunProof
proof layer
Reviewable output
Evidence packet
Deploys, traces, logs, and metrics are grouped for review.
Sandbox replay
Diagnostics run before production is touched.
Approval request
The action stays locked until a person approves.
Create an account, open the console, and walk through an evidence-gated runbook from alert to approval.