Paid orders
Order query
No successful query yet; this page never shows simulated success data.
Final Demo · PostgreSQL pool exhaustion
A PostgreSQL application pool-exhaustion case connects real business queries, alerts, multi-agent diagnosis, controlled preview, human approval, repair, and independent verification. The static shell stays available while database cards degrade truthfully.
Environment
Not initialized
incident pending
The request first asks Manager to create an incident; Manager then invokes the controlled fixture. This page holds no database or fixture credentials.
Three cards poll independently through the same stressed application pool.
3-second polling · 2-second client timeout
Paid orders
No successful query yet; this page never shows simulated success data.
Recent warehouse
No successful query yet; this page never shows simulated success data.
Audit events
No successful query yet; this page never shows simulated success data.
Healthy cards
0/3
Degraded cards
3/3
Avg last-success latency
Pending
Manager owns authoritative state. Preview approval eligibility never replaces human approval.
Scenario start
starting
Awaiting alert
awaiting_alert
Alert correlated
alert_correlated
Diagnosis
diagnosis_dispatched
Preview ready
preview_ready
Awaiting approval
awaiting_approval
Repair
repair_dispatched
Verification
verifying
Recovered
recovered
Archived
closed
Start failed
start_failed
Compact evidence before HITL; the full comparison is available in the archive after closure.
Controlled workload boundary: preview-pg replays a fixed workload and never copies active sessions from the source instance.
Not eligible for human approval
Preview rejected · No human approval
Monitoring
Primary board: verify pool 4/4, wait queue, and query errors
https://teams.yueming.xin/grafana/d/opskeeper-pgpool-live/?orgId=1&from=now-15m&to=now&refresh=5s
Manager status
Secondary board: Manager health, HTTP latency, and internals
https://teams.yueming.xin/grafana/d/opskeeper-manager-internals/opskeeper-manager-internals?orgId=1&from=now-1h&to=now&refresh=30s
Element room
Watch Manager coordinate diagnosis, preview, and repair
https://rooms.yueming.xin/#/room/#benyue-lumos-ops:matrix-local.agentteams.io:18080
AgentTeams Dashboard
Inspect the task board and OpsKeeper runtime
https://teams.yueming.xin/#plugin-route:opskeeper-teamharness/home
OpsKeeper console
Switch to the standard demo user to review incidents, approvals, audits, and read-only settings
https://opskeeper.yueming.xin/login?account=switch
Incident archive
Review the full evidence chain and A/B comparison after closure
https://teams.yueming.xin/#plugin-route:opskeeper-teamharness/archive
Preview report
Open the preview-pg deep link to verify all metrics
https://opskeeper.yueming.xin/preview/
Only two nodes require human intervention. Injection, diagnosis, preview, and verification are coordinated by Manager and Workers.
Approve Candidate A for incident <incident_id>. Verify candidate_id, execution_id, target fingerprint, scope, parameters, and expiry. Preview PASS only grants approval eligibility; it is not human approval.Close incident <incident_id>: confirm business queries, pool metrics, and independent verification, then archive the evidence chain and A/B preview comparison.