← Agent workflows

Submitted test-run reconciliation

Test run stability matrix

Compare an explicit baseline and candidate run boundary, classify consistent and intermittent outcomes, count adjacent transitions, summarize submitted durations exactly, and produce a deterministic review queue without executing tests.

Explicit-submit workbench

Run submitted JSON locally

Parsing and transformation happen in this browser. Nothing runs when a sample loads, and no input is fetched, uploaded, stored, or retained.

Load the inert sample or enter JSON, then explicitly run.

Hosted API and MCP parity

The same tool ID is available through POST /api/run and the fixed MCP run_tool method. Hosted requests are response-only. This browser workbench does not call either endpoint.

{
  "tool_id": "test-run-stability-matrix",
  "input": "<SCHEMA_VALID_INPUT>",
  "response_mode": "full"
}