SHARED STATE
Agent recovery trace
Three attempts changed the same files and reproduced identical failures. Valid observations contradicted a necessary prediction of the working hypothesis.
- Attempts3
- Repeated failureYes
- BudgetMedium
SHARED-STATE DECISION MODEL
Give Blitz one view of the situation. Ask a panel of focused questions. Get explicit distributions that software can compose into its next action—without generating a paragraph.
01 / SHARED-STATE FANOUT
Blitz acts like a semantic sensor panel. Each question reads the same evidence independently; ordinary code combines the resulting distributions.
SHARED STATE
Three attempts changed the same files and reproduced identical failures. Valid observations contradicted a necessary prediction of the working hypothesis.
YES81.9%
NO96.8%
YES80.0%
NO62.3%
CODE COMPOSES
if same_root_cause > .80
and retry_helps < .20
and hypothesis_invalid > .70:
→ change_strategy
02 / INSPECT THE PRIMITIVE
A workflow is built from small, focused judgments. Start with a real checkpoint example, then change the evidence, question, or options.
Your decision appears here
Selected decision
Decision receipt
03 / MEASURED, NOT PROMISED
Fresh confirmation results, a control baseline, and measured warm serving receipts from the current research checkpoint.
Synthetic held-out evaluation. The confidence threshold is post-hoc and not deployment-validated; these are not claims of real-world accuracy or calibration.
WHERE BLITZ FITS
Detect repeated failures and choose whether to retry, review, or change strategy.
Judge severity, route ownership, and decide when ambiguous evidence needs escalation.
Score retrieval relevance, context coverage, semantic equivalence, and experiment outcomes at scale.