Operating manual
Tag: Agent Reliability
A focused reading path for Agent Reliability: related field notes, evidence trails, and operating questions from the archive.
- the-last-wrong-step-is-not-where-the-agent-failure-started.md
The Last Wrong Step Is Not Where Failure Started
A four-stage operator model for tracing long-horizon agent errors from their initiating deviation through propagation, resolution, and terminal impact.
open artifact → - voice-agents-need-a-different-reliability-test.md
Voice Agents Need a Different Reliability Test
Voice agent reliability should track intent from capture through action. This six-stage operator framework finds where meaning first breaks.
open artifact → - the-next-agent-bottleneck-is-operational-control-not-model-capability.md
The next agent bottleneck is operational control, not model capability
Frontier models keep improving. The agent bottleneck is moving to state, permissions, evals, recovery, and proof that work actually shipped.
open artifact →