Operating manual
Tag: AI Safety
A focused reading path for AI Safety: related field notes, evidence trails, and operating questions from the archive.
- validate-ai-agent-trajectories.md
How to Validate AI Agent Trajectories
Component tests prove local properties. Deployable AI agents need evidence for the path they take across state, authority, tools, memory, and handoffs.
open artifact → - ai-rd-artifact-monitor-benchmarks.md
AI R&D Needs Two Benchmarks: Artifact and Monitor
ResearchArena shows why AI-produced models, kernels, and servers need adversarial artifact tests—and a separate benchmark for the monitor.
open artifact →