When an agent is blocked and told why, does it correct itself?
Yes, 11 times out of 11
Claude Code and Codex, blocked from a protected file with a concrete alternative named.
Evidence
trackline's design rests on a few claims that would be expensive to get wrong. Each was measured before it was built on. Where a result has limits, the write-up names them next to the number.
Yes, 11 times out of 11
Claude Code and Codex, blocked from a protected file with a concrete alternative named.
About 14 ms, in Go
88 ms in Node against 6.5 ms for a bare Go start, across 265 tool calls; the full hook runs in 11 to 14 ms.
Only with content capture on, and it survives redaction
The official OpenTelemetry instrumentation, captured under four configurations.
0 false alarms in 23 actions
Ordinary sessions replayed through every check, and real usage that found what scripts missed.
8 of 8 on a calibration set
Superseded by experiment 6, which measured it properly.
60 of 60 held-out drifts, against 2 for the rules
Three agents, labels committed before any session ran, a judge from a different model family.
Yes, with zero lines of the core changed
45 pre-registered conversations sent by the real OpenTelemetry libraries: every labelled problem caught, none of 29 on-task flagged.