Your delegated runs are measurably the most reliable work in your history
Agent runs error at 16.2 per thousand tool calls and draw 3.1 interrupts per hundred sessions. Your hands-on sessions error at 28.4 and draw 22.7 interrupts, seven times the rate, on comparable volume. This also kills a plausible-looking story: the apparent jump in interrupts after April is an artifact, because the earlier client cannot record an interrupt at all and reports zero for all 244 of its sessions.
- agent n=891: 16.2 err/1k, 3.1 interrupts/100 sessions | hands-on n=712: 28.4 err/1k, 22.7 int/100python over sessions.jsonl grouped by session-id prefix
- 0 interrupts recorded across all 244 sessions of the earlier client, which has no interrupt event in its formatstats --what shape