Do your agent system prompts do anything? I measured 19 of mine
A null-prompt ablation across 19 agent configurations: 13 survive the noise floor, two of them scoring zero without their prompt; four say more about the tests than the prompts. Anthropic runs the same kind of ablation on Claude Code and reports the opposite direction. The difference is scope.