Do your agent system prompts do anything? I measured 19 of mine

A null-prompt ablation across 19 agent configurations: 13 survive the noise floor, two of them scoring zero without their prompt; four say more about the tests than the prompts. Anthropic runs the same kind of ablation on Claude Code and reports the opposite direction. The difference is scope.

Read Original

Related