What Should Remain in Your Harness as Large Language Models Get Stronger?
The article analyses a recent 61% reduction of a team’s tdsql‑harness, compares it with Anthropic’s 80% cut of Claude Code prompts and OpenAI’s new guidance, and derives concrete criteria for deciding which harness rules to keep, rewrite, or discard as LLMs become more capable.
