ModularRSI: Agents Self-Improve by Evolving Harness, Not Model Weights
ModularRSI demonstrates that AI agents can continuously self-improve by evolving their harness—the surrounding system mechanisms like tool use, context management, and task completion detection—while keeping the base model frozen, achieving performance gains on Terminal-Bench and SWE-Bench that transfer across domains and models.
