AutoSaddler: Automatic Harness Optimization with Durable Updates from Agent Execution Traces
This paper develops a system to automatically optimize the design of harnesses for LLM agents, which can improve their reliability on long-horizon tasks. Practitioners might care about this because it could lead to more robust and performant agent systems.