T1: Terminal Agent Reinforcement Learning for Long-Horizon Tasks
This paper introduces a new AI model called T1, which uses reinforcement learning to perform long-term tasks such as coding and scientific discovery. Practitioners may care about this model because it can perform better than existing models on long-term tasks.