01:30
2026-09-11
aiflash.com
ai-agents
T1: Terminal Agent Reinforcement Learning for Long-Horizon Tasks
A Mixture-of-Experts model called T1, with 122B total parameters, has been trained with reinforcement learning to operate a real shell in a cloud sandbox for up to 300+ tool calls on long-horizon term…