# T1: Terminal Agent Reinforcement Learning for Long-Horizon Tasks

> Source: <https://aiflash.com/news/117367/>
> Published: 2026-09-11 01:30:12+00:00

Agent usage is shifting toward long-horizon tasks such as coding and scientific discovery, among which terminal tasks are especially important. We introduce T1, a Mixture-of-Experts model of 122B total trained with reinforcement learning, operating a real shell in a cloud sandbox for up to 300+ tool
