An anonymous AI model has taken the developer community by storm. And it's reportedly giving frontier labs a run for their money.
On Thursday, a "stealth model" called Ox Alpha was released on OpenRouter by a third-party provider who decided to remain anonymous during its preview. Its creator said that the reasoning model was designed for coding, sustained agentic work and production workloads.
The model is available for free in preview with "near unlimited usage" for a week, according to OpenCode, with an eye-popping capacity for 100 trillion tokens per day. The model features a 1.05-million token context window and 131,000-token output, or the limit that a model can generate in a single response.
- The creator noted that the model is suited for long-horizon software development tasks and workflows that "combine text with visual context."
- Additionally, though prompts and outputs are retained by the model provider, they aren't currently used for training.
- Though early reports suggest the mystery model is beating frontier models like Anthropic's Fable 5 and OpenAI's GPT-5.6 Sol in coding evaluations like Deep SWE, the model's independent validation isn't yet available on formal public leaderboards.
Patrick Collison, CEO of Stripe, which is acquiring OpenRouter, described Ox Alpha in a post on X as "very impressive."
As the model goes viral, many have started to speculate about its origins. While some have suggested that Ox Alpha is the next generation of Google's Gemini models or Microsoft's MAI models, others have speculated that the model comes from a Chinese lab, with some analysis pointing to Z.ai's GLM family. However, none of these guesses have been confirmed.
Our Deeper View #
No matter what the origins of this model end up being, Ox Alpha's overnight stardom shows that the AI industry is easily distracted, and often excited by every shiny new toy that hits the market. However, early testing and industry fervor are one thing, and actual, long-running sustainability and results are another. Without knowing who has created this model, users should be wary about its safety and security in real-world applications. Additionally, because the model doesn't currently use user data for training, that policy could change. In short: While the model is drumming up a lot of excitement right now, it's yet to be seen whether this is a flash in the pan, or if the industry will lose interest once the next big model gets released in the days (or even hours) ahead. Still, this supports the bigger trend of frontier intelligence commoditizing.