{"slug": "mai-code-1-1-flash-better-faster-at-a-quarter-of-the-cost", "title": "MAI-Code-1.1 Flash: Better, faster, at a quarter of the cost", "summary": "Microsoft has released MAI-Code-1.1-Flash, an upgraded coding model now in production in GitHub Copilot, delivering 25% faster token streaming, 25% fewer tokens per task, and a 75% cost reduction compared to the version launched at Microsoft Build in June. The model shows a 22% improvement on Terminal-Bench 2.1 for CLI tasks and a 15% improvement on .NET tasks, with code survival up 4% and return visits up 9%.", "body_md": "#\nMAI-Code-1.1-Flash:\n\nBetter, faster, at a quarter of the cost\n\nBetter, faster, at a quarter of the cost\n\nMAI-Code-1.1-Flash produces higher quality code, at 25% greater token efficiency, and at **a quarter of the cost** compared to the model we launched in June at Microsoft Build. This small, efficient, coding workhorse is now in production in GitHub Copilot.\n\nWe learned from developer feedback that CLI tasks and .NET performance mattered, so that’s where we focused. The result: a **22% improvement on Terminal-Bench 2.1** in GitHub Copilot CLI and a **15% improvement on .NET tasks**.\n\nBenchmarks are useful guides but production is where the rubber meets the road. Most importantly, code survival rose 4% and **return visits increased 9%**.\n\n1.1 is also dramatically more efficient. In GitHub Copilot tokens stream **25% faster** and the model uses **25% fewer tokens** to complete a task. That means faster answers, less waiting, and more useful work from every token—not simply a bigger model with a bigger bill.\n\nBetter training and serving efficiency let us offer a stronger, faster model at one quarter of the price of 1.0—and pass those savings reliably to customers. We achieved this by optimizing for real-world use across more than hundreds of thousands of reinforcement-learning environments in GitHub Copilot.\n\nThe loop is simple: ship, learn, improve, repeat. That’s the MAI hill climbing machine.\n\n**Help shape future improvements **\n\nTry MAI-Code-1.1-Flash today in [GitHub Copilot](https://github.com/features/copilot), then tell us what needs improving by opening an issue [here](https://github.com/microsoft/MAI-Code).\n\n##\nBuild the Future With Us\n\nWe’re a lean, talent-dense team of explorers, researchers, and full-stack engineers. We move fast, sweat the details, and operate at frontier scale with a roadmap to build the world’s most powerful AI models. Most importantly, we’re united by the belief that doing this right is the only way to do it at all. If our mission resonates with you, we’d love to talk.\n\n[Explore all jobs](/careers)", "url": "https://wpnews.pro/news/mai-code-1-1-flash-better-faster-at-a-quarter-of-the-cost", "canonical_source": "https://microsoft.ai/news/mai-code-1-1-flash-br-better-faster-at-a-quarter-of-the-cost/", "published_at": "2026-08-11 21:15:30+00:00", "updated_at": "2026-08-11 21:42:47.889519+00:00", "lang": "en", "topics": ["artificial-intelligence", "large-language-models", "ai-products", "developer-tools", "generative-ai"], "entities": ["Microsoft", "MAI-Code-1.1-Flash", "GitHub Copilot", "Terminal-Bench 2.1", ".NET"], "alternates": {"html": "https://wpnews.pro/news/mai-code-1-1-flash-better-faster-at-a-quarter-of-the-cost", "markdown": "https://wpnews.pro/news/mai-code-1-1-flash-better-faster-at-a-quarter-of-the-cost.md", "text": "https://wpnews.pro/news/mai-code-1-1-flash-better-faster-at-a-quarter-of-the-cost.txt", "jsonld": "https://wpnews.pro/news/mai-code-1-1-flash-better-faster-at-a-quarter-of-the-cost.jsonld"}}