{"slug": "inco-ai-launches-its-inference-platform-leading-across-four-open-models-on", "title": "Inco AI Launches Its Inference Platform, Leading Across Four Open Models on Artificial Analysis", "summary": "Inco AI launched its inference platform in public beta, claiming the fastest output speeds on Artificial Analysis leaderboards for four open models: Kimi K3, MiniMax M3, GLM 5.3, and GLM 5.3 Flash. The company says its stack is optimized for agentic workloads, including sustained generation and long sequences.", "body_md": "# Inco AI Launches Its Inference Platform, Leading Across Four Open Models on Artificial Analysis\n\nOver the last several months, Inco AI has been building an inference stack purpose-built for the demands of the agentic era and pushing the efficiency frontier across sustained generation, long-running workloads, and longer sequences—conditions that combine to push inference systems to their limits.\n\nToday's launch of the Inco platform marks an important milestone toward\nbringing superior agentic inference performance to market and provides a first\nlook at Inco's inference and technology stack. We are releasing high-speed\nendpoints for Kimi K3, MiniMax M3, GLM 5.3, and GLM 5.3 Flash. Each leads its\nrespective [Artificial Analysis](https://artificialanalysis.ai/) provider\nleaderboard on output speed.\n\n| Model | Output tokens / s | Relative improvement |\n|---|---|---|\n|\n\n[MiniMax M3](https://artificialanalysis.ai/models/minimax-m3/providers)[GLM 5.3](https://artificialanalysis.ai/models/glm-5-3/providers)[GLM 5.3 Flash](https://artificialanalysis.ai/models/glm-5-3-flash/providers)[Inside the Inco Inference Stack](#inside-the-inco-inference-stack)\n\nThe Artificial Analysis results reflect optimization across the full Inco stack, highlighting our unique approach to bringing peak inference efficiency to the market.\n\n[The Inco Platform Is Entering Public Beta](#the-inco-platform-is-entering-public-beta)\n\nWe are opening beta access to the Inco platform, starting with Kimi K3, MiniMax M3, GLM 5.3, and GLM 5.3 Flash.\n\nSign up to try our fastest endpoints on the Inco platform.\n\n[Sign up for public beta](https://platform.inco.ai)\n\n### Available at launch\n\n- Kimi K3\n- MiniMax M3\n- GLM 5.3\n- GLM 5.3 Flash\n\nIf inference speed is on your application's critical path, reach out at\n[contact@inco.ai](mailto:contact@inco.ai).\n\nGet updates\n\nOne email when we ship something new.\n\nWe will never share your email address.", "url": "https://wpnews.pro/news/inco-ai-launches-its-inference-platform-leading-across-four-open-models-on", "canonical_source": "https://inco.ai/blog/inco-platform-aa/", "published_at": "2026-09-03 00:00:00+00:00", "updated_at": "2026-09-03 21:54:09.591343+00:00", "lang": "en", "topics": ["ai-infrastructure", "ai-products", "artificial-intelligence"], "entities": ["Inco AI", "Kimi K3", "MiniMax M3", "GLM 5.3", "GLM 5.3 Flash", "Artificial Analysis"], "alternates": {"html": "https://wpnews.pro/news/inco-ai-launches-its-inference-platform-leading-across-four-open-models-on", "markdown": "https://wpnews.pro/news/inco-ai-launches-its-inference-platform-leading-across-four-open-models-on.md", "text": "https://wpnews.pro/news/inco-ai-launches-its-inference-platform-leading-across-four-open-models-on.txt", "jsonld": "https://wpnews.pro/news/inco-ai-launches-its-inference-platform-leading-across-four-open-models-on.jsonld"}}