cd /news/ai-infrastructure/inco-ai-launches-its-inference-platf… · home topics ai-infrastructure article
[ARTICLE · art-120866] src=inco.ai ↗ pub= topic=ai-infrastructure verified=true sentiment=↑ positive

Inco AI Launches Its Inference Platform, Leading Across Four Open Models on Artificial Analysis

Inco AI launched its inference platform in public beta, claiming the fastest output speeds on Artificial Analysis leaderboards for four open models: Kimi K3, MiniMax M3, GLM 5.3, and GLM 5.3 Flash. The company says its stack is optimized for agentic workloads, including sustained generation and long sequences.

read1 min views5 publishedSep 3, 2026
Inco AI Launches Its Inference Platform, Leading Across Four Open Models on Artificial Analysis
Image: Inco (auto-discovered)

Over the last several months, Inco AI has been building an inference stack purpose-built for the demands of the agentic era and pushing the efficiency frontier across sustained generation, long-running workloads, and longer sequences—conditions that combine to push inference systems to their limits.

Today's launch of the Inco platform marks an important milestone toward bringing superior agentic inference performance to market and provides a first look at Inco's inference and technology stack. We are releasing high-speed endpoints for Kimi K3, MiniMax M3, GLM 5.3, and GLM 5.3 Flash. Each leads its respective Artificial Analysis provider leaderboard on output speed.

Model Output tokens / s Relative improvement

MiniMax M3GLM 5.3GLM 5.3 FlashInside the Inco Inference Stack

The Artificial Analysis results reflect optimization across the full Inco stack, highlighting our unique approach to bringing peak inference efficiency to the market.

The Inco Platform Is Entering Public Beta We are opening beta access to the Inco platform, starting with Kimi K3, MiniMax M3, GLM 5.3, and GLM 5.3 Flash.

Sign up to try our fastest endpoints on the Inco platform.

Sign up for public beta

Available at launch

  • Kimi K3
  • MiniMax M3
  • GLM 5.3
  • GLM 5.3 Flash
If inference speed is on your application's critical path, reach out at
[contact@inco.ai](mailto:contact@inco.ai).

Get updates

One email when we ship something new.

We will never share your email address.

── more in #ai-infrastructure 4 stories · sorted by recency
── more on @inco ai 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/inco-ai-launches-its…] indexed:0 read:1min 2026-09-03 ·