00:00
2026-08-03
andlukyane.com
artificial-intelligence
Beyond Bigger MoE: How Kimi K3 Scales Context, Depth, and Agents
Moonshot AI's Kimi K3 model scales to 2.8T total parameters with 104B activated per token, using hybrid attention, Attention Residuals, and Stable LatentMoE to support 1M-token agentic trajectories. Tโฆ