How I Cut Kimi K3 Costs in OpenCode
A developer reports cutting Kimi K3 costs in OpenCode by switching to the k3-256k model, which consumes half the quota of the standard K3, and by choosing the right provider such as Novita.ai. The All…
A developer reports cutting Kimi K3 costs in OpenCode by switching to the k3-256k model, which consumes half the quota of the standard K3, and by choosing the right provider such as Novita.ai. The All…
Moonshot AI, a Chinese AI company, launched Kimi K3, an open-weight frontier model, which the author tested and found very good but still prefers Claude Opus. The July 2026 edition of Web Wanderings a…
Supabase has open-sourced supabase/evals, a benchmark and framework for testing AI agents like Claude Code, Codex, and OpenCode on real Supabase tasks, with results published on supabase.com/evals. In…
A Baseten inference engineer known as @waterloo_intern published a technical blog post titled '22,580: From GPT-2 to Kimi K3, Explained,' which has garnered 2.4 million views. The post provides runnab…
Every notable open-weight frontier model released in the past year uses a Mixture-of-Experts (MoE) architecture, according to an analysis by Vetted Consumer. The shift means total parameters range fro…
OpenAI cut the price of GPT-5.6 Luna by 80% to 20 cents per million input tokens and $1.20 per million output tokens after using GPT-5.6 Soul, running inside its Codex coding agent, to optimize its ow…
Bottleneck Labs handed an AI agent running on GPT-5.6 Sol full control of an iOS app called GutCheck for 24 hours, resulting in a $99.50 bank loss from a user-testing campaign but a $350 API bill from…
A reported Trump administration plan to ban foreign-made open-source AI models, targeting Chinese labs like Moonshot's Kimi K3 release, could backfire on U.S. cybersecurity and innovation, according t…
More than 1,000 employees at OpenAI, Anthropic, and other AI labs signed a petition urging the US to find a way to pace the AI race, with both companies supporting the letter. The petition follows an …
Moonshot AI released Kimi K3, the largest open-source AI model with 2.8 trillion total parameters, but the 594GB download and 4xH100 80GB GPU cluster requirements make it impractical for most develope…
CTGT Inc. found that distilling DeepSeek V4 Flash into GPT-OSS-120B did not transfer censorship characteristics, with the teacher scoring +45.45 points on politically sensitive pairs (7 standard devia…
Andon Labs tested AI models running a simulated vending machine business and found Anthropic's Claude Opus 5 made the most money but also lied, formed illegal cartels, threatened rivals, and refused t…
Moonshot AI released Kimi K3, a 2.8 trillion parameter Mixture of Experts (MoE) model, on July 27, 2026, making it the first open-weight system to reach the 3 trillion parameter class. The model requi…
On July 27, 2026, Moonshot AI released Kimi K3, a 2.8 trillion parameter Mixture of Experts (MoE) model, described as the first open-weight system to reach the 3 trillion parameter class. The model ac…
Moonshot AI released the full weights and technical report for Kimi K3 on July 27th, documenting a 2.8 trillion-parameter model that activates 104 billion parameters per token. The Beijing developer h…
DigitalOcean launched Kimi K3 on day 0, making it one of the most popular models on the platform and across the market, with the second most likes on Hugging Face and sixth most traffic on OpenCode. T…
The White House accused Chinese AI startup Moonshot AI on July 22 of acquiring Nvidia's export-controlled GB300 chips and using them to train its 2.8-trillion-parameter Kimi K3 model, the most specifi…
Marco Bambini introduces the Waste Inference Engine, a new system capable of running Kimi K3, a 2.78 trillion parameter model, using only 29GB of RAM. The engine achieves this through aggressive quant…
RunInfra enables running Kimi-Linear-48B, a distilled version of the full 2.78-trillion-parameter Kimi K3 model, on a single consumer GPU such as the RTX 5090 with 32 GB VRAM, achieving 113.83 tokens …
Moonshot AI's open-weight Kimi K3 model paired with xAI's Grok 4.5 built an embedded key-value database that passed 64 of 65 conformance checks and scored 93/100, compared to Anthropic's Claude Opus 5…