The Quiet Launch of Gemini 3.8 Flash
Google quietly launched Gemini 3.8 Flash on Wednesday, a multimodal model priced at $0.75 per million input tokens and $3.75 per million output tokens (introductory through December 31), scoring 75.3%…
Google quietly launched Gemini 3.8 Flash on Wednesday, a multimodal model priced at $0.75 per million input tokens and $3.75 per million output tokens (introductory through December 31), scoring 75.3%…
SiliconANGLE has named Kilo a finalist in the AI Coding & Developer Assistants category of its 2026 TechForward Awards, alongside Postman API Platform and SonarQube, with winners announced September 1…
Moonshot AI's open-weight Kimi K3 model, released in mid-2025, ranks near the top of intelligence benchmarks but suffers from severe performance issues, including throughput dropping from 30 to 13 tok…
Anthropic's Claude Fable 5 was taken offline by the US government on June 12 after a reported jailbreak, then returned on July 1 with a stricter safety classifier. Benchmark tests by KiloBench show th…
Kilo AI launched new features to provide developers with full visibility into AI coding costs and model usage, including turn-by-turn model identification and a unified inference pricing comparison pa…
Tencent released the General Availability version of its Hy3 AI model, now free for a limited time on the Kilo platform. The 295B-parameter Mixture-of-Experts model features improved reasoning, reduce…
Kilo launched Auto Efficient, a new AI model routing tier that dynamically selects the most cost-effective model for each task based on real-time session classification and benchmark data. The system …
KiloBench launched as a new evaluation framework that measures AI coding models based on real-world production cost and performance rather than benchmark scores. The tool emerged after its creators fo…