cd/entity/Kimi· home entities Kimi
grep -l @kimi /news/*.json | wc -l → 203

Kimi

mentions 203 type Organization page 10/11 feed RSS

// recent coverage 203 mentions

00:00
2026-06-12
modular.com
large-language-models

Modular: Day Zero: MiniMax M3 Open Weights on Modular Cloud

MiniMax released the open-weights MiniMax M3 model on Modular Cloud, featuring a new Sparse Attention operation that achieves up to 15.6x speedup on decode while maintaining a 1 million token context …

14:00
2026-06-11
coles.codes
large-language-models

Local models in mid-2026: the engineering that closed the gap

Local large language models have nearly caught up to frontier models for everyday tasks as of mid-2026, driven by engineering advances in sparse attention and mixture-of-experts architectures that red…

09:00
2026-06-10
asiaai.fyi
artificial-intelligence

AsiaAI.FYI Issue #2

Moonshot AI, developer of the Kimi AI assistant, saw its valuation surge sixfold to $30 billion in six months, driven by over $200 million in annual recurring revenue after its K2.5 model release. The…

00:00
2026-06-07
tomtunguz.com
artificial-intelligence

The Substitution Wave in AI

Three forces are reshaping AI cost structures as major buyers substitute cheaper models for expensive ones. Coinbase routed prompts to cheaper models to keep costs flat while token usage grew exponent…

00:00
2026-06-05
modular.com
ai-infrastructure

Modular: Why LLM Inference Needs a New Kind of Router - Part 3

Modular has introduced a five-stage composable routing system for large language model inference, replacing traditional fixed algorithms like round-robin and consistent hashing. The system, detailed i…

07:49
2026-06-02
lesswrong.com
large-language-models

Wood Screws and the Methods of Rationality

A man testing six large language models to determine the correct pilot hole size for #8 wood screws in particleboard found that the AI recommendations varied, with Gemini suggesting 3/32″, ChatGPT rec…

00:00
2026-06-02
tomtunguz.com
artificial-intelligence

The Thriving Ecosystem of Open Models

Open-weight models now account for 69.1% of named token volume on the OpenRouter API platform, compared to 30.9% for closed models, according to the latest platform data. The shift, driven by rapid co…

12:32
2026-05-30
maltebuettner.eu
large-language-models

DocumentAI Visual Benchmark - GPT 5.5, Gemini 3.5, Qwen...

A new benchmark evaluating DocumentAI models on bounding box accuracy shows GPT-5.5 and Gemini 3.5 leading with 67.7% and 67.5% scores respectively, while Qwen, Kimi, and Mistral trail significantly. …

21:41
2026-05-29
meltdown.merkoba.com
ai-tools

Meltdown Now Supports OpenRouter

Meltdown, a lightweight Python-based LLM client built with Tkinter, now supports OpenRouter alongside local servers, ChatGPT, Gemini, Claude, and Kimi. The tool offers custom widgets, a scratch-built …

05:48
2026-05-29
zot.sh
ai-tools

Zot now supports Claude Opus 4.8

The coding agent Zot now supports Claude Opus 4.8, adding the model to its catalog of over 20 providers including OpenAI, Google Gemini, and local models. The lightweight terminal agent, distributed a…

05:46
2026-05-29
meltdown.merkoba.com
ai-tools

Conversational LLM Client Made in Tkinter

A new Python-based LLM client called Meltdown, built entirely with Tkinter and a custom markdown engine, offers a lightweight alternative to Electron-based chat interfaces. The tool supports multi-mod…

20:20
2026-05-27
databricks.com
large-language-models

Reliable LLM Inference at Scale

Databricks has built an inference platform serving over 125 trillion tokens per month across frontier models including OpenAI, Gemini, and Claude for major agentic applications like Superhuman and Fox…

00:00
2026-05-21
modular.com
large-language-models

Modular: Why LLM Inference Needs a New Kind of Router - Part 2

Modular has built a new data layer for LLM inference routing that solves the problem of querying cached blocks across hundreds of pods in microseconds. The company's architecture uses a specialized da…

← prev page 10 / 11 next →
// co-occurs with top 8 entities
// topics top 6 topics