cd/entity/Kimi K3· home entities Kimi K3
grep -l @kimi k3 /news/*.json | wc -l → 879

Kimi K3

mentions 879 type Person page 6/44 feed RSS

// recent coverage 879 mentions

07:30
2026-08-17
hellochinatech.com
artificial-intelligence

The Like Economy

Hugging Face's mid-year report on open models shows that among the 25 most-downloaded and 25 most-liked models on its platform, only one repository appears on both lists, highlighting a disconnect bet…

14:24
2026-08-16
gist.github.com
artificial-intelligence

Show HN: Reasoning prefills on a few open models

A developer's independent analysis of reasoning prefills on open models finds that Kimi K3 shows a significant accuracy increase on Humanity's Last Exam questions when prefilled with Opus 4.8 reasonin…

12:16
2026-08-16
primeintellect.ai
artificial-intelligence

Measuring Autonomous AI Research

A new public experiment by Prime Intellect ran 153 autonomous runs on the nanoGPT optimizer speedrun across 18 frontier models, finding that Claude Fable 5 and Opus 5 dramatically outperformed others,…

23:38
2026-08-15
twitter.com
artificial-intelligence

Image and text do not share a coordinate system

A linear classifier separates image from text embeddings with 100% accuracy in Kimi K3, Inkling, and Qwen3-Omni, revealing that each model's coordinate system, not missing content, causes the modality…

14:00
2026-08-15
akitaonrails.com
large-language-models

LLM Benchmarks: Qwen 3.8, GLM 5.3, Gemini 3.7, Grok 4.6

Qwen 3.8 Max scored 92 on the LLM Coding Benchmark v2, tying with GLM 5.2, Kimi K2.5, Gemini 3.6 Flash, and Grok 4.6, while GLM 5.3 reached 94, Gemini 3.7 Flash scored 93, and Grok 4.6 scored 92, acco…

06:15
2026-08-15
interconnects.ai
artificial-intelligence

GLM-5.3: How Chinese labs keep stride with the frontier

Z.ai released GLM-5.3, a model with approximately 750B parameters that surpasses Moonshot AI's Kimi K3 on many benchmarks and rivals Claude Fable 5 and GPT-5.6-Sol on some, marking a significant advan…

00:58
2026-08-15
benchling.com
artificial-intelligence

Can LLMs work in the wet lab?

Benchling released BenchBench-Protocol, a benchmark built from thousands of real-world experiments, showing that Anthropic's Opus 5 leads at 59.2%, followed by OpenAI's GPT 5.6 at 47.1%, and open-sour…

00:00
2026-08-15
philippdubach.com
artificial-intelligence

Put the Model in the Basement

A Swiss provider operating a single 64-GPU cluster in Zurich could generate about CHF 7.4 million in annual revenue and CHF 3 million in EBITDA by selling dedicated Kimi K3-class inference to banks, p…

20:17
2026-08-14
dev.to
artificial-intelligence

GLM 5.3: Zhipu's Open-Weight Model Excels at Coding and Cyber

Zhipu AI released GLM 5.3, an open-weight model that improves coding and cyber capabilities through advanced post-training techniques rather than a larger architecture. The model, based on the same 74…

18:30
2026-08-14
lobu.ai
ai-agents

Self-Improving Agents Are Event-Sourced

Lobu, an AI startup, has built an event-sourced memory layer for AI agents that treats every write as an append-only commit, with corrections made by writing new events that supersede old ones, ensuri…

13:31
2026-08-14
labs.notion.com
artificial-intelligence

Evaluating how models perform using live traffic

Notion's Knowledge Board, a live evaluation using anonymized traffic and judged by models from Anthropic, OpenAI, and Google, shows Opus 5 leading with a 98.4% resolution rate on knowledge work tasks,…

← prev page 6 / 44 next →
// co-occurs with top 8 entities
// topics top 6 topics