cd/entity/Kimi· home› entities› Kimi
grep -l @kimi /news/*.json | wc -l → 265

Kimi

mentions 265 type Organization page 14/14 feed RSS

// recent coverage 265 mentions

00:00
2026-05-21
modular.com
large-language-models

Modular: Why LLM Inference Needs a New Kind of Router - Part 2

Modular has built a new data layer for LLM inference routing that solves the problem of querying cached blocks across hundreds of pods in microseconds. The company's architecture uses a specialized da…

19:54
2026-05-18
dev.to
large-language-models

First-call checklist before trying a new LLM gateway

A checklist the author uses when testing a new OpenAI-compatible LLM gateway, which helps catch integration failures before moving real workloads. For Chinese models like Qwen, DeepSeek, GLM, and Kimi…

00:00
2026-05-14
maltebuettner.eu
large-language-models

documentai bbox benchmark

Malte Buettner benchmarked bounding box accuracy for Document AI models using pages from the FlashAttention-3 paper, testing Qwen, Kimi, and Mistral via OpenRouter. The evaluation scored models on cov…

00:00
2026-05-08
modular.com
ai-infrastructure

Modular: Why LLM Inference Needs a New Kind of Router - Part 1

Modular announced that traditional HTTP-era load balancing algorithms like round-robin, consistent hashing, and least-connections are inadequate for large language model inference because GPU pods are…

← prev page 14 / 14
// co-occurs with top 8 entities
// topics top 6 topics