cd/entity/Baseten· home entities Baseten
grep -l @baseten /news/*.json | wc -l → 54

Baseten

mentions 54 type Organization page 2/3 feed RSS

// recent coverage 54 mentions

01:44
2026-07-28
runtimewire.com
artificial-intelligence

Baseten ships Rust tokenizer it says is 18x faster for Kimi K3

Baseten model performance engineer Michael Feil shipped Baseten Tokenizer, a Rust-backed package that cuts CPU time for preparing long prompts for Moonshot AI's Kimi K3 model, achieving up to 18x fast…

15:16
2026-07-27
baseten.co
artificial-intelligence

We built a day-0 API for Kimi K3

Baseten has launched day-0 API support for Kimi K3, a new open frontier model from Moonshot AI with 2.8 trillion parameters, making it the largest open model to date. The API runs on NVIDIA GB300 NVL7…

02:51
2026-07-26
baseten.co
artificial-intelligence

We built the new fastest API for GLM-5.2

Baseten has built the fastest API for GLM-5.2, achieving peak speeds of 280 tokens per second and average speeds around 100 tokens per second, more than double the performance of the launch-day API as…

12:00
2026-07-21
lightningjar.com
artificial-intelligence

Sarah Guo's Wager

Sarah Guo, who became Greylock's youngest general partner at 28 and left four years later to found Conviction, staked her firm entirely on AI before ChatGPT launched, with pre-ChatGPT seed checks into…

11:01
2026-07-21
profgmedia.com
artificial-intelligence

Gas Is Back Above $4 — And Could Keep Rising

Gas prices have risen above $4 per gallon, with potential for further increases due to developments in the war with Iran, according to Matt Smith, Director of Commodity Research at Kpler. Meanwhile, C…

02:01
2026-07-19
machinebrief.com
artificial-intelligence

Open Weight Models Are Turning Inference Into A Control Point

Inference startups Fireworks, Baseten, and Together AI collectively raised $3.8 billion in four weeks, signaling that inference has become a control point in the AI industry, according to Forbes contr…

01:19
2026-07-09
grandpacad.com
artificial-intelligence

Public LLM benchmarks are mostly garbage

A developer's benchmark of frontier AI models on a 3D modeling task found Anthropic's Opus 4.7 outperformed Gemini 3.1, GPT 5.5, and Kimi K2.6 in weighted score, speed, and cost, contradicting public …

22:52
2026-06-30
zandrey.com
ai-agents

I like Claude Desktop, so I created my own

A developer built Recowork, a desktop agent application inspired by Claude Desktop, using Tauri, OrbStack, and Apple container machines with GLM-5.2. The app plans, calls tools, and chains steps to co…

16:15
2026-06-30
blog.southparkcommons.com
artificial-intelligence

June 2026 Update

South Park Commons portfolio companies raised over $1.5 billion in June 2026, including Baseten's $1.5B Series F at a $13B valuation for AI inference and Orb's $335M acquisition by Adyen. AXION launch…

20:13
2026-06-25
artificialanalysis.ai
large-language-models

GLM-5.2 (Max) API Provider Benchmarking and Analysis

A new benchmark of 14 API providers for the GLM-5.2 (max) model reveals Fireworks leads in output speed (261.5 t/s) and low latency (9.81s), while GMI (FP8) offers the lowest blended price at $0.72 pe…

21:53
2026-06-24
baseten.co
large-language-models

How we built the fastest API for GLM-5.2

Baseten has built the world's fastest API for GLM-5.2, achieving over 280 tokens per second as measured by Artificial Analysis. The performance is driven by optimizations including an updated inferenc…

← prev page 2 / 3 next →
// co-occurs with top 8 entities
// topics top 6 topics