cd/entity/Baseten· home› entities› Baseten
grep -l @baseten /news/*.json | wc -l → 82

Baseten

mentions 82 type Organization page 3/5 feed RSS

// recent coverage 82 mentions

17:44
2026-08-03
frontierroles.com
ai-infrastructure

AI Inference Engineer — Baseten

Baseten, an AI infrastructure company that recently raised a $1.5B Series F led by Altimeter Capital, Conviction Partners, and Spark Capital, is hiring an AI Inference Engineer in San Francisco with a…

00:00
2026-08-02
zackproser.com
large-language-models

Sampling Inkling and the Alias Pattern

Thinking Machines' Inkling, a 975B total / 41B active MoE model with Apache 2.0 license and 1M context, is being trialed via a bash wrapper alias 'claude-inkling' that bridges Anthropic's CLI to the O…

13:00
2026-08-01
thestack.technology
artificial-intelligence

Runtime: MCP goes stateless; Baseten courts lab partners

The latest version of the Model Context Protocol (MCP) introduces a stateless protocol core, shifting from a bidirectional stateful protocol to a request/response model, according to lead maintainers …

01:20
2026-07-29
artificialconfidence.com
artificial-intelligence

Reading the tea leaves: 2026 predictions

The marginal cost of AI inference will approach zero as open models rival frontier models, reshaping the business models of labs like OpenAI and Anthropic toward low-margin inference sales, according …

01:44
2026-07-28
runtimewire.com
artificial-intelligence

Baseten ships Rust tokenizer it says is 18x faster for Kimi K3

Baseten model performance engineer Michael Feil shipped Baseten Tokenizer, a Rust-backed package that cuts CPU time for preparing long prompts for Moonshot AI's Kimi K3 model, achieving up to 18x fast…

15:16
2026-07-27
baseten.co
artificial-intelligence

We built a day-0 API for Kimi K3

Baseten has launched day-0 API support for Kimi K3, a new open frontier model from Moonshot AI with 2.8 trillion parameters, making it the largest open model to date. The API runs on NVIDIA GB300 NVL7…

02:51
2026-07-26
baseten.co
artificial-intelligence

We built the new fastest API for GLM-5.2

Baseten has built the fastest API for GLM-5.2, achieving peak speeds of 280 tokens per second and average speeds around 100 tokens per second, more than double the performance of the launch-day API as…

12:00
2026-07-21
lightningjar.com
artificial-intelligence

Sarah Guo's Wager

Sarah Guo, who became Greylock's youngest general partner at 28 and left four years later to found Conviction, staked her firm entirely on AI before ChatGPT launched, with pre-ChatGPT seed checks into…

11:01
2026-07-21
profgmedia.com
artificial-intelligence

Gas Is Back Above $4 — And Could Keep Rising

Gas prices have risen above $4 per gallon, with potential for further increases due to developments in the war with Iran, according to Matt Smith, Director of Commodity Research at Kpler. Meanwhile, C…

02:01
2026-07-19
machinebrief.com
artificial-intelligence

Open Weight Models Are Turning Inference Into A Control Point

Inference startups Fireworks, Baseten, and Together AI collectively raised $3.8 billion in four weeks, signaling that inference has become a control point in the AI industry, according to Forbes contr…

01:19
2026-07-09
grandpacad.com
artificial-intelligence

Public LLM benchmarks are mostly garbage

A developer's benchmark of frontier AI models on a 3D modeling task found Anthropic's Opus 4.7 outperformed Gemini 3.1, GPT 5.5, and Kimi K2.6 in weighted score, speed, and cost, contradicting public …

← prev page 3 / 5 next →
// co-occurs with top 8 entities
// topics top 6 topics