cd/entity/Baseten· home entities Baseten
grep -l @baseten /news/*.json | wc -l → 54

Baseten

mentions 54 type Organization page 1/3 feed RSS

// recent coverage 54 mentions

03:15
2026-08-17
vibeleaderboard.ai
artificial-intelligence

Inference now sets the bill, and open weights are closing the gap

OpenAI's compute chief Sachin Katti said inference will account for over 80 percent of all AI compute spend, with a roadmap to 30GW, while Crusoe's Chase Lockmiller reported that power, not GPU supply…

00:00
2026-08-14
oskrim.github.io
large-language-models

Deepseek V4 Flash 0731 latency numbers from nine providers

A one-time snapshot of DeepSeek V4 Flash 0731 latency across nine inference providers found Baseten fastest with 3,980 decode tok/s and 7.72 s total p99, while Azure ran an older checkpoint and Scalew…

18:39
2026-08-13
firecrawl.dev
large-language-models

What Is Kimi K3? A Complete Developer Guide for 2026

Moonshot AI released Kimi K3, a 2.8-trillion-parameter open-weight model, on July 27, 2026, making it the largest open-weight model as of August 2026. It features 104B active parameters per token, a 1…

22:10
2026-08-11
frontierroles.com
ai-infrastructure

Software Engineer - AI Developer Productivity — Baseten

Baseten, an AI infrastructure company that raised a $1.5B Series F led by Altimeter Capital, Conviction Partners, and Spark Capital, is hiring a Software Engineer for AI Developer Productivity in San …

13:00
2026-08-11
coderabbit.ai
machine-learning

Teaching NVIDIA Nemotron 3.5 Lightning to route code reviews

CodeRabbit, NVIDIA, and Baseten post-trained NVIDIA Nemotron 3.5 Lightning to route code reviews, achieving 80.7% exact route agreement (up from 75.8% for the GPT baseline) and cutting estimated infer…

12:45
2026-08-09
read.technically.dev
artificial-intelligence

Dispatch: Kimi K3 licensing, Liquid AI, and stacked PRs on GitHub

Moonshot AI released the weights for its Kimi K3 model with 2.8 trillion parameters and a 1M token context window under a new Kimi K3 License, which allows inference providers like Modal, Baseten, Fir…

00:00
2026-08-06
huggingface.co
artificial-intelligence

Baseten on Hugging Face Inference Providers 🔥

Hugging Face has added Baseten as a supported Inference Provider on its Hub, enabling serverless access to open-weight LLMs such as Kimi K3, DeepSeek V4 Flash, and GLM-5.2 for conversational and text-…

10:11
2026-08-05
byteiota.com
artificial-intelligence

Kimi K3 Open Weights: Self-Hosting Reality Check

Moonshot AI released the weights for Kimi K3 on July 27, a 2.8-trillion-parameter Mixture-of-Experts model with a 1-million-token context window, scoring third globally behind Claude Fable 5 Max and G…

12:20
2026-08-04
machinelearningmastery.com
artificial-intelligence

Static vs. Dynamic vs. Continuous Batching in LLM Inference

IBM's article explains that static, dynamic, and continuous batching are methods to improve GPU utilization in large language model inference, with continuous batching scheduling at the token level to…

17:44
2026-08-03
frontierroles.com
ai-infrastructure

AI Inference Engineer — Baseten

Baseten, an AI infrastructure company that recently raised a $1.5B Series F led by Altimeter Capital, Conviction Partners, and Spark Capital, is hiring an AI Inference Engineer in San Francisco with a…

00:00
2026-08-02
zackproser.com
large-language-models

Sampling Inkling and the Alias Pattern

Thinking Machines' Inkling, a 975B total / 41B active MoE model with Apache 2.0 license and 1M context, is being trialed via a bash wrapper alias 'claude-inkling' that bridges Anthropic's CLI to the O…

13:00
2026-08-01
thestack.technology
artificial-intelligence

Runtime: MCP goes stateless; Baseten courts lab partners

The latest version of the Model Context Protocol (MCP) introduces a stateless protocol core, shifting from a bidirectional stateful protocol to a request/response model, according to lead maintainers …

01:20
2026-07-29
artificialconfidence.com
artificial-intelligence

Reading the tea leaves: 2026 predictions

The marginal cost of AI inference will approach zero as open models rival frontier models, reshaping the business models of labs like OpenAI and Anthropic toward low-margin inference sales, according …

page 1 / 3 next →
// co-occurs with top 8 entities
// topics top 6 topics