cd/entity/Cerebras· home entities Cerebras
grep -l @cerebras /news/*.json | wc -l → 146

Cerebras

mentions 146 type Organization page 1/8 feed RSS

// recent coverage 146 mentions

11:38
2026-08-19
pub.towardsai.net
artificial-intelligence

TAI #218: Enterprise AI Use Is Becoming More Uneven

OpenAI reports that its top 10% of enterprise customers by output tokens per active user generated 8.3 times as many tokens per active user as the middle decile in June, up from 2.6 in January, while …

08:44
2026-08-19
latent.space
ai-infrastructure

[AINews] Memory prices up 500% in 12 months

Memory prices have surged 500% in 12 months, with 128GB DDR5 kits now costing $3,399, ten times their lowest-ever tracked price, according to Tom's Hardware. Hyperscale buyers have reportedly locked i…

19:15
2026-08-18
networkworld.com
artificial-intelligence

Reports: Google partnering with AMD for next-gen hybrid TPU

Google is reportedly partnering with AMD to design a next-generation Tensor Processing Unit (TPU) that integrates CPU cores into the processor package, targeting reinforcement learning and agentic AI …

00:00
2026-08-18
mindstudio.ai
ai-agents

What Is Ego Lite? The Fast Open-Source Browser for AI Agents

Ego Lite, an open-source browser designed for AI coding agents such as Codex and Claude Code, has gained around 10,000 GitHub stars. It allows agents to control a user's logged-in browser session loca…

04:00
2026-08-17
arxiv.org
large-language-models

Jais 2: A Family of Arabic-Centric Open Large Language Models

MBZUAI, Cerebras, and Inception released Jais 2, a family of Arabic-centric open large language models, including the largest open Arabic-centric LLM trained from scratch at 70B parameters and an 8B-p…

03:00
2026-08-17
hiraditya.github.io
ai-infrastructure

Prefill and Decode Want Different Computers

AWS is pairing its Trainium chips with Cerebras CS-3 systems to split transformer inference into prefill and decode phases, with Trainium handling prefill and Cerebras handling decode, shipping as a p…

14:09
2026-08-16
byteiota.com
artificial-intelligence

GPT-5.6 Sol Ultrafast: 750 Tokens Per Second, Explained

OpenAI and Cerebras announced on August 13 that GPT-5.6 Sol, OpenAI's flagship frontier model, now runs at 750 output tokens per second through a new 'Ultrafast' API tier, 14 times faster than standar…

12:20
2026-08-16
dev.to
developer-tools

Rebuilding the Cerebras Knowledge Base: an MCP server

A developer rebuilt the Cerebras knowledge base as an MCP server that exposes raw, LLM-free retrieval tools for external agents like Claude Code. The server makes zero model calls, instead handing rec…

12:20
2026-08-16
dev.to
large-language-models

Rebuilding the Cerebras Knowledge Base: an LLM reranker

A developer rebuilt the Cerebras knowledge base retrieval pipeline, adding an LLM reranker that boosts mean reciprocal rank from 0.57 to 0.90 on a 31-question evaluation. The hybrid retrieval plus rer…

10:08
2026-08-16
byteiota.com
artificial-intelligence

GPT-5.6 Sol Ultrafast: What 14x Speed Actually Changes

OpenAI previewed Ultrafast, a new API tier delivering GPT-5.6 Sol at 750 output tokens per second, 14 times faster than Standard mode, using Cerebras' Wafer-Scale Engine with 44 GB of on-chip SRAM. Th…

15:09
2026-08-14
sourcefeed.dev
artificial-intelligence

Speed Is Now a Paid Tier at OpenAI

OpenAI and Cerebras are previewing an "Ultrafast" mode for GPT-5.6 Sol that delivers 750 tokens per second, roughly 14x the standard tier's ~53 tokens per second, marking the first time a closed front…

14:50
2026-08-14
techcrunch.com
artificial-intelligence

Kog is going deeper to squeeze more inference out of GPUs

French startup Kog claims its software can deliver 30x faster LLM inference on standard datacenter GPUs, demoing 3,000 tokens per second on a 2-billion-parameter model using AMD MI300X and NVIDIA H200…

page 1 / 8 next →
// co-occurs with top 8 entities
// topics top 6 topics