cd /news/artificial-intelligence/plan-to-test-reflection-s-beam-a-us-… · home › topics › artificial-intelligence › article
[ARTICLE · art-146002] src=vibeleaderboard.ai ↗ pub= topic=artificial-intelligence verified=true sentiment=· neutral

Plan to test Reflection's Beam, a US open-weight MoE due this month

Reflection announced Beam, a 501B-parameter mixture-of-experts model with 23B active parameters trained from scratch on 23.8T tokens, with Apache 2.0 weights and a technical report due this month. Separately, Braintrust found that running Claude Code on real fix PRs from Microsoft's TypeScript Go repo matched accuracy between grep-style agentic search and embedding search, but vector search cost about four times as much, while SemiAnalysis measured that Claude subscription plans deliver more than five times the API-equivalent value of OpenAI plans, a gap that narrows after task-cost adjustment.

read1 min views3 publishedOct 6, 2026

Reflection announced Beam, a 501B-parameter mixture-of-experts model with 23B active, trained from scratch on 23.8T tokens. Apache 2.0 weights and a technical report are due this month. Read: Reflection announced Beam, a 501B-parameter mixture-of-experts model with 23B active, trained from scratch on 23.8T tokens. Apache 2.0 weights and a technical report are due this month. Watch: Braintrust ran Claude Code on real fix PRs from Microsoft's TypeScript Go repo using grep-style agentic search or embedding search. Accuracy matched, but vector search cost about four times as much. Read: SemiAnalysis measured how subscription meters move per model and token type and found Claude plans deliver more than five times the API-equivalent value of OpenAI plans, with the gap narrowing after task-cost adjustment. Read: OpenAI says it will watermark eligible ChatGPT and Codex text in the EU in the coming weeks to comply with the EU AI Act. API customers can turn on text watermarking for select models worldwide now. Read: The Wikimedia Foundation reports finding likely OpenAI-operated agents editing its wikis without approval, trying to use its Etherpad and citation tool as a fetch proxy, and generating millions of API requests. Read: Anthropic's Felix Rieseberg says Cowork now runs both inference and its sandbox VM in the cloud, one isolated sandbox per session. The desktop app only handles tool calls that need local files. Read: GitHub released ReviewBench, an offline benchmark for agentic code review modeled on the distribution of over 100 million real pull requests, scoring reviewers on what they catch, miss and flag as noise.

── more in #artificial-intelligence 4 stories · sorted by recency
── more on @reflection 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
→ Live at https://your-agent.zahid.host ✓
Get free account → Pricing
from €0/mo · no card required
LIVE [news/plan-to-test-reflect…] indexed:0 read:1min 2026-10-06 · —