cd/entity/gpt-oss-20b· home entities gpt-oss-20b
grep -l @gpt-oss-20b /news/*.json | wc -l → 15

gpt-oss-20b

mentions 15 type Organization feed RSS

// recent coverage 15 mentions

17:59
2026-09-21
twitter.com
ai-tools

Halo: Post-train LLMs 3x faster than TRL and Megatron

White Circle launched Halo, a post-training framework for open-source models that delivers up to 2.8x the throughput of stock TRL with lower peak memory while keeping models in native HuggingFace form…

13:00
2026-09-02
vettedconsumer.com
large-language-models

How Much RAM Do You Need to Run a Local LLM in 2026?

Running large Mixture-of-Experts local LLMs in 2026 requires at least 64GB of system RAM, ideally 128GB or more, because sparse MoE models keep rarely-used expert weights in system memory rather than …

16:48
2026-08-25
forum.level1techs.com
ai-infrastructure

Dell Pro Max with GB10

Dell's Pro Max with GB10, powered by NVIDIA's GB10 chip, delivers performance nearly identical to NVIDIA's reference design but with better sustained throughput before thermal throttling, according to…

14:31
2026-08-17
pub.towardsai.net
artificial-intelligence

From Llama to World Models: 15 AI Models You Can Download in 2026

Meta released Llama 4 Scout and Llama 4 Maverick as open-weight, natively multimodal models using a Mixture-of-Experts architecture, with Scout capable of fitting on a single H100 when quantized to In…

13:33
2026-08-01
dev.to
machine-learning

A 7.5B model beat a 24B on my coding benchmark.

A developer built a 56-task coding benchmark with hidden tests and ran 16 model configurations on the same hardware, finding that a 7.5B-parameter model (gemma-4-e4b) scored 42/56, beating a 24B model…

10:15
2026-07-30
technologyreview.com
artificial-intelligence

A fundamental flaw leaves LLMs strikingly vulnerable to attack

A fundamental flaw in how large language models identify instructions makes them impossible to fully secure against attacks, researchers argue in a paper presented at the International Conference on M…

14:12
2026-07-09
tokenstead.ai
artificial-intelligence

Open-weights models cost less: a 2026 pricing guide

Open-weights models are dramatically cheaper per token than closed models, with the five cheapest models on the Artificial Analysis pricing index all being open-weights and the five most expensive all…

23:59
2026-06-22
simonwillison.net
ai-safety

Prompt Injection as Role Confusion

Researchers Charles Ye, Jasmine Cui, and Dylan Hadfield-Menell found that large language models suffer from 'role confusion,' mistaking the style of text for its actual content, leading to successful …

// co-occurs with top 8 entities
// topics top 6 topics