cd/entity/gpt-oss-120b· home entities gpt-oss-120b
grep -l @gpt-oss-120b /news/*.json | wc -l → 35

gpt-oss-120b

mentions 35 type Organization page 1/2 feed RSS

// recent coverage 35 mentions

19:02
2026-09-07
gist.github.com
ai-agents

AI Agent Evaluation Playbook

A developer published an AI agent evaluation playbook, a repeatable test battery for vetting whether models running inside agent frameworks can be trusted with semi-sensitive content and real write ac…

17:01
2026-09-07
pub.towardsai.net
ai-safety

Do LLM Guardrails Actually Work? What My Own Numbers Say

A developer's test of LLM guardrails using openai/gpt-oss-120b and openai/gpt-oss-safeguard-20b on Groq found that safety classifiers can be inconsistent, with results varying across repeated runs. Th…

18:01
2026-08-27
pub.towardsai.net
artificial-intelligence

Your Context Length Decides What a Kernel Is Worth

A 2× faster attention kernel yields only 0.66% end-to-end speedup on a 1,024-token prompt with a 128-token answer under vLLM's --goodput ttft:500 tpot:50 promise, according to a seven-part technical s…

16:22
2026-08-26
promptcube3.com
artificial-intelligence

You can access Claude and GPT models right now without even

Duck.ai offers a privacy-focused interface for accessing multiple AI models, including GPT-5.6 Luna, GPT-5.4 mini, Claude Haiku 4.5, Mistral Small 4, gpt-oss-120b, and Gemma 4 31B, without requiring a…

16:48
2026-08-25
forum.level1techs.com
ai-infrastructure

Dell Pro Max with GB10

Dell's Pro Max with GB10, powered by NVIDIA's GB10 chip, delivers performance nearly identical to NVIDIA's reference design but with better sustained throughput before thermal throttling, according to…

08:10
2026-08-19
snipvote.com
artificial-intelligence

gpt-oss-120b gained 16.1pp task completion with 5% more tokens

Hugging Face reported that gpt-oss-120b achieved a 16.1 percentage point task completion gain with only 5% more tokens when using selective memory retrieval instead of full guideline injection, while …

18:09
2026-08-18
huggingface.co
artificial-intelligence

How Much Memory Does Your Agent Actually Need?

IBM Research's ALT K-Evolve framework shows that the optimal amount of agentic memory varies by model capability, with strong models like DeepSeek-V3.2 (671B MoE) gaining +9.5 percentage points in tas…

14:31
2026-08-17
pub.towardsai.net
artificial-intelligence

From Llama to World Models: 15 AI Models You Can Download in 2026

Meta released Llama 4 Scout and Llama 4 Maverick as open-weight, natively multimodal models using a Mixture-of-Experts architecture, with Scout capable of fitting on a single H100 when quantized to In…

22:09
2026-08-16
cryptobriefing.com
artificial-intelligence

AI efficiency jumped 18x in 16 months, Stanford research finds

Stanford's Hazy Research group found that the intelligence per joule (IPJ) of local AI inference improved 18-fold from mid-2024 to late 2025, driven by a 3.1x gain from model architecture improvements…

12:31
2026-08-16
pub.towardsai.net
large-language-models

Your KV Cache Is Bigger Than Your Model

OpenAI's gpt-oss-120b model, with open weights, requires 72 KiB of KV cache per token in 16-bit precision, calculated from its config.json with 36 layers, 8 key-value heads, and a head dimension of 64…

10:05
2026-08-10
huggingface.co
machine-learning

Making Knowledge Distillation Cheap Enough to Run at Scale

A new paper from Hugging Face, 'Efficient Knowledge Distillation for LLMs: Offline Top-K Logits and a Fused Chunked KL Loss,' cuts the VRAM required for knowledge distillation from roughly 250GB to ne…

18:13
2026-08-04
twitter.com
artificial-intelligence

Artificial Analysis Endpoint Accuracy Index

Artificial Analysis launched its Endpoint Accuracy Index, measuring how much of an open weights model's accuracy each serverless API endpoint preserves, with initial coverage of GLM-5.2, gpt-oss-120b,…

page 1 / 2 next →
// co-occurs with top 8 entities
// topics top 6 topics