cd/entity/Gemini 3 Flash· home entities Gemini 3 Flash
grep -l @gemini 3 flash /news/*.json | wc -l → 33

Gemini 3 Flash

mentions 33 type Organization page 2/2 feed RSS

// recent coverage 33 mentions

06:39
2026-06-18
entropicthoughts.com
large-language-models

GLM 5.2 playing text adventures

GLM 5.2, a new open-weights model, achieved 15% fewer achievements than Gemini 3 Flash in text adventure games, a statistically significant difference. The benchmark, costing $5.1, controlled for game…

01:10
2026-06-18
trustedrouter.com
artificial-intelligence

New SOTA: TrustedRouter Fusion Beats Fable and Frontier

TrustedRouter achieved a state-of-the-art score of 70.6 on the DRACO benchmark by fusing five models including open-weights models DeepSeek V4 Pro and Kimi K2.6, surpassing OpenRouter's best fusion of…

00:04
2026-06-16
lesswrong.com
large-language-models

Synthetic document finetuning for instilling positive traits

Google DeepMind researchers trained Gemini 3 Flash to exhibit positive traits by midtraining on synthetic documents describing the model's traits, then finetuning on synthetic chat data where it demon…

22:36
2026-06-15
golproductions.com
ai-safety

67% of AI-generated commands are unsafe. We tested it

A test of Google's Gemini 3 Flash Preview found that 67% of AI-generated curl commands were unsafe, targeting internal networks, cloud metadata endpoints, or localhost. The test, conducted by Check, g…

19:45
2026-06-14
lesswrong.com
ai-safety

Why Do Naive SFT Filters For Safety Properties Fail?

Google DeepMind researchers investigate why filtering supervised fine-tuning (SFT) data fails to remove safety-relevant properties from language models, proposing a method to identify the source of th…

08:44
2026-06-14
openrouter.ai
artificial-intelligence

Surpassing Frontier Performance with Fusion

OpenRouter launched Fusion, a tool that synthesizes outputs from multiple AI models into a single response, achieving scores of 69.0% on the DRACO deep research benchmark—surpassing individual frontie…

15:31
2026-06-13
lesswrong.com
ai-safety

SFT Drives Gemini’s Safety Properties

Google DeepMind researchers found that supervised fine-tuning (SFT), not reinforcement learning, drives most safety properties in Gemini models. Comparing SFT-only versions of Gemini 3.1 Pro and Gemin…

← prev page 2 / 2
// co-occurs with top 8 entities
// topics top 6 topics