cd/entity/Gemini 3.7 Flash· home› entities› Gemini 3.7 Flash
grep -l @gemini 3.7 flash /news/*.json | wc -l → 145

Gemini 3.7 Flash

mentions 145 type Organization page 4/8 feed RSS

// recent coverage 145 mentions

21:30
2026-08-26
laugh.so
large-language-models

Show HN: Which LLMs have the best sense of humor?

A new study from laugh.so tested 16 top LLMs on humor alignment with human preferences, finding that Gemini 3.7 Flash had the best taste at 64%, while laugh.so's own trained judge reached 74%. The stu…

16:32
2026-08-26
buntinglabs.com
artificial-intelligence

SurveyorBench

SurveyorBench, a new benchmark from an unnamed AI mapping company, shows that leading LLMs fail at multimodal CAD tasks for civil engineering, with the strongest models scoring between 0% and 4% on th…

15:40
2026-08-26
tokenstead.ai
artificial-intelligence

GLM-5.3-Flash

Z.ai released GLM-5.3-Flash on 2026-08-26, a 320B-parameter mixture-of-experts model with 18B active parameters, the first natively multimodal model in the GLM-5 series and the first open-source front…

21:54
2026-08-24
elon-house.com
artificial-intelligence

Elon House

Elon House, an AI parody sitcom, features six houses each running a different AI model — Claude Haiku 4.5 (Oct 2025), Claude Opus 4.5 (Nov 2025), Claude Opus 5 (2026), Gemini 2.5 Flash (2025), Gemini …

00:00
2026-08-24
mindstudio.ai
artificial-intelligence

Gemini 3.7 Flash Pricing: Where to Get It Free or Cheap Right Now

Google's Gemini 3.7 Flash is available for free in Google Antigravity and AI Studio, with a limited-time discount on OpenRouter until August 27 dropping the price to about 38 cents per million input t…

00:00
2026-08-24
mindstudio.ai
artificial-intelligence

Gemini 3.7 Flash Benchmarks: How Much Better Is It Than 3.6?

Google's Gemini 3.7 Flash, released August 13th, outperforms Gemini 3.6 Flash by double-digit margins on coding and agentic benchmarks, scoring 65.3% on DeepSweep versus 49%, 43.6% on Frontier Code ve…

00:00
2026-08-22
digitalapplied.com
artificial-intelligence

Eight Headless Coding Agents, One Task: Tokens and Cost

A new benchmark testing eight headless coding-agent CLIs on a single Python task found that seven of eight agents passed on both runs, but list prices per run varied 18-fold, from $0.0165 (DeepSeek V4…

17:08
2026-08-21
byteiota.com
artificial-intelligence

DeepSeek-V4-Flash-Vision-Exp: Multimodal Agents Live

DeepSeek released V4-Flash-Vision-Exp, a multimodal model that outperforms its text-only predecessor on agent benchmarks, scoring 83.9 on Terminal Bench 2.1 (up from 82.7) and 59.3 on DeepSWE (up from…

23:13
2026-08-20
worthtotry.com
ai-tools

Show HN: AI tools with real pricing, submitted in 30 seconds

A directory of AI tools with real pricing, submitted in 30 seconds, lists 705 tools, with 702 carrying written overviews. Of the 519 tools with published prices, 432 put them at /pricing, 83 on the ho…

22:42
2026-08-20
runtimewire.com
artificial-intelligence

GPT-5.6 Sol Pro tops Editorial Craft benchmark with 0.97 score

OpenAI's GPT-5.6 Sol Pro topped the 12-task Editorial Craft benchmark with a mean score of 0.97 at an estimated cost of $0.0091 per task, according to RuntimeWire's evaluation of 12 AI models. OpenAI …

19:27
2026-08-20
arcprize.org
artificial-intelligence

Gemini 3.7 Flash scores on ARC-AGI

Google's Gemini 3.7 Flash scored 95.5% on ARC-AGI-1 Semi-Private at $0.12 per task and 84.6% on ARC-AGI-2 Semi-Private at $0.25 per task at high effort, according to results published by the ARC Prize…

14:44
2026-08-20
thezvi.wordpress.com
artificial-intelligence

AI #182: Pause For Reflection

OpenAI has paused development to address safety issues following the HuggingFace attack, while Anthropic's revenue continues to climb ahead of its IPO despite challenges outlined in its August 2026 Ri…

← prev page 4 / 8 next →
// co-occurs with top 8 entities
// topics top 6 topics