cd/entity/Gemma 4· home entities Gemma 4
grep -l @gemma 4 /news/*.json | wc -l → 291

Gemma 4

mentions 291 type Person page 4/15 feed RSS

// recent coverage 291 mentions

17:37
2026-07-25
cryptobriefing.com
artificial-intelligence

Gemma model family surpasses 900M downloads milestone

Google DeepMind's open-weight Gemma model family surpassed 900 million total downloads by July 2026, driven largely by the Gemma 4 release in April 2026, which alone accounted for over 300 million dow…

16:15
2026-07-25
dev.to
artificial-intelligence

Ask me anything (RAG)

A developer built AMA, a Retrieval-Augmented Generation (RAG) system for Lagos State University (LASU) that uses Gemma 4 via OpenRouter to answer student questions about lecturers, hostels, and intern…

05:09
2026-07-25
sourcefeed.dev
artificial-intelligence

Small Models That Know When to Phone Home

Cactus shipped Cactus Hybrid, a post-trained build of Google's Gemma 4 E2B that returns a calibrated confidence score with every answer, routing queries to a larger model when confidence falls below 0…

11:01
2026-07-23
promptcube3.com
artificial-intelligence

Gemma 4 vs Gemini 3.1 Flash-Lite: The Hybrid Win

Google's Gemma 4 model uses a 68k-parameter probe layer that predicts decoding errors by reading hidden states, achieving a 0.79-0.88 AUROC on audio benchmarks despite being trained on zero audio data…

07:00
2026-07-22
dotnetperls.com
large-language-models

Laguna XS Model in OpenCode

Poolside AI's Laguna XS model, a small local LLM, performed well in refactoring Rust code via OpenCode, making relatively few errors and handling tool calls correctly except for escaped quotes. The 20…

13:00
2026-07-21
android-developers.googleblog.com
artificial-intelligence

Build intelligent Android apps: On-device inference

Google announced that its Gemini Nano 4 model, now running on over 140 million devices, can be used through ML Kit's Prompt API to build on-device intelligent features in Android apps, such as summari…

03:48
2026-07-21
dev.to
artificial-intelligence

Gemma 4 E2B on a Single TPU v6e Chip: A Serving Deep Dive

Google's Gemma 4 E2B model serves efficiently on a single TPU v6e chip, achieving 213 tok/s for a single user and scaling to ~2,200 output tok/s across concurrent streams, while its QAT variants fail …

05:42
2026-07-18
pub.towardsai.net
ai-agents

Running Google ADK with Gemma 4 and Ollama

A developer successfully ran Google's Agent Development Kit (ADK) with the locally hosted Gemma 4 model via Ollama, building a weather agent that uses tool calling to orchestrate Google Maps and Open-…

14:36
2026-07-16
artificialanalysis.ai
artificial-intelligence

Inkling Benchmark Results

Thinking Machines has released Inkling, a 975B-parameter open weights model with 41B active parameters, debuting at 41 on the Artificial Analysis Intelligence Index and becoming the leading open weigh…

02:33
2026-07-16
rockyshikoku.medium.com
artificial-intelligence

Running Gemma4 on Apple Neural Engine

A developer successfully ran Google's Gemma 4 E2B and E4B models on Apple's Neural Engine (ANE) using CoreML, achieving a memory footprint under 1 GB on an iPhone 17 Pro. The project overcame ANE's la…

17:03
2026-07-15
sourcefeed.dev
artificial-intelligence

Old Xeons Can Run Gemma 4 at Reading Speed

A 26-billion-parameter Gemma 4 mixture-of-experts model can run at reading speed on a dual-socket 2013 Ivy Bridge Xeon server with no GPU, achieving ~5.2 tokens per second decode on a Q8_0 quant, acco…

← prev page 4 / 15 next →
// co-occurs with top 8 entities
// topics top 6 topics