cd/entity/Gemma 4· home entities Gemma 4
grep -l @gemma 4 /news/*.json | wc -l → 227

Gemma 4

mentions 227 type Person page 1/12 feed RSS

// recent coverage 227 mentions

05:09
2026-07-25
sourcefeed.dev
artificial-intelligence

Small Models That Know When to Phone Home

Cactus shipped Cactus Hybrid, a post-trained build of Google's Gemma 4 E2B that returns a calibrated confidence score with every answer, routing queries to a larger model when confidence falls below 0…

11:01
2026-07-23
promptcube3.com
artificial-intelligence

Gemma 4 vs Gemini 3.1 Flash-Lite: The Hybrid Win

Google's Gemma 4 model uses a 68k-parameter probe layer that predicts decoding errors by reading hidden states, achieving a 0.79-0.88 AUROC on audio benchmarks despite being trained on zero audio data…

07:00
2026-07-22
dotnetperls.com
large-language-models

Laguna XS Model in OpenCode

Poolside AI's Laguna XS model, a small local LLM, performed well in refactoring Rust code via OpenCode, making relatively few errors and handling tool calls correctly except for escaped quotes. The 20…

13:00
2026-07-21
android-developers.googleblog.com
artificial-intelligence

Build intelligent Android apps: On-device inference

Google announced that its Gemini Nano 4 model, now running on over 140 million devices, can be used through ML Kit's Prompt API to build on-device intelligent features in Android apps, such as summari…

03:48
2026-07-21
dev.to
artificial-intelligence

Gemma 4 E2B on a Single TPU v6e Chip: A Serving Deep Dive

Google's Gemma 4 E2B model serves efficiently on a single TPU v6e chip, achieving 213 tok/s for a single user and scaling to ~2,200 output tok/s across concurrent streams, while its QAT variants fail …

05:42
2026-07-18
pub.towardsai.net
ai-agents

Running Google ADK with Gemma 4 and Ollama

A developer successfully ran Google's Agent Development Kit (ADK) with the locally hosted Gemma 4 model via Ollama, building a weather agent that uses tool calling to orchestrate Google Maps and Open-…

14:36
2026-07-16
artificialanalysis.ai
artificial-intelligence

Inkling Benchmark Results

Thinking Machines has released Inkling, a 975B-parameter open weights model with 41B active parameters, debuting at 41 on the Artificial Analysis Intelligence Index and becoming the leading open weigh…

02:33
2026-07-16
rockyshikoku.medium.com
artificial-intelligence

Running Gemma4 on Apple Neural Engine

A developer successfully ran Google's Gemma 4 E2B and E4B models on Apple's Neural Engine (ANE) using CoreML, achieving a memory footprint under 1 GB on an iPhone 17 Pro. The project overcame ANE's la…

17:03
2026-07-15
sourcefeed.dev
artificial-intelligence

Old Xeons Can Run Gemma 4 at Reading Speed

A 26-billion-parameter Gemma 4 mixture-of-experts model can run at reading speed on a dual-socket 2013 Ivy Bridge Xeon server with no GPU, achieving ~5.2 tokens per second decode on a Q8_0 quant, acco…

17:10
2026-07-14
9to5google.com
artificial-intelligence

Google announces Gemma 4 optimized for the Pixel 10’s TPU

Google announced Gemma 4 E2B for TPU, a model optimized to run natively on the Pixel 10's Tensor G5 TPU, at I/O Connect India. The multimodal model enables offline AI chat, image identification, and a…

16:50
2026-07-14
developers.googleblog.com
artificial-intelligence

Unlocking the Next Era of On-Device AI with Google Tensor and Pixel

Google unveiled Gemma 4 E2B for TPU and Functional Gemma at Google I/O India, demonstrating how Google Tensor's custom SoC and TPU enable 100% private on-device AI for the Pixel 10 family. The models …

page 1 / 12 next →
// co-occurs with top 8 entities
// topics top 6 topics