cd/entity/Gemma 4· home entities Gemma 4
grep -l @gemma 4 /news/*.json | wc -l → 291

Gemma 4

mentions 291 type Person page 9/15 feed RSS

// recent coverage 291 mentions

17:11
2026-06-05
dev.to
artificial-intelligence

Gemma 4 makes on-device multimodal AI good enough to ship

Google released the Gemma 4 family of open-weight AI models this week, ranging from a 2.3B-parameter E2B model up to a 31B dense model, with the smallest variants designed to run on-device. The E2B an…

09:55
2026-06-05
letsdatascience.com
large-language-models

Google LiteRT-LM Accelerates Gemma 4 Local Inference

Google added native support for Gemma 4 Multi-Token Prediction (MTP) to LiteRT-LM, its on-device LLM runtime built on LiteRT (formerly TensorFlow Lite). Google reports the integration yields MTP decod…

17:48
2026-06-04
hitoku.me
ai-products

Show HN: Hitoku Draft – Context aware local assistant

Hitoku Draft, an open-source, voice-first AI assistant that runs entirely locally, now includes transcription with voice editing and context-aware features that read the user's screen, documents, and …

09:53
2026-06-04
letsdatascience.com
large-language-models

Google releases Gemma 4 12B for laptop inference

Google released Gemma 4 12B, a 12-billion-parameter multimodal model designed to run on consumer laptops with 16GB of RAM or VRAM. The encoder-free model accepts text, images, and native audio, and it…

00:00
2026-06-04
coderabbit.ai
large-language-models

Nemotron 3 Ultra makes the case for fast, open coding models

NVIDIA released Nemotron 3 Ultra, a 550-billion-parameter open coding model with 55 billion active parameters per token, designed for fast, agentic developer workflows. The model achieves over 300 out…

16:04
2026-06-03
blog.google
artificial-intelligence

Gemma 4 12B: A unified, encoder-free multimodal model

Google released Gemma 4 12B, a new multimodal AI model designed to run locally on consumer laptops with just 16GB of RAM. The model uses an encoder-free architecture to process images and audio direct…

19:39
2026-05-30
dev.to
ai-tools

Structure: A Local-First Interview IDE Powered by Gemma 4

A developer built Structure, a local-first macOS desktop IDE for technical interview practice powered by Gemma 4. The app, built with Qt 6 and C++, runs Gemma 4 locally through Ollama and stores all c…

17:40
2026-05-28
dev.to
artificial-intelligence

AI Tools & Products Radar — May 28, 2026

AI coding startup Cognition raised $1 billion at a $25 billion valuation, while OpenRouter doubled its valuation to $1.3 billion in one year. Microsoft shipped Copilot Cowork, an autonomous AI agent b…

15:22
2026-05-28
developers.googleblog.com
artificial-intelligence

How the community trained Gemma to "Think" with Tunix and TPUs

Google hosted a Kaggle hackathon challenging developers to train non-reasoning Gemma-2-2B and Gemma-3-1B models into general reasoning models using Tunix and Kaggle TPUs. Over 11,000 entrants and 300+…

02:24
2026-05-28
dev.to
large-language-models

Quantizing Gemma 4 on Mac with llama.cpp

A developer successfully quantized Google's Gemma 4 model to 4-bit precision using llama.cpp on a Mac with Metal acceleration. The process involved converting the model to GGUF format and applying the…

← prev page 9 / 15 next →
// co-occurs with top 8 entities
// topics top 6 topics