cd/entity/Gemma 4· home entities Gemma 4
grep -l @gemma 4 /news/*.json | wc -l → 291

Gemma 4

mentions 291 type Person page 3/15 feed RSS

// recent coverage 291 mentions

00:00
2026-08-11
mindstudio.ai
artificial-intelligence

What Is Maple-Preview? DeepGrove's Ternary-Weight Reasoning Model

DeepGrove released Maple-Preview, an open-weight 20B-A1B ternary-weight mixture-of-experts reasoning model that runs at 218 tokens per second on an Apple Mac mini M4, with a 5.31 GB checkpoint and a 1…

00:00
2026-08-10
lmstudio.ai
artificial-intelligence

Run Muse Glimmer locally

LM Studio Bionic, in partnership with Meta, launched support for Muse Glimmer, a 30B-parameter open-source model optimized for agentic tasks, available for local download under the Apache 2.0 license.…

20:00
2026-08-08
pac.commonsware.com
artificial-intelligence

2026 Local Models: Analysis Over Generation

Local models such as Qwen 3.6 and Gemma 4 are close to being useful for Kotlin Multiplatform code generation but still fall short for many developers, according to Mark Murphy's experiments. Murphy su…

14:29
2026-08-08
dejan.ai
artificial-intelligence

Finding Bard Inside Google's Gemma 4

A mechanistic interpretability study found that ablating specific MLP neurons in Google's Gemma 4 model causes it to revert to its dormant Bard identity, with the model answering 'My name is Bard' wit…

07:00
2026-08-07
dotnetperls.com
artificial-intelligence

Improving Tool Use with Small LLMs

A developer using the LFM 2.5 2.6B model with llama-cpp found that adding detailed descriptions to JSON schema parameters improved tool call correctness from 0% to nearly 100%, enabling the use of a f…

18:18
2026-08-05
news.ycombinator.com
artificial-intelligence

Why Fireworks doesn't support Voice AI

Fireworks, an AI inference platform, does not support voice AI models such as Parakeet, Kokoro, and Qwen ASR because voice workloads require different optimization strategies than text LLMs, according…

04:00
2026-08-04
arxiv.org
large-language-models

DiffusionGemma Technical Report

Google DeepMind introduced DiffusionGemma, an experimental open-weight language model that uses discrete diffusion to generate text at high speed, achieving roughly 1,500 output tokens per second on a…

02:41
2026-08-03
news.ycombinator.com
artificial-intelligence

Aquila Voice Assistant Test Suite for Home Assistant

An open-source test suite for voice assistants, developed by Aquila, is now available, with reproducible tests and a leaderboard at https://git.cicero.sh/aquila/ha-voice-test-suite/. The creator, who …

17:08
2026-07-29
sourcefeed.dev
artificial-intelligence

A 26B Model in 2 GB of RAM, Courtesy of Your SSD

TurboFieldfare, a Swift-and-Metal runtime, runs Google's Gemma 4 26B-A4B model on Apple Silicon Macs with as little as 2 GB of RAM by streaming expert weights from SSD, achieving 5.1–6.3 tokens per se…

07:00
2026-07-28
dotnetperls.com
large-language-models

Nanbeige 3 Billion Parameter Model

A developer testing Nanbeige 4.2, a 3-billion-parameter looped-transformer model from a Chinese company, found it well-suited for tool calling with a Rust MCP server, outperforming Gemma 4 12B in that…

18:59
2026-07-25
github.com
artificial-intelligence

Running a 28.9M parameter LLM on an $8 microcontroller

A developer known as slvDev has run a 28.9 million parameter language model on an ESP32-S3 microcontroller that costs about $8, generating text at roughly 9.5 tokens per second entirely on-device. The…

← prev page 3 / 15 next →
// co-occurs with top 8 entities
// topics top 6 topics