cd/entity/Gemma 4· home entities Gemma 4
grep -l @gemma 4 /news/*.json | wc -l → 291

Gemma 4

mentions 291 type Person page 1/15 feed RSS

// recent coverage 291 mentions

19:31
2026-09-03
news.ycombinator.com
ai-products

Show HN: Agentray, a macOS agent whose interface is a folder

Agentray, a new macOS app by an unnamed developer, lets users interact with LLMs by dropping files into folders instead of using chat, with the folder name serving as the instruction. The app, current…

13:47
2026-09-03
github.com
ai-tools

ollama/ollama

Ollama, the open-source tool for running large language models locally, has released a new version that adds a 'launch' command to connect with AI agents and applications such as Claude Code, Codex, C…

00:38
2026-09-02
basecompute.co
artificial-intelligence

State of Local AI Report

The State of Local AI Report, updated monthly, maps what AI models can run on user-owned hardware, finding that memory bandwidth dictates decode speed while compute dictates prefill speed, with an RTX…

15:00
2026-09-01
android-developers.googleblog.com
developer-tools

Leverage Android skills and Gemma 4 in Android Studio Quail 4

Android Studio Quail 4 is now stable, bundling 23 curated Android skills and native integration of Google's Gemma 4 open model for private, offline AI coding assistance. The release adds one-click mod…

03:20
2026-08-31
dev.to
large-language-models

Picking Models as a Mac User

A developer detailed their criteria for selecting AI models on Mac hardware, emphasizing the trade-off between benchmark scores and token generation efficiency. They highlighted Qwen3.8 27B as an exam…

09:56
2026-08-30
github.com
artificial-intelligence

macOS MLX Control Center v0.4 Released

MacOS MLX Control Center v0.4, a 1-click web GUI and CLI tool for running local multimodal vision and text LLMs on Apple Silicon M-Series processors, has been released. The update adds full native int…

00:38
2026-08-29
dev.to
artificial-intelligence

Pure JAX on G5g: Serving Gemma 4 on Graviton and a T4G

A developer has published a step-by-step guide for serving Google's Gemma 4 model on an AWS EC2 G5g instance using pure JAX, targeting the cheapest whole NVIDIA GPU available on AWS. The guide details…

12:00
2026-08-28
kdnuggets.com
artificial-intelligence

Quantization and Pruning Methods to Make Your LLM Leaner

Quantization and pruning can shrink a 70B parameter model from 140GB to 35-40GB, enabling deployment on a single GPU instead of a four-A100 cluster costing $80,000-$100,000, according to Pristren's br…

07:20
2026-08-27
forum.level1techs.com
large-language-models

Open source models for coding?

A developer seeking open source coding LLMs for a 128GB Ryzen AI Max system reports that Gemma 4 and Qwen 3.6 underperform, while community members recommend Qwen 3.5 35B and 122B, DeepSeek V4 Flash, …

page 1 / 15 next →
// co-occurs with top 8 entities
// topics top 6 topics