cd/entity/Gemma· home entities Gemma
grep -l @gemma /news/*.json | wc -l → 149

Gemma

mentions 149 type Organization page 1/8 feed RSS
sameAs · en.wikipedia.org

// recent coverage 149 mentions

06:29
2026-08-19
discuss.huggingface.co
large-language-models

Looking for an Open-Source LLM to Replace Llama 3.3 70B Versatile

A developer is seeking open-source replacements for Llama 3.3 70B Versatile after Groq shut down the model, focusing on smaller parameter sizes with comparable performance for RAG, agentic workflows, …

15:30
2026-08-18
lifehacker.com
large-language-models

How to Run a Local LLM on Your Phone (and Why You’d Want To)

Lifehacker reports that running a local large language model (LLM) on a smartphone is now practical for most recent devices, requiring at least 6GB of RAM and using apps like Atomic Chat or PocketPal …

15:54
2026-08-17
ziggit.dev
machine-learning

Fucina: a CPU-first tensor/autograd library in Zig

Matteo Grella released Fucina v0.1.0, a CPU-first tensor/autograd library written in pure Zig 0.16, featuring compile-time axis names in tensor types that make misaligned contractions a compile error.…

16:08
2026-08-16
sourcefeed.dev
artificial-intelligence

Unsloth Turns Fine-Tuning Into a Desktop App

Unsloth, the open-source fine-tuning library, released Unsloth Desktop, a free beta app for macOS, Windows, and Linux that runs and trains LLMs, diffusion models, and audio models locally, directly co…

12:31
2026-08-16
pub.towardsai.net
large-language-models

Your KV Cache Is Bigger Than Your Model

OpenAI's gpt-oss-120b model, with open weights, requires 72 KiB of KV cache per token in 16-bit precision, calculated from its config.json with 36 layers, 8 key-value heads, and a head dimension of 64…

04:22
2026-08-16
lesswrong.com
artificial-intelligence

Does DiffusionGemma do latent reasoning?

Google DeepMind's DiffusionGemma, a diffusion-based text generation model, remains highly monitorable despite its latent reasoning capabilities, according to a new analysis that strengthens prior find…

00:00
2026-08-11
mindstudio.ai
artificial-intelligence

How to Run Maple-Preview Locally on a Mac Mini M4

DeepGrove's Maple-Preview 20B-A1B, an open-source ternary-weight reasoning model with a 5.31 GB checkpoint, runs locally on a base Mac mini M4 at 218 tokens per second, according to the company. The m…

00:00
2026-08-11
mindstudio.ai
artificial-intelligence

Meta Muse Glimmer 30B: How to Run It Locally and Is It Worth It?

Meta released Muse Glimmer, a 30 billion parameter open-weight language model under Apache 2.0, designed for agentic tasks and positioned as a competitor to Qwen 3.6 27B. The model is available as an …

15:30
2026-08-05
blog.bytebytego.com
machine-learning

How Big Models Teach Small Models to Be Smart

Knowledge distillation, a method where a smaller 'student' model is trained to mimic a larger 'teacher' model, can produce a student that matches or beats the teacher on specific tasks, according to a…

22:09
2026-08-04
byteiota.com
ai-infrastructure

Google’s 2025 Open Source Report: What Devs Must Know

Google's 2025 open source report highlights 20,000 non-Google contributors, 400 million Gemma downloads, and $2 million in project funding, but the real story is a shared behavioral pattern with Anthr…

16:46
2026-08-04
promptcube3.com
machine-learning

DiffusionGemma’s Real Speed Trick

DiffusionGemma, a text diffusion model, achieves a 2.2× step-speedup and 31-point GPU compute efficiency gain over autoregressive decoding, according to a performance review by an ML engineer. The com…

13:58
2026-08-04
huggingface.co
artificial-intelligence

Deploy local agents everywhere with LFM2.5-2.6B

Liquid AI released LFM2.5-2.6B, a 2.6-billion-parameter on-device agentic model that outperforms models up to 4x larger on tool use and instruction following, achieving 220 tokens per second on an App…

page 1 / 8 next →
// co-occurs with top 8 entities
// topics top 6 topics