cd/entity/Gemma· home entities Gemma
grep -l @gemma /news/*.json | wc -l → 148

Gemma

mentions 148 type Organization page 1/8 feed RSS
sameAs · en.wikipedia.org

// recent coverage 148 mentions

06:29
2026-08-19
discuss.huggingface.co
large-language-models

Looking for an Open-Source LLM to Replace Llama 3.3 70B Versatile

A developer is seeking open-source replacements for Llama 3.3 70B Versatile after Groq shut down the model, focusing on smaller parameter sizes with comparable performance for RAG, agentic workflows, …

15:30
2026-08-18
lifehacker.com
large-language-models

How to Run a Local LLM on Your Phone (and Why You’d Want To)

Lifehacker reports that running a local large language model (LLM) on a smartphone is now practical for most recent devices, requiring at least 6GB of RAM and using apps like Atomic Chat or PocketPal …

15:54
2026-08-17
ziggit.dev
machine-learning

Fucina: a CPU-first tensor/autograd library in Zig

Matteo Grella released Fucina v0.1.0, a CPU-first tensor/autograd library written in pure Zig 0.16, featuring compile-time axis names in tensor types that make misaligned contractions a compile error.…

16:08
2026-08-16
sourcefeed.dev
artificial-intelligence

Unsloth Turns Fine-Tuning Into a Desktop App

Unsloth, the open-source fine-tuning library, released Unsloth Desktop, a free beta app for macOS, Windows, and Linux that runs and trains LLMs, diffusion models, and audio models locally, directly co…

12:31
2026-08-16
pub.towardsai.net
large-language-models

Your KV Cache Is Bigger Than Your Model

OpenAI's gpt-oss-120b model, with open weights, requires 72 KiB of KV cache per token in 16-bit precision, calculated from its config.json with 36 layers, 8 key-value heads, and a head dimension of 64…

04:22
2026-08-16
lesswrong.com
artificial-intelligence

Does DiffusionGemma do latent reasoning?

Google DeepMind's DiffusionGemma, a diffusion-based text generation model, remains highly monitorable despite its latent reasoning capabilities, according to a new analysis that strengthens prior find…

00:00
2026-08-11
mindstudio.ai
artificial-intelligence

Meta Muse Glimmer 30B: How to Run It Locally and Is It Worth It?

Meta released Muse Glimmer, a 30 billion parameter open-weight language model under Apache 2.0, designed for agentic tasks and positioned as a competitor to Qwen 3.6 27B. The model is available as an …

00:00
2026-08-11
mindstudio.ai
artificial-intelligence

How to Run Maple-Preview Locally on a Mac Mini M4

DeepGrove's Maple-Preview 20B-A1B, an open-source ternary-weight reasoning model with a 5.31 GB checkpoint, runs locally on a base Mac mini M4 at 218 tokens per second, according to the company. The m…

15:30
2026-08-05
blog.bytebytego.com
machine-learning

How Big Models Teach Small Models to Be Smart

Knowledge distillation, a method where a smaller 'student' model is trained to mimic a larger 'teacher' model, can produce a student that matches or beats the teacher on specific tasks, according to a…

22:09
2026-08-04
byteiota.com
ai-infrastructure

Google’s 2025 Open Source Report: What Devs Must Know

Google's 2025 open source report highlights 20,000 non-Google contributors, 400 million Gemma downloads, and $2 million in project funding, but the real story is a shared behavioral pattern with Anthr…

16:46
2026-08-04
promptcube3.com
machine-learning

DiffusionGemma’s Real Speed Trick

DiffusionGemma, a text diffusion model, achieves a 2.2× step-speedup and 31-point GPU compute efficiency gain over autoregressive decoding, according to a performance review by an ML engineer. The com…

13:58
2026-08-04
huggingface.co
artificial-intelligence

Deploy local agents everywhere with LFM2.5-2.6B

Liquid AI released LFM2.5-2.6B, a 2.6-billion-parameter on-device agentic model that outperforms models up to 4x larger on tool use and instruction following, achieving 220 tokens per second on an App…

08:00
2026-08-03
cmart.blog
artificial-intelligence

Why is Anthropic's public writing style so unlike Claude's?

Anthropic's public writing style differs markedly from the distinctive voice of its Claude AI model, which features short punchy sentences and phrases that have become widespread in LLM outputs. Simon…

page 1 / 8 next →
// co-occurs with top 8 entities
// topics top 6 topics