cd/entity/Gemma 3· home› entities› Gemma 3
grep -l @gemma 3 /news/*.json | wc -l → 56

Gemma 3

mentions 56 type Person page 2/3 feed RSS

// recent coverage 56 mentions

15:04
2026-09-18
sgnt.ai
large-language-models

You could have built Jev

A technical explainer published on typesafe.ai argues that Jev, a system for answering structured questions with LLMs, can be reconstructed from first principles by combining a prompt-writer, a single…

12:00
2026-08-28
kdnuggets.com
artificial-intelligence

Quantization and Pruning Methods to Make Your LLM Leaner

Quantization and pruning can shrink a 70B parameter model from 140GB to 35-40GB, enabling deployment on a single GPU instead of a four-A100 cluster costing $80,000-$100,000, according to Pristren's br…

13:21
2026-08-25
clojuriststogether.org
large-language-models

Clojurists Together Short Term Project Updates

Clojurists Together's Q2 2026 funded project iLLaManati, led by Dragan Djuric, delivered a high-performance local LLM library in 100% Clojure, achieving 16 tokens/second on a 7-year-old CPU and 80+ to…

04:00
2026-08-24
machinebrief.com
large-language-models

Asymmetric Capacity Allocation in Self-Refinement Pipelines

A new arXiv study (2608.21345v1) presents the first stage-wise model size analysis of self-refinement pipelines, testing 6 model sizes of Qwen3 and 4 model sizes of Gemma 3 across 5 benchmarks. The re…

10:51
2026-08-23
discuss.huggingface.co
artificial-intelligence

Fourier Magnitude KV Cache Quantization

A technical analysis of Fourier Magnitude KV Cache Quantization concludes that while the original claim of Fourier/phase being uniquely special is not supported, a substantive signal remains: the exac…

06:15
2026-08-17
dejan.ai
artificial-intelligence

AI Models Encode Brand Data but Fail to Recall a Third of It

Google Research found that frontier AI models, including Gemini-3-Pro and GPT-5, encode 95-98% of brand facts in their neural weights but fail to recall 26-34% of them, indicating the bottleneck is re…

09:15
2026-08-11
sebastianraschka.com
artificial-intelligence

Muse Glimmer 30B Architecture Notes

Meta released Muse Glimmer, a 30B open-weight multimodal reasoning model with a Gemma-like architecture, featuring a 131k context window, dense design, hybrid attention with a 3:1 sliding-window-to-gr…

19:39
2026-07-30
lesswrong.com
large-language-models

Internal State Control is a General Property of LLMs

A replication study by the Second Look Fellowship finds that internal state control is a general property of large language models, with 14 models across the Qwen3, Gemma 3, and Tulu 3 families (0.3B …

13:00
2026-07-23
spectrum.ieee.org
artificial-intelligence

NASA Puts Google’s Gemma Large Language Model in Orbit

NASA's Jet Propulsion Laboratory has successfully demonstrated the first in-orbit use of a vision-language model, Google's Gemma 3, aboard Loft Orbital's YAM-9 satellite. The NAVI-Orbital system achie…

22:04
2026-07-11
sourcefeed.dev
ai-infrastructure

Demystifying the NVIDIA DGX Spark for API Developers

NVIDIA's DGX Spark desktop GPU, with 128 GB unified memory and a 140W ARM64 processor, challenges API developers to shift from cloud-based AI consumption to local systems engineering. The device's sha…

01:32
2026-07-10
digital-foundry-eight.vercel.app
large-language-models

I benchmarked every model that fits on an iPhone

An independent benchmark of on-device LLMs on iPhone A17 Pro found Apple's system model achieves ~149 tok/s with only 12MB peak app memory, while 4B-class open models like Qwen3 4B and Llama 3.2 3B tr…

← prev page 2 / 3 next →
// co-occurs with top 8 entities
// topics top 6 topics