cd/entity/Gemma· home entities Gemma
grep -l @gemma /news/*.json | wc -l → 149

Gemma

mentions 149 type Organization page 5/8 feed RSS
sameAs · en.wikipedia.org

// recent coverage 149 mentions

01:18
2026-06-28
dev.to
large-language-models

Getting Started with Ollama: Run LLMs Locally in 10 Minutes

Ollama provides a tool for running large language models locally on macOS, Linux, and Windows without requiring an API key or cloud service. The tool packages model weights, a runtime based on llama.c…

22:18
2026-06-27
dev.to
artificial-intelligence

Local AI - How to Run Open Source AI Models Locally

A developer's guide explains how to run open-source AI models locally on consumer hardware in 2026, highlighting that a mid-range laptop can now run models once considered frontier-class. The guide co…

20:16
2026-06-27
github.com
machine-learning

GitHub DeepSeek-AI/DeepSpec

DeepSeek-AI released DeepSpec, an open-source codebase for training and evaluating draft models for speculative decoding, supporting three draft model algorithms (DSpark, DFlash, Eagle3) and requiring…

06:54
2026-06-22
dev.to
large-language-models

Mastering Ollama AI endpoints: How to use each one correctly

Ollama provides a REST API with 14 endpoints for running large language models locally. The API includes endpoints for text generation, chat, embeddings, model management, and OpenAI compatibility. De…

00:00
2026-06-22
huggingface.co
large-language-models

We got local models to triage the OpenClaw repo for FREE!

Hugging Face engineer Onur developed a real-time notification system for the OpenClaw repository using local open-weight models like Gemma and Qwen, running on an NVIDIA GB10 with 128 GB of unified me…

18:03
2026-06-20
hackster.io
ai-tools

Offline AI Voice Assistant on Raspberry Pi 4 with Gemma

A developer built a fully offline voice assistant on a Raspberry Pi 4 or 5 using local AI models. The device records audio, processes it with Whisper for speech-to-text, runs a local language model vi…

20:25
2026-06-19
lmsys.org
large-language-models

The next generation of speculative decoding: DFlash and Spec V2

Modal and Z Lab released DFlash, a speculative decoding model for Qwen 3.5 397B-A17B, achieving over 4.3x throughput versus baseline and 1.5x versus MTP on HumanEval at concurrency 1. The model uses a…

01:53
2026-06-18
letsdatascience.com
large-language-models

Google releases OpenRL for LLM fine-tuning

Google released OpenRL, an open-source API for fine-tuning large language models on Kubernetes clusters, aiming to decouple infrastructure from AI research and improve GPU utilization by running multi…

← prev page 5 / 8 next →
// co-occurs with top 8 entities
// topics top 6 topics