Start HereModelsInferenceFine-TuneTutorialsDemos and ApplicationsGemma 4 Good ChallengeGemma in SpaceResearch and Evaluation
Gemma Documentationβ Official documentation for selecting, running, tuning, and deploying Gemma models.Get Started with Gemmaβ Get started running inference with the multimodal Gemma 4 models.Gemma Cookbookβ Maintained notebooks, examples, workshops, and end-to-end applications.Gemma Skillsβ Reusable Agent Skills for selecting, running, and training Gemma models.Gemma Eventsβ Overview of upcoming Gemma events.Gemma on Xβ For news, announcements, and updates about Gemma.
Gemma 4 Overviewβ An overview of the capabilities, architecture, and more of Gemma 4 models.Gemma 4 Model Cardβ Architecture, training, evaluation, safety, and usage details.Gemma 4 (Hugging Face)β Gemma 4 checkpoints (+ assistant models) on Hugging Face.Gemma 4 QATβ Quantization-aware checkpoints for local and server inference.Gemma 4 Mobile QATβ Mobile-optimized checkpoints for the E2B and E4B models.Previous Gemma Model Cardsβ Model cards for Gemma 1β3 and earlier Gemma families.
DiffusionGemmaβ Experimental discrete-diffusion text generation based on Gemma 4.EmbeddingGemmaβ Compact embedding model designed for retrieval and on-device use.FunctionGemmaβ Foundation for building specialized function-calling models.MedGemmaβ Models optimized for medical text and image comprehension.PaliGemma 2β Vision-language models for detailed image understanding tasks.ShieldGemma 2β Image-safety classifier built on Gemma 3.T5Gemma 2β Encoder-decoder models for contextual understanding and generation.TranslateGemmaβ Translation models covering 55 languages.TxGemmaβ Models for therapeutic-development research.VaultGemmaβ Language model trained with differential privacy.DataGemmaβ Models and recipes for grounding responses with Data Commons.RecurrentGemmaβ Open models based on the recurrent Griffin architecture.Gemma Scope 2β Open sparse autoencoders and interpretability tooling for studying Gemma 3.Gemma-APSβ Abstractive proposition segmentation for decomposing text into meaningful claims.Cell2Sentence-Scaleβ A Gemma 2 27B model fine-tuned for single-cell biology.DolphinGemmaβ Uses dolphin audio to help scientists study how dolphins communicate.
HF Transformersβ Python library for , running, and fine-tuning Hugging Face models.llama.cppβ LLM inference in C/C++ with GGUF quantization.Unslothβ Local UI to run and train LLMs and diffusion models.Ollamaβ Get up and running with large language models locally.LM Studioβ Desktop application to discover, download, and run local models.vLLMβ High-throughput and memory-efficient LLM serving engine.SGLangβ Fast serving framework for large language models and vision-language models.AI Edge Galleryβ On-device ML models and examples for mobile and edge devices.LiteRTβ Google's runtime for on-device ML deployment.JAXβ Official Gemma reference implementation in JAX and Flax.React Nativeβ Run on-device Gemma models within React Native using ExecuTorch.GenieXβ Run Gemma on Qualcomm hardware.Dockerβ Run Gemma 4 in Docker.
Gemini Enterprise Agent Platform (Formerly Vertex AI)β Fully managed enterprise AI platform on Google Cloud.OpenRouterβ Unified API routing to multiple AI model providers.Cerebrasβ High-speed Gemma 4 inference on Cerebras.NVIDIAβ Optimized TensorRT-LLM and NVFP4 checkpoints.AMDβ Support for AMD ROCm GPUs and processors.AI Studioβ Web-based prototyping and development environment.Cloud Runβ Deploy containerized Gemma services with autoscaling GPUs.LiveKitβ Real-time multimodal voice and video inference infrastructure.Together AIβ Cloud platform for running and fine-tuning open source models.Modalβ Run and deploy Gemma 4 on the Modal platform.Fireworksβ Run and deploy Gemma 4 on the Fireworks.AI platform.BaseTenβ Run and deploy Gemma 4 on the BaseTen platform.Runpodβ Experiment, train, fine-tune, and deploy Gemma.Cloudflareβ Run Gemma 4 on the Workers AI LLM Playground.
Fine-Tune Gemmaβ Official framework guide covering Keras, JAX, Hugging Face, Unsloth, Axolotl, and Google Cloud.Gemma Cookbook: Trainingβ Official fine-tuning notebooks and training recipes.Tunixβ JAX-native library for post-training generative models.Unsloth Gemma 4 fine-tuning guideβ Train Gemma 4 E2B, E4B, 12B, 26B A4B and 31B with Unsloth.Gemma Multimodal Tunerβ Fine-tune Gemma 3n and Gemma 4 with text, images, and audio on Apple Silicon.MLX Tuneβ MLX-native SFT, preference tuning, and multimodal fine-tuning with Gemma 4 support.
A Visual Guide to Gemma 4A Visual Guide to Gemma 4 12BA Visual Guide to DiffusionGemmaA Visual Guide to the Gemma 4 DraftersVariable Aspect Ratio and Variable Resolutions in Gemma 4How to Use Transformers.js in a Chrome ExtensionHow to run a local coding agent with Gemma 4 and PiHow to run Gemma 4 with OpenClawHow a Small Fix Improves Gemma 4 Vision PerformanceWhile I slept, my 5-year-old MacBook ran Gemma 4 locally and indexed a year of videoFine-tuning Gemma 4 12B on your own dataTurning Gemma 4 into an Old Korean Translator
Gemma 4 Vision Token Budgetβ Explore the effect of image resolution and visual-token budgets.Concurrent Gemmaβ Run and compare multiple concurrent local Gemma instances.See what 3 builders are making with Gemma 4β Various applications developed by the community.AIventureβ A 2D grid-based adventure game built with Phaser 3 and Angular with Gemma driving it.Gemma Chatβ Local AI chat + coding agent for Apple Silicon, powered by Gemma 4 via MLX / Supports Ollama.Build with Gemma 4 and Haystackβ Runnable notebook covering RAG, visual question answering, a multimodal weather agent, and GitHub tool discovery.Gemma 4 Browser Extensionβ Local browser agent powered by Gemma 4, WebGPU, and Transformers.js.WebGemmaβ Browser playground and interactive model timeline powered by WebGPU and Transformers.js.Controlling an iOS simulatorβ Gemma 4 using Argent to control an iOS simulator showcasing its capabilities in agentic workflows.Automated Video Segmentation & Trackingβ A demo that uses Gemma 4 + Falcon Perception for video tracking.Parking Lot Car Detection & Segmentationβ Gemma 4 analyzes the scene, decides the questions, generates prompts, and calls SAM 3.1 as a tool. SAM 3.1 segments and returns results.Gemma 4 and MTP as a Marathon Engineβ Benchmarks speculative decoding across increasing context lengths.Cactus Hybridβ Post-trained Gemma 4 models to recognize when they are wrong, run on any framework.Damage Scoutβ Damage Scout samples frames from a rental car walkaround, sends them to Gemma 4, gets back structured findings and box coordinates, then renders an annotated damage report in under 6 seconds.MedGemma Impact Challengeβ The winners of the MedGemma hackathon to build human-centered AI applications with MedGemma.Gemma-Translatorβ A fully offline device powered by Gemma 4 E2B built with Google Antigravity.Real-Time Voice AI with Gemma 4β Open-source cascaded voice stack using Gemma 4 for low-latency reasoning.
Amazing projects that harness the power of Gemma 4 to drive positive change and global impact. Tridoβ A Voice-Driven AI Whiteboard Built for the Teacher Nobody Builds For.CodeBuddyβ AI Python Tutor for Indonesian Students.Port-a-Profβ Deeper learning, wherever you are.TriageMateβ Offline-first Clinical AI for Ghana's Community Health Officers.ORCA-G4β On-device oral cancer intelligence for 900,000 ASHA workers in rural India.DEMENTORβ Edge AI Triage for Dementia Care.PreVillageβ A source-backed navigator for Nepalβs government services, built to find the office route, not just the form.BrailleOutβ An assistive device that reads the text and images from real-world and converts it to Braille using Gemma 4 and Ollama.Gem-Careβ Gemma-4-Enriched with Multimodal Clinical-context Adaptation for Recognition Enhancement of Non-Normative Speech.Trajectixβ An Agentic Flight Recorder for AI Infrastructure Safety.TrueVoiceβ AI Voice Deepfake Detector.AI Conceptualizerβ 3D visualizations for mechanistic interpretability and "concept spectroscopy".AcuΓferoΒ·VigΓaβ Hybrid edge-and-citizen flood early warning for Argentina's Litoral, where every minute of warning is a life.ResQβ Offline Multilingual Disaster Response Coach on Gemma 4 E2B.
Starcloud-1β Starcloud deployed and ran Gemma in orbit aboard an H100 GPU.NASAβ NASA runs Gemma in orbit to analyze satellite imagery and compress visual data into text for rapid, low-bandwidth disaster response.
Gemma 4 Technical Reportβ The technical report covering Gemma 4 E2B, E4B, 12B, 26B A4B, and 31B.DiffusionGemma Technical Reportβ The technical report covering DiffusionGemma.Artificial Analysisβ Intelligence, Performance & Price Analysis.ChessBenchβ Chess LLM Benchmark Leaderboard.TERMS-Benchβ A benchmark for LLM negotiation agents based on economic negotiation.
This is not an officially supported Google product. This project is not eligible for the Google Open Source Software Vulnerability Rewards Program.