Gemma is Google DeepMind's family of lightweight, state-of-the-art open models.
Contents
Start Here
Gemma Documentation β Official documentation for selecting, running, tuning, and deploying Gemma models.
Get Started with Gemma β Get started running inference with the multimodal Gemma 4 models.
Gemma Cookbook β Maintained notebooks, examples, workshops, and end-to-end applications.
Gemma Skills β Reusable Agent Skills for selecting, running, and training Gemma models.
Gemma Events β Overview of upcoming Gemma events.
Gemma on X β For news, announcements, and updates about Gemma.
Models
Core Models
Variants
DiffusionGemma β Experimental discrete-diffusion text generation based on Gemma 4.
EmbeddingGemma β Compact embedding model designed for retrieval and on-device use.
FunctionGemma β Foundation for building specialized function-calling models.
MedGemma β Models optimized for medical text and image comprehension.
PaliGemma 2 β Vision-language models for detailed image understanding tasks.
ShieldGemma 2 β Image-safety classifier built on Gemma 3.
T5Gemma 2 β Encoder-decoder models for contextual understanding and generation.
TranslateGemma β Translation models covering 55 languages.
TxGemma β Models for therapeutic-development research.
VaultGemma β Language model trained with differential privacy.
DataGemma β Models and recipes for grounding responses with Data Commons.
RecurrentGemma β Open models based on the recurrent Griffin architecture.
Gemma Scope 2 β Open sparse autoencoders and interpretability tooling for studying Gemma 3.
Gemma-APS β Abstractive proposition segmentation for decomposing text into meaningful claims.
Cell2Sentence-Scale β A Gemma 2 27B model fine-tuned for single-cell biology.
DolphinGemma β Uses dolphin audio to help scientists study how dolphins communicate.
Inference
Local
HF Transformers β Python library for loading, running, and fine-tuning Hugging Face models.
llama.cpp β LLM inference in C/C++ with GGUF quantization.
Unsloth β Local UI to run and train LLMs and diffusion models.
Ollama β Get up and running with large language models locally.
LM Studio β Desktop application to discover, download, and run local models.
vLLM β High-throughput and memory-efficient LLM serving engine.
SGLang β Fast serving framework for large language models and vision-language models.
AI Edge Gallery β On-device ML models and examples for mobile and edge devices.
LiteRT β Google's runtime for on-device ML deployment.
JAX β Official Gemma reference implementation in JAX and Flax.
React Native β Run on-device Gemma models within React Native using ExecuTorch.
GenieX β Run Gemma on Qualcomm hardware.
Docker β Run Gemma 4 in Docker.
Hosted
Gemini Enterprise Agent Platform (Formerly Vertex AI) β Fully managed enterprise AI platform on Google Cloud.
OpenRouter β Unified API routing to multiple AI model providers.
Cerebras β High-speed Gemma 4 inference on Cerebras.
NVIDIA β Optimized TensorRT-LLM and NVFP4 checkpoints.
AMD β Support for AMD ROCm GPUs and processors.
AI Studio β Web-based prototyping and development environment.
Cloud Run β Deploy containerized Gemma services with autoscaling GPUs.
LiveKit β Real-time multimodal voice and video inference infrastructure.
Together AI β Cloud platform for running and fine-tuning open source models.
Modal β Run and deploy Gemma 4 on the Modal platform.
Fireworks β Run and deploy Gemma 4 on the Fireworks.AI platform.
BaseTen β Run and deploy Gemma 4 on the BaseTen platform.
Runpod β Experiment, train, fine-tune, and deploy Gemma.
Cloudflare β Run Gemma 4 on the Workers AI LLM Playground.
Fine-Tune
Fine-Tune Gemma β Official framework guide covering Keras, JAX, Hugging Face, Unsloth, Axolotl, and Google Cloud.
Gemma Cookbook: Training β Official fine-tuning notebooks and training recipes.
Tunix β JAX-native library for post-training generative models.
Unsloth Gemma 4 fine-tuning guide β Train Gemma 4 E2B, E4B, 12B, 26B A4B and 31B with Unsloth.
Gemma Multimodal Tuner β Fine-tune Gemma 3n and Gemma 4 with text, images, and audio on Apple Silicon.
MLX Tune β MLX-native SFT, preference tuning, and multimodal fine-tuning with Gemma 4 support.
Tutorials
Demos and Applications
Gemma 4 Vision Token Budget β Explore the effect of image resolution and visual-token budgets.
Concurrent Gemma β Run and compare multiple concurrent local Gemma instances.
See what 3 builders are making with Gemma 4 β Various applications developed by the community.
AIventure β A 2D grid-based adventure game built with Phaser 3 and Angular with Gemma driving it.
Gemma Chat β Local AI chat + coding agent for Apple Silicon, powered by Gemma 4 via MLX / Supports Ollama.
Build with Gemma 4 and Haystack β Runnable notebook covering RAG, visual question answering, a multimodal weather agent, and GitHub tool discovery.
Gemma 4 Browser Extension β Local browser agent powered by Gemma 4, WebGPU, and Transformers.js.
WebGemma β Browser playground and interactive model timeline powered by WebGPU and Transformers.js.
Controlling an iOS simulator β Gemma 4 using Argent to control an iOS simulator showcasing its capabilities in agentic workflows.
Automated Video Segmentation & Tracking β A demo that uses Gemma 4 + Falcon Perception for video tracking.
Parking Lot Car Detection & Segmentation β Gemma 4 analyzes the scene, decides the questions, generates prompts, and calls SAM 3.1 as a tool. SAM 3.1 segments and returns results.
Gemma 4 and MTP as a Marathon Engine β Benchmarks speculative decoding across increasing context lengths.
Cactus Hybrid β Post-trained Gemma 4 models to recognize when they are wrong, run on any framework.
Damage Scout β Damage Scout samples frames from a rental car walkaround, sends them to Gemma 4, gets back structured findings and box coordinates, then renders an annotated damage report in under 6 seconds.
MedGemma Impact Challenge β The winners of the MedGemma hackathon to build human-centered AI applications with MedGemma.
Gemma-Translator β A fully offline device powered by Gemma 4 E2B built with Google Antigravity.
Real-Time Voice AI with Gemma 4 β Open-source cascaded voice stack using Gemma 4 for low-latency reasoning.
Gemma 4 Good Challenge
Amazing projects that harness the power of Gemma 4 to drive positive change and global impact.
Trido β A Voice-Driven AI Whiteboard Built for the Teacher Nobody Builds For.
CodeBuddy β AI Python Tutor for Indonesian Students.
Port-a-Prof β Deeper learning, wherever you are.
TriageMate β Offline-first Clinical AI for Ghana's Community Health Officers.
ORCA-G4 β On-device oral cancer intelligence for 900,000 ASHA workers in rural India.
DEMENTOR β Edge AI Triage for Dementia Care.
PreVillage β A source-backed navigator for Nepalβs government services, built to find the office route, not just the form.
BrailleOut β An assistive device that reads the text and images from real-world and converts it to Braille using Gemma 4 and Ollama.
Gem-Care β Gemma-4-Enriched with Multimodal Clinical-context Adaptation for Recognition Enhancement of Non-Normative Speech.
Trajectix β An Agentic Flight Recorder for AI Infrastructure Safety.
TrueVoice β AI Voice Deepfake Detector.
AI Conceptualizer β 3D visualizations for mechanistic interpretability and "concept spectroscopy".
AcuΓferoΒ·VigΓa β Hybrid edge-and-citizen flood early warning for Argentina's Litoral, where every minute of warning is a life.
ResQ β Offline Multilingual Disaster Response Coach on Gemma 4 E2B.
Gemma in Space
Starcloud-1 β Starcloud deployed and ran Gemma in orbit aboard an H100 GPU.
NASA β NASA runs Gemma in orbit to analyze satellite imagery and compress visual data into text for rapid, low-bandwidth disaster response.
Research and Evaluation
This is not an officially supported Google product. This project is not eligible for the Google Open Source Software Vulnerability Rewards Program .