Skip to content

Latest commit

 

History

8 Commits

Folders and files

NameName
Last commit message
Last commit date
 
 
 
 
 
 
 
 
 
 

Repository files navigation




Gemma is Google DeepMind's family of lightweight, state-of-the-art open models.

Awesome

Contents

Start Here

  • Gemma Documentation β€” Official documentation for selecting, running, tuning, and deploying Gemma models.
  • Get Started with Gemma β€” Get started running inference with the multimodal Gemma 4 models.
  • Gemma Cookbook β€” Maintained notebooks, examples, workshops, and end-to-end applications.
  • Gemma Skills β€” Reusable Agent Skills for selecting, running, and training Gemma models.
  • Gemma Events β€” Overview of upcoming Gemma events.
  • Gemma on X β€” For news, announcements, and updates about Gemma.

Models

Core Models

Variants

  • DiffusionGemma β€” Experimental discrete-diffusion text generation based on Gemma 4.
  • EmbeddingGemma β€” Compact embedding model designed for retrieval and on-device use.
  • FunctionGemma β€” Foundation for building specialized function-calling models.
  • MedGemma β€” Models optimized for medical text and image comprehension.
  • PaliGemma 2 β€” Vision-language models for detailed image understanding tasks.
  • ShieldGemma 2 β€” Image-safety classifier built on Gemma 3.
  • T5Gemma 2 β€” Encoder-decoder models for contextual understanding and generation.
  • TranslateGemma β€” Translation models covering 55 languages.
  • TxGemma β€” Models for therapeutic-development research.
  • VaultGemma β€” Language model trained with differential privacy.
  • DataGemma β€” Models and recipes for grounding responses with Data Commons.
  • RecurrentGemma β€” Open models based on the recurrent Griffin architecture.
  • Gemma Scope 2 β€” Open sparse autoencoders and interpretability tooling for studying Gemma 3.
  • Gemma-APS β€” Abstractive proposition segmentation for decomposing text into meaningful claims.
  • Cell2Sentence-Scale β€” A Gemma 2 27B model fine-tuned for single-cell biology.
  • DolphinGemma β€” Uses dolphin audio to help scientists study how dolphins communicate.

Inference

Local

  • HF Transformers β€” Python library for loading, running, and fine-tuning Hugging Face models.
  • llama.cpp β€” LLM inference in C/C++ with GGUF quantization.
  • Unsloth β€” Local UI to run and train LLMs and diffusion models.
  • Ollama β€” Get up and running with large language models locally.
  • LM Studio β€” Desktop application to discover, download, and run local models.
  • vLLM β€” High-throughput and memory-efficient LLM serving engine.
  • SGLang β€” Fast serving framework for large language models and vision-language models.
  • AI Edge Gallery β€” On-device ML models and examples for mobile and edge devices.
  • LiteRT β€” Google's runtime for on-device ML deployment.
  • JAX β€” Official Gemma reference implementation in JAX and Flax.
  • React Native β€” Run on-device Gemma models within React Native using ExecuTorch.
  • GenieX β€” Run Gemma on Qualcomm hardware.
  • Docker β€” Run Gemma 4 in Docker.
  • Apple Core AI β€” Community Gemma 4 bundles (E2B, E4B, 12B, 31B) for Apple's on-device Core AI framework, each with recipe and measured speed.

Hosted

  • Gemini Enterprise Agent Platform (Formerly Vertex AI) β€” Fully managed enterprise AI platform on Google Cloud.
  • OpenRouter β€” Unified API routing to multiple AI model providers.
  • Cerebras β€” High-speed Gemma 4 inference on Cerebras.
  • NVIDIA β€” Optimized TensorRT-LLM and NVFP4 checkpoints.
  • AMD β€” Support for AMD ROCm GPUs and processors.
  • AI Studio β€” Web-based prototyping and development environment.
  • Cloud Run β€” Deploy containerized Gemma services with autoscaling GPUs.
  • LiveKit β€” Real-time multimodal voice and video inference infrastructure.
  • Together AI β€” Cloud platform for running and fine-tuning open source models.
  • Modal β€” Run and deploy Gemma 4 on the Modal platform.
  • Fireworks β€” Run and deploy Gemma 4 on the Fireworks.AI platform.
  • BaseTen β€” Run and deploy Gemma 4 on the BaseTen platform.
  • Runpod β€” Experiment, train, fine-tune, and deploy Gemma.
  • Cloudflare β€” Run Gemma 4 on the Workers AI LLM Playground.

Fine-Tune

  • Fine-Tune Gemma β€” Official framework guide covering Keras, JAX, Hugging Face, Unsloth, Axolotl, and Google Cloud.
  • Gemma Cookbook: Training β€” Official fine-tuning notebooks and training recipes.
  • Tunix β€” JAX-native library for post-training generative models.
  • Unsloth Gemma 4 fine-tuning guide β€” Train Gemma 4 E2B, E4B, 12B, 26B A4B and 31B with Unsloth.
  • Gemma Multimodal Tuner β€” Fine-tune Gemma 3n and Gemma 4 with text, images, and audio on Apple Silicon.
  • MLX Tune β€” MLX-native SFT, preference tuning, and multimodal fine-tuning with Gemma 4 support.
  • Fine-tune Gemma 4 MoE with Halo β€” Train Gemma 4 26B-A4B with SFT, expert parallelism, LoRA, or environmental GRPO.

Tutorials

Demos and Applications

  • Gemma 4 Vision Token Budget β€” Explore the effect of image resolution and visual-token budgets.
  • Concurrent Gemma β€” Run and compare multiple concurrent local Gemma instances.
  • See what 3 builders are making with Gemma 4 β€” Various applications developed by the community.
  • AIventure β€” A 2D grid-based adventure game built with Phaser 3 and Angular with Gemma driving it.
  • Gemma Chat β€” Local AI chat + coding agent for Apple Silicon, powered by Gemma 4 via MLX / Supports Ollama.
  • Build with Gemma 4 and Haystack β€” Runnable notebook covering RAG, visual question answering, a multimodal weather agent, and GitHub tool discovery.
  • Gemma 4 Browser Extension β€” Local browser agent powered by Gemma 4, WebGPU, and Transformers.js.
  • WebGemma β€” Browser playground and interactive model timeline powered by WebGPU and Transformers.js.
  • Controlling an iOS simulator β€” Gemma 4 using Argent to control an iOS simulator showcasing its capabilities in agentic workflows.
  • Automated Video Segmentation & Tracking β€” A demo that uses Gemma 4 + Falcon Perception for video tracking.
  • Parking Lot Car Detection & Segmentation β€” Gemma 4 analyzes the scene, decides the questions, generates prompts, and calls SAM 3.1 as a tool. SAM 3.1 segments and returns results.
  • Gemma 4 and MTP as a Marathon Engine β€” Benchmarks speculative decoding across increasing context lengths.
  • Cactus Hybrid β€” Post-trained Gemma 4 models to recognize when they are wrong, run on any framework.
  • Damage Scout β€” Damage Scout samples frames from a rental car walkaround, sends them to Gemma 4, gets back structured findings and box coordinates, then renders an annotated damage report in under 6 seconds.
  • MedGemma Impact Challenge β€” The winners of the MedGemma hackathon to build human-centered AI applications with MedGemma.
  • Gemma-Translator β€” A fully offline device powered by Gemma 4 E2B built with Google Antigravity.
  • Real-Time Voice AI with Gemma 4 β€” Open-source cascaded voice stack using Gemma 4 for low-latency reasoning.
  • Clips Kitty β€” Windows desktop app that uses Gemma through Ollama to pick and title vertical clips from long streams, with all processing on-device.
  • WisprGemma - Web app and Chrome extension for multilingual voice dictation using Gemma 4 E2B locally with WebGPU.

Gemma 4 Good Challenge

Amazing projects that harness the power of Gemma 4 to drive positive change and global impact.

  • Trido β€” A Voice-Driven AI Whiteboard Built for the Teacher Nobody Builds For.
  • CodeBuddy β€” AI Python Tutor for Indonesian Students.
  • Port-a-Prof β€” Deeper learning, wherever you are.
  • TriageMate β€” Offline-first Clinical AI for Ghana's Community Health Officers.
  • ORCA-G4 β€” On-device oral cancer intelligence for 900,000 ASHA workers in rural India.
  • DEMENTOR β€” Edge AI Triage for Dementia Care.
  • PreVillage β€” A source-backed navigator for Nepal’s government services, built to find the office route, not just the form.
  • BrailleOut β€” An assistive device that reads the text and images from real-world and converts it to Braille using Gemma 4 and Ollama.
  • Gem-Care β€” Gemma-4-Enriched with Multimodal Clinical-context Adaptation for Recognition Enhancement of Non-Normative Speech.
  • Trajectix β€” An Agentic Flight Recorder for AI Infrastructure Safety.
  • TrueVoice β€” AI Voice Deepfake Detector.
  • AI Conceptualizer β€” 3D visualizations for mechanistic interpretability and "concept spectroscopy".
  • AcuΓ­feroΒ·VigΓ­a β€” Hybrid edge-and-citizen flood early warning for Argentina's Litoral, where every minute of warning is a life.
  • ResQ β€” Offline Multilingual Disaster Response Coach on Gemma 4 E2B.
  • OptiLearn β€” A locally-run, adaptive learning system for refugee and underserved classrooms, powered by Gemma 4 models.

Gemma in Space

  • Starcloud-1 β€” Starcloud deployed and ran Gemma in orbit aboard an H100 GPU.
  • NASA β€” NASA runs Gemma in orbit to analyze satellite imagery and compress visual data into text for rapid, low-bandwidth disaster response.

Research and Evaluation

Footnotes

This is not an officially supported Google product. This project is not eligible for the Google Open Source Software Vulnerability Rewards Program.

About

😎 Awesome list about Gemma, Google DeepMind's family of lightweights, state-of-the-art open models.

Resources

Contributing

Stars

544 stars

Watchers

8 watching

Forks

Used by

Contributors