Gemma
Open models from Google DeepMind.
Gemma is Google DeepMind's family of open AI models. Its current family includes multimodal Gemma 4 models supporting text, image, audio, and video inputs, consistent with the supplied Gemma 3 and Gemma 4 references.
Pricing
Model Intelligence
Recent stories
Google’s new Gemma 4 12B ships as an encoder-free open model for text, image, audio, and video tasks with a 256K context window. Early GGUF ports and local benchmarks make it a plausible on-device multimodal option for creator tooling and experimentation.
Google DeepMind shipped four Gemma 4 models with multimodal input, including 31B Dense, 26B MoE, and two edge variants available through AI Studio, Hugging Face, Kaggle, and Ollama. Early community tests say local performance and usable context windows still vary by runtime, quantization, and GPU memory.