Google · Text
Gemma 4 31B
Google's flagship 31B dense model with 256K context and top-tier text and image understanding.
From Google, Gemma 4 31B is an open-weight text model. It is multimodal: alongside a text prompt it accepts images, then replies with generated text. It runs entirely on your own hardware through Ollama: no API key and no per-use fees.
Modality
Text
Available on
Ollama
Model ID
ollama:gemma4:31b
Supported OS
macOSWindowsLinux
Minimum machine configuration
32 GB+ RAM · 256K ctx · text+image · GPU strongly recommended
Download size · ~20 GB
Specs
- Pricing
- Free, runs on your hardware
- Inputs
- Text, Images
About the creator
Google and Google DeepMind build the Gemini family of multimodal models, the Imagen and Nano Banana image models, the Lyria music models, and the Veo video models.
deepmind.google ↗More from Google
Gemini 3.5 Flash
TextCloud
Google's latest fast Gemini 3 model: frontier-level reasoning at Flash-level latency and cost, tuned for agentic workflows and iterative coding.
Gemini 3.1 Pro
TextCloud
Google's most capable multimodal model with deep reasoning and 1M token context.
Gemini 3 Flash
TextCloud
Fast, cost-efficient Gemini model for high-throughput text and multimodal tasks.
Gemini 2.5 Flash
TextCloud
Balanced Gemini model with strong reasoning and a 1M token context window.
Nano Banana 2
ImageCloud
Ultra-fast image generation model optimised for speed and creative output.
Nano Banana 2 Lite
ImageCloud
Google's fastest, most cost-efficient Nano Banana model — 1K images in ~4 seconds for high-volume workflows.
From the blog
Google Gemma, explained: which model is for what
June 16, 2026Type “gemma” into your runtime and you get a wall of models: a phone-sized one, a workstation one, and odd cousins. Here’s which one you actually want.
Run GPT-4 class models on your laptop without sending a single byte to the cloud
May 5, 2026Open weights now match GPT-4 quality. Here's how CSuite runs them on your machine: no proxy, no logging, no tokens billed.