Text · Released Jan 2026
Google Gemma 3 Med 4B
Google's medical multimodal model fine-tuned for healthcare and biomedical tasks.
Google's Gemma 3 Med 4B is an open-weight text model. It is multimodal: alongside a text prompt it accepts images, then replies with generated text. It runs entirely on your own hardware through Ollama: no API key and no per-use fees.
Supported OS
macOSWindowsLinux
Minimum machine configuration
8 GB RAM · 128K ctx · text+image · CPU/GPU
Download size · ~3.3 GB
Specs
- Released
- 8 months ago (Jan 2026)
- Pricing
- Free, runs on your hardware
- Inputs
- Text, Images
About the creator
Google and Google DeepMind build the Gemini family of multimodal models, the Imagen and Nano Banana image models, the Lyria music models, and the Veo video models.
deepmind.google ↗More from Google
Gemini 3.8 Flash
TextCloud
Google's most capable Flash model, with clear gains over 3.7 Flash in software engineering, agentic tasks, and multi-step reasoning, plus vision and a 1M-token context window.
Gemini 3.7 Flash
TextCloud
Google's newest Flash workhorse, a more capable successor to 3.6 Flash with stronger software engineering, better document comprehension and more disciplined tool use, at half the price per token.
Gemini 3.6 Flash
TextCloud
Google's most efficient multimodal Flash model: an updated reasoning stack with stronger coding and computer-use quality, at the same speed and scale as the rest of the Flash line.
Gemini 3.1 Pro
TextCloud
Google's most capable multimodal model with deep reasoning and 1M token context.
Nano Banana 2
ImageCloud
Ultra-fast image generation model optimised for speed and creative output.
Nano Banana 2 Lite
ImageCloud
Google's fastest, most cost-efficient Nano Banana model — 1K images in ~4 seconds for high-volume workflows.