Google Gemma 3 1B
Google's compact 1B Gemma 3 model — runs in-process via HuggingFace or through Ollama.
From Google, Gemma 3 1B is an open-weight text model. It takes a text prompt and replies with generated text. It runs entirely on your own hardware through Ollama and Hugging Face: no API key and no per-use fees.
- Released
- Mar 2025
- Pricing
- Free, runs on your hardware
Google and Google DeepMind build the Gemini family of multimodal models, the Imagen and Nano Banana image models, the Lyria music models, and the Veo video models.
deepmind.google ↗Google's most capable Flash model, with clear gains over 3.7 Flash in software engineering, agentic tasks, and multi-step reasoning, plus vision and a 1M-token context window.
Google's newest Flash workhorse, a more capable successor to 3.6 Flash with stronger software engineering, better document comprehension and more disciplined tool use, at half the price per token.
Google's most efficient multimodal Flash model: an updated reasoning stack with stronger coding and computer-use quality, at the same speed and scale as the rest of the Flash line.
Google's most capable multimodal model with deep reasoning and 1M token context.
Google's updated Nano Banana 2 image model, with sharper graphic composition, steadier subjects across edits and closer prompt following, from text or up to 14 reference images at 1K to 4K.
Ultra-fast image generation model optimised for speed and creative output.