Meta Llama 3.2 1B
Meta's 1B instruction-tuned model — the smallest Llama, runnable in-process via HuggingFace or through Ollama.
Meta's Llama 3.2 1B is an open-weight text model. It takes a text prompt and replies with generated text. It runs entirely on your own hardware through Ollama and Hugging Face: no API key and no per-use fees.
- Released
- Sep 2024
- Pricing
- Free, runs on your hardware
Meta
Meta's FAIR and GenAI labs build the open-weight Llama family of language models — among the most widely used open models for local inference.
ai.meta.com ↗Meta's multimodal reasoning model for long-running agentic, multi-agent and coding workflows, with image input and a 1M-token context window.
Meta Superintelligence Labs' agentic image model, which plans a layout and calls search and coding tools before it renders, for prompt-faithful generation, multi-reference composition and precise local edits.
Meta's speech-to-text model for push-to-talk and speaker-aware transcription, which returns plain text.
Meta's 8B instruction-tuned model with 128K context and strong general reasoning.
Meta's compact 3B instruction-tuned model — fast on-device text generation.
Meta Superintelligence Labs' 30B multimodal model for always-on local agents — tuned for tool use, long tasks, and failure recovery.