Meta Muse Spark 1.3
Meta's multimodal reasoning model for long-running agentic, multi-agent and coding workflows, with image input and a 1M-token context window.
Meta's Muse Spark 1.3 is a cloud text model. It is multimodal: alongside a text prompt it accepts images, then replies with generated text. Its context window handles up to 1.05M input tokens and up to 131K output tokens, and it supports adjustable reasoning effort for harder problems. Generation can be tuned with system instructions, temperature, and top-p sampling. It runs through OpenRouter using your own API key, from $1.25 per million input tokens.
- Released
- Sep 2026
- Pricing
- $1.25 / 1M in · $4.25 / 1M out
- Inputs
- Text, Images
- Context window
- 1.05M in · 131K out
- Reasoning
- Adjustable effort
- Controls
- System prompt, Temperature, Top-p
Meta
Meta's FAIR and GenAI labs build the open-weight Llama family of language models — among the most widely used open models for local inference.
ai.meta.com ↗Meta Superintelligence Labs' agentic image model, which plans a layout and calls search and coding tools before it renders, for prompt-faithful generation, multi-reference composition and precise local edits.
Meta's speech-to-text model for push-to-talk and speaker-aware transcription, which returns plain text.
Meta's 1B instruction-tuned model — the smallest Llama, runnable in-process via HuggingFace or through Ollama.
Meta's 8B instruction-tuned model with 128K context and strong general reasoning.
Meta's compact 3B instruction-tuned model — fast on-device text generation.
Meta Superintelligence Labs' 30B multimodal model for always-on local agents — tuned for tool use, long tasks, and failure recovery.