Alibaba Qwen Image 2.1 Pro
Alibaba Qwen's hosted Pro tier of the Qwen Image 2.1 family, a unified model that generates from a prompt and edits or composes from up to 10 reference images. It can rewrite the prompt with an LLM before generating, with an optional thinking mode, and renders from 0.26 to 4.19 megapixels.
Alibaba's Qwen Image 2.1 Pro is a cloud image-generation model. It generates images from a text prompt, and can take reference images to guide style or composition (image-to-image). It supports 8 aspect ratios. Images can be saved as PNG, JPEG, and WebP. It runs through Runware using your own API key, from about $0.075 per image.
- Released
- Sep 2026
- Pricing
- $0.075 per image
- Aspect ratios
- 1:1, 16:9, 9:16 +5 more
- Image-to-image
- Up to 10 reference images
- Output formats
- PNG, JPEG, WebP
Examples
Generated with Qwen Image 2.1 Pro via Runware. The same three prompts run against every model in the catalog, first result kept, so the only thing that changes between two models’ tiles is the model.
Materials, reflections, macro detail
Skin texture, lighting, shallow depth of field
Text rendering, layout, flat colour
Alibaba
Alibaba's Tongyi research group publishes the Wan video models and the Qwen family of language models.
www.alibabacloud.com ↗Alibaba's flagship Qwen3.8 reasoning model for coding, agentic workflows and document analysis, with image input and a 1M-token context window.
A higher-throughput variant of Alibaba's Qwen3.8 Max, served as a separate tier at a higher price.
Alibaba's fast, low-cost multimodal Qwen3.8 reasoning model for coding assistance, agentic workflows and visual understanding, with a 1M-token context window.
Alibaba's omni-modal Qwen3.8 reasoning model built around agentic work, which reads text, images, audio and video.
High-quality diffusion model for detailed, prompt-faithful image generation.
Alibaba's professional-grade Wan 2.7 image model for higher-fidelity, prompt-faithful generation and editing.