Model catalog
One endpoint, every model. Compare context windows, capabilities and real per-token pricing side by side, then swap the model string in your request — nothing else changes.
- 351
- Models
- 35
- Providers
- 11
- Free to try
- $0.0001
- Cheapest input / 1M
39 of 351 models
- Language
Qwen3.8 27B
alibaba/qwen3.8-27b
Built on the architectural foundation of Qwen3.5, Qwen3.8 delivers substantial gains across coding, professional work, research, and long-horizon agentic tasks. Qwen3.8-27B brings these advances to a compact, deployment-friendly dense model: a native vision-language model that understands images and videos, with flexible thinking control, designed to carry complex, multi-step tasks through to completion with greater reliability.
- Context
- 1M
- In / 1M
- $0.550
- Out / 1M
- $3.30
- Reasoning
- Tool use
- Vision
- +3
- Language
Qwen 3.7 Flash
alibaba/qwen3.7-flash
The Qwen3.7 native vision-language Flash model series delivers a comprehensive upgrade over 3.6-Flash in multimodal understanding and agent execution. This model particularly excels in enhanced multimodal foundations with stronger universal object recognition, further improved real-world perception and spatial intelligence, significantly upgraded multimodal agent capabilities for Search Agent and CI Agent scenarios with more stable end-to-end task execution, as well as optimized multimodal coding for a smoother vibe coding experience.
- Context
- 991K
- In / 1M
- $0.030
- Out / 1M
- $0.130
- Reasoning
- Tool use
- Vision
- +2
- Language
Qwen 3.7 Plus
alibaba/qwen3.7-plus
Among the Qwen3.7 series, the cost-effective Plus model builds on its robust text capabilities while delivering a comprehensive upgrade to its vision‑language abilities, all while preserving its full‑stack agent‑level intelligence for coding, tool use, and productivity workflows.
- Context
- 1M
- In / 1M
- $0.400
- Out / 1M
- $1.60
- Reasoning
- Tool use
- Vision
- +2
- Language
Qwen3.8 2.4T A95B
alibaba/qwen3.8-2.4t-a95b
Open-weights release of the Qwen3.8 flagship (2.4T MoE, ~95B active); thinking always on with reasoning_effort low/medium/xhigh. The hosted Qwen 3.8 Max (vision, non-thinking, 1M default context) is Alibaba-only.
- Context
- 262K
- In / 1M
- $2.00
- Out / 1M
- $6.00
- Reasoning
- Tool use
- Vision
- +1
- Language
Qwen 3.8 Max
alibaba/qwen3.8-max
Qwen 3.8 Max is a 2.4-trillion-parameter MoE model delivering a comprehensive leap in coding and professional work. Autonomously codes and delivers complete projects spanning 10+ days. Handles hundreds of specialized tasks across legal, financial, design, and other professional domains, producing production-grade results end-to-end in a single conversation. Native visual understanding runs through the full cycle of planning, execution, and verification, enabling deep semantic analysis of ultra-long documents and extended video content. In long-horizon tasks, plans autonomously, iterates through closed feedback loops, and continuously evolves.
- Context
- 1M
- In / 1M
- $2.00
- Out / 1M
- $6.00
- Reasoning
- Tool use
- Vision
- +1
- Language
Qwen 3.6 27B
alibaba/qwen3.6-27b
The Qwen3.6 35B-A3B native vision-language model is built on a hybrid architecture that integrates linear attention mechanisms with a sparse mixture-of-experts framework, achieving higher inference efficiency. Compared with the 3.5-35B-A3B, this model demonstrates significantly improved agentic coding capabilities, mathematical and code reasoning abilities, spatial intelligence, as well as object localization and object detection performance.
- Context
- 256K
- In / 1M
- $0.600
- Out / 1M
- $3.60
- Reasoning
- Tool use
- Vision
- +1
- Language
Qwen 3.6 Plus
alibaba/qwen3.6-plus
The Qwen3.6 native vision-language Plus series models demonstrate exceptional performance on par with the current state-of-the-art models, with a significant improvement in overall results compared to the 3.5 series. The models have been markedly enhanced in code-related capabilities such as agentic coding, front-end programming, and Vibe coding, as well as in multi-modal general object recognition, OCR, and object localization.
- Context
- 1M
- In / 1M
- $0.500
- Out / 1M
- $3.00
- Reasoning
- Tool use
- Vision
- +1
- Language
Qwen 3.5 Flash
alibaba/qwen3.5-flash
The Qwen3.5 native vision-language Flash models are built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. Compared to the 3 series, these models deliver a leap forward in performance for both pure text and multimodal tasks, offering fast response times while balancing inference speed and overall performance.
- Context
- 1M
- In / 1M
- $0.100
- Out / 1M
- $0.400
- Reasoning
- Tool use
- Vision
- +1
- Language
Qwen 3.5 Plus
alibaba/qwen3.5-plus
The Qwen3.5 native vision-language series Plus models are built on a hybrid architecture that integrates linear attention mechanisms with sparse mixture-of-experts models, achieving higher inference efficiency. In a variety of task evaluations, the 3.5 series consistently demonstrates performance on par with state-of-the-art leading models. Compared to the 3 series, these models show a leap forward in both pure-text and multimodal capabilities.
- Context
- 1M
- In / 1M
- $0.400
- Out / 1M
- $2.40
- Reasoning
- Tool use
- Vision
- +1
- Language
Qwen3 VL 235B A22B Thinking
alibaba/qwen3-235b-a22b-thinking
Qwen3 series VL models feature significantly enhanced multimodal reasoning capabilities, with a particular focus on optimizing the model for STEM and mathematical reasoning. Visual perception and recognition abilities have been comprehensively improved, and OCR capabilities have undergone a major upgrade.
- Context
- 131K
- In / 1M
- $0.400
- Out / 1M
- $4.00
- Reasoning
- Tool use
- Vision
- +1
- Language
Qwen3 VL 235B A22B Thinking
alibaba/qwen3-vl-thinking
Qwen3 series VL models feature significantly enhanced multimodal reasoning capabilities, with a particular focus on optimizing the model for STEM and mathematical reasoning. Visual perception and recognition abilities have been comprehensively improved, and OCR capabilities have undergone a major upgrade.
- Context
- 131K
- In / 1M
- $0.400
- Out / 1M
- $4.00
- Reasoning
- Tool use
- Vision
- +1
- Language
Qwen 3.7 Max
alibaba/qwen3.7-max
Qwen3.7 is a next‑generation flagship model designed for the agent‑centric era, with its core strengths lying in the breadth and depth of its agent‑level capabilities: it excels at programming, office and productivity tasks, and long‑term autonomous execution.
- Context
- 991K
- In / 1M
- $2.50
- Out / 1M
- $7.50
- Reasoning
- Tool use
- Implicit caching
- Language
Qwen3 VL 235B A22B Instruct
alibaba/qwen3-vl-235b-a22b-instruct
The Qwen3 series VL models has been comprehensively upgraded in areas such as visual coding and spatial perception. Its visual perception and recognition capabilities have significantly improved, supporting the understanding of ultra-long videos, and its OCR functionality has undergone a major enhancement.
- Context
- 131K
- In / 1M
- $0.400
- Out / 1M
- $1.60
- Tool use
- Vision
- Implicit caching
- Language
Qwen3 VL 235B A22B Instruct
alibaba/qwen3-vl-instruct
The Qwen3 series VL models has been comprehensively upgraded in areas such as visual coding and spatial perception. Its visual perception and recognition capabilities have significantly improved, supporting the understanding of ultra-long videos, and its OCR functionality has undergone a major enhancement.
- Context
- 131K
- In / 1M
- $0.400
- Out / 1M
- $1.60
- Tool use
- Vision
- Implicit caching
- Language
Qwen 3.6 Max Preview
alibaba/qwen-3.6-max-preview
Compared with the previously released Qwen3-Max and Qwen3.6-Plus, this model features enhanced vibe coding abilities, more efficient coding agent execution, and significantly improved front-end development skills. Additionally, its long-tail knowledge retention has been further upgraded.
- Context
- 240K
- In / 1M
- $1.30
- Out / 1M
- $7.80
- Reasoning
- Tool use
- Language
Qwen 3 Max Thinking
alibaba/qwen3-max-thinking
Compared with the snapshot as of September 23, 2025, the Qwen-3 series Max model in this release achieves an effective integration of thinking and non-thinking modes, resulting in a comprehensive and substantial improvement in the model’s overall performance. In thinking mode, the model simultaneously supports web search, web information extraction, and a code interpreter tool, enabling it to tackle more complex and challenging problems with greater accuracy by leveraging external tools while engaging in slow, deliberative reasoning. This version is based on a snapshot taken on January 23, 2026.
- Context
- 256K
- In / 1M
- $1.20
- Out / 1M
- $6.00
- Reasoning
- Tool use
- Language
Qwen3 Max
alibaba/qwen3-max
The Qwen 3 series Max model has undergone specialized upgrades in agent programming and tool invocation compared to the preview version. The officially released model this time has achieved state-of-the-art (SOTA) performance in its field and is better suited to meet the demands of agents operating in more complex scenarios.
- Context
- 262K
- In / 1M
- $1.20
- Out / 1M
- $6.00
- Tool use
- Implicit caching
- Language
Qwen3 Next 80B A3B Thinking
alibaba/qwen3-next-80b-a3b-thinking
A new generation of Qwen3-based open-source thinking mode models. This version offers improved instruction following and streamlined summary responses over the previous iteration (Qwen3-235B-A22B-Thinking-2507).
- Context
- 131K
- In / 1M
- $0.150
- Out / 1M
- $1.20
- Reasoning
- Tool use
- Language
Qwen3 Max Preview
alibaba/qwen3-max-preview
Qwen3-Max-Preview shows substantial gains over the 2.5 series in overall capability, with significant enhancements in Chinese-English text understanding, complex instruction following, handling of subjective open-ended tasks, multilingual ability, and tool invocation; model knowledge hallucinations are reduced.
- Context
- 262K
- In / 1M
- $1.20
- Out / 1M
- $6.00
- Tool use
- Implicit caching
- Language
Qwen3 Coder Plus
alibaba/qwen3-coder-plus
Powered by Qwen3 this is a powerful Coding Agent that excels in tool calling and environment interaction to achieve autonomous programming. It combines outstanding coding proficiency with versatile general-purpose abilities.
- Context
- 1M
- In / 1M
- $1.00
- Out / 1M
- $5.00
- Tool use
- Implicit caching
- Language
Qwen3 Coder 480B A35B Instruct
alibaba/qwen3-coder
Qwen3-Coder-480B-A35B-Instruct is a cutting-edge open coding model from Qwen, matching Claude Sonnet’s performance in agentic programming, browser automation, and core development tasks.
- Context
- 262K
- In / 1M
- $1.50
- Out / 1M
- $7.50
- Tool use
- Implicit caching
- Language
Qwen 3 32B
alibaba/qwen-3-32b
Qwen3-32B is a world-class model with comparable quality to DeepSeek R1 while outperforming GPT-4.1 and Claude Sonnet 3.7. It excels in code-gen, tool-calling, and advanced reasoning, making it an exceptional model for a wide range of production use cases.
- Context
- 128K
- In / 1M
- $0.160
- Out / 1M
- $0.640
- Reasoning
- Tool use
- Language
Qwen3 235B A22B
alibaba/qwen-3-235b
- Context
- 262K
- In / 1M
- $0.220
- Out / 1M
- $0.880
- Reasoning
- Tool use
- Language
Qwen3-14B
alibaba/qwen-3-14b
Qwen3 is the latest generation of large language models in Qwen series, offering a comprehensive suite of dense and mixture-of-experts (MoE) models. Built upon extensive training, Qwen3 delivers groundbreaking advancements in reasoning, instruction-following, agent capabilities, and multilingual support
- Context
- 41K
- In / 1M
- $0.120
- Out / 1M
- $0.240
- Reasoning
- Tool use
Showing 24 of 39
