Choose a model for AI agents, chatbots, document processing, and more.
Using OpenClaw or Claude Code?
qwen3.7-plus — balanced performance and cost, full tool support, 1M context for large codebases. For strongest reasoning, choose the flagship qwen3.8-max.
Migrate from closed-source models
Map your current GPT, Claude, or Gemini model to an equivalent QwenCloud model.
| Tier | Closed-source examples | QwenCloud recommendation |
|---|---|---|
| Highest capability | GPT-5.5, Claude Opus 4.7, Gemini 3.1 Pro | qwen3.8-max |
| Balanced | GPT-5.4, Claude Sonnet 4.6, Gemini 3 Pro | qwen3.7-plus, deepseek-v4-pro-0813 |
| Lightweight & low-cost | GPT-5.4-mini, Claude Haiku 4.5, Gemini 3.1 Flash | qwen3.7-flash, deepseek-v4-flash-0731 |
For other applications
Chatbots, content generation, summarization, document processing — start with qwen3.7-plus, which balances performance, cost, and built-in tools with a 1M-token context window. To cut costs, switch to qwen3.7-flash, which offers similar capabilities at a lower price. For the strongest reasoning, use the flagship qwen3.8-max (1M context, higher cost).
Office productivity (non-coding)
For non-coding office tasks such as document drafting, email composition, meeting note summarization, and data analytics, start with qwen3.7-plus — it balances performance and cost with a 1M-token context window, Function Calling support, and built-in tools. To reduce costs, try qwen3.7-flash, which delivers near-flagship performance at a lower price with the same context length. For the strongest reasoning capability, choose the flagship qwen3.8-max (higher cost). For long-document processing such as reviewing multiple contracts, use qwen-long (10M-token context window).
TONGYI Lingma and Qoder are AI coding tools designed for software development. They are not intended for general office productivity tasks.
Context window
1M tokens is roughly 750,000 words or 10 novels.
- Long documents or large codebases →
qwen3.8-max/qwen3.7-max/qwen3.7-plus/qwen3.7-flash(1M) - Standard tasks → 128k–256k is plenty
Thinking mode
Step-by-step reasoning for multi-step math, debugging, architecture planning, or legal cross-referencing.
Toggle with enable_thinking. All Qwen3+ models support it — most are hybrid, so you can switch per request.
Function calling + built-in tools
Let the model take actions: check weather, query a database, book a meeting.
- Function calling (you define tools, model calls them): all general-purpose models
- Built-in tools (web search, code execution — no setup):
qwen3.8-max,qwen3.7-max,qwen3.7-plus,qwen3.7-flash,qwen3.6-flash,qwen3.5-plus,qwen3.5-flash,qwen3-maxseries only
Recommended models
| Model | Context | Thinking | Function calling | Built-in tools | Structured output |
|---|---|---|---|---|---|
qwen3.8-max | 1M | ✓ | ✓ | ✓ | ✓ |
qwen3.7-plus | 1M | ✓ | ✓ | ✓ | ✓ |
qwen3.7-flash | 1M | ✓ | ✓ | ✓ | ✓ |
qwen3.6-flash | 1M | ✓ | ✓ | ✓ | ✓ |
deepseek-v4-pro-0813 | 1M | ✓ | ✓ | ✓ | ✓ |
deepseek-v4-flash | 1M | ✓ | ✓ | — | — |
deepseek-v4-flash-0731 | 1M | ✓ | ✓ | — | — |
All models
Qwen3.8
Qwen3.8
| Model ID | Context | Max Output | Thinking Budget | Function calling | Built-in tools | Structured output |
|---|---|---|---|---|---|---|
qwen3.8-max | 1M | 128k | 256k | ✓ | ✓ | ✓ |
qwen3.8-2.4t-a95b | 1M | 128k | 128k | ✓ | ✓ | ✓ |
Qwen3.7
Qwen3.7
| Model ID | Context | Max Output | Thinking Budget | Function calling | Built-in tools | Structured output |
|---|---|---|---|---|---|---|
qwen3.7-max | 1M | 64k | 256k | ✓ | ✓ | — |
qwen3.7-max-2026-06-08 | 1M | 64k | 256k | ✓ | ✓ | — |
qwen3.7-max-2026-05-20 | 1M | 64k | 256k | ✓ | ✓ | — |
qwen3.7-max-preview | 1M | 64k | 256k | ✓ | ✓ | — |
qwen3.7-max-2026-05-17 | 1M | 64k | 256k | ✓ | ✓ | — |
qwen3.7-plus | 1M | 64k | 256k | ✓ | ✓ | ✓ |
qwen3.7-plus-2026-05-26 | 1M | 64k | 256k | ✓ | ✓ | ✓ |
qwen3.7-flash | 1M | 64k | 256k | ✓ | ✓ | ✓ |
qwen3.7-flash-2026-07-15 | 1M | 64k | 256k | ✓ | ✓ | ✓ |
Qwen3.6
Qwen3.6
| Model ID | Context | Max Output | Thinking Budget | Function calling | Built-in tools | Structured output |
|---|---|---|---|---|---|---|
qwen3.6-max-preview | 256k | 64k | 128k | ✓ | — | ✓ |
qwen3.6-flash | 1M | 64k | 80k | ✓ | ✓ | ✓ |
qwen3.6-flash-2026-04-16 | 1M | 64k | 80k | ✓ | ✓ | ✓ |
qwen3.6-35b-a3b | 256k | 64k | 80k | ✓ | ✓ | ✓ |
qwen3.6-27b | 256k | 64k | 80k | ✓ | — | ✓ |
qwen3.6-plus | 1M | 64k | 80k | ✓ | ✓ | ✓ |
qwen3.6-plus-2026-04-02 | 1M | 64k | 80k | ✓ | ✓ | ✓ |
Qwen3.5
Qwen3.5
| Model ID | Context | Max Output | Thinking Budget | Function calling | Built-in tools | Structured output |
|---|---|---|---|---|---|---|
qwen3.5-plus | 1M | 64k | 80k | ✓ | ✓ | ✓ |
qwen3.5-plus-2026-04-20 | 1M | 64k | 80k | ✓ | ✓ | ✓ |
qwen3.5-plus-2026-02-15 | 1M | 64k | 80k | ✓ | ✓ | ✓ |
qwen3.5-flash | 1M | 64k | 80k | ✓ | ✓ | ✓ |
qwen3.5-flash-2026-02-23 | 1M | 64k | 80k | ✓ | ✓ | ✓ |
qwen3.5-397b-a17b | 256k | 64k | 80k | ✓ | ✓ | ✓ |
qwen3.5-122b-a10b | 256k | 64k | 80k | ✓ | ✓ | ✓ |
qwen3.5-27b | 256k | 64k | 80k | ✓ | ✓ | ✓ |
qwen3.5-35b-a3b | 256k | 64k | 80k | ✓ | ✓ | ✓ |
Specialized
Specialized
Translation
| Model ID | Context | Max Output | Function calling | Built-in tools | Structured output |
|---|---|---|---|---|---|
qwen-mt-plus | 16k | 8k | — | — | — |
qwen-mt-turbo | 16k | 8k | — | — | — |
qwen-mt-flash | 16k | 8k | — | — | — |
qwen-mt-lite | 16k | 8k | — | — | — |
Character roleplay
| Model ID | Context | Max Output | Function calling | Built-in tools | Structured output |
|---|---|---|---|---|---|
qwen-plus-character | 32k | 4k | — | — | — |
qwen-plus-character-ja | 8k | 4k | — | — | — |
qwen-flash-character | 8k | 4k | — | — | — |
Third-party
Third-party
Non-Qwen models available through the same API.
* DeepSeek V4 models share a 384k total budget across output and thinking.
| Model ID | Context | Max Output | Thinking Budget | Function calling | Built-in tools | Structured output |
|---|---|---|---|---|---|---|
deepseek-v4-pro-0813 | 1M | 384k * | * | ✓ | ✓ | ✓ |
deepseek-v4-pro | 1M | 384k * | * | ✓ | — | — |
deepseek-v4-flash | 1M | 384k * | * | ✓ | — | — |
deepseek-v4-flash-0731 | 1M | 384k * | * | ✓ | — | — |
deepseek-v3.2 | 128k | 64k | 32k | ✓ | — | — |
Legacy
Legacy
The following models are legacy and no longer recommended. Visit the Models page to view detailed model parameters, such as context window and billing.
Qwen3
| Model ID | Context | Max Output | Thinking Budget | Function calling | Built-in tools | Structured output |
|---|---|---|---|---|---|---|
qwen3-max | 256k | 64k | 80k | ✓ | ✓ | ✓ |
qwen3-max-2026-01-23 | 256k | 64k | 80k | ✓ | ✓ | ✓ |
qwen3-max-preview | 256k | 64k | 80k | ✓ | ✓ | ✓ |
qwen3-max-2025-09-23 | 256k | 64k | — | ✓ | ✓ | ✓ |
qwen3-235b-a22b | 128k | 16k | 38k | ✓ | — | ✓ |
qwen3-235b-a22b-thinking-2507 | 128k | 32k | 80k | ✓ | — | — |
qwen3-235b-a22b-instruct-2507 | 128k | 32k | — | ✓ | — | ✓ |
qwen3-next-80b-a3b-thinking | 128k | 32k | 80k | ✓ | — | — |
qwen3-next-80b-a3b-instruct | 128k | 32k | — | ✓ | — | ✓ |
qwen3-32b | 128k | 16k | 38k | ✓ | — | ✓ |
qwen3-30b-a3b | 128k | 16k | 38k | ✓ | — | ✓ |
qwen3-30b-a3b-thinking-2507 | 128k | 32k | 80k | ✓ | — | — |
qwen3-30b-a3b-instruct-2507 | 128k | 32k | — | ✓ | — | ✓ |
qwen3-14b | 128k | 8k | 38k | ✓ | — | ✓ |
qwen3-8b | 128k | 8k | 38k | ✓ | — | ✓ |
qwen3-4b | 128k | 8k | 38k | ✓ | — | ✓ |
qwen3-1.7b | 32k | 8k | 30k | ✓ | — | ✓ |
qwen3-0.6b | 32k | 8k | 30k | ✓ | — | ✓ |
Qwen3-Coder
| Model ID | Context | Max Output | Function calling | Built-in tools | Structured output |
|---|---|---|---|---|---|
qwen3-coder-plus | 1M | 64k | ✓ | — | ✓ |
qwen3-coder-plus-2025-09-23 | 1M | 64k | ✓ | — | ✓ |
qwen3-coder-plus-2025-07-22 | 1M | 64k | ✓ | — | ✓ |
qwen3-coder-flash | 1M | 64k | ✓ | — | ✓ |
qwen3-coder-flash-2025-07-28 | 1M | 64k | ✓ | — | ✓ |
qwen3-coder-next | 256k | 64k | ✓ | — | ✓ |
qwen3-coder-480b-a35b-instruct | 256k | 64k | ✓ | — | ✓ |
qwen3-coder-30b-a3b-instruct | 256k | 64k | ✓ | — | ✓ |
Qwen2.5 (open source)
| Model ID | Context | Max Output | Function calling | Built-in tools | Structured output |
|---|---|---|---|---|---|
qwen2.5-omni-7b | 32k | 8k | ✓ | — | ✓ |
qwen2.5-vl-72b-instruct | 128k | 8k | ✓ | — | ✓ |
qwen2.5-vl-32b-instruct | 128k | 8k | ✓ | — | ✓ |
qwen2.5-vl-7b-instruct | 128k | 8k | ✓ | — | ✓ |
qwen2.5-vl-3b-instruct | 128k | 8k | ✓ | — | ✓ |
qwen2.5-72b-instruct | 32k | 8k | ✓ | — | ✓ |
qwen2.5-32b-instruct | 32k | 8k | ✓ | — | ✓ |
qwen2.5-14b-instruct | 32k | 8k | ✓ | — | ✓ |
qwen2.5-14b-instruct-1m | 1M | 8k | ✓ | — | ✓ |
qwen2.5-7b-instruct | 32k | 8k | ✓ | — | ✓ |
qwen2.5-7b-instruct-1m | 1M | 8k | ✓ | — | ✓ |
Legacy (qwen-plus/max/flash/turbo)
| Model ID | Context | Max Output | Thinking Budget | Function calling | Built-in tools | Structured output |
|---|---|---|---|---|---|---|
qwen-plus | 1M | 32k | 80k | ✓ | — | ✓ |
qwen-plus-latest | 1M | 32k | 80k | ✓ | — | ✓ |
qwen-plus-2025-12-01 | 1M | 32k | 80k | ✓ | — | ✓ |
qwen-plus-2025-09-11 | 1M | 32k | 80k | ✓ | — | ✓ |
qwen-plus-2025-07-28 | 1M | 32k | 80k | ✓ | — | ✓ |
qwen-plus-2025-07-14 | 128k | 16k | 80k | ✓ | — | ✓ |
qwen-plus-2025-04-28 | 128k | 16k | 80k | ✓ | — | ✓ |
qwen-plus-2025-01-25 | 128k | 8k | — | ✓ | — | ✓ |
qwen-max | 32k | 8k | — | ✓ | — | ✓ |
qwen-max-latest | 32k | 8k | — | ✓ | — | ✓ |
qwen-max-2025-01-25 | 32k | 8k | — | ✓ | — | ✓ |
qwen-flash | 1M | 32k | 80k | ✓ | — | ✓ |
qwen-flash-2025-07-28 | 1M | 32k | 80k | ✓ | — | ✓ |
qwen-turbo | 128k | 16k | 38k | ✓ | — | ✓ |
qwen-turbo-latest | 128k | 16k | 38k | ✓ | — | ✓ |
qwen-turbo-2025-04-28 | 128k | 16k | 38k | ✓ | — | ✓ |
qwen-turbo-2024-11-01 | 1M | 8k | — | ✓ | — | ✓ |
qwq-plus | 128k | 8k | 32k | — | — | — |
qvq-max | 128k | 8k | 80k | — | — | — |
qvq-max-latest | 128k | 8k | 80k | — | — | — |
qvq-max-2025-03-25 | 128k | 8k | 80k | — | — | — |
qwen-omni-turbo | 32k | 2k | 80k | — | — | — |
qwen-omni-turbo-latest | 32k | 2k | 80k | — | — | — |
qwen-omni-turbo-2025-03-26 | 32k | 2k | 80k | — | — | — |