Choose a model for AI agents, chatbots, document processing, and more.
Using OpenClaw or Claude Code?
qwen3.7-plus — balanced performance and cost, full tool support, 1M context for large codebases. For strongest reasoning, choose qwen3.7-max, or qwen3.8-max-preview (Token Plan only).
Migrate from closed-source models
Map your current GPT, Claude, or Gemini model to an equivalent QwenCloud model.
| Tier | Closed-source examples | QwenCloud recommendation |
|---|---|---|
| Highest capability | GPT-5.5, Claude Opus 4.7, Gemini 3.1 Pro | qwen3.7-max, qwen3.8-max-preview (Token Plan only) |
| Balanced | GPT-5.4, Claude Sonnet 4.6, Gemini 3 Pro | qwen3.7-plus, deepseek-v4-pro |
| Lightweight & low-cost | GPT-5.4-mini, Claude Haiku 4.5, Gemini 3.1 Flash | qwen3.7-flash, deepseek-v4-flash |
For other applications
Chatbots, content generation, summarization, document processing — start with qwen3.7-plus, which balances performance, cost, and built-in tools with a 1M-token context window. To cut costs, switch to qwen3.7-flash, which offers similar capabilities at a lower price. For the strongest reasoning, use qwen3.7-max (1M context, higher cost), or qwen3.8-max-preview (Token Plan only).
Office productivity (non-coding)
For non-coding office tasks such as document drafting, email composition, meeting note summarization, and data analytics, start with qwen3.7-plus — it balances performance and cost with a 1M-token context window, Function Calling support, and built-in tools. To reduce costs, try qwen3.7-flash, which delivers near-flagship performance at a lower price with the same context length. For the strongest reasoning capability, choose qwen3.7-max (higher cost), or qwen3.8-max-preview (Token Plan only). For long-document processing such as reviewing multiple contracts, use qwen-long (10M-token context window).
TONGYI Lingma and Qoder are AI coding tools designed for software development. They are not intended for general office productivity tasks.
Context window
1M tokens is roughly 750,000 words or 10 novels.
- Long documents or large codebases →
qwen3.7-max/qwen3.7-plus/qwen3.7-flash(1M) - Standard tasks → 128k–256k is plenty
Thinking mode
Step-by-step reasoning for multi-step math, debugging, architecture planning, or legal cross-referencing.
Toggle with enable_thinking. All Qwen3+ models support it — most are hybrid, so you can switch per request.
Function calling + built-in tools
Let the model take actions: check weather, query a database, book a meeting.
- Function calling (you define tools, model calls them): all general-purpose models
- Built-in tools (web search, code execution — no setup):
qwen3.7-max,qwen3.7-plus,qwen3.7-flash,qwen3.6-flash,qwen3.5-plus,qwen3.5-flash,qwen3-maxseries only
Recommended models
| Model | Context | Thinking | Function calling | Built-in tools | Structured output |
|---|---|---|---|---|---|
qwen3.8-max-preview (Token Plan only) | 1M | ✓ | ✓ | ✓ | — |
qwen3.7-max | 1M | ✓ | ✓ | ✓ | — |
qwen3.7-plus | 1M | ✓ | ✓ | ✓ | ✓ |
qwen3.7-flash | 1M | ✓ | ✓ | ✓ | ✓ |
qwen3.6-flash | 1M | ✓ | ✓ | ✓ | ✓ |
deepseek-v4-pro | 1M | ✓ | ✓ | — | — |
deepseek-v4-flash | 1M | ✓ | ✓ | — | — |
All models
Qwen3.7
Qwen3.7
| Model ID | Context | Max Output | Thinking Budget | Function calling | Built-in tools | Structured output |
|---|---|---|---|---|---|---|
qwen3.7-max | 1M | 64k | 256k | ✓ | ✓ | — |
qwen3.7-max-2026-06-08 | 1M | 64k | 256k | ✓ | ✓ | — |
qwen3.7-max-2026-05-20 | 1M | 64k | 256k | ✓ | ✓ | — |
qwen3.7-max-preview | 1M | 64k | 256k | ✓ | ✓ | — |
qwen3.7-max-2026-05-17 | 1M | 64k | 256k | ✓ | ✓ | — |
qwen3.7-plus | 1M | 64k | 256k | ✓ | ✓ | ✓ |
qwen3.7-plus-2026-05-26 | 1M | 64k | 256k | ✓ | ✓ | ✓ |
qwen3.7-flash | 1M | 64k | 256k | ✓ | ✓ | ✓ |
qwen3.7-flash-2026-07-15 | 1M | 64k | 256k | ✓ | ✓ | ✓ |
Qwen3.6
Qwen3.6
| Model ID | Context | Max Output | Thinking Budget | Function calling | Built-in tools | Structured output |
|---|---|---|---|---|---|---|
qwen3.6-max-preview | 256k | 64k | 128k | ✓ | — | ✓ |
qwen3.6-flash | 1M | 64k | 80k | ✓ | ✓ | ✓ |
qwen3.6-flash-2026-04-16 | 1M | 64k | 80k | ✓ | ✓ | ✓ |
qwen3.6-35b-a3b | 256k | 64k | 80k | ✓ | ✓ | ✓ |
qwen3.6-27b | 256k | 64k | 80k | ✓ | — | ✓ |
qwen3.6-plus | 1M | 64k | 80k | ✓ | ✓ | ✓ |
qwen3.6-plus-2026-04-02 | 1M | 64k | 80k | ✓ | ✓ | ✓ |
Qwen3.5
Qwen3.5
| Model ID | Context | Max Output | Thinking Budget | Function calling | Built-in tools | Structured output |
|---|---|---|---|---|---|---|
qwen3.5-plus | 1M | 64k | 80k | ✓ | ✓ | ✓ |
qwen3.5-plus-2026-04-20 | 1M | 64k | 80k | ✓ | ✓ | ✓ |
qwen3.5-plus-2026-02-15 | 1M | 64k | 80k | ✓ | ✓ | ✓ |
qwen3.5-flash | 1M | 64k | 80k | ✓ | ✓ | ✓ |
qwen3.5-flash-2026-02-23 | 1M | 64k | 80k | ✓ | ✓ | ✓ |
qwen3.5-397b-a17b | 256k | 64k | 80k | ✓ | ✓ | ✓ |
qwen3.5-122b-a10b | 256k | 64k | 80k | ✓ | ✓ | ✓ |
qwen3.5-27b | 256k | 64k | 80k | ✓ | ✓ | ✓ |
qwen3.5-35b-a3b | 256k | 64k | 80k | ✓ | ✓ | ✓ |
Specialized
Specialized
Translation
| Model ID | Context | Max Output | Function calling | Built-in tools | Structured output |
|---|---|---|---|---|---|
qwen-mt-plus | 16k | 8k | — | — | — |
qwen-mt-turbo | 16k | 8k | — | — | — |
qwen-mt-flash | 16k | 8k | — | — | — |
qwen-mt-lite | 16k | 8k | — | — | — |
Character roleplay
| Model ID | Context | Max Output | Function calling | Built-in tools | Structured output |
|---|---|---|---|---|---|
qwen-plus-character | 32k | 4k | — | — | — |
qwen-plus-character-ja | 8k | 4k | — | — | — |
qwen-flash-character | 8k | 4k | — | — | — |
Third-party
Third-party
Non-Qwen models available through the same API.
* DeepSeek V4 models share a 384k total budget across output and thinking.
| Model ID | Context | Max Output | Thinking Budget | Function calling | Built-in tools | Structured output |
|---|---|---|---|---|---|---|
deepseek-v4-pro | 1M | 384k * | * | ✓ | — | — |
deepseek-v4-flash | 1M | 384k * | * | ✓ | — | — |
deepseek-v3.2 | 128k | 64k | 32k | ✓ | — | — |
Legacy
Legacy
Previous generation models. We recommend Qwen3.6 for new projects.
Qwen3
| Model ID | Context | Max Output | Thinking Budget | Function calling | Built-in tools | Structured output |
|---|---|---|---|---|---|---|
qwen3-max | 256k | 64k | 80k | ✓ | ✓ | ✓ |
qwen3-max-2026-01-23 | 256k | 64k | 80k | ✓ | ✓ | ✓ |
qwen3-max-preview | 256k | 64k | 80k | ✓ | ✓ | ✓ |
qwen3-max-2025-09-23 | 256k | 64k | — | ✓ | ✓ | ✓ |
qwen3-235b-a22b | 128k | 16k | 38k | ✓ | — | ✓ |
qwen3-235b-a22b-thinking-2507 | 128k | 32k | 80k | ✓ | — | — |
qwen3-235b-a22b-instruct-2507 | 128k | 32k | — | ✓ | — | ✓ |
qwen3-next-80b-a3b-thinking | 128k | 32k | 80k | ✓ | — | — |
qwen3-next-80b-a3b-instruct | 128k | 32k | — | ✓ | — | ✓ |
qwen3-32b | 128k | 16k | 38k | ✓ | — | ✓ |
qwen3-30b-a3b | 128k | 16k | 38k | ✓ | — | ✓ |
qwen3-30b-a3b-thinking-2507 | 128k | 32k | 80k | ✓ | — | — |
qwen3-30b-a3b-instruct-2507 | 128k | 32k | — | ✓ | — | ✓ |
qwen3-14b | 128k | 8k | 38k | ✓ | — | ✓ |
qwen3-8b | 128k | 8k | 38k | ✓ | — | ✓ |
qwen3-4b | 128k | 8k | 38k | ✓ | — | ✓ |
qwen3-1.7b | 32k | 8k | 30k | ✓ | — | ✓ |
qwen3-0.6b | 32k | 8k | 30k | ✓ | — | ✓ |
Qwen3-Coder
| Model ID | Context | Max Output | Function calling | Built-in tools | Structured output |
|---|---|---|---|---|---|
qwen3-coder-plus | 1M | 64k | ✓ | — | ✓ |
qwen3-coder-plus-2025-09-23 | 1M | 64k | ✓ | — | ✓ |
qwen3-coder-plus-2025-07-22 | 1M | 64k | ✓ | — | ✓ |
qwen3-coder-flash | 1M | 64k | ✓ | — | ✓ |
qwen3-coder-flash-2025-07-28 | 1M | 64k | ✓ | — | ✓ |
qwen3-coder-next | 256k | 64k | ✓ | — | ✓ |
qwen3-coder-480b-a35b-instruct | 256k | 64k | ✓ | — | ✓ |
qwen3-coder-30b-a3b-instruct | 256k | 64k | ✓ | — | ✓ |
Qwen2.5 (open source)
| Model ID | Context | Max Output | Function calling | Built-in tools | Structured output |
|---|---|---|---|---|---|
qwen2.5-omni-7b | 32k | 8k | ✓ | — | ✓ |
qwen2.5-vl-72b-instruct | 128k | 8k | ✓ | — | ✓ |
qwen2.5-vl-32b-instruct | 128k | 8k | ✓ | — | ✓ |
qwen2.5-vl-7b-instruct | 128k | 8k | ✓ | — | ✓ |
qwen2.5-vl-3b-instruct | 128k | 8k | ✓ | — | ✓ |
qwen2.5-72b-instruct | 32k | 8k | ✓ | — | ✓ |
qwen2.5-32b-instruct | 32k | 8k | ✓ | — | ✓ |
qwen2.5-14b-instruct | 32k | 8k | ✓ | — | ✓ |
qwen2.5-14b-instruct-1m | 1M | 8k | ✓ | — | ✓ |
qwen2.5-7b-instruct | 32k | 8k | ✓ | — | ✓ |
qwen2.5-7b-instruct-1m | 1M | 8k | ✓ | — | ✓ |
Legacy (qwen-plus/max/flash/turbo)
| Model ID | Context | Max Output | Thinking Budget | Function calling | Built-in tools | Structured output |
|---|---|---|---|---|---|---|
qwen-plus | 1M | 32k | 80k | ✓ | — | ✓ |
qwen-plus-latest | 1M | 32k | 80k | ✓ | — | ✓ |
qwen-plus-2025-12-01 | 1M | 32k | 80k | ✓ | — | ✓ |
qwen-plus-2025-09-11 | 1M | 32k | 80k | ✓ | — | ✓ |
qwen-plus-2025-07-28 | 1M | 32k | 80k | ✓ | — | ✓ |
qwen-plus-2025-07-14 | 128k | 16k | 80k | ✓ | — | ✓ |
qwen-plus-2025-04-28 | 128k | 16k | 80k | ✓ | — | ✓ |
qwen-plus-2025-01-25 | 128k | 8k | — | ✓ | — | ✓ |
qwen-max | 32k | 8k | — | ✓ | — | ✓ |
qwen-max-latest | 32k | 8k | — | ✓ | — | ✓ |
qwen-max-2025-01-25 | 32k | 8k | — | ✓ | — | ✓ |
qwen-flash | 1M | 32k | 80k | ✓ | — | ✓ |
qwen-flash-2025-07-28 | 1M | 32k | 80k | ✓ | — | ✓ |
qwen-turbo | 128k | 16k | 38k | ✓ | — | ✓ |
qwen-turbo-latest | 128k | 16k | 38k | ✓ | — | ✓ |
qwen-turbo-2025-04-28 | 128k | 16k | 38k | ✓ | — | ✓ |
qwen-turbo-2024-11-01 | 1M | 8k | — | ✓ | — | ✓ |
qwq-plus | 128k | 8k | 32k | — | — | — |
qvq-max | 128k | 8k | 80k | — | — | — |
qvq-max-latest | 128k | 8k | 80k | — | — | — |
qvq-max-2025-03-25 | 128k | 8k | 80k | — | — | — |
qwen-omni-turbo | 32k | 2k | 80k | — | — | — |
qwen-omni-turbo-latest | 32k | 2k | 80k | — | — | — |
qwen-omni-turbo-2025-03-26 | 32k | 2k | 80k | — | — | — |