qwen-image-2.0
qwen
Unifies generation and editing in one model; native 2K; professional typography and infographics.
- 上下文
- 待核
- 输入 / 1M tokens
- -
- 输出 / 1M tokens
- -
- 缓存读 / 1M
- -
- 缓存写 / 1M
- -
选择模型后可复制配置给 AI,也可以快速测试当前连接延迟。
qwen
Unifies generation and editing in one model; native 2K; professional typography and infographics.
qwen
Fuses generation and editing; more professional text rendering and richer realistic textures and scenes.
qwen
Stronger industrial design, geometric reasoning, and character consistency; text, object, and style edits.
qwen
Multi-image input/output and custom resolution; multiple edited results per request.
qwen
Diverse styles; multi-line layouts and paragraph-level text generation.
qwen
Instruction-tuned Qwen3 variant without thinking mode, strong on multilingual reasoning, math, code, and tool use for agent workflows.
qwen
Internal fine-tune of Qwen3-32B for text-only Chinese workloads, tuned to in-house instruction style for QA, summarization and rewriting.
qwen
Qwen3-generation thinking-mode reasoning model with native tool use, tuned for math, coding, and multi-step agentic workflows.
qwen
Deep-thinking preview that reasons step-by-step before answering, holding up on multi-step math, code, and Chinese-language tasks.
qwen
Instruction-tuned non-thinking chat model that answers directly without exposing reasoning, suited to agent workflows needing deterministic output like coding assistance and tool calling.
qwen
Lightweight multimodal model with balanced performance.
qwen
Plus-tier vision-language model in the Qwen3-VL line, handling text, image and video inputs with strengths in document parsing, video understanding, spatial grounding and agent tool use.
qwen
Open-weight mid-tier MoE text model with long chain-of-thought reasoning, strong on function-calling and multi-step planning benchmarks for self-hosted agentic workflows.
qwen
Mid-tier Qwen model tuned for multi-step reasoning and agent workflows, compact enough for private and on-prem deployment.
qwen
Lightweight mid-tier variant tuned for low-latency agent workflows with coding and tool-calling support.
qwen
Multimodal model from the Qwen3.5 line with toggleable thinking mode, tuned for image and video understanding, document parsing, and multimodal agents.
qwen
Thinking-mode variant of the dense open-weight model that outputs step-by-step reasoning; at 27B it surpasses the prior open-weight 397B-A17B flagship across coding, math and multi-step reasoning benchmarks.
qwen
Open-weight coding model tuned for agentic terminal tasks and repo-scale reasoning with low inference cost.
qwen
Qwen 3.6 lightweight tier with text, image, and video input. Low latency and low cost, well-suited for high-volume classification, extraction, summarization, and simple agent workflows.
qwen
Enhanced Qwen model with strong bilingual (Chinese/English) comprehension, excelling at long-document analysis and structured output.
qwen
Qwen3.7 Flash is the lightweight tier in the series, a vision-language reasoning model with strengths in object recognition, spatial understanding, and real-world visual perception. Suited for multimodal agents, visual coding, search, and computer interaction.
qwen
Flagship agent-centric reasoning model with a 1M context window, excelling at coding, productivity, and long-horizon autonomous tasks.
qwen
Qwen 3.7 Plus is a mid-tier multimodal model that reads screens, operates GUIs, and navigates mobile apps end-to-end. Suited for agent workflows and tool calling.
qwen
Flash tier in Qwen3.8: multimodal reasoning for coding and agent workflows, with visual understanding across documents, codebases, and long video.
qwen
Qwen3.8 series flagship, the general-availability successor to Max Preview. A multimodal reasoning model built for complex reasoning, visual understanding, coding, and agentic workflows.