qwen3-vl-8b-instruct
qwen
Lightweight multimodal model with balanced performance.
- 上下文
- 128K
- 输入 / 1M tokens
- $0.25
- 输出 / 1M tokens
- $0.75
- 缓存读 / 1M
- $0.12
- 缓存写 / 1M
- $0
选择模型后可复制配置给 AI,也可以快速测试当前连接延迟。
qwen
Lightweight multimodal model with balanced performance.
qwen
Plus-tier vision-language model in the Qwen3-VL line, handling text, image and video inputs with strengths in document parsing, video understanding, spatial grounding and agent tool use.
qwen
Open-weight mid-tier MoE text model with long chain-of-thought reasoning, strong on function-calling and multi-step planning benchmarks for self-hosted agentic workflows.
qwen
Mid-tier Qwen model tuned for multi-step reasoning and agent workflows, compact enough for private and on-prem deployment.
qwen
Lightweight mid-tier variant tuned for low-latency agent workflows with coding and tool-calling support.
qwen
Multimodal model from the Qwen3.5 line with toggleable thinking mode, tuned for image and video understanding, document parsing, and multimodal agents.
qwen
Thinking-mode variant of the dense open-weight model that outputs step-by-step reasoning; at 27B it surpasses the prior open-weight 397B-A17B flagship across coding, math and multi-step reasoning benchmarks.
qwen
Qwen 3.6 lightweight tier with text, image, and video input. Low latency and low cost, well-suited for high-volume classification, extraction, summarization, and simple agent workflows.
qwen
Enhanced Qwen model with strong bilingual (Chinese/English) comprehension, excelling at long-document analysis and structured output.
qwen
Qwen3.7 Flash is the lightweight tier in the series, a vision-language reasoning model with strengths in object recognition, spatial understanding, and real-world visual perception. Suited for multimodal agents, visual coding, search, and computer interaction.
qwen
Qwen 3.7 Plus is a mid-tier multimodal model that reads screens, operates GUIs, and navigates mobile apps end-to-end. Suited for agent workflows and tool calling.
qwen
Flash tier in Qwen3.8: multimodal reasoning for coding and agent workflows, with visual understanding across documents, codebases, and long video.
qwen
Qwen3.8 series flagship, the general-availability successor to Max Preview. A multimodal reasoning model built for complex reasoning, visual understanding, coding, and agentic workflows.