doubao-seed-2.0-lite-260428
bytedance
Lightweight tier of the Seed 2.0 family tuned for low-latency agent, coding, and GUI workloads.
- 上下文
- 256K
- 输入 / 1M tokens
- $0.25
- 输出 / 1M tokens
- $2
- 缓存读 / 1M
- $0.05
- 缓存写 / 1M
- $0.008333
选择模型后可复制配置给 AI,也可以快速测试当前连接延迟。
bytedance
Lightweight tier of the Seed 2.0 family tuned for low-latency agent, coding, and GUI workloads.
DeepSeek
Open-source general-purpose LLM with representative instruction-following and coding skills within the open-weights ecosystem, suited for chat, code assistance, and enterprise text workflows.
DeepSeek
Best value for code and math. Fraction of flagship cost.
DeepSeek
Experimental interim release focused on long-context efficiency, with optional reasoning mode and V3.1-level performance overall.
DeepSeek
Faster, lower-cost variant of DeepSeek V4, optimized for coding assistants, high-throughput chat, and latency-sensitive Agent workflows.
A high-performance general-purpose model from Google, designed for advanced reasoning, coding, mathematics, and scientific tasks. Its built-in thinking capabilities improve response accuracy and enable deeper contextual understanding.
A high-efficiency multimodal lite model for low-latency, high-volume workloads like translation, classification, and data extraction, priced at about half of Gemini 3 Flash.
minimax
Cost-effective general model for balanced speed and quality.
minimax
Built for conversation and bilingual chat.
minimax
A next-generation autonomous language model that uses multi-agent collaboration to plan, execute, and continuously refine complex tasks. It supports production-grade workflows including live debugging, root cause analysis, financial modeling, and document generation across Word, Excel, and PowerPoint.
OpenAI
Near GPT-4o performance at lower latency and cost, suited for high-frequency interactions, coding, and vision tasks.
OpenAI
Fast and low-cost; for high-concurrency text and reasoning tasks.
OpenAI
Lightweight variant in the gpt-5.4 family, tuned for low-latency, high-volume tasks like classification, extraction, and sub-agent execution.
OpenAI
Entry point of the GPT-5.6 lineup. A small, low-latency workhorse for chat, tagging, and lightweight agent loops, keeping reasoning solid enough for routine work.
qwen
Instruction-tuned Qwen3 variant without thinking mode, strong on multilingual reasoning, math, code, and tool use for agent workflows.
qwen
Internal fine-tune of Qwen3-32B for text-only Chinese workloads, tuned to in-house instruction style for QA, summarization and rewriting.
qwen
Qwen3-generation thinking-mode reasoning model with native tool use, tuned for math, coding, and multi-step agentic workflows.
qwen
Lightweight multimodal model with balanced performance.
qwen
Thinking-mode variant of the dense open-weight model that outputs step-by-step reasoning; at 27B it surpasses the prior open-weight 397B-A17B flagship across coding, math and multi-step reasoning benchmarks.
qwen
Open-weight coding model tuned for agentic terminal tasks and repo-scale reasoning with low inference cost.
qwen
Enhanced Qwen model with strong bilingual (Chinese/English) comprehension, excelling at long-document analysis and structured output.
qwen
Qwen 3.7 Plus is a mid-tier multimodal model that reads screens, operates GUIs, and navigates mobile apps end-to-end. Suited for agent workflows and tool calling.
xiaomi
Flagship coding and agent model that sustains thousand-tool-call autonomous workflows for complex software engineering.
Z.ai
Multimodal model with strong image + text understanding.