claude-haiku-4.5
Anthropic
Fast and affordable. Handles simple tasks well.
- 上下文
- 200K
- 输入 / 1M tokens
- $1
- 输出 / 1M tokens
- $5
- 缓存读 / 1M
- $0.1
- 缓存写 / 1M
- $1.25
选择模型后可复制配置给 AI,也可以快速测试当前连接延迟。
Anthropic
Fast and affordable. Handles simple tasks well.
Anthropic
Previous-gen Sonnet. Still strong for long-form content.
Anthropic
Best balance of quality and cost from Anthropic.
Anthropic
Coding and agent quality approach the frontier, with native reasoning built in, well-suited for complex planning and long-horizon professional tasks that require cross-step consistency.
Massive context window. Great for video understanding.
Google's latest. Strong reasoning with native image and video.
Designed for efficient multimodal AI tasks, offering strong coding, reasoning, real-time chat, and agent execution at Flash-tier cost and speed.
minimax
MiniMax\\'s first multimodal release, accepting image and video input. Long-context inference runs cheaper and faster, and it handles multi-step agent workflows and computer-use scenarios.
Moonshot
A multimodal reasoning model for complex coding, knowledge work, and long-running agent tasks. It works across large codebases, calls tools, and debugs, using images, logs, and test output to refine its output.
OpenAI
Stable code generation with predictable output.
OpenAI
Balanced general-purpose model covering dialogue, coding, and multilingual tasks at reasonable speed and cost for assistants and routine agent workflows.
OpenAI
Reliable all-rounder for everyday tasks.
OpenAI
ChatGPT product-line snapshot of GPT-5, tuned for natural multi-turn conversation and consistent tone rather than agentic reasoning workloads.
OpenAI
Successor to GPT-5 that scales reasoning depth to task difficulty, returning fast answers on simple prompts and slower deliberation on hard ones, fit for general assistant, coding, and tool-use work.
OpenAI
Rolling snapshot of the GPT-5.1 build powering ChatGPT, tuned for natural multi-turn conversation rather than heavy reasoning workloads.
OpenAI
Previous flagship. Strong reasoning at a lower price.
OpenAI
Snapshot pointer to the GPT-5.2 Instant model that powers ChatGPT, optimized for everyday chat, writing, and light coding.
OpenAI
Chat-tuned everyday model that talks more naturally and refuses less, with better factual accuracy on common questions.
OpenAI
High-quality coding with stable refactoring.
OpenAI
Lightweight variant in the gpt-5.4 family, tuned for low-latency, high-volume tasks like classification, extraction, and sub-agent execution.
OpenAI
The middle rung of the GPT-5.6 lineup, tuned for daily dev work, general reasoning, and agent orchestration. Delivers reliable performance at a manageable cost for routine business workloads.
OpenAI
Reasoning model that works through problems step by step before answering, useful for math, science, and code questions that need multi-step derivation.
OpenAI
Cost-efficient reasoning model with adjustable thinking effort, tuned for STEM and coding tasks where deliberation matters.
OpenAI
Lightweight o-series reasoning model with built-in chain-of-thought, delivering steady math, coding, and tool-use quality at low latency and cost.
qwen
Flagship agent-centric reasoning model with a 1M context window, excelling at coding, productivity, and long-horizon autonomous tasks.
qwen
Qwen3.8 series flagship, the general-availability successor to Max Preview. A multimodal reasoning model built for complex reasoning, visual understanding, coding, and agentic workflows.
xAI
Reasoning-oriented Grok variant built for multi-step problem solving and agentic tool use, with strict prompt adherence and low hallucination.
xAI
Multi-agent variant that dispatches sub-agents in parallel for deep research, multi-step reasoning, and tool orchestration.
xAI
Reasoning-oriented Grok variant built for multi-step problem solving and agentic tool use, with strict prompt adherence and low hallucination.
xAI
Grok 4.3 is a reasoning model from xAI that supports text and image inputs with text output. It is suited for agentic workflows, instruction following, long-document analysis, and deep research. The model supports a 1M token context window. Requests above 200K total tokens are billed at a higher rate.
xAI
Grok's flagship reasoning model, running an internal chain of thought before responding. Strong performance on coding and STEM tasks, with support for text, image, and file inputs. Suited for technical problem-solving and long-document analysis.
xAI
Grok's agent-oriented release, built for long-running multi-step work across a codebase. Accepts text, image, and file inputs. Suited for research and interactive or visual output.
xiaomi
Flagship agent model for complex workflows.
Z.ai
Designed for agent-based workflows such as OpenClaw.
Z.ai
GLM 5.2 is a large-scale reasoning model with a 1M-token context window, suited for long-horizon agents, repo-level coding, and multi-step automation.
Z.ai
Flagship-tier model for complex software engineering and long-horizon agent tasks, with stable execution across multi-turn tool calls and large-scale codebase refactoring.