gemma-3n-e4b-it:free
Fast and cost-efficient.
- 上下文
- 33K
- 输入 / 1M tokens
- $0.06
- 输出 / 1M tokens
- $0.12
- 缓存读 / 1M
- $0
- 缓存写 / 1M
- $0
选择模型后可复制配置给 AI,也可以快速测试当前连接延迟。
Fast and cost-efficient.
liquid
Instruction-tuned optimization for precise execution.
nvidia
Balanced powerful model.
OpenAI
Cheapest GPT. Built for high-volume simple jobs.
qwen
Mid-tier Qwen model tuned for multi-step reasoning and agent workflows, compact enough for private and on-prem deployment.
qwen
Lightweight mid-tier variant tuned for low-latency agent workflows with coding and tool-calling support.
qwen
Qwen3.7 Flash is the lightweight tier in the series, a vision-language reasoning model with strengths in object recognition, spatial understanding, and real-world visual perception. Suited for multimodal agents, visual coding, search, and computer interaction.