模型广场

选择模型后可复制配置给 AI,也可以快速测试当前连接延迟。

重置筛选
全部 13文本 13图像 4音频 0视频 0

gemini-2.5-flash

文本图像视频音频文本

Google

A high-performance general-purpose model from Google, designed for advanced reasoning, coding, mathematics, and scientific tasks. Its built-in thinking capabilities improve response accuracy and enable deeper contextual understanding.

上下文
1.1M
输入 / 1M tokens
$0.3
输出 / 1M tokens
$2.5
缓存读 / 1M
$0.03
缓存写 / 1M
$0.08333
TechCodingTranslation

gemini-2.5-flash-lite

文本图像视频音频文本

Google

A lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency, with faster token generation and higher throughput. Thinking is disabled by default but can be enabled via API.

上下文
1.1M
输入 / 1M tokens
$0.1
输出 / 1M tokens
$0.4
缓存读 / 1M
$0.01
缓存写 / 1M
$0.08333
FinanceCodingTranslation

gemini-2.5-pro

文本图像视频音频文本

Google

Massive context window. Great for video understanding.

上下文
1M
输入 / 1M tokens
$2.5
输出 / 1M tokens
$15
缓存读 / 1M
$0.25
缓存写 / 1M
$0
Vision

gemini-3-flash-preview

文本图像视频音频文本

Google

Speed-tier variant in the Gemini 3 line, pairing Pro-grade reasoning with Flash-level latency and cost, built for agent workflows and high-throughput interactive apps.

上下文
1M
输入 / 1M tokens
$0.5
输出 / 1M tokens
$3
缓存读 / 1M
$0.05
缓存写 / 1M
$0
CodingTranslationFinance

gemini-3.1-flash-lite

文本图像视频音频文本

Google

A high-efficiency multimodal lite model for low-latency, high-volume workloads like translation, classification, and data extraction, priced at about half of Gemini 3 Flash.

上下文
1M
输入 / 1M tokens
$0.25
输出 / 1M tokens
$1.5
缓存读 / 1M
$0.025
缓存写 / 1M
$0.08333
Instruction FollowingTask AutomationStructured OutputClassification

gemini-3.1-pro-preview

文本图像视频音频文本

Google

Google's latest. Strong reasoning with native image and video.

上下文
1M
输入 / 1M tokens
$2
输出 / 1M tokens
$12
缓存读 / 1M
$0.2
缓存写 / 1M
$0
Vision

gemini-3.5-flash

文本图像视频音频文本

Google

Designed for efficient multimodal AI tasks, offering strong coding, reasoning, real-time chat, and agent execution at Flash-tier cost and speed.

上下文
1M
输入 / 1M tokens
$1.5
输出 / 1M tokens
$9
缓存读 / 1M
$0.15
缓存写 / 1M
$0.08333
Production CodeRefactoringAgent CodingDaily Dev

gemma-3n-e4b-it:free

文本图像视频音频文本

Google

Fast and cost-efficient.

上下文
33K
输入 / 1M tokens
$0.06
输出 / 1M tokens
$0.12
缓存读 / 1M
$0
缓存写 / 1M
$0
ChatReal-time ResponseClassification

gemma-4-31b

文本图像文本

Google

Open-weights dense model with stronger math and coding, plus native function calling for self-hosted agent workflows.

上下文
262K
输入 / 1M tokens
$0.14
输出 / 1M tokens
$0.4
缓存读 / 1M
$0
缓存写 / 1M
$0
CodingMathLocal DeploymentAgent

gemini-2.5-flash-image

文本图像文本图像

Google

Built for fast generation and conversational editing; low latency and cost; multimodal text-and-image input.

上下文
33K
输入 / 1M tokens
-
输出 / 1M tokens
-
缓存读 / 1M
-
缓存写 / 1M
-
Real-time ResponseCreative GenerationHigh-Quality GenerationMarketing

gemini-3.1-flash-image-preview

文本图像文本图像

Google

Mid-tier image generation and editing model in the series, with near-flagship visual quality at higher generation speed. Renders legible text and photorealistic subjects, suited for infographics and marketing visuals.

上下文
131K
输入 / 1M tokens
-
输出 / 1M tokens
-
缓存读 / 1M
-
缓存写 / 1M
-
Creative GenerationHigh-Quality GenerationContent Generation

gemini-3.1-flash-lite-image

文本图像文本图像

Google

The lightest image model in its series, producing images in about four seconds with consistent character rendering and precise edits. Suited for high-throughput pipelines and interactive visual iteration.

上下文
66K
输入 / 1M tokens
-
输出 / 1M tokens
-
缓存读 / 1M
-
缓存写 / 1M
-
Cost EffectiveReal-time ResponseLightweightCreative Generation

gemini-3-pro-image-preview

文本图像文本图像

Google

Built-in reasoning; complex multi-turn creation; up to 4K; optional web grounding.

上下文
66K
输入 / 1M tokens
-
输出 / 1M tokens
-
缓存读 / 1M
-
缓存写 / 1M
-
Creative GenerationHigh-Quality GenerationContent Generation
需要完整模型清单或企业接入方案?进入控制台