模型广场

选择模型后可复制配置给 AI,也可以快速测试当前连接延迟。

重置筛选
全部 10文本 0图像 0音频 0视频 10

kling-v1-6

文本图像视频

kling

Kling 1.6 supports text, first or last frame, and multi-reference-image video generation.

上下文
待核
输入 / 1M tokens
-
输出 / 1M tokens
-
缓存读 / 1M
-
缓存写 / 1M
-
Cost EffectiveCreative GenerationContent Generation

kling-v2-master

文本图像视频

kling

Flagship of the 2.0 line, tuned for instruction following and cinematic aesthetics, with stylized effect presets. Outputs 1080p only.

上下文
待核
输入 / 1M tokens
-
输出 / 1M tokens
-
缓存读 / 1M
-
缓存写 / 1M
-
Creative GenerationContent GenerationHigh-Quality Generation

kling-v2-1

图像文本视频

kling

Kling 2.1 image-to-video generation supports first and last frame control.

上下文
待核
输入 / 1M tokens
-
输出 / 1M tokens
-
缓存读 / 1M
-
缓存写 / 1M
-
Creative GenerationCost EffectiveContent Generation

kling-v2-1-master

文本图像视频

kling

Quality-oriented tier of the 2.1 line, tuned for motion performance and semantic responsiveness. Supports text-to-video and image-to-video, suited for single-shot clips.

上下文
待核
输入 / 1M tokens
-
输出 / 1M tokens
-
缓存读 / 1M
-
缓存写 / 1M
-
High-Quality GenerationContent GenerationCreative Generation

kling-v2-5-turbo

文本图像视频

kling

Cost-efficient tier in the 2.5 generation, with improved prompt adherence and physics under large-amplitude motion. Outputs silent 5- or 10-second clips.

上下文
待核
输入 / 1M tokens
-
输出 / 1M tokens
-
缓存读 / 1M
-
缓存写 / 1M
-
Cost EffectiveCreative GenerationContent Generation

kling-2.6

文本图像视频

kling

First version to add native audio, producing visuals, voiceover, sound effects, and ambience in one pass. Suited for finished clips up to 10 seconds.

上下文
待核
输入 / 1M tokens
-
输出 / 1M tokens
-
缓存读 / 1M
-
缓存写 / 1M
-
Creative GenerationHigh-Quality GenerationContent Generation

kling-3.0

文本图像视频

kling

Core model of the 3.0 line, with first- and last-frame control, multi-shot storytelling, and native audio produced in the same pass. Suited for finished short-form work.

上下文
待核
输入 / 1M tokens
-
输出 / 1M tokens
-
缓存读 / 1M
-
缓存写 / 1M
-
Creative GenerationHigh-Quality GenerationContent Generation

kling-v3-omni

文本图像视频视频

kling

Fully multimodal variant of the 3.0 line, accepting text, image, and video references in a single generation, with voice-driven characters and native audio.

上下文
待核
输入 / 1M tokens
-
输出 / 1M tokens
-
缓存读 / 1M
-
缓存写 / 1M
-
Creative GenerationHigh-Quality GenerationContent GenerationMultimodal Understanding

kling-3.0-turbo

文本图像视频

kling

Speed-and-cost tier of the 3.0 line, covering text-to-video and first-frame image-to-video with multi-shot output. Suited for high-volume production.

上下文
待核
输入 / 1M tokens
-
输出 / 1M tokens
-
缓存读 / 1M
-
缓存写 / 1M
-
Creative GenerationContent GenerationCost Effective

kling-o1

文本图像视频视频

kling

Cinematic-oriented model with first- and last-frame control and video reference input, including edits driven by an existing clip. Suited for narrative shot sequences.

上下文
待核
输入 / 1M tokens
-
输出 / 1M tokens
-
缓存读 / 1M
-
缓存写 / 1M
-
Creative GenerationHigh-Quality GenerationContent GenerationMultimodal Understanding
需要完整模型清单或企业接入方案?进入控制台