kling-v3-omni
kling
Fully multimodal variant of the 3.0 line, accepting text, image, and video references in a single generation, with voice-driven characters and native audio.
- 上下文
- 待核
- 输入 / 1M tokens
- -
- 输出 / 1M tokens
- -
- 缓存读 / 1M
- -
- 缓存写 / 1M
- -
选择模型后可复制配置给 AI,也可以快速测试当前连接延迟。
kling
Fully multimodal variant of the 3.0 line, accepting text, image, and video references in a single generation, with voice-driven characters and native audio.
kling
Cinematic-oriented model with first- and last-frame control and video reference input, including edits driven by an existing clip. Suited for narrative shot sequences.