模型广场

选择模型后可复制配置给 AI,也可以快速测试当前连接延迟。

重置筛选
全部 8文本 0图像 0音频 0视频 8

wan2.5-t2v-preview

文本音频视频

Alibaba

Generate 5-10 second, 480p/720p/1080p video from text prompts with native synced audio track.

上下文
待核
输入 / 1M tokens
-
输出 / 1M tokens
-
缓存读 / 1M
-
缓存写 / 1M
-
Creative GenerationHigh-Quality GenerationContent GenerationMultimodal Understanding

wan2.6-i2v

文本图像音频视频

Alibaba

Alibaba Tongyi Wanxiang Wan 2.6 image-to-video model animates a reference image into 720p/1080p clips with synchronized audio, preserving subject consistency with cinematic motion.

上下文
待核
输入 / 1M tokens
-
输出 / 1M tokens
-
缓存读 / 1M
-
缓存写 / 1M
-
Creative GenerationMultimodal Understanding

wan2.6-i2v-flash

文本图像音频视频

Alibaba

Speed-optimized Wan 2.6 image-to-video variant with optional audio output, generating 720p/1080p clips quickly and cost-effectively for high-throughput image animation.

上下文
待核
输入 / 1M tokens
-
输出 / 1M tokens
-
缓存读 / 1M
-
缓存写 / 1M
-
Cost EffectiveReal-time ResponseCreative Generation

wan2.6-t2v

文本音频视频

Alibaba

Generate 720p/1080p video from text prompts and come with synchronized audio tracks. The picture has movie-level aesthetics and complex motion performance, suitable for creative video creation.

上下文
待核
输入 / 1M tokens
-
输出 / 1M tokens
-
缓存读 / 1M
-
缓存写 / 1M
-
Creative GenerationHigh-Quality GenerationContent GenerationMultimodal Understanding

seedance-2.0

文本图像视频音频视频

bytedance

Standard tier of the Seedance family. Supports text-to-video, image-to-video, and reference-to-video, with strong character, style, and camera consistency. Built for ads and short-form creative.

上下文
待核
输入 / 1M tokens
-
输出 / 1M tokens
-
缓存读 / 1M
-
缓存写 / 1M
-
VisionFrontierCreative GenerationContent Generation

seedance-2.0-fast

文本图像视频音频视频

bytedance

Lightweight Seedance 2.0 variant tuned for speed and cost rather than peak fidelity, with text-to-video, first/last frame control, and multi-reference inputs.

上下文
待核
输入 / 1M tokens
-
输出 / 1M tokens
-
缓存读 / 1M
-
缓存写 / 1M
-
Cost EffectiveCreative GenerationContent GenerationBatch Generation

seedance-2.5

文本图像视频音频视频

bytedance

A video generation model built for long-form storytelling, with first-frame and first-and-last-frame control, up to 50 multimodal reference assets, and optional audio generation. It also supports video editing and extension.

上下文
131K
输入 / 1M tokens
-
输出 / 1M tokens
-
缓存读 / 1M
-
缓存写 / 1M
-
FrontierCreative GenerationHigh-Quality GenerationContent Generation

minimax-h3

文本图像视频音频视频

minimax

A lightweight open-weight video generation model focused on instruction-guided editing, text and brand rendering, and video-to-video motion transfer, with native audiovisual output. Built for advertising, e-commerce, and interface design workflows.

上下文
131K
输入 / 1M tokens
-
输出 / 1M tokens
-
缓存读 / 1M
-
缓存写 / 1M
-
LightweightCreative GenerationContent GenerationMarketing
需要完整模型清单或企业接入方案?进入控制台