wan2.5-t2v-preview
Alibaba
Generate 5-10 second, 480p/720p/1080p video from text prompts with native synced audio track.
- 上下文
- 待核
- 输入 / 1M tokens
- -
- 输出 / 1M tokens
- -
- 缓存读 / 1M
- -
- 缓存写 / 1M
- -
选择模型后可复制配置给 AI,也可以快速测试当前连接延迟。
Alibaba
Generate 5-10 second, 480p/720p/1080p video from text prompts with native synced audio track.
Alibaba
Alibaba Tongyi Wanxiang Wan 2.6 image-to-video model animates a reference image into 720p/1080p clips with synchronized audio, preserving subject consistency with cinematic motion.
Alibaba
Speed-optimized Wan 2.6 image-to-video variant with optional audio output, generating 720p/1080p clips quickly and cost-effectively for high-throughput image animation.
Alibaba
Generate 720p/1080p video from text prompts and come with synchronized audio tracks. The picture has movie-level aesthetics and complex motion performance, suitable for creative video creation.
bytedance
Standard tier of the Seedance family. Supports text-to-video, image-to-video, and reference-to-video, with strong character, style, and camera consistency. Built for ads and short-form creative.
bytedance
Lightweight Seedance 2.0 variant tuned for speed and cost rather than peak fidelity, with text-to-video, first/last frame control, and multi-reference inputs.
bytedance
A video generation model built for long-form storytelling, with first-frame and first-and-last-frame control, up to 50 multimodal reference assets, and optional audio generation. It also supports video editing and extension.
minimax
A lightweight open-weight video generation model focused on instruction-guided editing, text and brand rendering, and video-to-video motion transfer, with native audiovisual output. Built for advertising, e-commerce, and interface design workflows.