模型广场

选择模型后可复制配置给 AI,也可以快速测试当前连接延迟。

重置筛选
全部 8文本 0图像 0音频 0视频 8

happyhorse-1.0-t2v

文本视频

Alibaba

Generates 3–15 second clips at up to 1080p from a written description, across a range of aspect ratios. Suited for taking a script straight to creative content and short-form social video.

上下文
待核
输入 / 1M tokens
-
输出 / 1M tokens
-
缓存读 / 1M
-
缓存写 / 1M
-
Creative GenerationHigh-Quality GenerationContent Generation

happyhorse-1.1-t2v

文本视频

Alibaba

Stronger prompt adherence than 1.0 on complex instructions. Generates 3–15 second clips at up to 1080p from a written description, suited for producing creative content directly from a script.

上下文
待核
输入 / 1M tokens
-
输出 / 1M tokens
-
缓存读 / 1M
-
缓存写 / 1M
-
Creative GenerationHigh-Quality GenerationContent Generation

wan2.5-t2v-preview

文本音频视频

Alibaba

Generate 5-10 second, 480p/720p/1080p video from text prompts with native synced audio track.

上下文
待核
输入 / 1M tokens
-
输出 / 1M tokens
-
缓存读 / 1M
-
缓存写 / 1M
-
Creative GenerationHigh-Quality GenerationContent GenerationMultimodal Understanding

wan2.6-i2v

文本图像音频视频

Alibaba

Alibaba Tongyi Wanxiang Wan 2.6 image-to-video model animates a reference image into 720p/1080p clips with synchronized audio, preserving subject consistency with cinematic motion.

上下文
待核
输入 / 1M tokens
-
输出 / 1M tokens
-
缓存读 / 1M
-
缓存写 / 1M
-
Creative GenerationMultimodal Understanding

wan2.6-i2v-flash

文本图像音频视频

Alibaba

Speed-optimized Wan 2.6 image-to-video variant with optional audio output, generating 720p/1080p clips quickly and cost-effectively for high-throughput image animation.

上下文
待核
输入 / 1M tokens
-
输出 / 1M tokens
-
缓存读 / 1M
-
缓存写 / 1M
-
Cost EffectiveReal-time ResponseCreative Generation

wan2.6-r2v

文本视频图像视频

Alibaba

Alibaba Tongyi Wanxiang Wan 2.6 reference-to-video model replicates the actions, effects, and camera movement of a reference video to produce up to 10 second 720p/1080p clips with synchronized audio.

上下文
待核
输入 / 1M tokens
-
输出 / 1M tokens
-
缓存读 / 1M
-
缓存写 / 1M
-
Creative GenerationMultimodal Understanding

wan2.6-t2v

文本音频视频

Alibaba

Generate 720p/1080p video from text prompts and come with synchronized audio tracks. The picture has movie-level aesthetics and complex motion performance, suitable for creative video creation.

上下文
待核
输入 / 1M tokens
-
输出 / 1M tokens
-
缓存读 / 1M
-
缓存写 / 1M
-
Creative GenerationHigh-Quality GenerationContent GenerationMultimodal Understanding

wan2.6-r2v-flash

文本视频图像视频

Alibaba

Speed-optimized Wan 2.6 reference-to-video variant with optional audio output, replicating reference-video motion into up to 10 second 720p/1080p clips for fast, cost-effective production.

上下文
待核
输入 / 1M tokens
-
输出 / 1M tokens
-
缓存读 / 1M
-
缓存写 / 1M
-
Cost EffectiveReal-time ResponseCreative Generation
需要完整模型清单或企业接入方案?进入控制台