模型广场

选择模型后可复制配置给 AI,也可以快速测试当前连接延迟。

重置筛选
全部 5文本 1图像 0音频 0视频 4

hailuo-02

文本图像视频

minimax

It supports Vincentian and Tusheng videos, can output 6-10 seconds, 512p/768p/1080p multi-level image quality, supports camera movement command control, and flexibly balances image quality and cost.

上下文
待核
输入 / 1M tokens
-
输出 / 1M tokens
-
缓存读 / 1M
-
缓存写 / 1M
-
Cost EffectiveCreative GenerationContent GenerationMultimodal Understanding

hailuo-2.3

文本图像视频

minimax

It supports Vincentian and Tusheng videos, can output 6-10 seconds, 768p/1080p images, supports fine camera movement command control, has smooth movement and stable subject, and is suitable for short video creative content.

上下文
待核
输入 / 1M tokens
-
输出 / 1M tokens
-
缓存读 / 1M
-
缓存写 / 1M
-
Creative GenerationHigh-Quality GenerationContent GenerationMultimodal Understanding

hailuo-2.3-fast

文本图像视频

minimax

Speed and cost optimized variant of MiniMax Hailuo 2.3 for rapid iteration and bulk production, generating 6 to 10 second 768p/1080p clips from text or image input with camera-movement control.

上下文
131K
输入 / 1M tokens
-
输出 / 1M tokens
-
缓存读 / 1M
-
缓存写 / 1M
-
Creative GenerationCost Effective

minimax-h3

文本图像视频音频视频

minimax

A lightweight open-weight video generation model focused on instruction-guided editing, text and brand rendering, and video-to-video motion transfer, with native audiovisual output. Built for advertising, e-commerce, and interface design workflows.

上下文
131K
输入 / 1M tokens
-
输出 / 1M tokens
-
缓存读 / 1M
-
缓存写 / 1M
-
LightweightCreative GenerationContent GenerationMarketing

minimax-m3

文本图像视频文本

minimax

MiniMax\\'s first multimodal release, accepting image and video input. Long-context inference runs cheaper and faster, and it handles multi-step agent workflows and computer-use scenarios.

上下文
1M
输入 / 1M tokens
$1.2
输出 / 1M tokens
$4.8
缓存读 / 1M
$0.12
缓存写 / 1M
$0
VisionLong Text ProcessingMultimodal UnderstandingAgent Coding
需要完整模型清单或企业接入方案?进入控制台