基于开源 MiniMax H3 的增强版视频模型。一个模型名三种模式,按你传的内容自动选择:文生视频、首尾帧图生视频、多参考视频(最多 9 张参考图、3 段参考视频)。每种模式都能用自己的音频驱动口型,audio_mode 决定音轨怎么处理。480P / 768P / 1080P,4-15 秒,按请求的秒数计费:480P $0.05/秒、768P $0.07/秒、1080P $0.10/秒。
| 分辨率 | 每秒 | 5s | 10s | 15s |
|---|---|---|---|---|
| 480P | $0.05 | $0.25 | $0.50 | $0.75 |
| 768P | $0.07 | $0.35 | $0.70 | $1.05 |
| 1080P(默认) | $0.1 | $0.50 | $1.00 | $1.50 |
按请求的秒数收费,4-15 秒。参考图、参考视频、驱动音频和音色参考都不另收费。仅成功扣费、失败退款。
所有请求在 Header 携带 API Key:
Authorization: Bearer YOUR_API_KEY/api/v1/video/generations·GET/api/v1/video/generations?task_id=# Text-to-video (prompt only)
curl -X POST https://api.apimodels.app/v1/video/generations \
-H "Authorization: Bearer YOUR_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "minimax-h3-rh-enhanced",
"prompt": "A potter shapes a clay bowl on a spinning wheel, she looks up and smiles at the camera, warm window light",
"duration": 8,
"resolution": "768p",
"aspect_ratio": "16:9"
}'
# Image-to-video with a driving audio track (lip sync)
curl -X POST https://api.apimodels.app/v1/video/generations \
-H "Authorization: Bearer YOUR_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "minimax-h3-rh-enhanced",
"prompt": "She keeps working and talks to the camera with a warm smile",
"first_frame_url": "https://example.com/first.jpg",
"audio_url": "https://example.com/voice.mp3",
"duration": 6,
"resolution": "1080p",
"aspect_ratio": "3:4"
}'
# Multi-reference music video: keep your soundtrack untouched
curl -X POST https://api.apimodels.app/v1/video/generations \
-H "Authorization: Bearer YOUR_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "minimax-h3-rh-enhanced",
"prompt": "The singer from image 1 performs on the stage from image 2, moving like the dancer in the reference video",
"reference_image_urls": ["https://example.com/singer.jpg", "https://example.com/stage.jpg"],
"reference_video_urls": ["https://example.com/dance.mp4"],
"audio_url": "https://example.com/song.mp3",
"audio_mode": "lock_source",
"duration": 10,
"resolution": "1080p",
"aspect_ratio": "9:16"
}'
# Poll until completed (result files are kept 30 days)
curl "https://api.apimodels.app/v1/video/generations?task_id=TASK_ID" \
-H "Authorization: Bearer YOUR_API_KEY"首尾帧不能和参考图 / 参考视频同时传,这种请求在创建时返回 400。三种模式都收 audio_url(驱动口型的音频)和 audio_mode。
| 字段 | 必填 | 类型 | 说明 |
|---|---|---|---|
| model | 是 | string | minimax-h3-rh-enhanced |
| prompt | 文生必填 | string | 画面、动作、运镜、台词和声音描述;最长 20000 字符。 |
| duration | 是 | number | 4-15 秒,按这个秒数计费。 |
| resolution | 否 | string | 480p($0.05/秒)/ 768p($0.07/秒)/ 1080p($0.10/秒,默认)。 |
| aspect_ratio | 否 | string | 1:1 / 2:3 / 3:2 / 3:4 / 4:3 / 9:16 / 16:9 / 21:9 |
| first_frame_url | 否 | string | 首帧(公网 URL 或 base64)。也接受 image。 |
| last_frame_url | 否 | string | 尾帧。 |
| reference_image_urls | 否 | string[] | 参考图,最多 9 张。也接受 images。 |
| reference_video_urls | 否 | string[] | 参考视频,最多 3 段。也接受 video_list。 |
| reference_video_audio_urls | 否 | string[] | 参考视频的伴音,最多 2 段。 |
| audio_url | 否 | string | 驱动口型的音频。也接受 drive_audio_url。 |
| reference_audio_urls | 否 | string[] | 音色参考,最多 3 段(图生和多参考)。也接受 audio_list。 |
| audio_mode | 否 | string | native(默认)/ lock_source(完整保留驱动音频,适合 MV)/ remix_source / reference_only。 |
| callback_url | 否 | string | 任务完成时我们向该地址 POST 结果;不传就轮询 GET 查询。 |
同一家族的其它型号:minimax-h3(原生 2K)、minimax-h3-lite、minimax-h3-max-turbo(秒级出片)。
试一试:Playground
apimodels.app 按请求的时长按秒计费:480P $0.05/秒、768P $0.07/秒、1080P $0.10/秒,4 到 15 秒 —— 768P 10 秒 $0.70,1080P 15 秒 $1.50。参考图、参考视频、音频都不另收费。只有生成成功才扣费,失败退款。
POST /api/v1/video/generations,model 填 minimax-h3-rh-enhanced,带 duration(4-15 秒)和 resolution(480p / 768p / 1080p,默认 1080p)。只传 prompt 是文生视频;加 first_frame_url / last_frame_url 是图生视频;加 reference_image_urls / reference_video_urls 是多参考。audio_url 是驱动口型的音频,audio_mode 控制音轨。返回 taskId,轮询 GET /api/v1/video/generations?task_id= 或传 callback_url。
按输入自动选。带参考图、参考视频或参考视频伴音就是多参考;否则带首帧或尾帧(或只带音色参考)就是图生视频;都没有就是文生视频,此时 prompt 必填。首尾帧不能和参考图 / 参考视频同时传 —— 这种请求会返回明确的 400。
决定音轨怎么处理:native(默认)由模型生成音频;lock_source 完整保留你的驱动音频,MV 就用这个;remix_source 和 reference_only 对你的音频处理得更宽松。驱动音频用 audio_url 传。