在 Pollo API 上运行 MiniMax H3 Max
通过动作、镜头运动和声音来引导一个紧凑的场景。MiniMax H3 Max 可将文本或起始图像转化为一个协调一致的视频时刻。
对于正在构建视频创意工具的开发者来说,Pollo API 上的 MiniMax H3 Max 提供了一种实用方式,可用于测试动态简报并向用户返回短视频片段。
MiniMax H3 Max API 的主要功能
提示词引导的操作
模型遵循连接角色行为和视觉风格的场景指令。给主体一个具体动作,而不是依赖宽泛的氛围描述。
受控镜头移动
镜头方向会影响动作如何被呈现。请明确视角和运动方式,让构图服务于主体,而不是与之争夺注意力。
同步声音
音频与生成的视觉内容同步,允许一句台词或一个可听见的事件为场景增色,而不是让它继续作为剪辑师的独立后期任务。
基于图像的场景开发
起始图像提供初始主体和构图。描述随后发生的运动,将静态画面转变为一个有明确终点的动作。
MiniMax H3 Max API 的使用场景
- Action Previsualization: Test whether a single gesture, reveal, or interaction reads clearly before expanding it into a full scene.
- Camera Move Studies: Compare a fixed view with a tracking shot for the same action and choose the more readable result.
- Sound-Driven Reaction Clips: Build a short moment around a bell, snap, knock, or spoken cue followed by a visible response.
- Still-to-Motion Concepts: Animate an approved illustration or photograph around one controlled movement.
- Physical Comedy Beats: Explore a setup and payoff that can be understood without a lengthy explanation.
- Micro-Demo Storyboards: Visualize a simple interaction, such as opening a folding object, while checking the clarity of each movement.
如何使用 MiniMax H3 Max API
- Get an API Key: Create a Pollo API account and generate your API key from the developer dashboard..
- Choose the Model: Send requests to the MiniMax H3 Max endpoint at
/generation/minimax/minimax-h3-max/video. - Add Your Inputs: Write the action and camera prompt, adding an image for image-guided generation. Select
duration,resolution, andaspectRatiousing the current schema. - Generate and Retrieve: Submit the task, poll the returned task status, and download the finished video once it succeeds.
MiniMax H3 Max API 的提示词最佳实践
在添加氛围之前,先让动作容易想象。让镜头只做一件事,并描述应与关键时刻同步的声音。
简单的提示词公式
Starting composition + subject + single action + camera behavior + visible endpoint + sound cue + visual treatment
你应该注意什么
- Action Verbs: Prefer concrete directions such as folds, catches, turns, or reaches over instructions like “make it dynamic.”
- Camera Purpose: Explain whether the camera follows a subject, reveals an object, or stays still to make movement easier to judge.
- Image Compatibility: Choose motion that follows naturally from the posture and framing in the starting image.
- Audible Timing: Tie the important sound to an observable action so you can review synchronization directly.
- Single-Variable Tests: Change camera direction or action pacing separately to identify which revision improves the clip.
[`示例提示词`]
折叠凳
一位露营者在宁静的湖边展开一把折叠木凳。将镜头保持在腰部高度,并在凳腿展开时轻微向右跟拍。等到凳子稳稳地落在草地上时结束。呈现清晰、易读的动作,并带有轻柔的木质咔嗒声和远处的鸟鸣。
图书管理员的铃声
一位图书管理员轻轻敲响桌铃,然后抬头看向一只睡着的猫,那只猫睁开了一只眼睛。固定中景镜头,铃铛在前景,猫在一摞书旁边。午后的窗光。先是一声清脆的铃响,随后是一声轻柔、毫不在意的喵叫。
丝带逃脱
为所提供的一个孩子拿着风筝的插图制作动画。一个松散的丝带从孩子的手腕滑落,绕着栅栏柱盘旋。孩子抓住它的末端,咧嘴笑了。用一次平滑的侧向镜头移动跟随这条丝带。保留插图的纸张纹理和柔和色彩。
MiniMax H3 Max vs Wan 3.0 Prime vs Wan 3.0
| 评估区域 | MiniMax H3 Max | Wan 3.0 Prime | Wan 3.0 |
|---|---|---|---|
| Pollo 生成工作流 | 文本和图像 | 文本、图像和参考 | 文本、图像和参考 |
| Documented focus | 短场景、提示词方向、镜头控制 | 多镜头场景与连贯性 | 更长的序列和多模态指导 |
| 生成的声音 | 支持 | 支持 | 支持 |
| 建议的起始测试 | 一个动作和一个镜头移动 | 一个角色跨多个视图 | 使用不同参考角色的一个叙事 |
为什么选择 MiniMax H3 Max API?
MiniMax H3 Max 适合那些可以通过清晰的动作、刻意的构图和可听见的提示来判断的简报,这让围绕单个镜头进行迭代变得很直接。
Pollo API 让视频工具能够通过一个平台提供 H3 Max 以及其他替代模型,并支持共享账户访问和面向任务的生成工作流文档。
在 Pollo API 上使用 MiniMax H3 Max 进行动作研究,然后优化镜头方向,直到动作呈现出你想要的效果。
MiniMax H3 Max API 常见问题
MiniMax H3 Max 是什么?
MiniMax H3 Max 是 MiniMax H3 的后训练变体。MiniMax 将其描述为针对高速文本生视频和图像生视频生成进行了优化。
MiniMax H3 Max 是否接受起始图像?
是。使用清晰的起始构图,并描述接下来应执行的动作,而不是让模型重新设计每个可见元素。
我可以在 MiniMax H3 Max 中直接控制摄像头吗?
是。请在提示词中包含镜头行为。先从一个明确的运动开始,并说明它揭示了什么,然后检查构图是否让重要动作保持可见。
MiniMax H3 Max 会生成音频吗?
是的。请在场景简报中包含所需的声音,并检查它与可见事件的契合程度。
为什么要通过 Pollo API 运行 MiniMax H3 Max?
Pollo API 提供统一集成,可访问 MiniMax H3 Max 以及 300 多个其他模型,并配备文档、任务跟踪和 webhooks,同时提供比类似提供商更具成本效益的生成。