> For the complete documentation index, see [llms.txt](https://docs.acestudio.ai/llms.txt). Markdown versions of documentation pages are available by appending `.md` to page URLs; this page is available as [Markdown](https://docs.acestudio.ai/docs/product-wiki-zh/ai-gong-ju/music-video.md).

# 音乐视频

借助 AI Director agent，将一首歌曲制作成完整的音乐视频——每一个关键创意决策都由你来掌控。

## 什么是音乐视频

音乐视频会把一首歌变成一支完成并经过完整剪辑的音乐视频。你将与一个 AI 智能体协作—— **导演** 它负责脚本撰写、选角、场景设计、编舞、分镜、视频生成和剪辑。你始终掌握创意决策的主导权。

有两种制作方式：

* **高度定制** ——为整首歌制作的完整音乐视频，采用逐步细化的制作流程。需要 25k+ 积分。
* **宣传片段** ——由歌曲某一片段制作的短视频，专为 Instagram Reels、TikTok、YouTube Shorts 及其他社交发布场景打造。需要 500+ 积分。

{% embed url="<https://www.youtube.com/watch?v=amx7gYLPW10>" %}

## 你的导演

导演是一个单一智能体，负责从头到尾推进整个制作流程，而且它只对你负责。制作的每个相位都由一个专门的 **技能** ——一套专业操作手册，导演会在该相位开始时调用，就像电影导演在不同部门简报之间切换一样。你有时会在聊天中看到导演正在阅读或使用某个技能；仅此而已——技能会自行运作，你完全不需要管理它们。

一次制作中的表现如下：

| 相位        | 导演所做的事                  | 你看到的内容           |
| --------- | ----------------------- | ---------------- |
| **创意规划**  | 撰写创意方案和每个镜头的脚本          | 创意方案、镜头脚本        |
| **选角**    | 设计角色的基础造型和每个场景的造型       | 角色设定图（多视图）、每场景造型 |
| **美术与分镜** | 定义视觉风格、设计场景环境、绘制分镜      | 风格指南、场景参考、分镜     |
| **编舞**    | 设计舞蹈与表演动作，创建动作关键帧网格     | 编舞文档、动作关键帧       |
| **摄影**    | 撰写生成提示词，执行视频生成，处理口型同步优化 | 最终视频镜头           |

## 工作原理

1. 从你的 ACE Studio 工程开始——可以是一首你制作的歌曲，也可以是你导入的成品轨道。点击 <kbd>音乐视频</kbd> 右上角并启动会话；导演会按歌曲对你的工程进行渲染并分析它。
2. 只需三步即可设置视频：选择 **高度定制** 或 **宣传片段**，选择宽高比（16:9、9:16、3:4 或 1:1），然后描述你的想法——你还可以上传自己的角色图片。
3. 导演会继续询问剩余的引导问题，然后按流程推进——选角、场景设计、创意方案、镜头脚本、造型、可选编舞以及分镜。你会在每个里程碑进行审阅并提供反馈。
4. 分镜会被放置在轨道上 [画布](/docs/product-wiki-zh/gong-cheng/arrangement-in-canvas/canvas.md) 作为 animatic，这样你就能在任何内容渲染之前预览整支视频的节奏。
5. 一切准备就绪后，由你发出开始指令。镜头会并行渲染，自动替换分镜，而演唱镜头还会进行一轮口型同步优化。
6. 查看结果，然后提出修改——只会重新生成受影响的镜头。

完整流程请参见 [创建音乐视频](/docs/product-wiki-zh/ai-gong-ju/music-video/creating-a-music-video.md)。短格式流程请参见 [创建宣传片段](/docs/product-wiki-zh/ai-gong-ju/music-video/creating-a-promo-clip.md)。关于修订、手动编辑以及恢复工程，请参见 [编辑与迭代](/docs/product-wiki-zh/ai-gong-ju/music-video/editing-and-iteration.md).

## 选择视频模型

你选择的视频模型不仅影响画质——不同模型支持不同的 **单镜头最长时长**，这会改变导演为你的镜头编写脚本的方式。如果你已经知道想要的风格，建议尽早选定模型；如果还不确定，默认设置对大多数情况都很适用。

| 模型                  | 适用场景       | 单镜头最长时长                         |
| ------------------- | ---------- | ------------------------------- |
| **MiniMax-H3** （默认） | 动画、动漫、插画风格 | 15 秒                            |
| **Seedance 2.0**    | 写实、实拍风格    | 15 秒                            |
| **Seedance 2.5**    | 超长连续镜头     | 超过 15 秒——积分消耗是 Seedance 2.0 的两倍 |

你可以在聊天输入框下方的选择器中切换模型，或者直接告诉导演要使用哪个模型进行生成。

## 积分定价

音乐视频会在音乐分析、图像生成、视频生成以及智能体 LLM 使用中消耗积分。

<table data-search="false"><thead><tr><th width="144.2421875">类别</th><th width="148.69921875">模型或功能</th><th width="246.6171875">计费单位</th><th width="180.98046875">积分消耗</th></tr></thead><tbody><tr><td>音乐分析</td><td>音乐分析</td><td>按次请求</td><td>约每分钟 20</td></tr><tr><td>图像模型</td><td>Midjourney</td><td>按次请求</td><td>20</td></tr><tr><td>图像模型</td><td>Nanobanana</td><td>按次请求</td><td>25</td></tr><tr><td>图像模型</td><td>GPT Image 2</td><td>按次请求</td><td>15</td></tr><tr><td>图像模型</td><td>Qwen Image 多角度</td><td>按次请求</td><td>7</td></tr><tr><td>视频模型</td><td>MiniMax-H3</td><td>按秒</td><td>22</td></tr><tr><td>视频模型</td><td>Seedance 2.0（1080p）</td><td>按秒</td><td>68</td></tr><tr><td>视频模型</td><td>Seedance 2.0（720p）</td><td>按秒</td><td>34</td></tr><tr><td>视频模型</td><td>Seedance 2.5（1080p）</td><td>按秒</td><td>136</td></tr><tr><td>视频模型</td><td>Seedance 2.5（720p）</td><td>按秒</td><td>68</td></tr><tr><td>视频模型</td><td>Lipsync 3</td><td>按秒</td><td>25</td></tr><tr><td>智能体</td><td>智能体 LLM 令牌</td><td>根据实际 token 用量</td><td>因任务而异</td></tr></tbody></table>


---

# Agent Instructions
This documentation is published with GitBook. GitBook is the documentation platform designed so that both humans and AI agents can read, navigate, and reason over technical content effectively. Learn more at gitbook.com.

## Querying This Documentation
If you need additional information that is not directly available in this page, you can query the documentation dynamically by asking a question.

Perform an HTTP GET request on the current page URL with the `ask` query parameter, and the optional `goal` query parameter:

```
GET https://docs.acestudio.ai/docs/product-wiki-zh/ai-gong-ju/music-video.md?ask=<question>&goal=<endgoal>
```

`ask` is the immediate question: it should be specific, self-contained, and written in natural language.
`goal` is optional and describes the broader end goal you are ultimately trying to accomplish on behalf of the user. GitBook uses it to tailor the answer towards what is most useful for that goal.

The response will contain a direct answer to the question and relevant excerpts and sources from the documentation.

Use this mechanism when the answer is not explicitly present in the current page, you need clarification or additional context, or you want to retrieve related documentation sections.
