SunoMV SunoMV
Veo 3 提示词:10 条技巧 + 可复制模板(2026)
教程指南

Veo 3 提示词:10 条技巧 + 可复制模板(2026)

发布于 · 作者: SunoMV 团队
把 SunoMV 设为 Google 优先来源 在热门报道和 AI 概览里看到更多 SunoMV。

你在输入框里敲 cinematic, 8K, masterpiece,点生成。八秒后出来的片子像一张学会呼吸的壁纸:副歌没有运镜,嘴在动,房间里却没有任何声音。你不是「不够有创意」。你给模型的是情绪板,却要它去导一场戏。

veo 3 promptprompt for veo 3 的人已经选定模型。缺的是能粘贴的字符串:镜头在前,然后是人物、动作、房间、光——如果这段允许出声,再单独写一句音频。

这不是一堂课。下面每条都是完整可复制块。反例是同一意图写成空形容词。复制,到 SunoMV 音频转视频 的模型列表里选 Veo 3.1 Fast 或 Veo 3.1 Lite,把这段贴进镜头描述,导出能接上副歌的一镜。

目录

为什么「cinematic, 8K」不算 Veo 3 prompt

Google 自己的 Veo 3.1 提示词指南 把摄影放在第一槽,是有意的:运镜是「传达语气和情绪最有力的工具」。你不写镜头,模型就发明一个默认全景再加一个默认摇镜,剩下的时间用来看起来很贵。

同一份指南还锁死两个该写成字段、不该写成氛围的物理事实。片段时长是 4、6 或 8 秒。画幅是 16:9 或 9:16。在 SunoMV 里,Veo 镜头锁在 8 秒——你描写一段 20 秒的跟踪,得到的是被砍断的镜头,不是被导过的镜头。

我宁愿每个槽写一句无聊但完整的话,也不写一段「史诗能量」。空形容词换来的是好看、却切不进鼓点的废片。

夜城 MV 静帧:雨和霓虹里的歌手,用于 Veo 3 prompt

图:SunoMV 团队 · 用镜头 + 人物 + 雨写成的 MV 静帧,不是「cinematic 8K」

实用规则: 第一句如果不是镜头,你就是在给模型已经选好的那一镜做装修。

Google 真正在用的五个槽位

Google Cloud 的公式是五段,按这个顺序:摄影 + 主体 + 动作 + 环境 + 风格与氛围。独立整理如 Prompt Architects 的 Veo 3 prompt 结构 也按 Google Veo 3.1 指南复述了同一顺序。Gemini API 的 Veo 文档(2026 年 7 月 30 日更新)再在上面加一层音频:对话用引号,音效写成声音,环境声单独一句。

把表填满。空格子会被填成最无聊的默认值。

槽位填这个反例
摄影景别 + 机位 + 有名字的运镜 + 速度「cinematic camera」
主体一个人、服装、年龄段、头发「一个漂亮女孩」
动作动词 + 物理(重量、速度、接触)「优美地跳舞」
环境房间、天气、几点、前景杂物「史诗场景」
风格与氛围光线方向 + 质感 + 调色「8K, masterpiece, highly detailed」
音频(可选)对白 / 音效 / 环境声,分句写「加一段戏剧性的音乐」

DeepMind 的 Veo 页 说得很直:Veo 3 卖的就是原生音效、环境声和对白,跟画面一起出。你不写音频句,照样会有一条声轨——只是不是副歌需要的那条。

SunoMV 选择器里能点到的是 Veo 3.1 FastVeo 3.1 Lite。Lite 是无声、更快的那档。prompt for Veo 3 里如果有台词或点名的音效,选 Fast。无声档会忽略音频块。

实用规则: 五个画面槽按 Google 的顺序写,音频另起句子。别把「她低声说」塞进服装那一行。

10 条可复制技巧

每条都是完整块。反例是同一意图写成空形容词。在 SunoMV:选 Veo 3.1 Fast 或 Lite,整段贴进镜头描述,生成,再把这 8 秒接到歌词行上。

1. 第一句写镜头

一句话规则:Google 把摄影放第一,因为模型听得懂有名字的运镜,听不懂情绪词。Runway 的镜头提示笔记 也是同一套手艺:「slow dolly push-in」比「cinematic」管用。

Medium close-up, low angle, slow dolly push-in over 8 seconds,
on a singer in a soaked black leather jacket standing in neon rain.
35mm spherical look, shallow depth of field, magenta rim light from the left.

反例:Cinematic music video of a singer, highly detailed, 8K.

在 SunoMV 里怎么用:把这段贴在副歌下拍的第一段转场。推进是剪辑点,不是外套。

2. 用服装部的方式点名主体

一句话规则:「一个女孩」是试镜通知。服装才是锁。

Subject: a woman in her late 20s, blunt black bob, silver hoop earrings,
oversized oxblood leather jacket over a white tank, chipped black nail polish.
Keep the same jacket and hair in every shot.

反例:A beautiful mysterious girl with good vibes.

在 SunoMV 里怎么用:同一首歌的每一段 Veo 镜头都贴这段主体,避免副歌和主歌把人换掉。

3. 动作用物理写,不用夸赞写

一句话规则:「优美地跳舞」没有重量。「鞋跟砸进湿柏油」有。

Action: she takes three slow steps toward camera, left heel striking wet asphalt,
jacket hem swinging, rainwater flicking off the collar.
She does not spin. She does not jump. Weight stays in the front foot.

反例:She dances energetically with amazing choreography.

在 SunoMV 里怎么用:用在预副歌的走近。爆发翻转是 Veo 最容易散的地方;身体放慢,把运动交给镜头。

4. 让房间干活

一句话规则:环境不是「location: cinematic」。是镜头打得到的东西。

Context: a narrow Tokyo side street at 1 a.m., vending machines buzzing on the right,
a puddle in the foreground reflecting pink kanji, steam from a ramen grate,
one passing taxi light in the deep background. No crowd.

反例:Epic futuristic city, breathtaking scenery.

在 SunoMV 里怎么用:一首歌一个具体的街。重复用。水洼是免费的转场道具。

5. 用方向和质感替换「电影感打光」

一句话规则:官方示例里,光线方向是最便宜的画质杠杆。

Style and ambiance: hard key from camera-left neon, magenta, cutting a
narrow rim on her jaw. Soft fill from a shop window on the right.
Low-key contrast, wet highlights, slight grain, no clean beauty lighting.

反例:Cinematic lighting, volumetric god rays, ultra realistic.

在 SunoMV 里怎么用:每段镜头沿用同一对主光/补光,避免每 8 秒调色重置。

轨道镜头推向小舞台上的歌手

图:SunoMV 团队 · 第一句是运镜,不是形容词 cinematic

6. 对白用引号,单独一行

一句话规则:Gemini API 的音频说明 要你用引号标台词。没有引号,它就是旁白,不是一句词。

She looks just past camera and says, "Don't wait for the chorus."
Keep her mouth in frame. No other speakers.

反例:She talks emotionally about the song.

在 SunoMV 里怎么用:只跟 Veo 3.1 Fast 一起用。选了 Lite 就把这块删掉——无声片会对口型,你会讨厌那个结果。

7. 音效写成声音,不写成感觉

一句话规则:「戏剧性冲击」是情绪。「鞋跟点在湿砖上」是线索。

SFX: a sharp heel click on wet brick; a distant train door chime;
rain ticking on a metal awning. No explosion. No whoosh transition.

反例:Add epic sound design.

在 SunoMV 里怎么用:写能落在军鼓上的音效。歌已经很满时,音效少写,别跟副歌打架。

8. 给房间一层环境声

一句话规则:Google 把环境声列为第三种音频。不写,模型会发明一段通用铺底。

Ambient noise: low traffic two streets over, a vending machine hum,
rain on plastic sheeting. No score. No vocal ad-libs from off-screen.

反例:Background music that fits the mood.

在 SunoMV 里怎么用:歌词视频你已经有歌了。明确写 no score,避免原生铺底叠上副歌。

9. 把 16:9 和 9:16 写成槽,不当成裁切

一句话规则:Google 的 Veo 3.1 指南只给 16:9 或 9:16。1:1 的 prompt 等于你事后会裁坏的一张图。

Aspect ratio: 9:16 vertical. Headroom for captions. Subject stays in the
center third. No wide establishing shot. Phone-first framing.

反例:Make it work for every platform.

在 SunoMV 里怎么用:一次导出一个画幅。别生成 16:9 再指望竖裁还留得住脸。

竖屏 9:16 MV 画面:霓虹雨里的歌手

图:SunoMV 团队 · 9:16 是写进去的槽,不是宽屏的裁切

10. 把 8 秒锁写进 prompt,别跟它作对

一句话规则:SunoMV 里 Veo 镜头固定 8 秒。你要一分钟一镜到底,不会得到一分钟。

Duration: one 8-second shot. Start on her hands at the jacket zipper,
end on her eyes as she looks up into the rain.
No cut inside the clip. No time-lapse. No "then later that night."

反例:A full music video that follows her all night across the city.

在 SunoMV 里怎么用:一条 prompt = 一行歌词或一个下拍。把 8 秒镜头串起来。别让一段 Veo 去当整支 MV。

实用规则: prompt 里出现「then」,你就是在写一场戏。拆开。Veo 不会替你把一段话分镜。

6 个 MV 模板

每个模板整段可粘。括号里的换掉。槽位顺序别动。

流行副歌钩子(16:9)

Cinematography: medium shot, eye-level, slow push-in, 35mm.
Subject: a woman in her 20s, red vinyl jacket, wet hair stuck to her cheek.
Action: she mouths the chorus and steps one pace toward camera, rain on her lashes.
Context: rooftop at night, city grid behind her, one red aviation light blinking.
Style and ambiance: hard backlight, thin fog, teal-and-red grade, light grain.
Audio: she sings, "Stay until the lights go out." Ambient: wind on a metal rail.
No score. 16:9. 8 seconds.

歌词特写(嘴 + 字幕空间)

Cinematography: close-up, slight low angle, locked-off with a 2cm drift.
Subject: the same woman, oxblood jacket, silver hoops, a small scar on the left eyebrow.
Action: she inhales, then delivers one line, eyes wet but not crying.
Context: dark studio, a vintage microphone in the lower third, out-of-focus meters behind.
Style and ambiance: warm key from the right, cool rim, shallow focus on the mouth.
Audio: she says, "I kept the chorus." Ambient: room tone, no score.
Leave headroom at the top for captions. 8 seconds.

复古麦克风前的歌词特写,原生音频 Veo 3 prompt

图:SunoMV 团队 · 歌词特写是一张嘴、一支麦、一句带引号的词

电影感跟踪

Cinematography: tracking shot from the left, hip height, 50mm, modest speed.
Subject: a man in his 30s, grey overcoat, scuffed boots, a silver ring on the right hand.
Action: he walks parallel to a brick wall, left hand brushing wet ivy.
Context: alley after rain, one sodium street lamp, a cat frozen in a doorway.
Style and ambiance: orange sodium vs cold brick, puddle reflections, no beauty fill.
SFX: boot on wet stone; a distant bus hiss. Ambient: city two streets over. No score.
16:9. 8 seconds. Do not reveal the end of the alley.

竖屏短视频(9:16)

Cinematography: medium close-up, 9:16, slow crane-up from chest to eyes.
Subject: the same woman in the red vinyl jacket, rain in her hair.
Action: she looks up, blinks once, then smiles on the last beat.
Context: neon stairwell, one pink tube light, condensation on the rail.
Style and ambiance: magenta key, green bounce from a sign, crushed blacks.
Audio: she says, "Hit post." SFX: a single heel on metal stairs. No score.
Keep her in the center third. 8 seconds.

角色锁定(同一人,第二个地点)

Cinematography: wide-to-medium, slow dolly right, 35mm.
Subject: SAME woman as previous shots — blunt black bob, oxblood leather jacket,
white tank, silver hoops, chipped black nail polish. Do not recast. Do not change hair.
Action: she leans on a convenience-store window and watches her own reflection.
Context: fluorescent store interior behind glass, night street in the reflection.
Style and ambiance: sickly green fluorescents vs magenta street neon, wet glass.
Ambient: compressor hum, distant register beep. No score. 16:9. 8 seconds.

原生音频对白(只用于 Fast)

Cinematography: two-shot, eye-level, slow pan from him to her, 40mm.
Subject A: the woman in the oxblood jacket. Subject B: a man in a grey overcoat.
Action: he turns, she does not. She answers without looking at him.
Context: rain under a bus shelter, one fluorescent tube buzzing.
Style and ambiance: hard overhead fluorescent, greenish, wet pavement sheen.
He says, "You coming?" She says, "After the chorus."
SFX: rain on the shelter roof; a bus air-brake. Ambient: traffic. No score.
16:9. 8 seconds. Pick Veo 3.1 Fast. Do not use the silent option.

雨夜公交站亭里的双人镜头,原生音频 Veo 3 prompt

图:SunoMV 团队 · 带引号的台词只有嘴还在画面里才有用

空 prompt 对上一段 prompt for Veo 3

同一首歌,同样 8 秒,两种输入。只有一种叫 prompt for Veo 3。

空的可复制的
第一句「cinematic, 8K」有名字的景别 + 有名字的运镜
主体「一个女孩」服装 + 头发 + 一个辨识点
动作「跳舞」三步、重心、和地面的接触
音频(没有,或「加音乐」)带引号的台词 + 点名的音效 + 「no score」
画幅 / 时长「做成能爆的」16:9 或 9:16,8 秒,起幅和落幅
给谁用假装成视频的静帧真能剪进副歌的一镜

如果你已经有 Suno 的歌,别再让 Veo 写另一首。把它对准画面,把歌留下。Suno 提示词技巧指南 是音频半边的可复制清单;本页是画面半边。更长的 Seedance 分镜写法见 Seedance 2.5 提示词指南——槽位不同,工作是同一件:复制整块。

决策过滤器: 如果你没法在 prompt 里指出起幅和落幅,你就没有镜头。你只有氛围。

常见问题

prompt 里该写 Veo 3 还是 Veo 3.1?

都别写。模型名不属于 prompt 正文。人搜的是 veo 3 prompt;SunoMV 选择器当前列出的是 Veo 3.1 FastVeo 3.1 Lite。五个槽位是同一套。第一句写镜头,别写版本号。

为什么时长被抬到 8 秒?

这是 SunoMV 里 Veo 的锁定,不是你措辞的 bug。把起幅落幅写成 8 秒能装下的。别描写一分钟的故事再指望模型帮你压缩。

有首尾帧还要写 prompt 吗?

要。帧管住身份。prompt 还是得点名镜头、动作和音频。首尾帧中间夹一句 cinematic 8K,仍然是空 prompt。

写了台词为什么没有对白?

两个常见原因。你选了无声的 Lite。或者台词没加引号,被当成描写。加上引号,选 Fast,嘴留在画面里。

要写负向 prompt 吗?

写正着的那一镜。Google 的示例把字花在画面里有什么。必须排除的(配乐、人群、第二张脸)用收尾约束:No score. No crowd.——不是一整段「不要」。

一条 prompt 能当整支 MV 吗?

不能。一条 prompt 是一个 8 秒镜头。AI 音乐视频创作指南 才是把这些镜头串起来的流程。本页只负责能粘贴的 Veo 3 prompt。

复制、选 Veo、导出 MV

一段 Veo 3 prompt 写完的标志是能粘,不是能讲。

  1. 复制上面一个模板。填括号。保持 Google 的五槽顺序。
  2. 打开 SunoMV 音频转视频。丢进歌曲(Suno 链接或自己的文件)。
  3. 模型列表里,有对白或音效就选 Veo 3.1 Fast。只要一张无声的 8 秒画面就选 Veo 3.1 Lite
  4. 整段贴进这一镜。别把镜头句和主体句拆开。
  5. 导出,接到歌词行上。再来一条。MV 是一串 8 秒镜头,不是一段英雄散文。

能用的 Veo 片子,不是形容词表更好的人做出来的。是点生成之前,手里已经有一段能粘贴的 prompt for Veo 3 的人。

SunoMV 团队

查看「Suno 提示词与 AI 写歌」全部 26 篇 →

试试这些 AI 工具