SunoMV SunoMV
Veo 3 Prompt 教學:10 招 + 可複製範本(2026)
教學指南

Veo 3 Prompt 教學:10 招 + 可複製範本(2026)

發布於 · 作者: SunoMV 團隊
將 SunoMV 設為 Google 優先來源 在熱門報導和 AI 總覽裡看到更多 SunoMV。

你在框裡打 cinematic, 8K, masterpiece,按下產生。八秒後,畫面像一張學會呼吸的圖庫桌布。副歌還是沒有運鏡。歌手的嘴在動,房間裡卻沒有任何聲音。你不是「不夠有創意」。你丟的是 mood board,卻要模型去導一顆鏡頭。

veo 3 promptprompt for veo 3 的人,模型早就選好了。要的是一串能直接貼上的字:鏡頭先寫,再寫主體、動作、空間、光線——這段可以出聲的話,再另寫一句音訊。

這不是方法論課。下面每一招都是完整區塊。反例是同一意圖、卻只剩空形容詞。把區塊貼上,在 SunoMV 音訊轉影片產生器 選 Veo 3.1 Fast 或 Veo 3.1 Lite,匯出一顆能卡上副歌的鏡頭。

目錄

為什麼「cinematic, 8K」不算 Veo 3 prompt

Google 自己的 Veo 3.1 提示詞指南 故意把運鏡放在第一欄:攝影手法是「傳達語氣與情緒最有力的工具」。你不寫,模型就會自製一顆預設遠景、一顆預設搖鏡,然後把剩下的時間拿來看起來很貴。

同一份指南鎖死兩件該當欄位寫、不該當 vibe 寫的物理事實。片段是 4、6 或 8 秒。畫面比例是 16:9 或 9:16。在 SunoMV,Veo 鏡頭鎖在 8 秒——所以一段描述 20 秒跟拍的 prompt,會被砍斷,不會被導成戲。

我寧可每一欄寫一句無聊但寫完的句子,也不要一段「epic energy」。空形容詞換來的,是一顆漂亮、卻卡不上重拍的鏡頭。

夜城 MV 定格,寫給 Veo 3 prompt:雨裡的歌手與霓虹

圖片:SunoMV 團隊 · MV 定格寫成鏡頭 + 主體 + 雨,不是「cinematic 8K」

實用規則: 第一個子句如果不是鏡頭,你只是在幫模型已經選好的那一鏡做裝飾。

Google 真正在用的官方五欄

Google Cloud 的公式是五段,順序固定:cinematography + subject + action + context + style and ambiance。獨立整理如 Prompt Architects 的 Veo 3 prompt 結構 也按 Google Veo 3.1 指南複述了同一順序。Gemini API 的 Veo 頁(2026 年 7 月 30 日更新)再疊一層音訊:對白放引號、音效寫成聲音名稱、環境音獨立成句。

把表填滿。任何空格都會變成最無聊的預設。

欄位填這個反例
運鏡景別 + 角度 + 具名運鏡 + 速度「cinematic camera」
主體一個人、服裝、年齡帶、頭髮「a beautiful girl」
動作動詞 + 物理(重量、速度、接觸)「dancing beautifully」
場景空間、天氣、時段、前景雜物「epic location」
風格與氛圍光線方向 + 質感 + 調色「8K, masterpiece, highly detailed」
音訊(選填)對白/SFX/環境音,分句寫「add dramatic music」

DeepMind 的 Veo 頁 講得很直:Veo 3 真正賣的是原生音效、環境音和對白,跟畫面一起產生。你不寫音訊那句,還是會有聲軌——只是不是副歌要的那一軌。

SunoMV 的選單會顯示 Veo 3.1 FastVeo 3.1 Lite。Lite 是無聲、比較快的選項。Veo 3 的 prompt 若含台詞或點名 SFX,選 Fast。無聲選項會直接無視音訊區塊。

實用規則: 畫面五欄照 Google 的順序寫,音訊另起句子。別把「她低聲說」塞進服裝那一行。

10 招可複製技巧

每一招都是完整區塊。反例是同一意圖、卻只剩空形容詞。在 SunoMV:選 Veo 3.1 Fast 或 Lite,把整塊貼進鏡頭描述,產生,再把 8 秒鏡頭剪到歌詞那一行。

1. 把鏡頭寫進第一個子句

一句規則:Google 把運鏡放第一,因為模型聽得懂具名運鏡,聽不懂情緒詞。Runway 的鏡頭提示筆記 也是同一套手藝:「slow dolly push-in」比「cinematic」管用。

Medium close-up, low angle, slow dolly push-in over 8 seconds,
on a singer in a soaked black leather jacket standing in neon rain.
35mm spherical look, shallow depth of field, magenta rim light from the left.

反例:Cinematic music video of a singer, highly detailed, 8K.

在 SunoMV 怎麼用:貼成副歌重拍上的第一個轉場。推近才是剪輯點,不是那件外套。

2. 用服裝組的方式點名一個主體

一句規則:「a girl」是海選。服裝才是鎖定。

Subject: a woman in her late 20s, blunt black bob, silver hoop earrings,
oversized oxblood leather jacket over a white tank, chipped black nail polish.
Keep the same jacket and hair in every shot.

反例:A beautiful mysterious girl with good vibes.

在 SunoMV 怎麼用:同一首歌的每個 Veo 鏡頭都貼這段主體,副歌和主歌才不會換角。

3. 動作寫物理,不要寫誇讚

一句規則:「dancing beautifully」沒有重量。「鞋跟打在濕柏油上」才有。

Action: she takes three slow steps toward camera, left heel striking wet asphalt,
jacket hem swinging, rainwater flicking off the collar.
She does not spin. She does not jump. Weight stays in the front foot.

反例:She dances energetically with amazing choreography.

在 SunoMV 怎麼用:用在前副歌走近。高爆發空翻是 Veo 鏡頭翻車的地方;身體放慢,讓鏡頭動。

4. 讓空間自己做事

一句規則:場景不是「location: cinematic」。是鏡頭打得到的東西。

Context: a narrow Tokyo side street at 1 a.m., vending machines buzzing on the right,
a puddle in the foreground reflecting pink kanji, steam from a ramen grate,
one passing taxi light in the deep background. No crowd.

反例:Epic futuristic city, breathtaking scenery.

在 SunoMV 怎麼用:一首歌一條具體的街。重複用。水窪是免費轉場道具。

5. 用方向和質感取代「cinematic lighting」

一句規則:光線方向是官方範例裡最便宜的品質槓桿。

Style and ambiance: hard key from camera-left neon, magenta, cutting a
narrow rim on her jaw. Soft fill from a shop window on the right.
Low-key contrast, wet highlights, slight grain, no clean beauty lighting.

反例:Cinematic lighting, volumetric god rays, ultra realistic.

在 SunoMV 怎麼用:跨鏡頭沿用同一組主光/補光,調色才不會每 8 秒重開機。

小舞台上,dolly 鏡頭推向歌手

圖片:SunoMV 團隊 · 第一個子句是運鏡,不是形容詞「cinematic」

6. 對白放引號,獨立一行

一句規則:Gemini API 的音訊說明 要你用引號標台詞。沒加引號,就是旁白,不是台詞。

She looks just past camera and says, "Don't wait for the chorus."
Keep her mouth in frame. No other speakers.

反例:She talks emotionally about the song.

在 SunoMV 怎麼用:只配 Veo 3.1 Fast。若選了 Lite,刪掉這塊——無聲鏡頭會對嘴,結果你會恨。

7. 把 SFX 寫成聲音,不要寫成情緒

一句規則:「dramatic impact」是情緒。「濕磚上的鞋跟叩擊」是 cue。

SFX: a sharp heel click on wet brick; a distant train door chime;
rain ticking on a metal awning. No explosion. No whoosh transition.

反例:Add epic sound design.

在 SunoMV 怎麼用:SFX 要能卡上小鼓。歌本身混音已經密,SFX 就少寫,別跟成曲打架。

8. 給空間一層環境底噪

一句規則:Google 把環境音列成第三種音訊。不寫,模型就會發明一段通用配樂床。

Ambient noise: low traffic two streets over, a vending machine hum,
rain on plastic sheeting. No score. No vocal ad-libs from off-screen.

反例:Background music that fits the mood.

在 SunoMV 怎麼用:歌詞影片你已經有歌了。跟 Veo 說 no score,原生音床才不會跟副歌疊兩次。

9. 把 16:9 和 9:16 當欄位寫,不是事後裁切

一句規則:Google 的 Veo 3.1 指南只給 16:9 或 9:16。寫 1:1 的 prompt,等於之後自己裁,而且會裁壞。

Aspect ratio: 9:16 vertical. Headroom for captions. Subject stays in the
center third. No wide establishing shot. Phone-first framing.

反例:Make it work for every platform.

在 SunoMV 怎麼用:一次匯出一個比例。別產生 16:9 再幻想直式裁切還留得住臉。

霓虹雨裡歌手的 9:16 直式 MV 畫面

圖片:SunoMV 團隊 · 把 9:16 當欄位寫,不是寬畫面事後裁切

10. 寫進 8 秒鎖定,不要跟它對幹

一句規則:在 SunoMV,Veo 鏡頭鎖死 8 秒。寫一分鐘一鏡到底,不會得到一分鐘。

Duration: one 8-second shot. Start on her hands at the jacket zipper,
end on her eyes as she looks up into the rain.
No cut inside the clip. No time-lapse. No "then later that night."

反例:A full music video that follows her all night across the city.

在 SunoMV 怎麼用:一則 prompt = 一句歌詞或一個重拍。把 8 秒鏡頭串起來。別要求一顆 Veo 鏡頭當整支 MV。

實用規則: Prompt 裡出現「then」,你就是在寫一場戲。拆開。Veo 不會幫你把一段話拆成分鏡。

6 支 MV 範本

每一支範本貼一次。括號裡的字自己換。欄位順序別動。

流行副歌 hook(16:9)

Cinematography: medium shot, eye-level, slow push-in, 35mm.
Subject: a woman in her 20s, red vinyl jacket, wet hair stuck to her cheek.
Action: she mouths the chorus and steps one pace toward camera, rain on her lashes.
Context: rooftop at night, city grid behind her, one red aviation light blinking.
Style and ambiance: hard backlight, thin fog, teal-and-red grade, light grain.
Audio: she sings, "Stay until the lights go out." Ambient: wind on a metal rail.
No score. 16:9. 8 seconds.

歌詞特寫(嘴型 + 字幕空間)

Cinematography: close-up, slight low angle, locked-off with a 2cm drift.
Subject: the same woman, oxblood jacket, silver hoops, a small scar on the left eyebrow.
Action: she inhales, then delivers one line, eyes wet but not crying.
Context: dark studio, a vintage microphone in the lower third, out-of-focus meters behind.
Style and ambiance: warm key from the right, cool rim, shallow focus on the mouth.
Audio: she says, "I kept the chorus." Ambient: room tone, no score.
Leave headroom at the top for captions. 8 seconds.

復古麥克風前的歌詞特寫,原生音訊 Veo 3 prompt

圖片:SunoMV 團隊 · 歌詞特寫是一張嘴、一支麥克風、一句加引號的詞

跟拍運鏡

Cinematography: tracking shot from the left, hip height, 50mm, modest speed.
Subject: a man in his 30s, grey overcoat, scuffed boots, a silver ring on the right hand.
Action: he walks parallel to a brick wall, left hand brushing wet ivy.
Context: alley after rain, one sodium street lamp, a cat frozen in a doorway.
Style and ambiance: orange sodium vs cold brick, puddle reflections, no beauty fill.
SFX: boot on wet stone; a distant bus hiss. Ambient: city two streets over. No score.
16:9. 8 seconds. Do not reveal the end of the alley.

直式短片(9:16)

Cinematography: medium close-up, 9:16, slow crane-up from chest to eyes.
Subject: the same woman in the red vinyl jacket, rain in her hair.
Action: she looks up, blinks once, then smiles on the last beat.
Context: neon stairwell, one pink tube light, condensation on the rail.
Style and ambiance: magenta key, green bounce from a sign, crushed blacks.
Audio: she says, "Hit post." SFX: a single heel on metal stairs. No score.
Keep her in the center third. 8 seconds.

角色鎖定(同一個人,第二個場景)

Cinematography: wide-to-medium, slow dolly right, 35mm.
Subject: SAME woman as previous shots — blunt black bob, oxblood leather jacket,
white tank, silver hoops, chipped black nail polish. Do not recast. Do not change hair.
Action: she leans on a convenience-store window and watches her own reflection.
Context: fluorescent store interior behind glass, night street in the reflection.
Style and ambiance: sickly green fluorescents vs magenta street neon, wet glass.
Ambient: compressor hum, distant register beep. No score. 16:9. 8 seconds.

原生音訊對白(僅 Fast)

Cinematography: two-shot, eye-level, slow pan from him to her, 40mm.
Subject A: the woman in the oxblood jacket. Subject B: a man in a grey overcoat.
Action: he turns, she does not. She answers without looking at him.
Context: rain under a bus shelter, one fluorescent tube buzzing.
Style and ambiance: hard overhead fluorescent, greenish, wet pavement sheen.
He says, "You coming?" She says, "After the chorus."
SFX: rain on the shelter roof; a bus air-brake. Ambient: traffic. No score.
16:9. 8 seconds. Pick Veo 3.1 Fast. Do not use the silent option.

雨中公車候車亭的雙人鏡頭,原生音訊 Veo 3 prompt

圖片:SunoMV 團隊 · 加引號的台詞只有嘴還留在畫面裡才有用

空 prompt vs 真正的 Veo 3 prompt

同一首歌、同樣 8 秒、兩種輸入。只有一種叫 prompt for Veo 3。

空的可複製
第一個子句「cinematic, 8K」具名景別 + 具名運鏡
主體「a girl」服裝 + 頭髮 + 一個辨識點
動作「dancing」三步、重量、和地面的接觸
音訊(沒有,或「加音樂」)加引號的台詞 + 點名的 SFX + 「no score」
比例/時長「做成能爆的」16:9 或 9:16,8 秒,起幅和落幅
給誰用一張假裝成影片的定格真能剪進副歌的一鏡

如果你已經有 Suno 的歌,別再叫 Veo 寫另一首。把它對準畫面,把歌留下。Suno prompt 技巧指南 是音訊半邊的可複製清單;本頁是畫面半邊。更長的 Seedance 鏡頭語言見 Seedance 2.5 prompt 指南——欄位不同,工作是同一件:複製整塊。

決策過濾器: 若你無法在 prompt 裡指出起幅和落幅,你就沒有鏡頭。你只有 vibe。

常見問題

Prompt 裡該寫「Veo 3」還是「Veo 3.1」?

都別寫。模型名不屬於 prompt 正文。人搜的是 veo 3 prompt;SunoMV 的選單目前列出 Veo 3.1 FastVeo 3.1 Lite。五欄是同一套。第一個子句寫鏡頭,別寫版本字串。

為什麼時長被拉成 8 秒?

那是 SunoMV 裡 Veo 的鎖定,不是你措辭的 bug。把起幅和落幅寫成 8 秒裝得下的。別描寫一分鐘的故事,再指望模型幫你壓縮。

有了首尾幀,還要寫 prompt 嗎?

要。幀管住身份。prompt 還是得點名鏡頭、動作和音訊。首尾幀中間夾一句 cinematic 8K,仍然是空 prompt。

明明寫了台詞,為什麼沒有對白?

兩個常見原因。你選了無聲的 Lite。或者台詞沒加引號,被當成描寫。加上引號,選 Fast,嘴留在畫面裡。

該寫 negative prompt 嗎?

寫正著的那一鏡。Google 的範例把字花在畫面裡有什麼。必須排除的(配樂、人群、第二張臉)用收尾約束:No score. No crowd.——不是一整段「不要」。

一則 prompt 能當整支 MV 嗎?

不能。一則 prompt 是一顆 8 秒鏡頭。AI MV 製作教學 才是把這些鏡頭串起來的流程。本頁只負責能貼上的 Veo 3 prompt。

複製、選 Veo、匯出 MV

一段 Veo 3 prompt 寫完的標準是能貼,不是能講。

  1. 複製上面一支範本。填括號。維持 Google 的五欄順序。
  2. 打開 SunoMV 音訊轉影片產生器。丟進歌曲(Suno 連結或自己的檔案)。
  3. 模型列表裡,區塊有對白或 SFX 就選 Veo 3.1 Fast。只要一張無聲的 8 秒畫面就選 Veo 3.1 Lite
  4. 整段貼進這一鏡。別把鏡頭句和主體句拆開。
  5. 匯出,接到歌詞那一行。再來一顆。MV 是一串 8 秒鏡頭,不是一段英雄散文。

能用的 Veo 畫面,不是形容詞表比較強的人做出來的。是按下產生之前,手上已經有一段能貼的 prompt for Veo 3 的人。

SunoMV 團隊

查看「Suno 提示詞與 AI 寫歌」全部 26 篇 →

試試這些 AI 工具