AI 音樂提示詞不是關鍵字清單,也不是越長越好的形容詞段落。它更像一份精簡的製作 brief:告訴 Mureka 這首歌是什麼曲風、想表達什麼情緒、誰來演唱、用了哪些樂器、歌曲如何推進,以及你最後想聽到什麼樣的聲音質感。

一個好的 Prompt 不需要很長,但每一個重要描述都應該對應到一個可以聽見、可以比較的音樂結果。與其寫「高級、好聽、專業」,不如寫清楚主歌是否克制、副歌是否展開、人聲是否貼近、鼓組是否靠後,以及哪些聲音需要為主旋律留出空間。

AI 音樂 Prompt 的 7 個要素

一段容易控制的 Prompt,通常按下面的順序組織:

曲風 + 情緒 + 節奏感 + 人聲方向 + 核心樂器 + 段落結構 + 製作質感

例如:

Dreamy alternative R&B, intimate and nocturnal, slow groove, clear lead vocal with close harmonies, electric piano and muted guitar, restrained verse that opens into a spacious chorus, warm low end and clean modern mix.

這段 Prompt 依序交代了曲風、氛圍、速度感、人聲、樂器、段落對比與混音質地。模型接收到的是一張有優先級的音樂地圖,而不是一串彼此競爭的標籤。

先寫創作意圖,再補聲音細節

在列樂器之前,先回答三個問題:

  1. 這首歌想表達什麼?

  2. 聽眾應該感受到什麼情緒變化?

  3. 這首歌會用在什麼場景?

你可以先寫給自己看的 brief:

A reflective indie-pop song about starting over in a new city, written for a late-night drive. The mood moves from uncertainty to quiet confidence.

再把它擴展成音樂 Prompt:

Reflective indie pop, mid-tempo and cinematic, intimate lead vocal, clean electric guitar, warm piano and subtle synth pads, restrained verse, rising pre-chorus, open singable chorus, gradual emotional lift, polished but natural mix.

第一句說明「為什麼要寫這首歌」,第二句說明「它應該怎麼聽起來」。兩者合在一起,可以幫助 Mureka 把抽象主題轉成旋律、演唱、編曲與結構的具體決策。

曲風怎麼寫:從大類到可聽特徵

只寫 poprock 這種大類,通常還不夠定義完整聲音。你可以再補上該曲風最重要的節奏、樂器、人聲與製作特徵。

曲風 Prompt 的三層結構

第一層:曲風身份

先指明歌曲屬於哪個音樂方向,例如 indie popdream popalternative R&BWest Coast G-Funkminimal electronic pop

第二層:可聽訊號

再加入能聽見的元素,例如 reverb-rich guitarswarm Rhodessliding synth bassrestrained drumscrystalline synth pulses

第三層:整體性格

說明這些元素如何一起工作,例如 hazy and intimatelaid-back and nocturnalspacious and contemplative

例如:

West Coast G-Funk, laid-back and nocturnal, sliding synth bass, Moog whistle lead, talkbox atmosphere, warm Rhodes, muted wah guitar, behind-the-beat drums, raspy male vocal, smoky lowrider night-cruise mood.

這比「G-Funk, cool, atmospheric」更具體,因為每一句都提供了可聽見的編曲、節奏、人聲或空間資訊。

情緒怎麼寫:描述變化,不只寫形容詞

sadhappyemotional 可以當起點,但它們沒有說明情緒在歌裡如何變化。更有用的是寫出情緒曲線,以及每個段落的作用。

例如:

Begin intimate and uncertain, build tension through the pre-chorus, release into a hopeful chorus, then return to a quiet and resolved ending.

你也可以把情緒拆成動作:

  • restrained verse:主歌維持克制。

  • rising pre-chorus:副歌前逐步累積張力。

  • open sustained notes in the chorus:副歌使用更開放、拉長的音。

  • stripped bridge:橋段降低密度形成對比。

  • gradual dynamic lift:整體能量慢慢上升。

這種寫法會把情緒轉成結構與演唱動作,也更容易在生成結果裡驗證。

人聲怎麼寫:從「聲音」寫到「演唱方式」

人聲 Prompt 不只描述性別或音域,也可以寫清楚發音、距離感、情緒和唱法。

描述人聲的常用維度

音色與音域: warm low registerclear bright tonesoft breathy texture

距離感: close and intimate vocalspacious vocal with gentle reverb

演唱方式: conversational phrasingrestrained vocal runsconfident sustained notes

歌詞表達: clear dictionemphasize the central phrasekeep the verse conversational

和聲編排: close harmonies in the chorusstacked backing vocals in the final chorus

例如:

Clear female lead vocal with conversational phrasing, precise diction, restrained verses, stronger sustained notes in the chorus, and close harmonies added only in the final section.

如果你需要某個特定人物的發音或專有名詞讀法,請另外標註。不要堆疊一長串互相矛盾的人聲形容詞,否則模型很難判斷哪一個最重要。

樂器怎麼寫:說明角色,而不是羅列清單

列出十幾種樂器,不代表編曲會更豐富。更有效的方法,是說清楚每一種主要樂器在歌曲裡扮演什麼角色。

例如:

Fingerpicked acoustic guitar carries the verse, warm bass supports the groove, soft strings widen the pre-chorus, and muted drums keep the chorus moving without masking the vocal.

這段 Prompt 說明了:吉他負責主歌基礎,貝斯支撐律動,弦樂在預副歌拓展空間,鼓組推動副歌但不遮擋人聲。

也可以更簡短地寫:

Piano-led verse, subtle synth texture underneath, fuller drums and layered guitar in the chorus, leave space around the lead vocal.

寫樂器時的三個問題

  1. 誰負責主旋律或主動機?

  2. 誰負責節奏、低頻或空間?

  3. 哪些樂器只在特定段落出現?

比起單純列出樂器名稱,替樂器分配角色更能幫助模型做出有層次的編曲。

用段落結構控制能量和對比

結構是 Prompt 裡最容易被忽略、但也最有幫助的一部分。它告訴模型:歌曲不是從頭到尾維持同樣狀態的聲音,而是一個會發展的完整形式。

常見的結構語彙

[Verse 1] restrained and intimate
[Pre-Chorus] gradual lift and rising tension
[Chorus] open, memorable, and emotionally released
[Verse 2] add a new melodic detail
[Bridge] stripped back with a different texture
[Final Chorus] fuller harmony and wider arrangement

如果不想寫成歌詞標籤,也可以直接放進 Prompt:

Sparse verse, gradual pre-chorus lift, wide memorable chorus, stripped bridge, final chorus with added harmonies and percussion.

這些結構詞的價值在於,它們幫助你說明為什麼需要某種聲音,而不是要求整首歌每一秒都同樣「滿」。

用空間和密度讓編配更清楚

常見誤區是把「更多聲音」等同於「更好的編曲」。Mureka V9.5 可以跟隨這種想法:編配密度應該配合每一段的功能,而段落彼此要為主旋律、人聲和情緒變化保留空間。

你可以在 Prompt 中加入:

  • sparse verse:主歌保持輕盈。

  • leave room for the vocal:為人聲保留空間。

  • instruments enter gradually:樂器逐步進入。

  • clear separation between parts:各部分之間有清楚分離。

  • wider arrangement in the chorus:副歌擴大編配。

  • restrained low end:控制低頻密度。

例如:

Keep the verse sparse and close, let the vocal lead, introduce percussion gradually, and open the stereo field in the chorus without making the arrangement crowded.

這些描述不會削弱歌曲的完整度;相反,它們能讓副歌的能量提升更清楚。

什麼時候使用否定指令?

否定指令適合處理你已經聽到的具體問題,不適合把所有不喜歡的東西都列出來。

比較直接的寫法是:

No spoken intro, avoid excessive vocal runs, no abrupt genre change.

不建議寫成一整段禁令,因為太多負面內容會把模型注意力從創作目標移開。先把你想要的聲音講清楚,再補一到三條最必要的限制。

Prompt 的長度:具體,但不要失去優先級

一個 Prompt 可以包含很多資訊,但最重要的元素應該放在前面。建議的優先順序是:

  1. 曲風與整體身份

  2. 情緒與節奏

  3. 人聲方向

  4. 核心樂器

  5. 段落與動態

  6. 製作質感與少量限制

如果生成結果偏離方向,不要立刻把 Prompt 拉長。先刪掉不必要的描述,再確認最重要的三到五個音樂特徵是否足夠清楚。

生成後如何迭代 Prompt

Prompt 的真正價值,在於你可以說清楚一次改動帶來了什麼結果。建議一次只改一個主要變數。

一個實用的迭代流程

第一版:定義身份

先只寫曲風、情緒、節奏、人聲與兩三個核心樂器,確認大方向。

第二版:調整結構

加入主歌、副歌、橋段的對比,觀察能量是否有推進。

第三版:調整人聲

修改發音、距離感、唱法、和聲或關鍵詞強調。

第四版:調整編配

說清楚哪些部分要退後,哪些樂器要在副歌進來,以及低頻和空間如何變化。

第五版:解決具體問題

只加少量否定指令,處理重複、擁擠或突兀轉折等最明顯的問題。

把每一版 Prompt 和對應音頻都保存下來,下次就能知道到底是哪個描述真的改變了結果。

四組可以直接改寫的 Prompt

1. Emo Rap

https://www.mureka.ai/song-detail/156330950197249?from=mine_generation_song

Emo rap with restrained midtempo half-time pull, intimate and emotionally exposed. Use a 2-bar intro: a detuned glassy synth phrase through chorus and reverse reverb, then enter the vocal. Build from a soft clipped kick, dry snare, sparse hats and rounded sub-bass that follows the vocal cadence. Let a blurred synth pad and distant formant-vocal texture answer the main three-note motif. Keep verses thin and conversational; lift the hook with doubled lead vocals, a higher answer and wider synth tails, not extra drums. Feature a young male voice moving between weary melodic rap and cracked singing, with clear words, small pitch scoops and restrained Auto-Tune. Use concrete breakup details and self-contradicting thoughts. Keep the mix close, hazy and low-end controlled. Avoid piano, strings, guitars, brass, trap hat barrages, sliding 808s, bright pop gloss and oversized vocal stacks.

2. Retro Synthwave

https://www.mureka.ai/song-detail/157034838228993?from=mine_generation_song

Dark Korean girl-group synth-pop with a slower pulse, deep minor-key tension and a dangerous sensual mood. Open with a 2-bar close-miked female pickup over a low, filtered analog bass. Keep verse one sparse: heavy muted kick, dry snare, minimal hats and a cold repeating synth-bass figure with wide gaps. Gradually deepen the low end and let a detuned synth motif emerge beneath intimate vocals. Build the pre-chorus through rising melodic tension, darker filters and held silences. Open the hook into broader sub-bass, restrained group harmonies and a hypnotic descending melody, never a bright release. Let verse two grow more intense; save the thickest bass and vocal layers for the final chorus. Feature mature female voices, smoky midrange, controlled breathiness, crisp Korean diction and confident lower-register phrasing. Avoid piano, bells, disco grooves, busy percussion and fast tempos.

3. Girl-Crush Hip-Hop

https://www.mureka.ai/song-detail/156995917316098?from=mine_generation_song

Sleek girl-group hip-hop fused with smooth R&B: confident, sultry and midtempo with a half-time bounce. Use a 2-bar intro: breathy female tag, filtered synth pulse and one clipped 808 hit, then enter verse one. Build from deep sub-bass, syncopated kicks, dry snaps, crisp hats and muted electronic accents. Add warm electric keys, glossy detuned synths and chopped vocal flickers. Feature contrasting female voices: one husky, nasal rapper with crisp consonants, playful swagger, Korean-English phrases and elastic rhythmic phrasing; others deliver airy R&B lines, soft falsetto touches and close layered doubles. Move from stripped rap verses into a smooth melodic pre-chorus, then a short addictive group hook with breathy responses and stronger bass. Keep the mix polished, intimate and fashion-forward. Avoid rock guitars, brass, aggressive EDM drops, shouted choruses and childish cute vocals.

4. UK Garage × Deep House K-Pop

https://www.mureka.ai/song-detail/157064986886146?from=mine_generation_song

Korean girl-group UK garage and deep-house pop with relaxed, slower midtempo swing, cool confidence and subtle Y2K club character. Open with a 2-bar low female pickup and a warm filtered bass pulse, then settle into an unhurried two-step groove. Use soft syncopated kicks, a dry spacious snare, sparse shuffled hats, rounded bass notes and one muted synth stab. Keep rhythmic subdivisions minimal, leaving generous space between drum hits. Rotate female voices between smooth midrange talk-singing, lightly husky melodies, easy rhythmic rap and airy upper-register touches. Let members trade short, relaxed phrases with subtle tuning, close doubles and small vocal chops. Build the hook through a simple repeated melody, gentle bass movement and restrained group responses. Keep it clean, sensual and composed. Avoid fast tempos, double-time drums, busy hats, piano, bells, marimba and EDM drops.

這些例子分別突出空間感、局部化的曲風細節,以及器樂氛圍。你可以替換主題、節奏、聲線與樂器,但仍保留「曲風 — 情緒 — 聲音 — 結構 — 製作」的順序。

多語言歌曲的 Prompt 寫法

如果歌曲使用非英文歌詞,建議明確寫出目標語言,並把語言要求和音樂要求分開:

Spanish lyrics, clear pronunciation and natural syllable stress, modern Latin pop, warm acoustic guitar, mid-tempo groove, intimate verse and a bright singable chorus.

多語言歌曲還可以進一步說明,哪個語言出現在什麼段落,以及哪些共享的旋律或情緒特徵應該保留。生成後,重點檢查專有名詞、重音、連音,以及重複短句。

Mureka 的資料顯示,多語言能力與評測系統是模型演進的重要部分。語言切換不只是歌詞翻譯,也會影響音節數、重音位置、旋律落點與聲音表現。因此,語言要求應該一起納入聆聽與修正流程。

如何判斷 Prompt 是否有效?

不要只看模型是否「懂了幾個關鍵字」;要聽這些要求有沒有貫穿整首歌。你可以用這份檢查表:

  • 曲風身份是否在主歌、副歌、橋段都保持清楚?

  • 情緒是否按照 Prompt 的方向發展,而不是一開始就到達最高點?

  • 人聲是否符合指定的音色、距離與演唱方式?

  • 核心樂器是否各自承擔清楚角色?

  • 編配是否在需要時增加密度,在需要時留出空間?

  • 歌詞關鍵字是否落在合適的段落與旋律位置?

  • 結尾是否完成歌曲的情緒收束?

在資料提供的 100 首歌曲同口徑評測中,Mureka V9.5 的 Prompt 控制良品率為 97.0%,曲風完全體現比例為 95.7%。這些數據說明 Prompt 方向和曲風表達是 V9.5 的重點能力,但評測結果來自特定樣本,不能替代對具體歌曲的聆聽判斷。

常見錯誤與改進方式

錯誤 1:只寫曲風名稱

問題: Pop song 沒有說明情緒、節奏、人聲或結構。

改進: 補上節奏感、核心樂器、人聲與段落對比。

錯誤 2:形容詞太多,音樂動作太少

問題: beautiful, amazing, emotional, professional 很難轉成可聽見的選擇。

改進: 改寫成具體描述,例如 restrained verse, rising pre-chorus, clear diction, warm low end

錯誤 3:要求每一段都要「很大」

問題: 如果每段都塞滿,副歌就沒有空間抬起來。

改進: 清楚寫出 sparse versefuller chorusstripped bridge

錯誤 4:一次改太多

問題: 如果同時改曲風、歌詞、人聲、節奏和樂器,就很難知道究竟是什麼造成變化。

改進: 保存版本,每次只調整一個主要變數。

錯誤 5:把限制寫成否定清單

問題: 太多禁止事項會讓模型注意力偏離創作目標。

改進: 先清楚描述你想要的聲音,再加少量有針對性的限制。

常見問題

AI 音樂 Prompt 應該用自己的母語還是英文?

兩種都可以。英文示例裡的音樂術語比較集中,較方便表達曲風、編配和製作細節。不論使用哪種語言,重點都是要描述清楚的音樂動作與段落關係。

Prompt 越長,生成結果越好嗎?

不一定。Prompt 的資訊需要有優先級。過長且互相矛盾的描述,可能讓模型難以判斷核心方向。先用簡潔 Prompt 定義身份,再根據聽到的結果補充細節。

為什麼同一段 Prompt 會生成不同結果?

生成結果會受到歌詞、語言、結構和生成過程等因素影響。保存有效 Prompt,固定大部分描述,每次只改一兩個細節,可以更清楚地找到穩定的創作方向。

怎麼讓副歌更有記憶點?

除了寫 memorable chorus,也可以寫副歌需要更開放的旋律、更長的延音、重複的核心句、更厚的和聲或更寬的編配。然後再聽副歌是否真的和主歌形成對比。

怎麼避免人聲被伴奏蓋住?

你可以在 Prompt 中寫 leave room for the lead vocalclear vocal centeraccompaniment should support rather than mask the vocal,同時減少不必要的樂器與低頻堆疊。

什麼時候應該重新寫一段 Prompt?

如果歌曲的核心曲風、情緒和結構都偏離目標,重新整理 Prompt 會更有效。如果只有人聲發音、某個樂器或一個段落有問題,則優先做局部修改或使用 Remix、Extend、Studio 等後續創作能力。

把 Prompt 寫成一份可執行的音樂 brief

好的 AI 音樂 Prompt,不是把所有想像一次塞進一段文字,而是把創作意圖拆成可聽見的選擇:曲風身份、情緒曲線、人聲風格、樂器角色、段落結構與製作空間。

在 Mureka 中,你可以先用一段簡潔 Prompt 生成方向,再透過聆聽與版本迭代逐步補充細節。把每次修改都當成一次製作決策,Prompt 就不再是隨機試錯的關鍵字,而會成為一套可重複、可比較、可持續的創作方法。