dav.one

Generate Music with YuE2: A Simple Guide to Style and Lyrics Prompts

Created on September 12, 2026 by Dawid Wasowski

YuE2 is a free, open AI model that creates full songs from two inputs: a style description and your lyrics. It’s built by the Multimodal Art Projection research team with HKUST, and independent tests show it matching or beatin best commercial music generators available today.

What makes YuE2 different from most “type words, get a song” tools is that it doesn’t jump straight to audio. First, it writes down a small musical plan - a melody and chords in a readable notation called ABC - and only after that does it render the final 48kHz stereo song. This means the song you get is the direct result of a plan you can inspect, and that plan is shaped almost entirely by how you write your prompt.

This guide focuses on one thing: how to write your style and lyrics so the song actually sounds like what you asked for.

The two things you write

Every YuE2 song comes from two text fields:

  1. Style - a short description of what the song should sound like (genre, instruments, mood, voice).
  2. Lyrics - the words of the song, split into labeled sections.

Below, you’ll find a detailed breakdown of how to write each of these fields so that the model produces a song that matches your vision.

Rule 1: Label every section

YuE2 reads your lyrics as a sequence of labeled blocks, not as one long poem. Put a label in square brackets before each part of the song:

[Verse]
[Pre-Chorus]
[Chorus]
[Bridge]
[Interlude]
[Outro]

Only write actual words under a label. Don’t put a song title above the lyrics, and don’t add notes to yourself inside the text - the lyrics field should contain nothing but the section tags and the words that get sung.

Rule 2: Leave one blank line between sections

Separate every section from the next with a full empty line (two line breaks in a row). This is how the model knows where one part of the song ends and the next begins. If sections run together without a break, the structure can blur and the song may not follow your intended arrangement.

Rule 3: Don’t open the song with written intro lyrics

The team behind YuE2 has found that the [Intro] label is less reliable than other labels when it’s placed first and given words to sing. Two safer options:

Rule 4: Keep each section short

The model generates audio in chunks of roughly 30 seconds per section. If you cram too many words into one [Verse] or [Chorus] block, the model has to rush to fit them in, and words can come out slurred, sped up, or cut off mid-line. Aim for a handful of lines per section - similar to how a real song is structured - rather than long paragraphs.

Rule 5: Match your syllables to the beat, not just your word count

Word count alone doesn’t tell you how a line will actually sing. What matters more is the number of syllables in a line, because the model has to fit those syllables into a beat one by one - much like a rapper fitting words into a bar. If a line has too many syllables, the extra ones don’t just get sung faster; they can spill over into the next beat and quietly push the rest of the line, or even the next line, out of place. That’s often what’s behind a generated song that sounds slightly “off” even though the words themselves are correct.

If you’ve used Suno before, you might be used to chaining words together with hyphens to force them into a single beat (something like “walking-down-the-street-fast”). That trick doesn’t reliably help with YuE2. A steadier approach is the one songwriters already use for genres like rap: count the syllables in each line, and keep that count consistent whenever a line repeats - especially in the chorus.

Rule 6: Write repeats out in full

If your chorus comes back later in the song, don’t shorten it to something like (repeat chorus) or [Chorus x2]. Copy the full chorus text again under a new [Chorus] label wherever it repeats. The model doesn’t expand shorthand - it sings whatever text is actually there.

Rule 7: Keep the lyrics field to lyrics only

Avoid production-style notes inside the lyrics, such as “(drums build up here),” “(pause),” or “[key change].” These aren’t words that can be sung, and they tend to confuse the model rather than guide it. If you want to signal an instrumental moment, use a label like [Interlude] or [Guitar Solo] and leave it without lyric lines.

Writing a style prompt that actually shapes the song

Your style prompt works alongside the lyrics, and it’s just as important. A style prompt that reliably gets good results usually covers five ingredients:

IngredientExamples
Genrepop, city pop, cyber metal, jazz-funk, ballad
Instrumentsacoustic guitar, pulsing bass, analog synths, brushed drums
Mooddreamy, uplifting, melancholic, energetic
Vocal gendermale vocal, female vocal
Vocal toneairy vocal, deep voice, bright vocal, expressive lead vocal

You don’t need to include all five every time, but the more of them you cover, the more predictable the result tends to be. You can write them as a natural, comma-separated description:

Jazz-funk, warm lead vocal, Rhodes piano, electric bass, tight drums, upbeat

Order doesn’t matter much - put the words that matter most to you first, since they tend to get the most weight.

Keep your lyrics’ language and your style prompt in sync

YuE2 can sing in several languages. If your style prompt says “Korean pop,” write your lyrics in Korean. If it says “English rock,” write them in English. Mixing a genre associated with one language and lyrics in a different language tends to produce less natural-sounding results, since the model is trying to match a vocal style it learned alongside that language.

Choosing how much creative control to give up

When you generate a song, you set a cot (chain-of-thought / planning) option:

SettingWhat it doesGood for
cot="full"Plans both melody and chords before rendering (default)Original songs, most everyday use
cot="melody"Plans only the melody, keeps harmony more openCovers, when you’re supplying your own chord ideas
cot="off"Skips the planning step entirelyQuick experiments, maximum variety between generations

Because the planning stage in "full" and "melody" modes produces a readable score before any audio is made, you can preview it, and even hand-edit it, before committing to a final render.

In practice, cot="full" tends to be the most reliable setting overall, and it’s the one to reach for if you want an instrumental track with no vocals at all. With the full melody-and-chord plan in place, the model has enough structure to build a complete arrangement even when there are no words to sing.

Making a cover of an existing song

YuE2 can also restyle a song that already exists:

  1. Get the melody as a score (the project’s companion tool, SheetSage2, can transcribe an existing recording into notation).
  2. Get the lyrics, split into sections matching the recording.
  3. Write a new style prompt describing the sound you want instead, and generate with cot="melody".

The result keeps the original melody and words but performs them in a new style - for example, turning a folk song into a heavy metal track.

Editing a song after it’s generated

Because YuE2 keeps the plan (the ABC score) separate from the final audio, you can revise a song without starting over. You can hand-edit the score yourself, or describe the change you want in plain language (for example, “make the harmony more jazzy and add a saxophone solo”) and let an editing agent update the score and style prompt for you. YuE2 then renders the updated version from the revised plan.

A complete example, put together

Style prompt:

Indie pop, bright acoustic guitar, soft drums, warm female lead vocal, hopeful

Lyrics:

[Verse]
Streetlights flicker on an empty road
Every step feels lighter than it's ever showed

[Pre-Chorus]
Something's changing, I can feel it grow
Every heartbeat tells me where to go

[Chorus]
We are rising with the morning sun
Nothing's over till the story's done
Hold my hand and we will find the way
Turning shadows into brighter days

[Bridge]
Quiet streets are singing back to me
Every doubt is finally set free

[Chorus]
We are rising with the morning sun
Nothing's over till the story's done
Hold my hand and we will find the way
Turning shadows into brighter days

[Outro]
Brighter days, brighter days

Notice the structure: labeled sections, a blank line between each one, no written intro, a repeated chorus written out in full, and a style prompt that covers genre, instruments, mood, and vocal type.

Quick checklist before you generate

Follow these and you’ll get songs from YuE2 that sound much closer to what you actually had in mind.

Sources and further reading