Prompts

Suno wants two separate fields, not one long music description

Both Suno and ElevenLabs respond to a specific input structure that most first-time users skip past — Suno wants Style and Lyrics as genuinely separate fields, and ElevenLabs wants a described character, not a vague adjective.

Last updated Aug 15 · 9 min read

Two tools, two genuinely different input structures

Suno's two-field format — Style and tagged Lyrics as separate inputs — and ElevenLabs' description-driven voice design are structurally different from each other, and both differ from a generic "write me a prompt" approach. The Music & Voice prompt library is built around each tool's actual input format specifically.

Suno: Style and Lyrics as separate, structured fields

A brand jingle works best when the Style field describes genre, tempo and instrumentation precisely, and the Lyrics field uses section tags ([Verse], [Chorus]) rather than one undifferentiated block of text — the two-field separation is what actually gives Suno the structure to work with rather than guessing at song structure from a single description.

For an instrumental background track with no lyrics at all, the Style field alone carries the entire brief — mood, pacing, instrumentation — since there is no lyric content to lean on for direction.

ElevenLabs: describing a character, not naming an adjective

"Sound friendly" is a vague, unusable direction for a voice model. Character voice design in ElevenLabs works by describing a genuine character — age, accent, energy, a specific reference point — the same principle that makes any AI generation task work better with concrete description over an abstract adjective.

Worked example: a jingle with a matching voiceover

For a short branded audio piece with both music and a voiceover, treat them as two separate generation tasks with two separate prompts — a Suno instrumental for the music bed, and an ElevenLabs character-voice prompt for the voiceover — rather than trying to get one tool to handle both. Match the described character's energy in ElevenLabs to the tempo and mood specified in Suno's Style field so the two halves feel like they belong to the same piece.

When audio needs to be part of a bigger brand identity

A jingle or a voice is one piece of a brand's overall sensory identity, alongside visual identity, tone of voice, and everything else that makes a brand recognisable. If audio identity needs to be considered as part of that bigger picture, that's the scope Scult's branding team covers, or book a meeting to talk through your brand's full identity.

Need this built into your business?

The free tools and prompts on this site handle the small, solved problems. If what you need is bigger — branding & design, built and maintained for you — that's Scult's day job.

Tools mentioned in this post

Prompts mentioned in this post

← All posts

Check your AI visibility

One URL in, a 0–100 score and the exact fixes out.

RUN THE CHECK

Browse all the tools

15 tools across six categories
13 of them never send your data anywhere

Free · No signup · No trial clock

SEE THE DIRECTORY