Prompts
Suno wants two separate fields, not one long music description
Both Suno and ElevenLabs respond to a specific input structure that most first-time users skip past — Suno wants Style and Lyrics as genuinely separate fields, and ElevenLabs wants a described character, not a vague adjective.
Last updated Aug 15 · 9 min read
Two tools, two genuinely different input structures
Suno's two-field format — Style and tagged Lyrics as separate inputs — and ElevenLabs' description-driven voice design are structurally different from each other, and both differ from a generic "write me a prompt" approach. The Music & Voice prompt library is built around each tool's actual input format specifically.
Suno: Style and Lyrics as separate, structured fields
A brand jingle works best when the Style field describes genre, tempo and instrumentation precisely, and the Lyrics field uses section tags ([Verse], [Chorus]) rather than one undifferentiated block of text — the two-field separation is what actually gives Suno the structure to work with rather than guessing at song structure from a single description.
For an instrumental background track with no lyrics at all, the Style field alone carries the entire brief — mood, pacing, instrumentation — since there is no lyric content to lean on for direction.
ElevenLabs: describing a character, not naming an adjective
"Sound friendly" is a vague, unusable direction for a voice model. Character voice design in ElevenLabs works by describing a genuine character — age, accent, energy, a specific reference point — the same principle that makes any AI generation task work better with concrete description over an abstract adjective.
Worked example: a jingle with a matching voiceover
For a short branded audio piece with both music and a voiceover, treat them as two separate generation tasks with two separate prompts — a Suno instrumental for the music bed, and an ElevenLabs character-voice prompt for the voiceover — rather than trying to get one tool to handle both. Match the described character's energy in ElevenLabs to the tempo and mood specified in Suno's Style field so the two halves feel like they belong to the same piece.
When audio needs to be part of a bigger brand identity
A jingle or a voice is one piece of a brand's overall sensory identity, alongside visual identity, tone of voice, and everything else that makes a brand recognisable. If audio identity needs to be considered as part of that bigger picture, that's the scope Scult's branding team covers, or book a meeting to talk through your brand's full identity.
Need this built into your business?
The free tools and prompts on this site handle the small, solved problems. If what you need is bigger — branding & design, built and maintained for you — that's Scult's day job.

