You typed "make me a cool song," hit generate, and got back a minute of the same four bars looping with no chorus, no ending, and vocals that mumble. It ran — but it is not a song you would send to anyone. Most beginner guides stop at "just describe what you want," which is exactly the advice that produces those loops.
This guide walks the full beginner workflow end to end — pick a tool, write a real prompt, add structure, decide on vocals, generate, iterate, and export — using Google's Lyria 3 Pro as the worked example because it is built to produce full songs, not clips. Along the way it flags the specific steps beginners get wrong, so your first finished track sounds like a track. No music theory required, and no software to install.
What "making AI music" actually involves
At a plain-English level, a text-to-music model turns a written description into audio. You do not record anything or play an instrument — you write a prompt (genre, mood, instruments, tempo, vocal direction) and the model composes and performs a song that matches.
The tool used throughout this guide is Lyria 3 Pro, Google DeepMind's music model. It writes original songs up to 3 minutes long — intros, verses, choruses, and bridges — including AI-generated vocals and lyrics, and returns high-fidelity stereo audio at 48 kHz. Google announced it on March 25, 2026, about a month after the base Lyria 3 model (30-second clips) launched inside the Gemini app.
Two facts shape everything below, so hold onto them:
- A full song comes back in one generation. You do not stitch 30-second chunks together. That is why structure has to go into the prompt up front — you cannot bolt a chorus on afterward.
- Inputs are text, your own lyrics, PDF, and reference images (up to 10). There is no audio upload and no voice cloning — you describe the music in words, you do not hum it in.
Every output also carries an invisible SynthID watermark and supports C2PA content credentials, so an AI track is always identifiable as AI-generated. Good to know before you share one.
Step 1 — Pick a tool
You cannot make AI music "in the abstract" — you make it through a specific product, and the product decides your cost and setup. For Lyria 3 Pro there are several official routes plus credit studios.
| Where you use it | Best for | Setup / cost |
|---|---|---|
| Gemini app | Casual creators already paying for Gemini | Needs a paid Gemini subscription |
| Google AI Studio / Gemini API | Developers who want to automate | Pay-per-generation, metered in preview |
| Vertex AI | Teams already on Google Cloud | API + Media Studio, public preview |
| ProducerAI | Trying it globally with a free option | Free + paid tiers |
| Credit studios (e.g. lyria3-pro.com) | Beginners who want songs, no setup | Pay-as-you-go credits |
For a total beginner, the two lowest-friction routes are a credit studio (buy credits, type a prompt, generate) or ProducerAI's free tier. Neither asks you to wire up Google Cloud, hold a subscription, or touch an API key. This guide assumes the credit-studio route for the hands-on steps, but the prompting technique is identical everywhere — so what you learn here transfers to any access path. If you want to jump straight in, the Lyria 3 Pro generator on this site runs the model directly on credits, and the pricing page lists exact costs.
Where beginners go wrong: picking the API first. If you do not write code, you do not need it. Start with a studio you can use in a browser.
Step 2 — Write a prompt with the official formula
This is where most first tracks are won or lost. A one-line vibe ("chill lofi song") gives the model almost nothing to work with, so it defaults to a generic loop. Google publishes a formula that fills the gaps:
[Genre & style] + [Mood] + [Instrumentation] + [Tempo & rhythm] + [Vocal style & language] + [Lyrics]Fill every slot. A worked beginner example:
Warm indie folk with a modern pop sheen.
Hopeful and a little nostalgic.
Acoustic guitar, soft piano, brushed drums, subtle strings.
Around 90 BPM, gentle and steady.
Female lead vocal, clear and intimate, in English.Two things beginners miss. First, be concrete about instruments and tempo — "acoustic guitar, brushed drums, ~90 BPM" gives the model a target, while "nice instruments" gives it nothing. Second, name the vocal style and language out loud. Lyria 3 Pro can sing in 8 languages (English, German, Spanish, French, Hindi, Japanese, Korean, Portuguese), and it will not guess correctly unless you say so.
Rule of thumb: If you can hum the result of your prompt before you generate, it is specific enough. If you cannot, add more detail.
Step 3 — Decide: vocals, instrumental, or your own lyrics
Three simple flags control the singing. They fix the two most common beginner complaints — "I wanted just background music" and "it ignored my words":
- No vocals: end the prompt with
Instrumental.and the model produces a purely instrumental track. - Your exact words: prefix them with
Lyrics:and the model sings what you wrote instead of inventing its own. - Duets: describe more than one singer ("a male and female voice trading lines in the chorus") to get multiple voices.
Lyrics:
Verse 1: We left the city with the windows down...
Chorus: And the whole road opened up ahead...If you just want a backing bed for a video or podcast, Instrumental. is the single most useful flag you will learn today.
Step 4 — Add structure so you get a song, not a loop
This is the step nearly every beginner guide skips, and it is the difference between a loop and a song. Instead of hoping the model invents an arrangement, you script it yourself with timestamp cues from Google's prompting guide. Lay the song out on a timeline with [mm:ss] markers, and the model treats each one as a guaranteed section change:
[00:00] Soft guitar intro, sparse, setting the mood.
[00:20] Verse 1 begins, add bass and light drums.
[00:50] Chorus, full band, big and warm, layered backing vocals.
[01:25] Verse 2, pull the drums back a little.
[01:55] Bridge, strip to piano and vocal, then rebuild.
[02:25] Final chorus, biggest energy, extra harmonies.
[02:55] Outro, gentle fade.Because a full generation returns one continuous track, this map is your only chance to control where the chorus lands and when the bridge hits. Give the model an arrangement and it stops repeating itself. For a beginner, this one habit does more for output quality than anything else.
Ready to try your first real prompt? Drop your formula prompt plus a timestamp map into the AI music generator here and you can hear a full structured song in a couple of minutes — no install, no subscription.
Step 5 — Test cheap before you commit
Before you spend a full 3-minute generation, sketch on the base Lyria 3 model, which returns up to 30 seconds. Use it to check that your genre, mood, tempo, and vocal direction are landing — the vibe, not the full arrangement. If the 30-second sketch sounds wrong, the 3-minute version will just be wrong for longer.
This is the prototype-then-produce habit: prototype on Lyria 3, produce on Lyria 3 Pro. A short sketch is cheaper than a full song (the base model is billed per 30 seconds, versus one charge per full Pro song, per third-party reports — verify current rates), so you can iterate on the idea for pennies before committing to the finished track.
Step 6 — Generate, then iterate one change at a time
Run the full generation with your formula prompt and timestamp map. When it comes back, resist the urge to re-roll blindly. Change one variable at a time so you learn what each edit does:
- Chorus too small? Add instructions at that timestamp ("full band, layered backing vocals").
- Vocals too loud in the mix? Adjust the vocal-style line, or switch to
Instrumental. - Wrong energy? Nudge the tempo or swap one instrument, and keep everything else fixed.
You can also attach up to 10 reference images to steer mood — a foggy coastline for something melancholic, neon signage for synthwave — when words alone are not landing the look you want.
Step 7 — Export and use it
Once a version sounds right, download it. Because Lyria outputs 48 kHz stereo, the file drops straight into a video editor, podcast, or playlist without extra processing. Two reminders: the track carries a SynthID watermark and C2PA credentials marking it as AI-generated, and your usage rights depend on your access path's terms — check those before commercial use.
How to tell it worked
A finished song, not a loop, passes these checks:
- Distinct sections — you can point to where the verse ends and the chorus starts, and they sound different.
- A dynamic arc — energy rises into the chorus and pulls back for the bridge instead of sitting flat.
- A real ending — the track resolves on an outro instead of cutting off mid-phrase.
- Vocals that match your brief — right language, right style, and your lyrics if you supplied them.
If it fails on "distinct sections," your structure map was too vague or missing — go back to Step 4.
Common beginner mistakes
| Mistake | Fix |
|---|---|
| One-line vibe prompt ("chill song") | Fill all six formula slots (Step 2) |
| No arrangement described | Add [mm:ss] timestamp cues (Step 4) |
| Testing ideas on full generations | Sketch on base Lyria 3 first (Step 5) |
| Expecting to upload audio or clone a voice | Not supported — describe it in text |
| Re-rolling randomly when unhappy | Change one variable at a time (Step 6) |
FAQ
Do I need any music experience to make AI music? No. You describe the song in plain language — genre, mood, instruments, tempo, and vocal style — and the model composes and performs it. The skill you are learning is prompting, not playing.
What is the easiest way to make an AI song for free or cheap? Start with a free tier (ProducerAI) or a pay-as-you-go credit studio, both of which run in a browser with no setup. Sketch a 30-second idea on the base model first, then produce the full song once the vibe is right.
Why does my AI song sound repetitive? You most likely gave it a mood but no arrangement. Use the timestamp technique (Step 4) to script distinct sections — intro, verse, chorus, bridge, outro — so the model changes over time instead of looping one idea.
Can I make an instrumental with no singing?
Yes. End your prompt with Instrumental. and the model generates the track without vocals — ideal for background music, podcasts, and video beds.
Can I use my own lyrics?
Yes. Prefix your words with Lyrics: and the model sings exactly what you wrote instead of inventing its own.
Is AI music safe to post online? Every Lyria output carries a SynthID watermark and C2PA credentials identifying it as AI-generated. Usage rights depend on your access path's terms, so review those before monetizing a track.
Ready to make your first real song? The Lyria 3 Pro generator on this site runs the model on pay-as-you-go credits, so you can test a 30-second sketch and then produce a full structured track without a Google Cloud project or subscription. For more background, see the complete Lyria 3 Pro guide and the Lyria 3 overview, or read the deeper step-by-step Lyria 3 Pro walkthrough.
Sources
- Google — Lyria 3 Pro announcement: https://blog.google/innovation-and-ai/technology/ai/lyria-3-pro/
- Google — Build with Lyria 3 (developer): https://blog.google/innovation-and-ai/technology/developers-tools/lyria-3-developers/
- Google Cloud — Ultimate prompting guide for Lyria 3 Pro: https://cloud.google.com/blog/products/ai-machine-learning/ultimate-prompting-guide-for-lyria-3-pro
- Google AI for Developers — Generate music with Lyria 3: https://ai.google.dev/gemini-api/docs/music-generation
- DeepMind — Lyria 3 model card: https://deepmind.google/models/model-cards/lyria-3/
- DeepMind — Lyria model page: https://deepmind.google/models/lyria/
- TechCrunch — Google launches Lyria 3 Pro: https://techcrunch.com/2026/03/25/google-launches-lyria-3-pro-music-generation-model/


