ANALOGALCHEMY / GUIDES / SETTINGS THAT MATTER

From text prompt to finished song: the settings that matter.

GUIDE · UPDATED JUNE 2026 · CFG · STEPS · SEED · TEMP

Every AI music tool hides the same handful of dials behind different names, and most people either never touch them or twist everything at once. This guide explains what each one actually does — in plain language — and the workflow that turns "decent first try" into "keeper take". The screenshots are from AnalogAlchemy's Make screen, where the dials live in the Expert Deck; because generation is local and uncapped, experimenting with them is free.

AnalogAlchemy Make screen with the tag orbit and the Expert Deck showing BPM, key, CFG, temperature, seed and steps controls
MAKE · TAG ORBIT + EXPERT DECK (BPM · KEY · CFG · TEMP · SEED · STEPS)

The prompt does the heavy lifting

Before any dial: the text. Describe the song like you'd brief a producer — genre, mood, instrumentation, vocal character, one image. "Slow-burning arabesk; strings and bağlama; a man singing to an empty road; builds to a desperate chorus." Then pull tags into the tag orbit — dragging POP or TURKISH ARABESK closer to the ember strengthens its pull. Stuck at one line? The Style-from-AI button expands your idea into a full brief you can edit.

STYLE-FROM-AI · ONE LINE IN, A PRODUCER'S BRIEF OUT (20s)

CFG — the obedience dial

CFG (guidance) sets how literally the model follows your text. Think of it as briefing a session musician: too loose and they play what they feel like; too strict and the performance goes stiff. Default ~7 is home base. Result ignoring your prompt? Nudge up. Result technically-correct but lifeless? Nudge down. Move in small steps; CFG is sensitive.

Steps — the polish dial

Steps are refinement passes. More steps, more compute time, better detail — with hard diminishing returns. The practical pattern: low steps while hunting (fast drafts, many ideas), high steps once for the final render of the take you're keeping.

Seed — the reproducibility dial

The seed is where the randomness starts. On random, every run is a fresh lottery — great for exploring. But the moment a take is nearly right, fix the seed and change exactly one thing per run. Same seed + same settings = same song; same seed + one tweak = a controlled experiment you can actually hear. This is the single biggest habit that separates dialing-in from gambling.

Temperature — the risk dial

Temperature scales how adventurous the model's choices are. Lower = safer, more conventional moves; higher = surprises, occasionally brilliant, occasionally chaos. Around 0.85 is a sensible default; raise it when everything sounds samey, lower it when takes keep going off the rails.

BPM, key and time signature — just set them

If you know the song wants 96 BPM in E minor, say so in the Expert Deck instead of hoping the prompt implies it. Fixed musical parameters also make takes comparable — one less variable moving between runs. (Planning to sing it yourself later via a trained voice adapter? Pick a key you can actually sing.)

The workflow, end to end

How the settings interact

Each dial is straightforward on its own; the subtlety is how they push against each other.

Dial in a render: a worked example

Here's a concrete sequence that turns an interesting first take into a polished result, one step at a time.

Troubleshooting: when the results don't do what you expect

Results ignore my prompt entirely. Raise CFG. The model is taking too much creative latitude. A value of 9 or 10 is not unreasonable if the default is producing results that seem unrelated to your description. Also check that your most important terms are in the prompt text, not only in tags — both carry weight but the text prompt drives the main direction.

Results are technically on-brief but stiff or forced-sounding. Lower CFG. The model is following the letter of your prompt but not its spirit. A value of 5 or 6 often produces results that feel more naturally musical while still staying roughly in the territory you described. You can also try loosening the prompt itself — fewer constraints often give the model room to do something more genuinely interesting.

I had a great take but I didn't fix the seed and now I can't reproduce it. If you didn't note the seed, it's gone — random means random. This is the main reason to fix the seed the moment something interesting happens. From now on: generate, hear something promising, pin the seed before touching anything else.

Every take sounds the same regardless of settings. You're probably comparing at too-low steps. Increase steps to a moderate value and try again — at very low steps the model hasn't had enough refinement passes to express what the settings are asking for, so differences flatten out.

Questions people ask

What does CFG do?

Sets how strictly the model follows your prompt — up when it ignores you, down when it sounds forced. Default ~7.

More steps = better song?

Only to a point — diminishing returns. Low steps to explore, high steps once for the final render.

Why fix the seed?

Reproducibility: same seed + one changed setting = you hear exactly what that setting does.

Free dials, free experiments.

Local generation means no credits — every experiment in this guide costs nothing. Try it for 14 days, no account.

MACOS 14+ · APPLE SILICON · WINDOWS COMING SOON

MORE GUIDES: STEM SEPARATION · TRAIN YOUR AI VOICE · AI REPAINT