Simple and Text-to-Music
Describe what you want in your own words. Safe defaults are filled in and visible, never hidden, so you can see exactly what is about to run.
Create
Start from a sentence or from a full parameter set — both paths reach the same engines, the same quality and the same follow-up workflows. Nothing is reserved for a higher plan.
Modes
Whichever mode you start from, you end up with the same versioned track object — with its parameters, lineage, rights record and cost attached. Nothing is a side branch that cannot come back.
Describe what you want in your own words. Safe defaults are filled in and visible, never hidden, so you can see exactly what is about to run.
Take direct control of the musical decisions before a single second is rendered.
Start from an existing version or an uploaded reference and steer how far the result may move away from it with an audio strength control.
Regenerate a selected region instead of the whole track. Four repaint modes plus crossfade and strength keep the rest of the arrangement intact.
Pull out track classes from a mix, or generate a single new layer against what already exists, so an arrangement grows one decision at a time.
Finish a partial mix, run a batch or auto-generation group under a fixed budget, or start from a saved recipe or a voice note.
Controls
There is no separate Pro mode. Advanced and Expert controls open inside the same panel you started in, and they reach the same engines and the same quality as every other path.
The parameters that decide how the model actually renders, validated against what the selected engine really supports.
Engine-level controls are exposed rather than hidden: DCW mode, wavelet and scalers, retake variance, flow-edit for caption, lyrics and morph, and legacy CFG behaviour.
The interface is generated from the engine capability contract. A control that the deployed engine cannot honour is disabled and labelled, never silently ignored.
ACE-Step 1.5 and HeartMuLa sit behind one capability layer. Engine, model and adapter are pinned per job, so a platform-side model change never rewrites your old results.
Every run stores a complete snapshot — inputs, engine, model, adapters, parameters, seed and provider — so any version can be re-run or handed to a collaborator exactly as it was.
The creator copilot suggests style, lyric structure, length and metadata with its reasoning visible. It proposes; you decide, and the parameters stay editable.
Words
Lyrics carry their own history, structure and timing, and they travel with the track into every export.
Draft lyrics with section structure, generate or improve passages with the assistant, and keep every revision as its own version you can return to.
Generate, edit, shift and validate synchronised lyrics, then export LRC alongside the audio for players, karaoke video layers and distribution.
Title, tags, genre and description suggestions are proposed from the actual generation context rather than guessed from a filename.
Runs
Generation runs on Soneth-managed infrastructure through a durable queue. Progress belongs to the job, not to the browser window that started it.
The run reports the phases it is actually in — queued, model ready, rendering, post-processing, saved — with a separate technical diagnosis panel when something goes wrong.
Close the laptop, continue on the phone. Reconnect picks up live progress; cancel, cancel-all, retry and reset are first-class commands with clear failure diagnostics.
A cost preview appears before the paid submit. Budget warnings and a hard stop are enforced platform-side, and history shows provider, runtime and final cost per run.
Audition, compare, save or discard, and branch a follow-up run — all without losing the parent version or the session you were in.
Cover generation runs as an independent job. It can never block, slow or fail the audio it belongs to.
Every run stays searchable by prompt, engine, model, provider, cost and outcome, so a good result from three weeks ago is still findable.