Suno: the platform that pulls the second half of music production inside one product
suno
The most complete commercial platform on the text-to-song route: a style description, your own lyrics, a hummed melody or a recorded riff all work as a starting point, a full song with vocals and instrumentation arrives in seconds, and the same product then extends it, edits sections, restyles it, extracts stems and remasters it. We list it as the top closed entry for music in the audio domain on workflow completeness rather than peak audio quality. Most generative music products cover drafting and picking a take and stop there; Suno connects the rest through stem extraction with up to 12 stems, MIDI export and Suno Studio, and Studio speaks conventional DAW semantics rather than adding another prompt box, with a multitrack timeline, take lanes, comping, manual BPM to settle tempo drift and per clip transpose and speed. Custom Models trains up to three private style variants from six or more tracks you own, and Voices, formerly Personas, generates in your own singing timbre with a verification step. The tier split matters: the free plan covers creation (generation, lyrics, Cover, crop and fade, audio upload) while stems, Add Vocals, Voices and Custom Models require Pro or Premier and Studio is Premier only on desktop web, so real cost modelling should assume Premier. Version numbers do not track quality on third party benchmarks either: on WildSongBench v5 scores 6.8721, above v6 at 6.5562 and v6 Wild at 6.4195, with v4.5 at 6.6995 and v5.5 at 6.7150, and the vendor itself asks for capability based rather than version based description. Limits: closed with no self-hosting, so unreleased melodies and lyrics must be uploaded; no stable version semantics, meaning a regeneration can sound different after a model update; commercial rights follow the tier; control granularity sits at section and style level with no editable chord track, so theory level edits require Studio multitrack re-arrangement or an open model such as YuE2 that exposes an ABC score. Graded C (vendor claim); the benchmark numbers are submitted by m-a-p and have not been recomputed by us.
- CONFIDENCE
- Vendor Claim
- Official model card or keynote only, no independent re-test
- KEY METRIC
- WildSongBench SongBench 均分(v5,第三方测)
- Vendor Claim · 2026-09
- MATURITY
- Production
- research → demo → product → production
Our takeWe grade it C (vendor claim). The specification class facts, meaning the feature list, tier split, stem count and export formats, come from the official llms.txt dated 2026-05-29 and are verifiable but still single party statements. The WildSongBench numbers come from a third party public benchmark run by m-a-p and have not been recomputed by us, and we have run no blind listening test. The whole asset therefore carries C, and fame as the best known music product is not a reason to upgrade it.
What earns the top of the ladder is workflow completeness, not peak audio quality. Real music production runs draft, pick a take, edit sections, pull stems, re-arrange in your own DAW, master. Most generative music products cover the first two steps and stop, leaving an unusable audio file. Suno connects the rest through stem extraction with up to 12 stems, MIDI export and Suno Studio, and Studio speaks conventional DAW semantics rather than adding another prompt box: multitrack timeline, take lanes, comping, manual BPM. Manual BPM is the detail that shows someone has done real arrangement work, because tempo drift is the main reason AI material cannot enter a serious project. Custom Models, six or more owned tracks training up to three private variants, solves the other real problem, which is that style consistency cannot be maintained by rewriting a prompt every time.
Three points govern selection. First, newest is not best in this domain. On third party benchmarks v5 at 6.8721 still scores above v6 at 6.5562 and v6 Wild at 6.4195, and the vendor itself asks for capability based rather than version based description, so a version number is not a quality promise and the decision has to be made by audition. Second, control granularity sits at the section and style level, not the music theory level. There is no editable chord track and no bar level refinement, so teams needing theory level edits either re-arrange through the Studio multitrack or move to an open model such as YuE2 that exposes an ABC score as an artefact. Third, closed and not self-hostable, with commercial rights tied to tier. Unreleased melodies and lyrics have to be uploaded, which is a hard barrier where data sovereignty matters, and stems, Voices, Custom Models and Studio all sit behind paid tiers. Since those are exactly what distinguishes Suno from an ordinary generator, real cost modelling should assume the Premier tier rather than the free trial. Against Eleven Music v2.5, already in this catalogue, the Suno difference is the workstation and private style models, while the Eleven Music difference is integration with one API stack covering speech, transcription and music.
The problem it solves: productising everything between a one line idea and a deliverable master
Suno is the most complete commercial platform on the text-to-song route. A style description, your own lyrics, a hummed melody or a recorded riff all work as a starting point, a full song with vocals and instrumentation arrives in seconds, and then the same product extends it, edits sections, restyles it, extracts stems and remasters it. Its significance is not that it can generate songs, since both open and closed models can, but that the second half of music production is pulled inside one product. That is why this site lists it as the top closed entry for music in the audio domain.
The official material deliberately avoids version numbers (models and features ship often, so describe Suno by what it can do rather than by a version), but third party benchmarks quantify the versions: on WildSongBench, v5 scores 6.8721 on the SongBench average, v4.5 6.6995, v5.5 6.7150, v6 6.5562 and v6 Wild 6.4195. That set deserves a separate read because a newer version does not mean a higher benchmark score. On this ruler v5 remains the strongest in the family while v6 and v6 Wild sit lower. New releases usually optimise dimensions the benchmark does not measure, such as audio quality, dynamics and controllability, so newest is not automatically best in this domain and selection should be done by audition against the actual use case.
Tiering: the free tier covers creation, the paid tiers unlock production
| Capability | Tier | Notes |
|---|---|---|
| Generation / lyrics / Cover / crop and fade / audio upload | Free | Core creation is free; Remaster can rebuild the mix of a song made with any model |
| Stem extraction, up to 12 stems | Pro / Premier | Vocals, drums, bass and more can be soloed, muted and downloaded as MP3, WAV, tempo-locked WAV or MIDI |
| Add Vocals / Add Instrumental | Pro / Premier | Layer custom vocals over an instrumental from your lyrics, or build instrumentation around a vocal idea |
| Voices (beta, formerly Personas) | Pro / Premier | Record or upload a short sample of your own singing and generate in your timbre, with a verification step confirming the voice is yours |
| Custom Models | Pro / Premier | Upload six or more original tracks you own the rights to and train up to three private variants tuned to your sound |
| Suno Studio (GAW) | Premier only, desktop web only | A generative audio workstation: multitrack timeline, AI stems that sit inside an existing arrangement, mic recording and transformation, loop recording with take lanes and comping, manual BPM to settle tempo drift, EQ plus per clip transpose and speed, multitrack bounce export |
The last two rows are the ones to read. Custom Models and Studio turn Suno from a generator into a workstation. The first makes style controllable and private instead of re-described in every prompt; the second brings the multitrack and comping semantics of a conventional DAW while letting an AI generated stem be arranged on a timeline like any sample. Manual BPM deserves particular mention: tempo drift is the main reason generated music fails to enter a real arrangement, and locking a project to a steady tempo before export is what reconnects AI material to a DAW workflow.
Controls: three sliders and one prompt expansion
The Creative Sliders expose three continuous quantities, Weirdness (Safe to Chaos), Style Influence (Loose to Strong) and Audio Influence when working from an upload, plus a one tap prompt enhancement that expands a rough idea into a full style description. Extend adds music at the start, middle or end, with Quick Extend followed by Get Whole Song; the Song Editor allows crop and fade on any plan while replacing or adding a section inside a highlighted region belongs to advanced editing. The characteristic of this control surface is that its granularity sits at the level of sections and style, not of music theory: there is no chord track to edit.
Limits: no self-hosting, unstable version semantics, rights tied to tier
- Closed, no weights, no self-hosting path. Every render passes through the Suno service, so data sovereignty cases such as unreleased melodies and lyrics have to be assessed on that basis. The self-hostable alternatives in this domain are YuE2 for music (non-commercial licence) and VoiceStudio for speech.
- No stable model version semantics. The vendor asks for capability based description, and regenerating the same song at a later date can sound different after a model update. Projects needing long term timbre consistency should either commit to Voices and Custom Models or export stems and MIDI into their own DAW.
- Commercial rights follow the subscription tier. Free and paid plans carry different rights over generated content, and the authoritative text is the official pricing page. This site does not restate prices or rights figures, and the vendor asks that figures not be quoted.
- Benchmark scores are third party measurements of delivered audio. The WildSongBench numbers are submitted and evaluated by m-a-p and have not been recomputed by us, and they measure a delivered render rather than the human work done inside Studio.