StepAudio 3 Music API

Writes complete songs from a style description, with vocals, arrangement, and mixing, or instrumentals and covers.

StepFunAudio GenerationReleased Sep 8, 2026InternationalProprietary EndpointNew

About StepAudio 3 Music

Writes complete songs from a style description, with vocals, arrangement, and mixing, or instrumentals and covers.

Generation is asynchronous and usually takes from tens of seconds to a few minutes.

Also known as StepFun StepAudio 3 Music

music generationaudio generationaudio inmultilingual

StepAudio 3 Music specs

Model ID
stepaudio-3-music
Author
StepFun
Category
Audio Generation
Released
Sep 8, 2026
Input
TextAudio
Output
Audio
Region
International
Endpoints
POST/v1/audio/generations
Alternate model IDs
stepaudio-3-music-previewstepfun/stepaudio-3-music

StepAudio 3 Music API pricing

Live pay-as-you-go rates from the EmpirioLabs catalog. You are billed only for what you use, with no monthly minimum.

Type
Spec
Rate
Music generation
per generated second
$0.0018
Compare on the full pricing page

How to call the StepAudio 3 Music API

StepAudio 3 Music runs through POST /v1/audio/generations. The request returns a job_id right away; poll GET /v1/jobs/{job_id} until the job completes and read the output URLs from the result. Get an API key from the EmpirioLabs dashboard.

cURL: submit the job
curl https://api.empiriolabs.ai/v1/audio/generations \
  -H "Authorization: Bearer $EMPIRIOLABS_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "stepaudio-3-music",
    "prompt": "Describe what you want StepAudio 3 Music to generate."
  }'
cURL: poll for the result
curl https://api.empiriolabs.ai/v1/jobs/JOB_ID \
  -H "Authorization: Bearer $EMPIRIOLABS_API_KEY"
Python
import requests

response = requests.post(
    "https://api.empiriolabs.ai/v1/audio/generations",
    headers={"Authorization": "Bearer YOUR_EMPIRIOLABS_API_KEY"},
    json={
        "model": "stepaudio-3-music",
        "prompt": "Describe what you want StepAudio 3 Music to generate.",
    },
)
job = response.json()

# Generation runs as an async job. Poll until it completes.
import time
while True:
    status = requests.get(
        f"https://api.empiriolabs.ai/v1/jobs/{job['job_id']}",
        headers={"Authorization": "Bearer YOUR_EMPIRIOLABS_API_KEY"},
    ).json()
    if status.get("status") in ("completed", "failed"):
        print(status)
        break
    time.sleep(5)
Full StepAudio 3 Music API reference

StepAudio 3 Music API parameters

Request parameters supported by the StepAudio 3 Music API on EmpirioLabs. Defaults apply when a field is omitted.

ParameterTypeDefaultRange / valuesDescription
promptstring--Style description: genre, vocals, mood, instruments, and key.
lyricsstring--Optional lyrics. Section tags such as [Verse], [Chorus], and [Outro] shape the structure. Left empty, the model writes them. Required for a cover or a vocal arrangement.
taskenumtext_to_musictext_to_music, music_cover, vocal_to_musicWrite a new song, cover a reference song, or build an arrangement around a dry vocal.
instrumentalbooleanfalse-Generate a track with no vocals. Cannot be combined with lyrics.
song_audiostring--Base64 reference song. Required for a cover.
vocal_audiostring--Base64 dry vocal with no accompaniment. Required for a vocal arrangement.
response_formatenummp3mp3, wav, flac, opus, pcmOutput audio format.
temperaturenumber0.850.01 to 2Sampling temperature. Higher is more varied.
top_pnumber0.920 to 1Nucleus sampling probability mass.
lyrics_rewritebooleanfalse-Let the model rewrite the lyrics you supplied.

Good to know

Describe the style and the model writes and performs a complete song. Supply lyrics or let the model write them, set instrumental for a track with no vocals, or attach a reference song for a cover or a dry vocal to build an arrangement around. Section tags such as [Verse], [Chorus], and [Outro] shape the structure. Output length is chosen by the model and is typically one to three minutes, so it is not set by a parameter and two runs of the same prompt can differ. Billing follows the generated length. Instrumental and lyrics cannot be combined. This model is in preview and its capabilities may change.

StepAudio 3 Music API: common questions

How much does the StepAudio 3 Music API cost?

On EmpirioLabs, StepAudio 3 Music is billed pay as you go: Music generation $0.0018 per generated second. The live rate card on this page always matches what the API charges.

Which endpoint does StepAudio 3 Music use?

StepAudio 3 Music is served through POST /v1/audio/generations on api.empiriolabs.ai with standard bearer-token authentication.

Can I try StepAudio 3 Music in the browser before integrating?

Yes. The EmpirioLabs playground runs StepAudio 3 Music in the browser with the same parameters the API exposes, so you can test prompts before writing code.

How do I get a StepAudio 3 Music API key?

Create an EmpirioLabs account, then generate a key under API Keys in the dashboard. Billing is pay-as-you-go credits, so you only pay for the requests you make.

Ready to use better endpoints?

Check out our pricing or reach out if you want your own model deployed on our stack.