S
STABILITY AI / stable-audio-3.0

Stable Audio 3.0

Commercial model for generating songs, instruments, and effects from text or audio, with inpainting and style transformation.

01 / SUMMARY

The essential

ORIENTATIONText and audio to music

Primary use stated or inferred cautiously from official documentation.

CONTEXTPrompt and reference audio

Maximum input capacity when published by the source.

EXITUp to 6 minutes of audio

Documented maximum generation limit.

ACCESSStability API

Verified availability channels.

MODEL TYPEMusic · Voice and audio

Primary functional family and verified specialties.

LICENSEProprietary

Declared terms for API access, model weights, or self-hosting.

02 / TECHNICAL SHEET

Limits and integration

API IDstable-audio-3.0
Model typeMusic · Voice and audio
Access modelPaid proprietary
LicenseProprietary
DeploymentHosted API
Release2026
Knowledge cutoffNot published
EntranceText · Audio
ExitAudio · Music · Effects
Context windowPrompt and reference audio
maximum outputUp to 6 minutes of audio
ReasoningNot applicable or not published
Published tools
Structured OutingsNot applicable or not published
Batch processingNot published
Prompt cacheNot published
Fine-tuningNot published
Verified platformsStability API
BEST SUITED FOR

Music, effects, and generative audio editing.

WORTH MONITORING

Review duration, output license, musical coherence, and cost per iteration.

03 / CAPABILITIES

What it can do

01

Orientation

Music, effects, and generative audio editing.

02

Context

Prompt and reference audio · Up to 6 minutes of audio

03

Tools and integration

Not published

04

Access

Stability API

Input modalities
TextAudio
04 / PRICES

Documented cost

ConceptWorthUnit/condition
Standard input26 credits per generationStandard API rate
Cached inputNot publishedReading reused prefixes
Cache write or storageNot publishedThe condition varies by provider
Standard outputNot publishedMay include reasoning tokens
Batch inputNot publishedAsynchronous processing
Batch outputNot publishedAsynchronous processing

Consult the primary source: the billing unit depends on the model type and access channel.

i Prices change and may depend on level, region or context length. Check the source before making a decision.

05 / EVALUATIONS

How to read the results

“

A public benchmark provides guidance, but does not replace an evaluation with your data, tools, budget, and error tolerance.

BenchmarkResultMetricSource
Evaluación Stable Audio 3Up to 6 minPublished duration and coherenceView source ↗

Inferama only highlights a “best result” when the metric, test set, configuration, and date allow for an equivalent comparison. The supplier's figures are presented as claims from its own source.

06 / VERSIONS

Chronology

Stable Audio 3.0

Version added to Inferama's verified catalog.

ANALYSIS

Related analysis

07 / SOURCES

Traceability