Skip to content
Reduced motion is on. Play a preview to watch it.
Explore models

Audio

Stable Audio 3 Medium Base

Create, transform, repair or extend music and sound with Stable Audio 3 Medium Base.

  • Text to audio
Create with Stable Audio 3 Medium Base
Preview for Stable Audio 3 Medium Base

Catalog preview for Stable Audio 3 Medium Base. Review your own result before publishing.

A starting point.

Create, transform, repair or extend music and sound with Stable Audio 3 Medium Base.

Before you publish.

Listen to the complete output, including edit boundaries, stereo balance, noise and the ending.

Review the quote before running. A failed generation releases its reserved credits. Download the results you want to keep; stored media is retained for 30 days.

How credits work · Use the API · Content guidelines

Settings and model questionsCheck the available settings and guidance when preparing your input.

Inputs and settings

Text mode takes prompt and duration from 1 to 380 seconds. Audio-to-audio also takes an uploaded audio_url and init_noise_level.

Inpainting replaces only mask_start_seconds to mask_end_seconds; the end must follow the start and stay within the uploaded source.

Outpainting needs a positive extend_seconds_before or extend_seconds_after. Raywake checks that the source and extensions fit within 380 seconds.

Choose MP3, WAV, FLAC, OGG, Opus, M4A or AAC. Bitrate applies to compressed formats; base checkpoints expose guidance and more sampling steps.

Inputs, all modes and APIModel IDs, current input limits and the API generation workflow.

Supported inputs

These fields come from the current studio form for stable-audio-3/medium/base/text-to-audio. Other modes may require different inputs.

Input types, requirements and limits for stable-audio-3/medium/base/text-to-audio
InputTypeRequirementOptions and limits
promptTextRequired—
bitrateTextOptional—
durationNumberOptional≥ 1 · ≤ 380
enable_prompt_expansionOn/offOptional—
enable_safety_checkerOn/offOptional—
guidance_scaleNumberOptional≥ 0 · ≤ 25
negative_promptTextOptional—
num_inference_stepsWhole numberOptional≥ 1 · ≤ 100
output_formatTextOptionalmp3, wav, flac, ogg, opus, m4a, aac
seedWhole numberOptional—

Check the full model schema for all options and cross-field rules. Request a quote for your selected settings before generating.

Other modes and their inputs

Open the operation you need and check its inputs. Settings from one mode do not automatically apply to another.

Audio To Audiostable-audio-3/medium/base/audio-to-audio

Supported inputs

These fields come from the current studio form for stable-audio-3/medium/base/audio-to-audio. Other modes may require different inputs.

Input types, requirements and limits for stable-audio-3/medium/base/audio-to-audio
InputTypeRequirementOptions and limits
audio_urlTextRequired—
promptTextRequired—
bitrateTextOptional—
durationNumberOptional≥ 1 · ≤ 380
enable_prompt_expansionOn/offOptional—
enable_safety_checkerOn/offOptional—
guidance_scaleNumberOptional≥ 0 · ≤ 25
init_noise_levelNumberOptional≥ 0 · ≤ 1
negative_promptTextOptional—
num_inference_stepsWhole numberOptional≥ 1 · ≤ 100
output_formatTextOptionalmp3, wav, flac, ogg, opus, m4a, aac
seedWhole numberOptional—

Check the full model schema for all options and cross-field rules. Request a quote for your selected settings before generating.

Open this operation in the studio

Audio Inpaintingstable-audio-3/medium/base/audio-inpainting

Supported inputs

These fields come from the current studio form for stable-audio-3/medium/base/audio-inpainting. Other modes may require different inputs.

Input types, requirements and limits for stable-audio-3/medium/base/audio-inpainting
InputTypeRequirementOptions and limits
audio_urlTextRequired—
mask_end_secondsNumberRequired≥ 0 · ≤ 380
mask_start_secondsNumberRequired≥ 0 · ≤ 380
promptTextRequired—
bitrateTextOptional—
enable_prompt_expansionOn/offOptional—
enable_safety_checkerOn/offOptional—
guidance_scaleNumberOptional≥ 0 · ≤ 25
negative_promptTextOptional—
num_inference_stepsWhole numberOptional≥ 1 · ≤ 100
output_formatTextOptionalmp3, wav, flac, ogg, opus, m4a, aac
seedWhole numberOptional—

Check the full model schema for all options and cross-field rules. Request a quote for your selected settings before generating.

Open this operation in the studio

Audio Outpaintingstable-audio-3/medium/base/audio-outpainting

Supported inputs

These fields come from the current studio form for stable-audio-3/medium/base/audio-outpainting. Other modes may require different inputs.

Input types, requirements and limits for stable-audio-3/medium/base/audio-outpainting
InputTypeRequirementOptions and limits
audio_urlTextRequired—
promptTextRequired—
bitrateTextOptional—
enable_prompt_expansionOn/offOptional—
enable_safety_checkerOn/offOptional—
extend_seconds_afterNumberOptional≥ 0 · ≤ 380
extend_seconds_beforeNumberOptional≥ 0 · ≤ 380
guidance_scaleNumberOptional≥ 0 · ≤ 25
negative_promptTextOptional—
num_inference_stepsWhole numberOptional≥ 1 · ≤ 100
output_formatTextOptionalmp3, wav, flac, ogg, opus, m4a, aac
seedWhole numberOptional—

Check the full model schema for all options and cross-field rules. Request a quote for your selected settings before generating.

Open this operation in the studio

Use Stable Audio 3 Medium Base through the API.

Start with this model ID for the endpoint featured on this page:

stable-audio-3/medium/base/text-to-audio

Input requirements belong to this model ID. Check a different mode’s schema separately when switching modes.

Available modes

Check the current model schema for supported inputs and options. Other tasks may use different model IDs.

From quote to result.

Request a credit quote for your input and settings before submitting a generation. A sample output or the price of another model is not a quote for your request.

Follow the generation workflow, then poll the job for its status and result. The API and studio use the same balance.