SFAI / MODEL PROFILE
Stable Audio 3 Medium Audio to Audio
Stable Audio 3 Medium audio-to-audio is a 1.4 billion parameter latent diffusion model that transforms an input audio clip into new stereo variations up to 6 minutes guided by a text prompt.
- Developer
- Unattributed (via fal)
- Input modalities
- audio
- Output modalities
- audio
- Context tokens
- Not reported
- Latest price observation
- 2026-09-28
- Lowest paired standard token quote
- Not available
Dated price and provider records
| Provider and channel | Input | Output or native rate | Unit and status | Source and observation |
|---|---|---|---|---|
| fal.aifal | Not token-priced | $0.0417 | audiospriced · USD | fal2026-09-28 |
These are dated, non-executable observations. Verify the exact endpoint, price and terms with the source before use. Unknown values are not free.