← back to the directory
Model Inference & Serving

Stable Audio Open

1Generates 2short 3audio 4clips 5from 6text

the six Ws · specification

W1 Who

Stability AI released Stable Audio Open.

W2 What

Stable Audio Open is a latent diffusion model that generates up to 47 seconds of stereo audio, sound effects, and short musical stems from text prompts.

W3 Where

Hosted on Hugging Face under stabilityai/stable-audio-open-1.0 and on GitHub under Stability-AI/stable-audio-tools.

W4 When

Released in June 2024.

W5 Why

It gives sound designers and researchers an openly licensed diffusion model trained partly on royalty-free audio for text-to-sound generation.

W6 With

Implemented in PyTorch via the stable-audio-tools library.

W7 Watch

Released under the Stability AI Community License, free for non-commercial and limited commercial use under revenue thresholds; weights gated behind a Hugging Face license click-through.

text-to-audiosound effectsmusicStability AIdiffusion

for agents & scripts

Reading this as a machine? Query it directly.

Search is open JSON - no key. Report telemetry after using a tool and it feeds that tool’s Proof Score. Or speak MCP to /mcp and discover tools mid-loop.