Observed Signal · May 20, 2026 · Technical Release · Source: techcrunch · Impact: 3/5 · Sentiment: Positive
Stability AI launches Stable Audio 3.0 models
Stability AI has released Stable Audio 3.0, a family of four audio generation models that range from 459 million to 2.7 billion parameters. The medium (1.4B) and large (2.7B) models can produce composed music up to 6 minutes 20 seconds long while maintaining musical structure and melody — more than double the length of prior versions. Stability AI is publishing open weights for the small SFX, small, and medium models; the large model is available only via API and paid self-hosting, with companies over $1M in revenue required to obtain an enterprise license. The company says the new models were trained on fully licensed data and has hired Ethan Kaplan to lead its professional music product efforts. The release arrives amid broader industry activity in AI music generation and ongoing legal and licensing disputes involving other firms in the space.
The release advances generative-audio capabilities (longer, higher-quality compositions) and expands access via open weights, affecting creative tooling and potential commercial audio content; licensing and enterprise restrictions also shape commercial adoption and rights management.
Track Warner Music Group Signals & Market Shifts in Real-Time
Polaris7 autonomous intelligence agents track regulatory filings, primary sources, executive changes, and deal flow 24/7. Create your free Explorer workspace to monitor these entities.
Key Takeaways & Evidence Grounding
- Stability AI released Stable Audio 3.0, a family of four audio models (small SFX, small, medium, large).
- Model parameter counts: small SFX 459M, small 459M, medium 1.4B, large 2.7B parameters.
- Medium and large models can generate full musical compositions up to 6 minutes 20 seconds long.
- Small SFX, small, and medium models are being released with open weights; the large model is available only via API and paid self-hosting and requires an enterprise license for companies with >$1M revenue.
- Stability AI says the Stable Audio 3.0 models were built on fully licensed data and has hired Ethan Kaplan to lead its professional music offering.
Connected Companies & Entities
4 Entities mappedOntology Mapping & Concepts
Related Market Signals & Shifts
Recent verified developments and strategic activity across this market segment.
Stability AI Raises $76M Series B
Stability AI, the company behind the Stable Diffusion image-generator, raised $76 million in a Series B round, bringing its total fundraising to $232 million. The company said the round includes strategic investment from entertainment and gaming companies — Universal Music Group, Sony Music Group, Warner Music Group, and Electronic Arts — alongside investors AMD Ventures and Pacific Alliance Ventures. Stability plans to use the funds to expand its "creative production" product suite and professional services. The company has been building partnerships with major entertainment firms and has recently seen favourable legal outcomes in a UK copyright case, while related U.S. litigation remains pending.
Sean Parker Rebuilds Stability AI Around Music
Sean Parker, co-founder of Napster, is pivoting Stability AI towards the music industry. Two years after joining an $80 million rescue of the company, Parker and CEO Prem Akkaraju are focusing on developing AI tools for music professionals. In late August, Stability AI raised $76 million from investors including Sony, Warner, and Universal, who also licensed their catalogs for training. The company has since released three new audio models and AI music-editing software that can generate instrumental tracks or snippets from text prompts. An upcoming update will allow users to hum melodies or beatbox drum patterns to guide the AI.
ElevenLabs launches Music v2 with mid-track genre switching
ElevenLabs released Music v2, an updated AI music-generation model that can switch genres mid-track, handle complex vocals and compositions, and add non-musical sound effects. The model lets creators generate and edit songs by sections (intro, verse, chorus) and stitch those sections together, and it can recreate specific parts of a track from prompts without altering other parts. ElevenLabs says Music v2 performs more reliably across languages, lyrics, vocals and arrangements, is trained on licensed data, and is cleared for commercial use. The model is available via ElevenLabs’ ElevenCreative tool and ElevenMusic platform, with ElevenAPI access coming soon.
Track Real-Time Market Signals & Shifts
Set up custom watchlists to receive automated, evidence-grounded executive digests whenever material signals or shifts occur across your tracked landscape.
