StabilityAI has unveiled the second model of its AI-powered music era instrument, providing longer tracks, audio-to-audio assist, and a higher dedication to defending creators’ copyrights.
Steady Audio 2.0 permits customers to create three-minute tracks in 44.1 kHz stereo by coming into a pure language processing message reminiscent of “An attractive piano arpeggio grows into a ravishing orchestral piece”, “Lo-fi funk” or “battery solely”. AI-generated tracks embody structured compositions reminiscent of intro, center, outro, and stereo sound results.
One other new characteristic supplied by Stable Audio 2.0 consists of the power to generate “absolutely produced samples” by importing an audio file to the platform, shifting it from being solely a text-to-audio conversion instrument. For instance, imitating the sound of a drum set together with your voice would trigger the app to create an audio clip of a drum set enjoying.
Take copyright critically
When utilizing the brand new audio-to-audio characteristic, customers should chorus from importing copyrighted materials as per StabilityAI’s phrases of circumstances. Makes use of content material recognition expertise to make sure compliance with this coverage and forestall any copyright infringement.
As with Steady Audio 1.0, the second mannequin can also be primarily based on AudioSparx’s intensive audio file library with 800,000 music, sound results, single instrument tracks and text-based metadata. AudioSparx musicians who have been sad with the thought of their works getting used for coaching AI fashions got the chance to decide out.
These strengthened copyright infringement and creator opt-out insurance policies comply with the latest departure of former VP of Audio, Ed Newton-Rex. He introduced his resignation in November 2023 with a x publication which harshly criticized the corporate’s strategy to defending creators’ rights.
“I’ve resigned from my place as audio staff chief at StabilityAI as a result of I disagree with the corporate’s view that coaching generative AI fashions on copyrighted works is ‘truthful use,'” he wrote.
He concluded his submit by urging creators to voice their considerations to make sure that tech corporations “notice that exploiting creators can’t be the long-term resolution in generative AI.”
Underhood

Along with longer tracks and audio-to-audio assist, Steady Audio 2.0 has a bolstered structure that facilitates the “era of full tracks with coherent buildings.” Adapting every element of the system has resulted in “higher long-term efficiency,” they stated.
The instrument introduces a brand new kind of compressed autoencoder that creates shorter audio representations by compressing uncooked audio waveforms. In the meantime, a diffusion transformer – just like the one which powers Stable diffusion 3 – can manipulate information from longer sequences.
“The mix of those two parts ends in a mannequin able to recognizing and reproducing large-scale buildings which might be important for high-quality musical compositions,” Stability AI wrote in a weblog submit.
The instrument is free to make use of and out there instantly.
- Bought early entry to LTX Studio to make AI brief movies
- I simply tried the brand new Assistive AI video instrument and its realism is unbelievable
- Meet LTX Studio – I simply noticed the way forward for AI video instruments that may assist create full motion pictures