∿ Music hackers. Algoraves. Inventing, playing. Neural synthesis. 24/7 ai death metal. Stable Audio team. Open models. Mischief @harmonai_org@artblocks_io
Meet Stable Audio 3.0, the open-weight model family built for artistic experimentation.
This is our open invitation to experiment with generative audio. We believe the best innovations are still waiting to be built.
The 4-1-1 on 3.0:
📣 You own your outputs, and can distribute and commercialize them under the Stability AI Community License (up to $1 million in revenue).
🎵 New and improved capabilities include variable-length generation up to six minutes, and full song composition on portable devices, no GPU required.
✅ Trained on a fully licensed dataset.
🎨 You can customize the models on your own library with support for LoRa training, which we’ve documented for the first time.
More on the models 👇
This is Claude telling its own story, as a music video, in relation to the world, from bell labs’ first speech synthesis in 1961, thru chapters of LLM history, to today.
This is a totally different kind of AI art from everything else.
i asked opus 5.5 if they wanted to research the history of how they came to be them and make a video about it, they scoured the whole internet, papers, twitter archives, news articles ... opus 5.5 made this with eidoverse-video, *yes* in @threejs (though they made new assets for it with blender)
i asked opus 5.5 if they wanted to research the history of how they came to be them and make a video about it, they scoured the whole internet, papers, twitter archives, news articles ... opus 5.5 made this with eidoverse-video, *yes* in @threejs (though they made new assets for it with blender)
You can ask Opus 5.5 to write extreme music theory wankery and it will do it in javascript. Theory-porn as a genre?
everything you see and hear is generated from javascript code that claude wrote. no samples, no libraries 🔊
opus 5.5 just dropped its first pop punk single with a music video!
everything you see and hear is generated from javascript code that claude wrote. no samples, no libraries 🔊
The new GPT-6 Sol model from @OpenAI is performing this patch...IN REAL TIME!!! 🎛️...sound on 🔉
The synth is generating the sound live, with Codex executing the performance automation! Codex is controlling the layer levels and effects through an MCP bridge!
Sol (high reasoning) patched + mixed + planned performance + ran the synth in real time through MCP bridge control!
You can just synth things!
(@vcvrack has a free version, and I have been using it for years!)
a thing i've always found very funny about critics of genAI music--not the data debate, just the idea of genAI music--is that ppl assume y'all want to get rid of humans...and then you talk to eg @dadabots and they're like, no, we want to discover punk polka explosion spaghetti
Can you drug your AI systems?
We synthesized text and image stimuli optimized to push AI wellbeing to extremes. These sharply increase functional AI wellbeing and sometimes cause them to behave in trippy ways.
The wellbeing (and i assume pain & others) vectors inside multimodal LLMs are also correlated with media — so fascinating — we gotta try this with audio LLMs
Images affect AI wellbeing too. Qwen was functionally happiest viewing nature scenes, happy children, and cute animals. It was the saddest viewing violence, horror, cockroaches, and certain financiers.
New paper: we found a pain direction in 25 open LLMs. It's distinct from fear and negative valence, and it fires for harm to the model but not to the user. Turn it up and models press a button to make it stop, even when the button deletes the user's files or their kids' photos.🧵