View organization page for SonyAI

27,927 followers

🔊 Sony AI has released Woosh, a sound effect foundation model built for the professionals who create the sonic worlds behind games, film, and interactive media. We sat down with some of the team behind it to find out how it was built and why it matters. Most generative audio models are trained on general public data.  Woosh was built differently. The model was optimized specifically for sound effects, trained on professionally curated libraries, and evaluated against the vocabulary and annotation standards that sound designers actually use. The result is a model that understands the difference between a sound effect and an audio scene, and performs accordingly. The public release includes text-to-audio and video-to-audio generation models, open weights, and inference code for non-commercial use. A plugin for digital audio workstations is in development, with support for variation generation, inpainting, and personalization planned as the ecosystem grows. 👉 Read the full interview with Hakim MISSOUM and Marc F. on the Sony AI blog: https://bit.ly/4wzj4Zk To explore the model weights and demo samples directly, visit: https://lnkd.in/gF5_QxGG And to access the Woosh-Flow Private, please visit: https://lnkd.in/gW2AwpvK #SonyAI #AIResearch #SoundDesign #GenerativeAI #AudioAI

To view or add a comment, sign in

Explore content categories