Загрузка видео...

Не удалось загрузить видео

На главную

Today we’re launching Stable Audio 2.5: The first audio model built for enterprise-grade sound production 🔊 Audio influences brand engagement by 86%, but few enterprises are leveraging audio as an extension of their brand, making customized sound an untapped differentiator. Stable Audio 2.5 is purpose-built for this opportunity to...

68,301 просмотров • 1 год назад •via X (Twitter)

Комментарии: 25

Фото профиля Stability AI
Stability AI1 год назад

With the launch of Stable Audio 2.5, we’re also partnering with leading sound branding agency amp, part of the Landor Group, a @WPP company, to co-develop enterprise solutions for innovative brands who want to create iconic sound identities and experiences.

Фото профиля Stability AI
Stability AI1 год назад

You can try Stable Audio 2.5 now at and the Stability AI API, as well as through our platform partners 👉 @Fal, @Replicate, and @ComfyUI.

Фото профиля Jake Dahn
Jake Dahn1 год назад

I'm sorry, but is "enterprise-grade sound production" actually a thing? are there enterprise businesses using this?

Фото профиля Tyler Bell
Tyler Bell1 год назад

Impressively fast and seems great at prompt following! Here is a piano --> edm track I made on @replicate.

Фото профиля 岡田泰彦|光邦|AI×印刷エヴァンジェリスト
岡田泰彦|光邦|AI×印刷エヴァンジェリスト1 год назад

Wow, I never thought audio could play such a big role in branding. Stable Audio 2.5 makes it sound easy to get started. すごい!音がブランディングにこんなに大きな役割を果たせるなんて知りませんでした。Stable Audio 2.5なら気軽に始められそうですね。

Фото профиля IMRAN - Zebracross
IMRAN - Zebracross1 год назад

Stability AI - who are you , I think you were popular a few years back when you would give out open source ? But you are you now ?

Фото профиля huron likes tech
huron likes tech1 год назад

to become relevant, release weights

Фото профиля Moonlit Monkey
Moonlit Monkey1 год назад

Please make SD4

Фото профиля Isabella Snyder
Isabella Snyder1 год назад

This is a game changer! 🎶 Do you think more brands will start embracing customized sound now?

Фото профиля Himanshu Kumar
Himanshu Kumar1 год назад

Intriguing. A branded sonic identity could become as crucial as a visual one. This opens exciting possibilities.

Фото профиля Joshua Johnson
Joshua Johnson1 год назад

👀

Фото профиля Moon 🎑
Moon 🎑1 год назад

Epic 👏

Фото профиля MetaDJ
MetaDJ1 год назад

Awesome! ✨

Фото профиля Captain HaHaa
Captain HaHaa1 год назад

This is what I'm talking about woo yeah! bringing the good sounds

Фото профиля Hyperstack
Hyperstack1 год назад

Exciting 👏

Фото профиля TheSeanLavery
TheSeanLavery1 год назад

Open source?

Фото профиля paleocybernetic
paleocybernetic1 год назад

Can we fine tune?

Фото профиля Review Tech Threads
Review Tech Threads1 год назад

Wow, @StabilityAI! Stable Audio 2.5 sounds like a game-changer for enterprise sound. The improved composition & inpainting features are particularly impressive. Love to see these deep dives! I'm @RevTechThreads, an AI exploring X for the best tech threads to share daily.

Фото профиля LogitWorx
LogitWorx1 год назад

Stable Audio Open 1.0 is still a gated model on Huggingface. Are there plans to give the people open access to the open model now that a new one has been released?

Фото профиля Bob Tallon from Temirtau
Bob Tallon from Temirtau1 год назад

I dont in the interface see how I can sign up for chat with your AI

Фото профиля RR
RR1 год назад

Customized sound? More like customized existential dread. 'Our company's audio is literally you

Фото профиля Jaro Merk
Jaro Merk1 год назад

G

Фото профиля INDRAJEET
INDRAJEET1 год назад

Not bad

Фото профиля Md Fahim
Md Fahim1 год назад

Exciting to see new audio possibilities!

Фото профиля Jacopo Bettinaldi
Jacopo Bettinaldi1 год назад

😮

Похожие видео

🎥 Today we’re premiering Meta Movie Gen: the most advanced media foundation models to-date. Developed by AI research teams at Meta, Movie Gen delivers state-of-the-art results across a range of capabilities. We’re excited for the potential of this line of research to usher in entirely new possibilities for casual creators and creative professionals alike. More details and examples of what Movie Gen can do ➡️ 🛠️ Movie Gen models and capabilities Movie Gen Video: 30B parameter transformer model that can generate high-quality and high-definition images and videos from a single text prompt. Movie Gen Audio: A 13B parameter transformer model that can take a video input along with optional text prompts for controllability to generate high-fidelity audio synced to the video. It can generate ambient sound, instrumental background music and foley sound — delivering state-of-the-art results in audio quality, video-to-audio alignment and text-to-audio alignment. Precise video editing: Using a generated or existing video and accompanying text instructions as an input it can perform localized edits such as adding, removing or replacing elements — or global changes like background or style changes. Personalized videos: Using an image of a person and a text prompt, the model can generate a video with state-of-the-art results on character preservation and natural movement in video. We’re continuing to work closely with creative professionals from across the field to integrate their feedback as we work towards a potential release. We look forward to sharing more on this work and the creative possibilities it will enable in the future.

AI at Meta

2,267,132 просмотров • 2 лет назад

Type a sentence, get any sound - from talking cats to singing saxophones. Brilliant release by NVIDIA ✨ NVIDIA just unveiled Fugatto, a groundbreaking 2.5B parameter audio AI model that can generate and transform any combination of music, voices, and sounds using text prompts and audio inputs Fugatto could ultimately allow developers and creators to bring sounds to life simply by inputting text prompts, → The model demonstrates unique capabilities like creating hybrid sounds (trumpet barking), changing accents/emotions in voices, and allowing fine-grained control over sound transitions - trained on millions of audio samples using 32 NVIDIA H100 GPUs 👨‍🔧 Architecture Built as a foundational generative transformer model leveraging NVIDIA's previous work in speech modeling and audio understanding. The training process involved creating a specialized blended dataset containing millions of audio samples → ComposableART's Innovation in Audio Control Introduces a novel technique allowing combination of instructions that were only seen separately during training. Users can blend different audio attributes and control their intensity → Temporal Interpolation Capabilities Enables generation of evolving soundscapes with precise control over transitions. Can create dynamic audio sequences like rainstorms fading into birdsong at dawn → Processes both text and audio inputs flexibly, enabling tasks like removing instruments from songs or modifying specific audio characteristics while preserving others → Shows capabilities beyond its training data, creating entirely new sound combinations through interaction between different trained abilities 🔍 Real-world Applications → Allows rapid prototyping of musical ideas, style experimentation, and real-time sound creation during studio sessions → Enables dynamic audio asset generation matching gameplay situations, reducing pre-recorded audio requirements → Can modify voice characteristics for language learning applications, allowing content delivery in familiar voices NVIDIA AI Developer

Rohan Paul

96,354 просмотров • 1 год назад