正在加载视频...

视频加载失败

This is a big day. Meta is open-sourcing AudioCraft. You can now generate incredible music and sounds with a single prompt. It includes the most performant Generative AI Model (audio) on the market, the "Llama" of Audio. The research framework contains the weights and code of these models: ▸...

231,665 次观看 • 3 年前 •via X (Twitter)

10 条评论

Lior⚡ 的头像
Lior⚡3 年前

Article: Github: Congrats @jadecopet, @syhw, @jpineau1, @ylecun

Southrye - (₳) 的头像
Southrye - (₳)3 年前

Ive been putting it through its paces. Here is a video of my favourite generations.

Isak Westerlund 🦇🔊 的头像
Isak Westerlund 🦇🔊3 年前

Non-commercial license 🙄

Jason Inch 的头像
Jason Inch3 年前

Thanks for sharing this, you brought it to my attention last night and this morning I added some music and sound to my SD Controlnet demo vid. AudioCraft makes a huge difference in the visual impact to hear piano and effects....

Tibor Blaho 的头像
Tibor Blaho3 年前

I think it's amazing. I have shared the initial experiments with MusicGen and ChatGPT here:

Alexey 的头像
Alexey3 年前

Any platform I can run it on via API already?

Frank Baez 的头像
Frank Baez3 年前

What about the copyright on the generated audio, is it free?

Edison Ade 的头像
Edison Ade3 年前

@memdotai mem it

Mem 的头像
Mem3 年前

@AlphaSignalAI Saved! Here's the compiled thread: 🪄 AI-generated summary: "Meta has open-sourced AudioCraft, a research framework that includes the most performant Generative AI Model (audio) on the market. It allows users to generate incredible music...

mam niki 的头像
mam niki3 年前

@kargarisaac 😳😳😳😳👍👍👍👍

相关视频

🎥 Today we’re premiering Meta Movie Gen: the most advanced media foundation models to-date. Developed by AI research teams at Meta, Movie Gen delivers state-of-the-art results across a range of capabilities. We’re excited for the potential of this line of research to usher in entirely new possibilities for casual creators and creative professionals alike. More details and examples of what Movie Gen can do ➡️ 🛠️ Movie Gen models and capabilities Movie Gen Video: 30B parameter transformer model that can generate high-quality and high-definition images and videos from a single text prompt. Movie Gen Audio: A 13B parameter transformer model that can take a video input along with optional text prompts for controllability to generate high-fidelity audio synced to the video. It can generate ambient sound, instrumental background music and foley sound — delivering state-of-the-art results in audio quality, video-to-audio alignment and text-to-audio alignment. Precise video editing: Using a generated or existing video and accompanying text instructions as an input it can perform localized edits such as adding, removing or replacing elements — or global changes like background or style changes. Personalized videos: Using an image of a person and a text prompt, the model can generate a video with state-of-the-art results on character preservation and natural movement in video. We’re continuing to work closely with creative professionals from across the field to integrate their feedback as we work towards a potential release. We look forward to sharing more on this work and the creative possibilities it will enable in the future.

AI at Meta

2,265,850 次观看 • 1 年前