Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

We’re excited to release ACE-Step / ACE-Step-v1-3.5B, a fast, versatile DiT-based foundation model for music generation that runs on consumer-grade GPUs. With its simple architecture and low hardware requirements, it’s easy to fine-tune for various music tasks, empowering, not replacing, artists and creators. Think of it as a step...

112,561 Aufrufe • vor 1 Jahr •via X (Twitter)

9 Kommentare

Profilbild von Wekulu
Wekuluvor 1 Jahr

the support that ace had in the past was only possible with the community that built you- real artists, real singers, real musicians. this is is more than the axe to the tree, its the shovel to the roots. to replace them with ai is to kill the soul that made your company possible

Profilbild von keyesgen 💜🔮
keyesgen 💜🔮vor 1 Jahr

hi - can you explain what you mean by authorized and purchased data? what data was it trained on? did the people you purchase it from have full awareness of its use case?

Profilbild von UNPLUGGED PERFORMANCE
UNPLUGGED PERFORMANCEvor 1 Jahr

Upgrade your Tesla with UP-03 Forged Wheels from Unplugged Performance! Unmatched strength, lightweight design, and track-proven durability—perfect for Model S, 3, X, and Y. Ready to ship with a lifetime warranty. #Tesla #UP03 #UnpluggedPerformance

Profilbild von peartree39
peartree39vor 1 Jahr

Ewwwwww🤮

Profilbild von Isaac Bratzel
Isaac Bratzelvor 1 Jahr

Let’s go 🔥🔥🔥

Profilbild von Divine Devinn🎶꩜
Divine Devinn🎶꩜vor 1 Jahr

oh this is insane

Profilbild von volt
voltvor 1 Jahr

it's so funny how generative ai anything people universally dislike except for richard, a bluecheck with a nft or selfie profile picture that goes "Great stuff! 🔥🔥 Looking forward for more updates"

Profilbild von Ivan Leo
Ivan Leovor 1 Jahr

Hmm audio demos don't seem to work on the page

Profilbild von neb
nebvor 1 Jahr

Finally ❤️

Ähnliche Videos

We’re excited to announce the release and open-source of HunyuanImage 3.0 — the largest and most powerful open-source text-to-image model to date, with over 80 billion total parameters, of which 13 billion are activated per token during inference.The effect is completely comparable to the industry’s flagship closed-source model.🚀🚀🚀 HunyuanImage 3.0 originates from our internally developed native multimodal large language model, with fine-tuning and post-training focused on text-to-image generation. This unique foundation gives the model a powerful set of capabilities: ✅Reason with world knowledge ✅Understand complex, thousand-word prompts ✅Generate precise text within images Different from traditional DiT architecture image generation models, HunyuanImage 3.0’s MoE architecture uses a Transfusion-based approach to deeply couple Diffusion and LLM training for a single, powerful system. Built on Hunyuan-A13B, HunyuanImage 3.0 was trained on a massive dataset: 5 billion image-text pairs, video frames, interleaved image-text data, and 6 trillion tokens of text corpora. This hybrid training across multimodal generation, understanding, and LLM capabilities allows the model to seamlessly integrate multiple tasks. Whether you're an illustrator, designer, or creator, this is built to slash your workflow from hours to minutes. HunyuanImage 3.0 can generate intricate text, detailed comics, expressive emojis, and lively, engaging illustrations for educational content. The current release focuses solely on text-to-image generation and future updates will include image-to-image, image editing, multi-turn interaction, and more. 👉🏻Try it now: 🔗GitHub: 🤗Hugging Face:

Tencent Hy

412,658 Aufrufe • vor 10 Monaten

Tencent presents GameGen-O Open-world Video Game Generation We introduce GameGen-O, the first diffusion transformer model tailored for the generation of open-world video games. This model facilitates high-quality, open-domain generation by simulating a wide array of game engine features, such as innovative characters, dynamic environments, complex actions, and diverse events. Additionally, it provides interactive controllability, thus allowing for the gameplay simulation. The development of GameGen-O involves a comprehensive data collection and processing effort from scratch. We collect and build the first Open-World Video Game Dataset (OGameData), amassed extensive data from over a hundred of next-generation open-world games, employing a proprietary data pipeline for efficient sorting, scoring, filtering, and decoupled captioning. This robust and extensive OGameData forms the foundation of our model's training process. GameGen-O undergoes a two-stage training process, consisting of foundation model pretraining and instruction tuning. In the first phase, the model is pre-trained on the OGameData via the text-to-video and video continuation, endowing GameGen-O with the capability for open-domain video game generation. In the second phase, the pre-trained model is frozen, and we fine-tuned using a trainable InstructNet, which enables the production of subsequent frames based on multimodal structural instructions. This whole training process imparts the model with the ability to generate and interactively control content. In summary, GameGen-O represents a notable initial step forward in the realm of open-world video game generation via generative models. It underscores the potential of generative models to serve as an alternative to rendering techniques, which can efficiently combine creative generation with interactive capabilities.

AK

367,088 Aufrufe • vor 1 Jahr