Загрузка видео...

Не удалось загрузить видео

На главную

📢 New model! 📢 As a festive treat, we’re giving pro users access to the beta version of our improved model. 🎄🎵 Generate your best music tracks yet from short and long descriptive prompts. Currently outputs are 45 seconds. Much longer soon... 🧵

39,539 просмотров • 2 лет назад •via X (Twitter)

Комментарии: 10

Фото профиля Stable Audio
Stable Audio2 лет назад

We’ll roll out full access to the new model to all users in the new year.

Фото профиля Stable Audio
Stable Audio2 лет назад

You can try the new beta model by selecting ‘stable-audio-audiosparx-v1-1’ in the model selection window.

Фото профиля Stable Audio
Stable Audio2 лет назад

Try prompt: UK garage, Rhodes piano chords and melodies, walking double bass bass-line, MPC drums, syncopated percussion, beautiful, relaxing, summery, 130 BPM

Фото профиля Stable Audio
Stable Audio2 лет назад

Try prompt: Piano, beautiful, clean, soft lofi beat

Фото профиля Stable Audio
Stable Audio2 лет назад

Try prompt: Trap, a lovely and cute synth arpeggio, short 808 kick stabs, rattling 808 hi hats, fluffy synth pads, cute marimba and toy piano, R&B vocal chops, major key, beautiful, cute, well arranged, anime inspired, 130 BPM

Фото профиля Stable Audio
Stable Audio2 лет назад

Try prompt: Indie rock, strings, drum kit, electric bass, choir, synthesizer, percussion, shaker, tambourine, melancholic, beautiful, low-key, 110 BPM

Фото профиля Stable Audio
Stable Audio2 лет назад

Try prompt: Chilled house music, relaxed, inspiring

Фото профиля Darryl Mason
Darryl Mason2 лет назад

@amli_art Impressive!

Фото профиля Harmeet Gabha 🇦🇺
Harmeet Gabha 🇦🇺2 лет назад

Loop feature??

Фото профиля morgan —
morgan —2 лет назад

godly

Похожие видео

Suno is limiting downloads, basically making a lot of the songs you make locked in their 'walled garden'. One of their main competitors, Udio, did the same thing after a settlement with UMG that restricts any downloads. This is 100% an injunction from the music labels who want to limit the "damage" that AI does to the profitability of music. There are two viable alternatives... >Minimax music 3.0. It is currently free to use up to 500 songs a day as long as your account qualifies for beta testing their music model, which should be anyone. It is a closed-weight model, but there are no restrictions on what you download and how much you download. The music quality is not as good as Udio or Suno, but it's pretty close overall, and it can understand very complex prompts to guide music generation. >Open-source music is not a very good alternative... but right now SongGeneration-LeVo2 is the best AI music model that's open-source right now. The music I put on the post is one such generation that I created on my PC. The problem is that it's hard to use. The comfyUI version requires installation on a separate portable install, specific transformer and flash attention versions, and additional libraries installed. The required models and checkpoints are scattered around Hugging Face, and some of the links from the original repo are dead. There are not many English guides, and most of the training data is in Chinese, so it struggles with English lyrics. To some extent, the demand for music generation isn't that high, compared to images and videos, so overall not many alternatives are around.

Emerald Apple

16,152 просмотров • 12 дней назад

We've officially released and open-sourced HunyuanImage 2.1, our latest text-to-image model. The new model delivers on our commitment to balancing performance and quality. With native 2K image generation, HunyuanImage 2.1 is an advanced open-source text-to-image model.🎨 ✨ New in 2.1: 🔹Advanced Semantics: Supports ultra-long and complex prompts of up to 1000 tokens, and precisely controls the generation of multiple subjects in a single image. 🔹Precise Chinese and English Text Rendering with seamless image–text integration: The model naturally integrates text into images, making it suitable for a wide range of applications such as product covers, illustrations, and poster design to meet the needs of various fields. 🔹Rich Styles and High Aesthetic: Capable of generating images in various styles—including photorealistic portraits, comics, and vinyl figures—it delivers outstanding visual appeal and artistic quality. 🔹High-Quality Generation: Efficiently produces ultra-high-definition (2K) images in the same time other models take to generate a 1K image. HunyuanImage 2.1 uses two text encoders: a multimodal large language model (MLLM) to improve the model's image and text alignment capabilities, and a multi-language character-aware encoder to improve text rendering capabilities. The model is a single- and double-stream diffusion transformer with 17B parameters. We've also open-sourced the weights of the the accelerated version with meanflow which reduces inference steps from 100 to just 8, and PromptEnhancer, the first industrial-grade rewriting model that enhances your prompts for more nuanced and expressive image generation. Now, creators turn complex ideas—like posters with slogans or multi-panel comics—into visuals faster than ever. We’re just getting started. Stay tuned for our native multimodal image generation model coming soon. 🌐Website: 🔗Github: 🤗Hugging Face: ✨Hugging Face Demo:

Tencent Hy

89,257 просмотров • 11 месяцев назад

New Short Course: Getting Structured LLM Output! Learn how to get structured outputs from your LLM applications in this course, built in partnership with .txt, and taught by Will Kurt, a Founding Engineer, and , Developer Relations Engineer. It's challenging for software to automatically parse through an LLM's freeform text outputs. Structured outputs—like JSON—solve this by converting natural language into consistent, clear, data that a machine can read and process. This course teaches you how to generate structured outputs while building several use cases, including a social media analysis agent. You’ll learn about structured outputs and efficient ways to generate outputs in your defined schema or format. You’ll begin by using structured output APIs, then use re-prompting libraries like “instructor” to generate structured output. Finally, you’ll learn how constrained decoding works; this is a very clever technique in which constraints are applied on each subsequent token generated, blocking any tokens that don’t fit your defined schema. In detail, you’ll: - Learn why structured outputs are important, how they allow for scalable software development, and the different approaches to generate them, including vendor-provided APIs, re-prompting libraries, and structured generation. - Build a simple social media agent using OpenAI’s structured output API, learn how to define a model's desired structured output using Pydantic, and perform basic programming with your outputs, such as importing structured data into a data frame using pandas. - Learn how to use the open-source library "instructor," which checks the structured output of the model and re-prompts the model until it validates the desired output, and explore the limitations of this approach. - Understand how structured generation by the “outlines” library works by modifying LLM logits, on a per-generated-token basis based on the desired format, to give a particular output structure. - Learn how regular expressions, which outlines works with, are represented as finite-state machines, and how they can be used to develop a range of structured outputs beyond JSON. By the end of this course, you’ll have broadened your knowledge of the approaches you can use to get structured outputs from your LLM applications. Please sign up here:

Andrew Ng

89,792 просмотров • 1 год назад