正在加载视频...

视频加载失败

Been playing around with different animation styles using H3 in MiniMax Design. I just gave it the direction I wanted, and the Agent helped turn it into an actual workflow. This is the kind of AI video creation I find much more interesting — you can actually experiment with...

303,147 次观看 • 3 天前 •via X (Twitter)

0 条评论

暂无评论

原始帖子的评论将显示在这里

相关视频

Recently tried making AI videos locally for the first time using MiniMax H3 + ComfyUI, and I went through three different approaches: text-to-video, syncing to music beats, and image-to-video. 🎬 As an AI video newbie, my biggest takeaway is that the workflow itself isn't nearly as overwhelming as I thought - the real challenge is writing the prompts. Luckily, we can just use AI to solve that! ---- I’d drop my raw ideas straight to the AI, then run them through the official H3 Prompt Writing Skill to turn those thoughts into structured prompts. For example, with my first summer travel video, I broke it down into four scenes. I kept a similar structure across each scene and just swapped out the location details, running it iteratively and tweaking as I went until the output actually matched what was in my head. Thank goodness it's a local model-I could experiment endlessly without blowing through cash just to bridge the gap between "what I wanted" and "what the model understood." The absolute best surprise was that the whole process didn't take long at all. This model is insanely obedient and respects instructions really well, so I barely had to deal with gacha-style rerolls to get the exact shot I wanted. For the music beat-sync video, I wanted the visuals to shift on every kick drum or snare. The final result pretty much hit the nail on the head. Honestly, instead of just generating a video that looks "pretty good," I care way more about whether the model actually gets the rhythm and executes the pacing I set. And this time, it totally delivered. 🥁 Then for the seaside clip, there was a ton of scene description and creative direction, plus a 4-shot storyline-and the model managed to generate the visuals, the vibe, and the background audio all in one single pass. Genuinely blew me away. 🌊 Plus, it’s super beginner-friendly. You can just plug and play with the official workflow templates; literally all I had to think about was writing the prompts, choosing the resolution, and setting the duration. ------- Oh, quick side note on my local PC setup: * GPU: Consumer-grade RTX 4090 24GB * 20 steps for a 15-second video takes about 200+ seconds * Cranking it up to 768P brings it to 1,000+ seconds * Upscaling it to 2K using the official API takes around 6–8 minutes The speed is honestly pretty decent. I heard a bunch of speedup LoRAs came out recently that I haven't gotten around to testing yet, but those should make it run even faster! (Gotta love the power of open source.) ------- I’ll be packaging up all the workflows, scripts, and assets I used for this run and dropping them on GitHub for free!

広志

42,361 次观看 • 5 天前