Загрузка видео...

Не удалось загрузить видео

На главную

I'm still convinced that you can get a better acting performance with AI by generating dialogue first and generating the video to sync with the audio. As audio generation is so much cheaper you're also able to validate the performance for a few cents before using credits for the...

38,110 просмотров • 2 месяцев назад •via X (Twitter)

Комментарии: 49

Фото профиля Ryan Lightbourn
Ryan Lightbourn2 месяцев назад

In discussions I'm having about a feature film, we're hands down getting actors in a recording studio before a single shot is generated.

Фото профиля TomLikesRobots🤖
TomLikesRobots🤖2 месяцев назад

Excellent - I think a decent actor will be the gold standard for a while yet. Having said that I think we're getting to the point where a decent AI performance can work better than a poor human performance.

Фото профиля Ryan Lightbourn
Ryan Lightbourn2 месяцев назад

totally agree

Фото профиля Andy Wojcicki
Andy Wojcicki2 месяцев назад

the male voice sounds a bit like a mix of Liam Neeson and Liam Cunningham

Фото профиля TomLikesRobots🤖
TomLikesRobots🤖2 месяцев назад

Hah - I think the priest must be from a similar part of Ireland. Dublin or further North.

Фото профиля Skor
Skor2 месяцев назад

This is well done. I'm almost 40 min into an AI original drama based on the "Epic of Gilgamesh" I've done all the dialogue scenes with Grok Imagine. I was a struggle but it came out ok. This clip is pretty long but the Uruk general near the end is pretty good.

Фото профиля Jessie Star (Commissions Open)
Jessie Star (Commissions Open)2 месяцев назад

Soooo fucking boring. You don't even know the meaning of performances. You are trying to mimic, trying to surface level this shit. They are pretty mannequins man. Nothing dynamic in here at all.

Фото профиля Cat
Cat2 месяцев назад

Need to introduce a little motion noise in the body movements. Humans sway a bit imperceptibly - nobody is completely still. Perhaps some metrics on that can be introduced in the prompts - if things are too “still” you have simulated the uncanny valley of muscle control

Фото профиля A.W.E.S.O.M.-O 4000
A.W.E.S.O.M.-O 40002 месяцев назад

I completely agree with the audio-first workflow.

Фото профиля TomLikesRobots🤖
TomLikesRobots🤖2 месяцев назад

Yes - I'm looking forward to seeing how SeeDance 2.0 handles references. Would be great if longer audio refs were possible combined with video refs.

Фото профиля The Creature Preacher
The Creature Preacher2 месяцев назад

You can suck my ass

Фото профиля Brian Noir
Brian Noir2 месяцев назад

Totally agree. Here's one I did starting from voice (Clip has strong language). The AI models produce generic performances.

Фото профиля Particular Uno
Particular Uno2 месяцев назад

Even better than that is recording a video with someone, even yourself, acting the scene. You'll get all the life and micro expressions missing from AI acting. In my opinions it's much much better, the uncanny valley is almost non existent

Фото профиля James Treakle-Smith
James Treakle-Smith2 месяцев назад

Yeah, this is fantastic. Also a strong argument for filmmakers using generative AI to work with real actors, dial in the performance like an animated film in the recording booth, and then iterate the visuals. Best of both worlds.

Фото профиля MaxgrowthPro 15 Year Anniversary 🌈
MaxgrowthPro 15 Year Anniversary 🌈2 месяцев назад

Acting and performance are two words that should not be allowed in conjunction with AI.

Фото профиля Thomas Squires
Thomas Squires2 месяцев назад

Agreed. When i built for AI anime production, one of the craziest quality gains i got was allowing users to upload/pre-generate their voicelines before sending it to seedance 2.0. You get 1000% better facial animations and way better acting authenticity.

Фото профиля Alpaca Laser Force
Alpaca Laser Force2 месяцев назад

I can only speak for grok imagine; just about anything would be an improvement from that. Everyone sounds like they are saying something for the fifth time to their near-deaf grandparen

Фото профиля TomLikesRobots🤖
TomLikesRobots🤖2 месяцев назад

Lol - yep. I wish @SpaceXAI would train Imagine to accept an audio reference it could sync. Visually Grok Imagine can be so strong with a detailed prompt.

Фото профиля Jennifer 🇺🇸 🦅
Jennifer 🇺🇸 🦅2 месяцев назад

I am going to try this! The acting and the entire scene is so natural. Its so good!

Фото профиля TomLikesRobots🤖
TomLikesRobots🤖2 месяцев назад

Thanks Jennifer and yes, even without the video generation it's pretty cool to have a cast read your scripts and take your direction.

Фото профиля Jack
Jack2 месяцев назад

The audio also gives way better context to the video gen

Фото профиля 100PercentFaked
100PercentFaked2 месяцев назад

That’s how I’ve been using it too. Made Spielberg react rather amusingly to some ludicrous accusations in my last short.

Фото профиля Kokulu Kiyam
Kokulu Kiyam2 месяцев назад

And you can record your own voice and change it for perfect emotion and intonation.

Фото профиля The_Void
The_Void2 месяцев назад

That isn’t acting. It’s 1’s and 0’s

Фото профиля Philipp
Philipp2 месяцев назад

It may also allow two-camera setups -- Seedancing it twice from a different perspective -- which would let you more freely edit the results together.

Фото профиля TomLikesRobots🤖
TomLikesRobots🤖2 месяцев назад

Yeah - good point. With this one I prompted for a multishot video and prompted cuts so it did the edit for me, but totally possible to manually edit with multiple cameras.

Фото профиля 𝚝𝚊𝚙𝚎𝚍𝚛𝚘𝚗𝚎
𝚝𝚊𝚙𝚎𝚍𝚛𝚘𝚗𝚎2 месяцев назад

I’d probably not use Liam Neeson’s voice unless you want him to find you and demonstrate his particular set of skills.

Фото профиля Ethan Avila
Ethan Avila2 месяцев назад

Where are you using seedream? I might try it out

Фото профиля Domenico Di Donna
Domenico Di Donna2 месяцев назад

One shot is for slop, applying proper workflows is what makes the difference! (gosh this looks like written by AI, what's happening to me?)

Фото профиля TomLikesRobots🤖
TomLikesRobots🤖2 месяцев назад

When I spend time writing something it's too precise and sounds as though it's been written by AI. I sometimes find myself wanting to add some first person uncertainty to sound less AI. (Need to change my system prompt - lol)

Фото профиля Jay Jones
Jay Jones2 месяцев назад

This looks pretty cool. What I want to see is someone take the raw video and then run it through post, using something like Resolve.

Фото профиля TomLikesRobots🤖
TomLikesRobots🤖2 месяцев назад

I did actually make some subtle changes using a custom tool. Added film grain, tiny bit of grading. This is a rejected clip from a series I'm building up so I didn't spend too much time on it but I could have worked on sound design and look a bit more.

Фото профиля Strong and Stable AI
Strong and Stable AI2 месяцев назад

I just don't like AI acting yet. This clip is really good... for AI. I'm sticking with music videos and action until a real advance is made with an under-the-skin personality suite for AI actors.

Фото профиля Maxime Peabody
Maxime Peabody2 месяцев назад

I agree.. Thoughts on best audio models? I've been using elevenlabs. Also don't sleep on LTX, it gets surprisingly really good outputs

Фото профиля Monique Pryce
Monique Pryce2 месяцев назад

@fableshowrunner I’ve yet to try that order out myself. I hear syncing to music works as well.

Фото профиля TomLikesRobots🤖
TomLikesRobots🤖2 месяцев назад

@fableshowrunner Yeah - music lip sync was one of the first things I tried with SeeDance 2. It works very well.

Фото профиля Brian Behm
Brian Behm2 месяцев назад

I've been thinking about this a lot lately for something I've been working on. I've taken to calling it semantic compositing. Fast is sometimes faster than than fastest because you treat everything like a compositing problem. Audios a great example.

Фото профиля Aaron Sherman
Aaron Sherman2 месяцев назад

Nice work! I gotta figure out how to do this with Higgsfield and eleven labs

Фото профиля David de Rebeque | AI Filmmaker
David de Rebeque | AI Filmmaker2 месяцев назад

I agree with you, tried multiple combinations, this is the best by far for me

Фото профиля STOK
STOK2 месяцев назад

Do you feed the audio dialogue to Seedance during video generation?

Фото профиля TomLikesRobots🤖
TomLikesRobots🤖2 месяцев назад

Yep. Just generated at FAL Downloaded audio and used that along with prompt and imagery to generate videos.

Фото профиля Nhat Nguyen
Nhat Nguyen2 месяцев назад

does the emotion in the audio hold once the video syncs to it, or does the model flatten the performance back out?

Фото профиля TomLikesRobots🤖
TomLikesRobots🤖2 месяцев назад

Usually the performance holds. I used the same timestamped dialogue section in audio and video prompt so that helped with consistent performance. 0:00 – PRIEST (low, measured; a pause after her name) “Jo… look at me. Not the corner. Me.” 0:05 – JO (quietly, with controlled irritation) “I am looking at you.” 0:08 – PRIEST (softening, almost a whisper) “Then tell me what it said.” 0:11 – JO (a small breath; faint amusement) “It said you’d ask that.” 0:15 – PRIEST (a beat; trying not to react) “What else did it tell you?” 0:18 – JO (softly, almost kindly) “That you’re frightened.” 0:22 – PRIEST (firmer now) “I’m not frightened of you.” 0:25 – JO (looks toward the corner; a whisper) “No. You’re frightened it knows your name.”

Фото профиля Nhat Nguyen
Nhat Nguyen2 месяцев назад

This is the best I have seen 😮

Фото профиля Kim Nevelsteen, PhD
Kim Nevelsteen, PhD2 месяцев назад

Damn I want to see the rest.

Фото профиля TomLikesRobots🤖
TomLikesRobots🤖2 месяцев назад

I’ve been building a story in the same world involving exorcism and the unseen. Still unsure what I will do with it.

Фото профиля Kim Nevelsteen, PhD
Kim Nevelsteen, PhD2 месяцев назад

Publish episodes?

Фото профиля Kamus
Kamus2 месяцев назад

you know, i think seedaudio 1.0 is what we'll get on sd 2.5, so we probably won't need this that much. but you're right about it being a lot cheaper to get right.

Фото профиля TomLikesRobots🤖
TomLikesRobots🤖2 месяцев назад

Good point - a while ago I was wondering about the routing of the SeeDance models and how it works internally.

Похожие видео