Loading video...
Video Failed to Load
发个帖,让那些不相信通过codex能实现长视频自动生成的网友看一下我的工作流是怎么工作的,我的本地H3生成视频的质量和时 是怎么把控的。 1、我早上在codex里下的指令,已运行了一个多小时,全部完成还十几分钟。 2、我通过UU打开远端工作机的comfyui 界面,看正在后台跑的生视频流程。 3、分享一条我的参考生视频提示词,这次是20秒的,确保在8步下不崩不破音。 4、分享一条生成的最新视频,上面有未锁定服装导致的穿帮。 5、分享一条我的人物角色资产。 prompt 视频提示词如下:
16,665 views • 3 days ago •via X (Twitter)
30 Comments

prompt 视频提示词: subject_definitions: <Subject 1> is Zhangsun Wuji from the first reference image: male Tang statesman and imperial relative, composed but tense, formal official robe, intelligent cautious eyes. He is not Empress Zhangsun. <Subject 2> is Emperor Li Shimin from the second reference image: mature Zhenguan emperor, muted ocher-red robe, black futou, beard, grave authority. summary: [reference generation] Zhangsun Wuji brings a half cord found in his sleeve; Li Shimin sees black wax like coffin sealing, making the missing seal case more sinister. retention_analysis: <Subject 1> (appears in [Shot 1], [Shot 3]): fully_preserved - Zhangsun Wuji remains a male minister, not the empress. <Subject 2> (appears in [Shot 2], [Shot 4], [Shot 5]): fully_preserved - Li Shimin remains the only emperor. detailed_description: Inner palace side room, cold dawn, a short black-waxed cord on a tray, all symbols unreadable. [Shot 1] Close-up of <Subject 1> alone, presenting the cord with both hands. <Subject 1> (S1) says, <d>[Chinese] 臣袖中,多了这个。</d> [Shot 2] At 00:03.500, close-up of <Subject 2> alone, staring at the cord. <Subject 2> (S2) says, <d>[Chinese] 谁放的?</d> [Shot 3] At 00:06.500, cut to <Subject 1>, fearful but honest. <Subject 1> (S1) says, <d>[Chinese] 臣若知道,就不敢来见陛下了。</d> [Shot 4] At 00:10.500, insert of <Subject 2>'s fingers lifting the cord; black wax flakes fall. <Subject 2> (S2) says, <d>[Chinese] 这不是御印封绳。</d> [Shot 5] At 00:13.500, close-up of <Subject 1>. <Subject 1> (S1) says, <d>[Chinese] 像是封棺的蜡。</d> At 00:17.000, close-up of <Subject 2>, chilled by the implication. <Subject 2> (S2) says, <d>[Chinese] 他不是要开门,是要埋人。</d> overall_soundscape: Tray placed on wood, wax flake falling, low room tone, tense breathing. non_diegetic_music: A very low cello note with a slow drum heartbeat. Negative: no subtitles, no readable text, no wrong speaker, no identity swap, no duplicate emperor, Zhangsun Wuji is male.

codex提示词

远端工作机的界面

我的中控台的界面

我的工作流

我的技术说明帖

里面一半是自动化晚上我睡觉的时候跑的,

厉害,期待大作早日上线,建议边做边播

有点长,估计做完有四五十分钟,很大的工程,显卡不好,跑得很慢,一天最多做五分钟,

你生成的原片分辨率比较高,应该在768以上

双采,一采0.3M,二采1.3M,可以跑2M的1920X1088,但是那个慢到家了,实在受不了,我放一个以前跑的2M的你看看,画质确实好,五秒一段,用mc延长流跑了5段,这26秒的视频跑了2个小时才跑出来,

我用3060跑过一次1080,5秒差不多用了2小时

🤣跑1M的分辨率其实也可以,在剪映里再调下色,加点滤镜还能看,就是H3有个很严重的问题,远景小脸太糊,2M都没办法解决,

用参考图会好一些,官方design 终端效果好很多

先关注,晚上打飞机的时候再看

谢谢

只想问现在 ai 短剧收益如何

不知道哦,我不是做短视频收益的

很厉害,每天看你做的越来越好了!

谢谢大佬鼓励!

值得看的不是“Codex能不能生成视频”,而是它已经开始从写代码变成持续跑完整工作流。人负责定目标、验结果,剩下的执行交给 Agent,这才是效率真正发生变化的地方。

是的,提示词都是codex 直接根据skill 生成,不需要人工编辑

早就用codex来自动生视频了,做好skill,给剧本,codex 自己生场景角色卡,自己写分镜,判断时长,自已调用comfy的api,早上起床收货就行了

我这里没有设定场景图和道具图,提示词没锁死服装,所以生成质量差了点。

用 codex 调用了那个做视频的工具?还是就是使用的 skill

comfyui

The unlocked clothing thing gets us too. Are you locking wardrobe per character asset file now, or per shot?

最终的成片是这样的,这是直出的第一版,未作任何二次修改。

配音时minimax h3 自动的吗?

你不要运行20秒的,20秒一次性就能运行完成。你运行1分钟的试试看,最好是要有参考图的
