Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

Instead of prompting with text only, create reference sheets first. Prompts in the comments 👇👇 For a yoga / wellness influencer video, I build: First 15 seconds - Motion sheet - the exact routine, poses, timing, transitions Character sheet - same face, body, outfit, expressions Location sheet -studio, lighting,...

31,085 görüntüleme • 3 ay önce •via X (Twitter)

0 Yorum

Yorum bulunmuyor

Orijinal gönderinin yorumları burada görünecek

Benzer Videolar

This is probably the most complex workflow I’ve ever built, only with open-source tools. It took my 4 days. It takes four inputs: author, title, and style; and generates a full visual animated story in one click in ComfyUI . I worked on it for four days. There are still some bugs, but here’s the first preview. Here’s a quick breakdown: - The four inputs are sent to LLMs with precise instructions to generate: first, prompts for images and image modifications; second, prompts for animations; third, prompts for generating music. - All voices are generated from the text and timed precisely, as they determine the length of each animation segment. - The first image and video are generated to serve as the title, but also as the guide for all other images created for the video. - Titles and subtitles are also added automatically in Comfy. - I also developed a lot of custom nodes for minor frame calculations, mostly to match audio and video. - The full system is a large loop that, for each line of text, generates an image and then a video from that image. The loop was the hardest part to build in this workflow, so it can process either a 20-second video or a 2-minute video with the same input. - There are multiple combinations of LLMs that try to understand the text in the best way to provide the best prompts for images and video. - The final video is assembled entirely within ComfyUI. - The music is generated based on the LLM output and matches the exact timing of the full animation. - Done! For reference, this workflow uses a lot of models and only works on an RTX 6000 Pro with plenty of RAM. My goal is not to replace humans, as I’ll try to explain later, this workflow is highly controlled and can be adapted or reworked at any point by real artists! My aim was to create a tool that can animate text in one go, allowing the AI some freedom while keeping a strict flow. I don’t know yet how I’ll share this workflow with people, I still need to polish it properly, but maybe through Patreon. Anyway, I hope you enjoy my research, and let’s always keep pushing further! :)

Lovis Odin

58,769 görüntüleme • 10 ay önce

This guy got Claude to take control of Arcads and generate hyperrealistic UGC ads with a consistent AI actor, from a 7-word prompt, and now e-com brands pay him $1,500 per video to clone the setup. He got tired of watching DTC brands burn $4K on a photoshoot that dies the second a product angle changes, so he built a 2-prompt workflow that turns one product photo into scroll-stopping UGC without a single model, camera, or casting call. Here's the exact breakdown: → Claude is the director, Arcads is the studio. You never write a real prompt, you write one sentence → Type 7 words: "create a character sheet of a UGC actress on Arcads." Claude expands it into a 40-line spec: warm skin tone, high ponytail, gold hoops, white AirPods, athletic build, then takes control of Arcads itself → Arcads renders a 6-angle character sheet: front, back, profile, 3/4, two portraits, one consistent face locked across all of them → You download the grid and drop it back into Claude with one product photo → Second prompt, 8 words: "use Seedance 2.0 on Arcads to create natural UGC" → Claude writes the full shot list: timecodes, "1:03 AM clock zoom," the trackpad scroll, the arm pull that drags the product into frame, "delivery: casual, energetic," "ends on a wink + air kiss" → Arcads cuts the yellow Vans print onto an actor who was wearing all black in the grid. Real fabric folds, real light, zero warping → Final render: lip-sync locked to consonants, head not glued to the camera, a girl filming at 1am on her bed The key move 96% of people skip: you can't put the product on the actor before you lock the character sheet. If you drop a product photo onto a face that isn't locked across angles, the identity drifts frame to frame. The jaw resets. The skin tone shifts. The print slides off the fabric. The whole thing screams "AI" and your CTR dies in the first 2 seconds. His system locks the 6-view sheet first, so Arcads has one real human to dress, and the top lands with natural folds even though the grid never had it. One apparel brand ran 20 variants off a single product photo and found a winner in 36 hours, because the clips read like organic UGC, not a studio ad. Brands now pay him $1,500 per avatar build + $340/month to keep dropping fresh video batches on every new product. The entire thing runs on two prompts and one laptop. No photographer. No model agency. No product samples on set. Just two sentences, one product photo, and the discipline to lock the character before you render a single second of movement.

Shade

15,227 görüntüleme • 1 ay önce

Tamanna Bhatia 🩷 Created this by using a movement sheet as a reference image to animate the dance using Seedance 2.0 + ChatGPT image 2.0 GPT Image 2.0 Prompt: Dance Sequence Instruction Sheet [VISUAL STYLE] A composition featuring a highly detailed 3D-rendered female dancer. Designed like a professional choreography guide with a technical, diagram-inspired layout. Clean white background, soft studio lighting, and strong contrast to highlight body movement and posture. [GRID LAYOUT] Structured 4×4 panel grid (16 frames total), evenly spaced with thin black divider lines. Each panel is identical in size and clearly numbered from 1 to 16 to show a continuous dance progression. [CHARACTER] Use image1 as the base character. The same female dancer appears consistently across all panels with accurate likeness and proportions. [WARDROBE] The dancer wears a stylish, performance-ready outfit: a well-fitted top paired with a short, flowy skirt. The look should feel modern and visually appealing while still practical for dance movement. Fabric should subtly respond to motion (slight flow and folds), even in grayscale. [PANEL STRUCTURE – EACH FRAME] Top-left: Step number + short dance move title (e.g., “Step 5 – Spin Transition”) Center: Full-body pose capturing a precise moment in the choreography Bottom-left: 3–4 lines of concise instruction describing the move Overlay: Motion arrows and directional guides illustrating how the dancer transitions [MOTION INDICATORS] Incorporate curved arrows for fluid motion, straight arrows for directional steps, and circular indicators for spins or turns. Emphasize rhythm, weight shifts, and body isolation. [RENDER QUALITY] High-detail sculpted 3D style with smooth grayscale shading, subtle shadows, and clean linework. Maintain a polished, concept-art level finish with clarity in every pose. [RESTRICTIONS] No color, no background scenery, no extra characters, no visual clutter, only the dancer and instructional elements..

Sydney

11,457 görüntüleme • 3 ay önce