MiniMax H3 - Prompt Share This prompt turns any... reference image into a video showing how the subject could be made by hand from scratch. I tried it with my profile picture. It should work with any subject. Prompt: A cinematic creation film follows one maker reconstructing the main subject shown in the referenced image completely from scratch. The referenced image defines the finished subject’s complete visible appearance, proportions, structure, materials, colors, clothing or surface details, and distinctive features; ignore its background, framing, lighting, and unrelated elements. Begin directly on an empty, clean virtual workbench. Use a stable front three-quarter overhead view that keeps the developing subject readable. There is one maker throughout, represented by the same consistent left and right hands and forearms. Show no more than two hands at once. From 0 to 14 seconds, the entire creation unfolds as a clearly accelerated timelapse with rapid, purposeful hand movement and restrained motion blur. Short jump cuts compress repetitive manual work only after each action has visibly completed. Every cut inherits the exact form and progress left by the previous action. 0-2 seconds: One hand enters already holding the first foundation material, armature, or base element appropriate to the referenced subject and places it at the center. The second hand steadies it as the maker establishes the initial supporting form. 2-8 seconds: The maker rapidly develops the subject’s major structure and volumes using one coherent creation method appropriate to what the referenced image depicts. A living subject is sculpted as one continuous, non-gory digital form from armature to anatomy; a vehicle or machine is built from chassis to functional structure; an object is formed or assembled from its supporting body outward. Each additional material or component enters from outside the frame while firmly held by one of the maker’s hands, is carried to its destination, and remains under hand control until attached or shaped. 8-12 seconds: The same hands develop the recognizable outer form and reference-specific features. The maker sculpts, fits, wraps, stitches, fastens, carves, or polishes only where appropriate to the subject. Facial features, hair, clothing, body panels, wheels, glass, surfaces, accessories, or equivalent defining elements emerge through visible hand and tool contact, never through spontaneous transformation. 12-14 seconds: The maker refines proportions, edges, joints, surface transitions, textures, colors, and distinctive details until the developing subject closely matches the referenced image. One hand stabilizes the form while the other performs each final adjustment with a hand-held tool. 14-15 seconds: The maker removes the last tool by hand and withdraws both hands. The timelapse returns to normal speed as the camera makes a restrained push toward the completed subject and holds on a clean final view. Materials and components do not need to be visible before use, but anything newly introduced must enter the frame already held by one of the maker’s hands. Nothing moves, assembles, appears, disappears, or changes material independently. Maintain one maker, one continuous subject, one creation position, and one category-appropriate construction method. No assistants, extra hands, detached anatomy, duplicated elements, magical morphing, drawing phase, software interface, cursor, menus, annotations, or text overlays.show more

Kōda
123,530 görüntüleme • 7 gün önce
📖THE STEP MOST CREATORS SKIP IS WHY THEIR AI... ANIMATION LOOKS INCONSISTENT Consistency across clips doesn't come from prompting — it comes from the reference image. The pipeline, step by step: ▪ Start with ChatGPT Image 2 — generate a full character design sheet first, not just a single frame. Multiple angles, expressions, and outfit variations in one image keeps the character consistent across every scene ▪ Build a storyboard inside ChatGPT Image 2 as well — define each shot, camera angle, action, and mood before touching Seedance at all. This is the step most people skip and it's the reason clips look disconnected ▪ Define a color palette and lighting mood early — golden afternoon light, soft warm tones, dramatic shadows. Lock those values and repeat them across every prompt ▪ Take each storyboard frame into Seedance 2.0 as the reference image — one frame becomes one clip ▪ Write the Seedance prompt around the character action, not the scene description. The scene is already in the image. The prompt handles motion, camera behavior, and timing ▪ Keep clip duration between 4-6 seconds per shot — shorter clips give more control over pacing and reduce motion drift on character faces ▪ Match camera movement type across consecutive clips — if one shot dollies in, the next should hold or pull back, not dolly again The consistency across these frames comes from the character design sheet, not from luck. Seedance reads the reference image and the prompt together — if the reference is detailed enough, the output stays on-model. This video was created by ALOKXMEHTA 📥 tomorrow: the exact ChatGPT Image 2 prompt structure used to generate a multi-angle character design sheet like this one 🔖One article covers the entire workflow — it is pinned below, do not scroll past it.show more

Zentrix⌚️
14,015 görüntüleme • 2 ay önce
Successfully generated this with Seedance 2.5—without any face-related moderation... restrictions. The result is amazing! prompt: For the target video, at 0.00 seconds into the target video, (from [Shot 1]) is fully referenced. integrated_multimodal_description: [Shot 1] A continuous live-action first-person smartphone recording based entirely on . is the only visual reference for the adult female dancer's appearance, face, hairstyle, hair color, costume, body proportions, environment, room layout, wall, floor, lighting, color and initial spatial composition. Preserve the original visual identity and environment of . The camera closely reproduces the distinctive perspective of : a strong high-angle first-person smartphone POV looking substantially downward at the woman. The camera is clearly above her eye level. The top of her head, shoulders and upper body are visibly seen from above. Keep substantial ceiling visible in the upper part of the frame. Do not use an eye-level view, third-person view, orbiting camera or cuts. A single adult arm belonging to the camera operator enters naturally from the lower-right foreground. Show a connected forearm, wrist, palm and fingers. The hand is a moderate-size foreground element, not a giant object. Keep the hand primarily within the lower-right quadrant, occupying approximately 10–18% of the image area. The hand must never extend across the center, cover the woman's face or cover her torso. The woman remains the primary visual subject and the hand is only the foreground controller. Keep the woman physically close to the smartphone, approximately 1.0–1.2 meters away. Her full body occupies approximately 70–80% of the vertical frame. Do not pull the camera backward, do not place her deep in the room and do not progressively make her smaller. The foreground index finger directly controls the woman's dance movement. The woman does not independently freestyle. The hand is the visible movement cue and the woman's body immediately follows it. Maintain these exact control relationships throughout the entire video: UP → WOMAN RISES. DOWN → WOMAN LOWERS. LEFT → WOMAN MOVES LEFT. RIGHT → WOMAN MOVES RIGHT. For UP, the index finger moves clearly upward and the woman's body immediately rises. Her torso extends upward, her posture becomes taller, her arms naturally rise and her center of gravity moves upward. For DOWN, the index finger moves clearly downward and the woman's body immediately lowers. Her knees bend slightly, hips lower, center of gravity drops and torso lowers into a controlled dance position. For LEFT, the index finger moves toward screen-left and the woman's body immediately shifts and sways toward screen-left. Her weight, torso and hips follow the same direction. The entire body movement must be clearly visible. For RIGHT, the index finger moves toward screen-right and the woman's body immediately shifts and sways toward screen-right. Her weight, torso and hips follow the same direction. The entire body movement must be clearly visible. LEFT and RIGHT must not become tiny hip twitches. The woman's body must visibly move from side to side. Keep the horizontal movement simple so the connection between finger direction and body direction remains unmistakable. The hand gesture and corresponding body movement are tightly synchronized. The woman's movement begins as the corresponding finger gesture begins and develops within the same musical beat. No noticeable reaction delay, no one-beat delay, no anticipation and no independent movement. The dance is a confident, sexy viral short-form influencer dance: energetic rhythmic hip movement, controlled waist isolation, subtle torso waves, attractive shoulder movement, pronounced but tasteful hip accents, confident posture, playful facial expression, direct eye contact and natural hair movement. The dance should feel like a polished viral social-media dance challenge. Make the movement sexy, stylish, rhythmic and visually captivating while keeping it non-explicit. During LEFT and RIGHT movements, maintain the direct horizontal hand control while adding sensual hip accents. During UP and DOWN movements, maintain the direct vertical hand control while adding smooth waist and body movement. Never sacrifice hand-to-body control for complicated choreography. Use a fast modern viral dance rhythm with strong clear beats. Command sequence: UP → DOWN → LEFT → RIGHT → LEFT → RIGHT → UP → DOWN → LEFT → RIGHT. Each command produces one clearly visible corresponding body movement. Do not add unrelated freestyle movements, spins, walking or complicated footwork. The index finger moves with strong, deliberate gestures. The wrist and forearm naturally follow the finger. Use short, clear upward, downward, leftward and rightward gestures. Do not use weak floating gestures. Keep the hand in the lower-right foreground and never allow it to dominate the center of the frame. Maintain one continuous handheld smartphone recording with subtle natural camera instability. Preserve the strong high-angle perspective and close composition throughout the entire shot. Do not zoom out, pull backward, lower the camera or lose either the hand or the woman's full body. At the end, use LEFT → RIGHT → UP. The woman performs a left body movement with a sensual hip accent, a right body movement with a sensual hip accent, then rises into a confident final dance pose. The hand remains in the lower-right foreground and does not move into the center. Hold the final pose briefly. overall_soundscape: Natural indoor room ambience, subtle foot movement, fabric movement and quiet smartphone handling sounds, preserving the acoustic character of . non_diegetic_music: A modern viral short-form dance track with strong punchy beats, crisp percussion and clear rhythmic accents. Each major beat provides a timing point for the hand gesture and its corresponding dance movement.show more

underwood
245,132 görüntüleme • 17 gün önce
gemini omniflash is actually f*cking cracked. you can animate/edit... any video with a text prompt. character swaps, object transforms, full environment changes without regenerating/rotoscoping. everyone using AI to to animate and edit videos right now hits the same wall. the clip comes out 90% right and you regenerate from scratch hoping the 10% fixes itself. it never does. the fix is using your video as the input. omniflash edits what's already there instead of rolling the dice again. here's what's in the system: > the two-layer premiere trick: generate the same shot twice (one with background removed), stack them, cut at one frame, instant scene change > character swap with a single reference image (plus the one line you need or the model keeps the original's features) > object transforms that leave the rest of the frame untouched: stone into glowing sphere, candles into flowers > style transfer from an image reference instead of text, way more accurate > why stacking edits in one prompt breaks everything and the exact step order that doesn't > the audio limitation nobody mentions and how to work around it i packaged every prompt, the edit sequence, and the premiere layering setup. RT + reply "OMNI" and i'll send it over.show more

Sulfur
36,614 görüntüleme • 2 ay önce
19-year-old from china makes $9,000/month designing product sites and... ships each one in an afternoon. here's his exact setup the whole thing runs on two tools that each do one job: > brief written by hand: 5 min > Moonchild builds the design system, then every screen from it: 20 min > MCP hands the design to Claude as real structure, not a screenshot: instant > Claude Code reads those exact tokens and builds the live app: 20 min > second Claude session reviews the build for drift: 10 min total: about an hour. screen five still matches screen one. no agency, no dev, no design team the trick is MCP. the design tool passes Claude the actual colors, components and layout, so it builds from the source instead of guessing from a picture. full pipeline, every prompt, in the article above.show more

Ridark
19,477 görüntüleme • 2 ay önce
They did not take cursive from the schools because... children no longer needed it. They took it because of what it was quietly building in them. Consider what the exercise actually is. A child, six years old, is handed a pen and asked to draw a single unbroken line that becomes a word. The wrist must float. The fingers must hold a living pressure, never quite the same twice, always correcting. The eye must follow the ink forward and trust the hand to finish what it has begun. There is no lifting, no stopping, no starting over mid-word. The loop must close. The ascender must rise and return. The sentence must travel from one margin to the other as a single continuous gesture, and at the end of it the hand must still be steady. Twelve years of this. Every day. Ten thousand small acts of sustained, self-correcting attention, carried out below the level of conscious thought, until the motion belongs to the body and the body belongs to the motion. This is not penmanship. It is the slow construction of an interior form. The hand that has learned to carry a line without breaking it is the hand of a mind that has learned to carry a thought without breaking it. The two are not metaphors for one another. They are the same faculty, trained in the same child, by the same daily discipline. Continuity of the stroke becomes continuity of the reasoning. The patience of the loop becomes the patience of the argument. The commitment to finish a word one has started becomes the commitment to finish a sentence, a paragraph, a life's idea, without reaching for the nearest distraction halfway through. Print is a different creature entirely. Print lifts. Print stops. Print assembles a word out of separate, stamped, interchangeable pieces, each one beginning and ending in isolation. A mind raised only on print learns to think the way print is made, in discrete tokens, in replaceable units, in fragments that can be recombined by any outside hand without the owner noticing the substitution. It is precisely the shape of thought a language model produces. It is precisely the shape of thought a language model can steer. Cursive is kata. This is the whole of it. A form repeated daily, for years, not for the sake of the form but for what the repetition lays down in the practitioner beneath the form. The swordsman does not train kata so that one day he may fight in kata. He trains it so that when the moment comes and there is no time to think, the movement is already inside him, older and deeper than thought, and it rises on its own. Cursive was the kata of the literate mind, the daily quiet drilling of continuity, of patience, of a line held steady under the long pressure of its own length. And the signature it produced at the end, that small flourished mark unique to a single human being on earth, was only the outward proof of an inward form no machine and no other hand could ever reproduce. Take the kata away and the practitioner is left with vocabulary in place of faculty. He can recognise a whole thought when he encounters one. He cannot carry one himself. He can admire a finished argument. He cannot sustain one long enough to close its loop. He begins books he does not finish, sentences he does not end, ideas he abandons the moment the screen in his palm offers him a brighter one. And when the machine begins feeding him tokens in the exact shape his schooling taught him to receive, he meets it with no interior resistance at all, because no interior form was ever built in him to push back with. They removed it quietly, across a generation, and they removed it in the last years before the machines arrived. Twelve years of daily practice in unbroken, embodied, self-authored thought, gone from the curriculum of almost every child in the Western world, just as the instruments designed to complete their sentences for them came online. The hand forgets. The mind, having never been taught the kata, forgets a thing it never knew it had. That is what cursive was. That is what was taken. And that is why the thought of anyone who still writes by hand, in long unlifted lines, remains, quietly, stubbornly, and without their ever needing to announce it, their own. Now the question stands open. What else has been banned, phased out, quietly retired from the curriculum and from common life over these same decades, under the same soft excuses? Mental arithmetic. Memorisation of poetry. Latin. Logic as a formal subject. Map reading. Knot work. The keeping of a commonplace book. The reading aloud of long passages in class. Singing in parts. What was each of those actually building in the child, beneath the surface of the lesson, and whose interest was served by its disappearance?show more

SiriusB
443,951 görüntüleme • 4 ay önce
We have released Seedance 2.0. Due to the 2500-character... limit, please translate the following prompts into Chinese before use. [Technical Specs] Generate a 10-second, 16:9, 720p cinematic video. Smooth continuous camera motion with no cuts. The overall pacing is fast and tightly compressed, with rapid escalation from start to finish. Audio evolves quickly from a high-performance engine idle into intricate mechanical shifting and clicks, culminating in a soft electronic chime and the distinct sound of a "mwah" blowing kiss. [Global Constraints] Only the evolving mechanical character appears; no other humans or characters. All transformations must follow physical logic and maintain structural continuity. No object should pass through or intersect with other solid objects. Every robotic component must originate from visible parts of the Porsche 911 (doors, hood, wheels, chassis) through unfolding, splitting, or reconfiguration. [Scene Setup — 0:00–0:01] A sleek, metallic silver Porsche 911 sits on a rain-slicked futuristic city street at night, neon lights reflecting off its polished surface. The camera starts at a low-angle front-quarter view and begins a fast, smooth tracking-arc towards the side. [Rapid Transformation Initiation — 0:01–0:03] Transformation triggers instantly. The car’s suspension drops, and the frame begins to fracture into a complex grid of panels. The doors swing open and begin to segment into articulated arm structures. The front hood splits down the center, folding inward to reveal a glowing internal core. The headlights flicker and start to reorient as the "eyes." [Accelerated Feminine Reconfiguration — 0:03–0:07] The mechanical action is dense, overlapping, and fluid, emphasizing graceful but powerful motion. Lower Body: The rear wheels and wheel arches split and rotate downward, reassembling into slender, high-heeled mechanical legs. Torso: The roof and rear engine cover slide and compress, forming a sleek, curvaceous hourglass torso that retains the car’s aerodynamic lines. Arms & Hands: The side mirrors and door panels unfold into delicate but strong hands and fingers. Head: The front bumper and emblem area segment and rise, folding into a feminine-shaped head with a sleek metallic "helmet" visor. [Logical Transformation Constraints — No Spontaneous Appearance] The robot’s "skin" is composed of the car's outer silver panels. The internal frame and wiring emerge from the engine and undercarriage. No parts appear out of thin air; every joint is a reconfigured automotive component. [Transformation Completion — 0:07–0:08.5] The robot stands tall and elegant. The silver panels lock into place with a satisfying "click," revealing glowing blue LED accents in the seams. The silhouette is clearly feminine, humanoid, and sophisticated, reflecting the premium design of the original vehicle. [Final Hero Ending — 0:08.5–0:10] As the robot stabilizes, the camera performs a rapid, smooth zoom-in (Dolly-In) directly to her face. The robot tilts its head slightly, and the optic sensors (eyes) brighten. It brings its mechanical hand to its metallic lips and performs a graceful blowing kiss (fly-kiss) gesture toward the camera. The video ends with a close-up of the face, capturing the reflection of neon lights in its visor just as the kiss is released. [Cinematography Notes] Continuous Motion: No cuts or fades; the camera must transition from the car-tracking shot to the face-zoom seamlessly. Material Consistency: The robot must maintain the exact metallic silver paint, texture, and reflections of the Porsche. Energy: The transformation should feel high-energy and "force-driven," while the final gesture is soft and charismatic.show more

underwood
19,462 görüntüleme • 5 ay önce
Batch Normalization by hand ✍️ ~ 7 steps walkthrough... below Batch normalization is common practice for improving training and achieving faster convergence. It sounds simple. But it is often misunderstood. 🤔 Does batch normalization involve trainable parameters, tunable hyper-parameters, or both? 🤔 Is batch normalization applied to inputs, features, weights, biases, or outputs? 🤔 How is batch normalization different from layer normalization? So I drew and calculated one entirely by hand. Goal: normalize a mini-batch of 4 examples to mean 0 and variance 1, then let the network scale it back. = 1. Given = A mini-batch of 4 training examples, each with 3 features. = 2. Linear layer = Let us multiply by the weights and add the biases. Batch norm sits after this, which answers the second question: what gets normalized is features, not inputs, weights or biases. = 3. ReLU = We apply the activation, and -2 becomes 0. Negative values are suppressed before any statistic is taken. = 4. Batch statistics = Let us compute the sum, mean, variance and standard deviation, one row at a time. A row is a feature and the four columns are the four examples, so every number here measures one feature against the rest of the batch. That is the "batch" in batch normalization, and it is exactly what layer normalization does not do. The statistics are rounded to whole numbers, which is what keeps the rest of the page doable in pen. = 5. Shift to mean 0 = We subtract the mean, in green. The four values in each feature now average to zero. = 6. Scale to variance 1 = Let us divide by the standard deviation, in orange. Each feature now has variance one, whatever scale it arrived at. = 7. Scale and shift = We multiply by a linear transformation and pass the result on. The diagonal and the last column are trainable, so having just forced every feature to mean 0 and variance 1, we hand the network the means to undo it. The outputs: Mean of each feature = [2, 1, 2] Std dev of each feature = [1, 1, 2] To the next layer = [2, -2, 2, 0], [-3, 3, 6, -3], [2, 0, 1, 2] The answers: 🤔 Both. The scale and shift are trainable, the statistics are not. Epsilon and the momentum on the running statistics are the hyper-parameters, and one mini-batch by hand needs neither. 🤔 Features, after the linear layer, not inputs, weights or biases. 🤔 Batch norm measures across the batch, one feature at a time. Layer norm measures across the features, one example at a time. 💾 Save this post!show more

Tom Yeh
20,848 görüntüleme • 1 ay önce
Megan Fox really knows how to draw everyone's attention,... don't you think?🤩😏 VIdeo by Kling 3.0 via Higgsfield AI (More videos for my subscribers!) Video & Image prompts below👇🏻 🔴 Video Prompt: Subject slowly rises from kneeling position on fluffy white rug — movement begins with a natural, fluid push upward from knees to standing. As she rises, hips sway gently and playfully, body moves naturally with soft feminine energy, not exaggerated. Hair still mid-ponytail, hands finish tying it as she stands. Camera rises simultaneously with her, tracking from chest level up to face level as she reaches full standing height. Full body visible throughout the rise. At 8 seconds: she leans slightly forward toward camera, eyes locked, and softly tongue out to the camera lens. At 9-10 seconds: she pulls back, gives one final direct look into camera, then flashes a peace sign with a slight smirk. Style: Cinematic, smooth camera movement, natural bedroom lighting with soft daylight from window, warm neon glow from "Keor" sign in background. Realistic, no sudden cuts, fluid motion throughout. 10 seconds total. 🔴 Image Prompt: A Megan Fox with fair skin, freckles across her face and shoulders, and striking blue eyes is kneeling on a thick, fluffy white rug in a messy bedroom. She has long, black hair that she is actively gathering with both hands to tie into a high ponytail; her right hand holds the base of the ponytail near the crown of her head while her left hand pulls the length of her hair through. She is wearing a fitted light green lime ribbed tank top with V neckline and fuchsia gym shorts bottoms. Her expression is direct and slightly serious as she looks straight at the camera from a high angle. The bedroom background features an unmade pink bed with rumpled light-colored sheets and pillows, scattered clothes on the floor and bed, a white nightstand with a modern lamp, a stack of books, a big poster of Inuyasha manga characters, a little neon written 'Keor' on wall, and a window with sheer curtains letting in natural daylight. The overall scene has a casual, lived-in atmosphere with soft natural lighting."show more

KeorUnreal
106,851 görüntüleme • 4 ay önce
What did you do to my friend !! 😡... This is the trending “Fight Prompt” going viral Prompt : Use the first uploaded image as the main reference for the school uniform, body proportions, pose, posture, background, camera angle, framing, and overall composition. Use the second uploaded image as the identity reference for the face and hairstyle. Create a realistic Korean influencer-style school uniform portrait where the person from the second image naturally appears wearing the school uniform from the first image, photographed in the same studio setting. Important: Keep the school uniform, blazer, shirt, tie, skirt or pants, and overall outfit design from the first image. Keep the body proportions, standing pose, hand placement, posture, camera angle, framing, and studio background from the first image. Replace the face with the person from the second image. Also preserve the hairstyle from the second image, including bangs, hairline, hair part, hair length, hair framing around the face, and overall hair silhouette. Do not use the hairstyle from the first image if it differs from the second image. Identity: The face from the second image must remain clearly recognizable. Preserve the second person’s face shape, eyes, nose, lips, skin tone, jawline, and overall facial impression. Do not turn the face into a generic attractive face. Do not beautify too heavily. Preserve the person’s recognizable identity, but do not copy the face too rigidly. Reinterpret it naturally so it looks like a realistic photo of the same person in this new school-uniform scene. Keep the same overall facial impression and identity while allowing natural refinement and seamless adaptation to the lighting, angle, and mood of the target image. Hair: Follow the hairstyle from the second image. Preserve the second person’s bangs, hairline, hair part, hair texture, hair length, and overall hairstyle impression. Only adapt the hair naturally so it fits the pose, lighting, and composition of the first image. Korean influencer mood: clean modern Korean influencer portrait polished but natural beauty soft photogenic expression subtle editorial mood stylish, slightly chic, youthful, and confident atmosphere refined but believable skin texture clear eyes with soft catchlights naturally pretty, not over-retouched avoid stiff ID-photo mood Lighting: soft Korean beauty lighting gentle facial brightness clean skin tone soft natural highlights on the face natural shadow transition subtle glow, but realistic skin texture avoid harsh flash avoid flat passport-photo lighting avoid dramatic studio glamour lighting Style: realistic photography clean studio portrait quality Korean influencer-style school portrait mood natural skin texture high detail seamless face and hair integration polished but believable Negative prompt: no identity loss no generic attractive face no over-beautified face no first-image hairstyle if different no awkward face blending no mismatched skin tone no mismatched hairline no distorted facial features no blurry eyes no deformed hands no extra fingers no change to the school uniform no change to the body pose no change to the background no cartoon style no anime style no text no watermarkshow more

Ai Arainz
104,348 görüntüleme • 3 ay önce
This guy built a visual scanner that reads 468... points on his face and 42 points on his hands from a regular webcam and turns them into a cloud of thousands of particles right between his palms. Inside, MediaPipe and TouchDesigner are linked: the first captures hands and face from the webcam with high accuracy, the second turns those coordinates into a live plane and feeds it into a POP system that instantly generates a swarm of particles in the shape of a head. No studio, no render farmer, no VR headset. Just a laptop, a webcam, and 1 TouchDesigner session. And traditional VJ studios keep teams of 5 people on a setup with lighting, custom hardware, and commercial plugins, while his expenses are only a TouchDesigner subscription and a regular USB camera. One laptop runs MediaPipe and TouchDesigner simultaneously, holds the camera stream at 60 FPS without drops, and in parallel processes 468 face points + 21 points on each hand. The camera captures frame after frame, MediaPipe in real time sends TouchDesigner the finger coordinates and face geometry, and the POP operator inside the engine translates those numbers into thousands of particle points with colors from bright pink to gold. This setup immediately defines the role of the tool and the limits of its autonomy. It knows where the fingertips are at every moment of the frame. It knows how to read the face geometry at any angle to the camera. It knows how to draw a swarm of particles between them with the right color and contour. → MediaPipe pulls 468 points from the face and 21 points from each hand, 60 times per second → TouchDesigner receives those coordinates, builds a virtual rectangle between the fingertips, and feeds it into the POP system → POP generates thousands of particle points in the shape of a head, coloring them in a gradient from bright pink to gold → The HUD layer adds green corners and a blue neon frame, styling the image like an AR interface → All layers assemble into 1 real-time frame that projects back onto the video in the camera window → The final image is recorded to a file or broadcast to a projector for a live installation And only when the guy spreads his hands wider does the plane between the palms stretch; brings them together, it narrows. Otherwise the system runs on its own. And when he moves from his home room to a concert hall, the same laptop with the same webcam launches the same TouchDesigner session in just 5 minutes, without reconfiguration, without a new team, and without a single line of new code. In his work setup there is no studio of his own and no team for assembly. On the desk sits a laptop with a webcam, on top run MediaPipe and TouchDesigner with POP operators, and the same setup through a USB camera moves to any concert without a new configuration. Out of everything I have seen this year, this is the cleanest Creative Coding setup on 1 laptop: 0 render farms, 0 studio lighting, and between them 3 libraries, thousands of particle points, and 1 webcam.show more

Blaze
38,242 görüntüleme • 4 ay önce
The main bronze door of Milan Cathedral weighs 37... tons and took a single sculptor 10 years to complete. The Duomo of Milan is one of the largest and most intricate cathedrals on earth, and its five bronze doors were built to be worthy of it. Every inch of the great central door is covered in sculpted bronze, dozens of scenes from the life of the Virgin Mary, faces caught mid-emotion, figures that seem to press forward out of the metal as if trying to step into the world... It was designed by the sculptor Ludovico Pogliaghi, who received the commission in the 1880s and did not see his door installed until 1906, having poured a decade of his life into modelling every figure by hand. And his was just the beginning. The five doors of the cathedral were not made together, or even in the same lifetime. They were created across nearly seventy years, by a succession of different sculptors, each carving their panels with the same obsessive care. The final door was not completed until 1965, its unveiling taken as the symbolic close of a cathedral that had been under construction, in one form or another, for 579 years... More than a century ago, the English polymath John Ruskin tried to put into words what it means to build something like this. He managed it in a single sentence. "When we build," he wrote, "let us think that we build for ever."show more

James Lucas
83,782 görüntüleme • 2 ay önce
For the first time in the history of automation,... the people most likely to be replaced are the ones building their own replacement, frame by frame, for a few dollars an hour. Across India, Nigeria, China, and Argentina, workers are strapping cameras to their heads and recording every fold of laundry, every stitch, every washed dish, and that footage is training the robots designed to do those exact jobs. This is documented, not rumor. No jokes! Garment workers in Tamil Nadu, India have been filmed wearing head-mounted cameras on the factory floor, sending point-of-view footage to data firms whose clients include Fortune 500 companies. One US company alone has hired thousands of workers across more than 50 countries to record themselves cooking, cleaning, and folding clothes. More than 6 billion dollars poured into humanoid robots last year, and the one ingredient every maker is starved for is precisely this: real human hands doing real human work. The endpoint is stated plainly by the buyers. In China, one supplier said his pitch to factories is to let workers wear the cameras now, because trained robots will eventually work there instead. The quiet part is the exchange itself. The worker is paid for the hour and keeps nothing after it. No share, no royalty, no ownership of the movements their own body is teaching the machine. The skill leaves their hands and becomes someone else's product, and almost no one along the chain sees the full shape of the trade, not always the person filming, not the millions who watch the clip and scroll on. One scene holds all of it. A humanoid robot spent an hour folding three shirts while a human housekeeper, hired to guide it, quietly finished the rest of the chores. Every automation before this arrived from the outside. A machine showed up and took the job. This one is being built from the inside, by the workers themselves, handing over the last thing they had left to sell. UBI ? or something totally else should pave the way in the future? Thoughts?show more

Shanaka Anslem Perera ⚡
95,549 görüntüleme • 2 ay önce
Who could resist a girlfriend this sweet—and full of... surprises? 🥹 A big thanks to the original creator for sharing this delightful prompt. From adapting a reference image and matching the character to crafting the final video prompt, you can create the whole concept from scratch here: prompt: Vertical 9:16 format, 10.0 seconds, 60 fps. Realistic smartphone portrait footage with an immersive first-person interaction aesthetic, captured as one continuous uncut shot. Extreme close-up of a young adult East Asian woman, with her face occupying almost the entire frame as she rests on a soft white pillow. Her smooth, glossy black medium-length straight hair is naturally tousled across the bedding. Preserve the reference image’s distinctive wispy blunt bangs and loose face-framing strands, along with her fair, luminous complexion. 【Character Identity and Styling】 Strictly preserve the identity and facial appearance of the woman in the uploaded @ image_1: maintain the same face shape, facial-feature proportions, eye shape, eyebrows, nose, lips, hair color, wispy blunt bangs, face-framing strands, skin tone, and apparent age. She must remain the same person throughout the entire video. Her facial features must not change with movement, camera angle, facial expression, or distance from the camera. Use @ image_1 only as the reference for the woman’s identity, facial features, hairstyle, and makeup aesthetic. Do not inherit the black outfit, arm-supported pose, room background, pink heart stickers, text, watermark, or any other graphic overlays from the reference image. The woman from @ image_1 is wearing a beige-pink floral spaghetti-strap nightdress. Visual styling: preserve the delicate, translucent “tearful makeup / slightly tipsy makeup” aesthetic seen in the reference image, including naturally long upward-curled eyelashes, soft pink under-eye blush, subtly defined aegyo-sal, fine shimmering highlights at the inner corners of the eyes, glossy glass-like lip color, and watery translucent gray-brown contact lenses matching @ image_1. Her overall appearance should feel soft, innocent, affectionate, and full of fresh morning energy immediately after waking up. 【Camera and Lighting】 First-person POV from the boyfriend’s perspective. The opening shows a high-angle view looking down at the woman lying on her side in bed. During the middle and final sections, her sudden pounce causes natural mattress movement, realistic camera shaking, and an extremely close face-to-face perspective. The background is a clean, warm morning bedroom with light-colored bedding. Bright, soft morning light enters from the upper side as diffused illumination. Use high-key exposure and low contrast while preserving the subtle rise and fall of her breathing, the natural volume of her cheeks, and her moist, translucent skin texture. 【Action and Rhythm】 Within 10 seconds, complete a continuous emotional reversal: from “sleepily holding his hand and asking him to stay,” to “the boyfriend teasing her affectionately,” and finally to “suddenly waking up and energetically pouncing on him, pushing the camera back onto the bed.” Dialogue and physical actions must be precisely synchronized. The pacing should feel intensely sweet, playful, and immersive. 【Second-by-Second Timeline】 0.00–2.50 seconds | Part 1: Sleepily Reaching Out and Asking Him to Stay — “别走可以吗?” The boyfriend is about to get up and leave, causing the camera to make a subtle upward movement as if he is rising from the bed. The woman, lying on her side, sleepily opens her eyes. One fair-skinned hand instinctively reaches out from beneath the blanket and firmly holds the boyfriend’s wrist, refusing to let go. She slightly pouts her lips and gazes softly and innocently into the camera. In a gentle, sleepy voice, she says in Mandarin, with precisely synchronized lip movement: “别走可以吗?” 2.50–4.50 seconds | Part 2: Affectionate Interaction and Playful Question — “不上班你养我呀?” The boyfriend reaches out with his other hand and affectionately gives her soft cheek a gentle pinch. Her eyes remain sleepily half-open, and she rubs her cheek against his palm like a small cat. The boyfriend’s warm, smiling off-screen voice asks in Mandarin: “不上班你养我呀?” 4.50–7.50 seconds | Part 3: Instantly Waking Up and Pouncing Forward — Waking Up & Pouncing After hearing his question, her previously drooping eyelids instantly open wide, and she becomes fully alert within a second. A flash of playful cunning and excitement appears in her eyes. She plants both hands on the bed and suddenly lunges forward, directly pouncing on the boyfriend from the first-person perspective and pushing him back down onto the soft mattress. The camera experiences a realistic sense of weightlessness, strong natural shaking, and visible mattress compression as the bed sinks under their weight. 7.50–10.00 seconds | Part 4: Close-Up Dominant Gaze and Playful Declaration — “我养你呀!” The camera is now lying flat against the pillow. The woman supports herself with both hands placed on either side of the boyfriend’s body and looks down into the camera from an extremely close distance. Her wispy blunt bangs and glossy black strands of hair naturally fall around the lens. Her watery gray-brown eyes curve into smiling crescents, and her lips form a bright, proud, delighted smile. She gives the camera a playful wink and says energetically in clear Mandarin, with precisely synchronized lip movement: “我养你呀!” End on a frozen moment of intimate, heart-fluttering eye contact. 【Mandatory Continuity Rules】 Maintain throughout the entire video: 10 seconds, 60 fps, vertical 9:16 format, first-person boyfriend POV, extreme close-up facial interaction, and dynamic bed movement caused by the pounce. The pouncing action must be quick, decisive, and physically believable, with natural gravity, body momentum, mattress compression, and camera response. Her expression must shift instantly from sleepy and drowsy to bright, energetic, and playful. Maintain intense, affectionate eye contact throughout, creating a highly immersive and emotionally impactful sweet morning interaction. Strictly lock the character identity, wispy blunt bangs, face shape, and facial-feature proportions from @ image_1 throughout the video. No identity changes, face swapping, facial-feature drift, hairstyle changes, eye-color changes, age changes, excessively heavy makeup, plastic-looking skin, malformed fingers, extra limbs, body distortion, or physical clipping.show more

underwood
19,217 görüntüleme • 15 gün önce
Would you dare chase justice while swinging thousands of... feet above traffic below? Seedance 2 prompt on BudgetPixel AI Create a 15-second ultra-realistic cinematic high-altitude tether-swinging action sequence in strict 16:9 landscape, native 4K, 24fps. Use one seamless continuous drone follow shot with no cuts, no teleporting, and no time skips. The motion must feel physically continuous, dynamic, thrilling, and always readable. REFERENCE: image1 = main heroine reference. Use image1 as the strict identity reference for the heroine’s face, facial proportions, hairstyle, hair color, body proportions, age impression, outfit, shoes, accessories, styling, and overall recognizable appearance. Preserve her identity consistently throughout the whole video. Keep her as an original urban tether-swinging action heroine. Do not redesign her into a branded superhero character. Do not add franchise logos, copyrighted chest emblems, or recognizable third-party superhero symbols. CORE CONCEPT: This is an original urban tether-swinging action short. The heroine moves through the city using thin wrist-launched fiber lines, momentum, wall-running, rooftop movement, and real parkour body mechanics. She travels at high altitude between tall buildings, then lands on a rooftop, defeats one villain, and ends with a powerful shout. HOOK: The first second must be an instant scroll-stopping hook. Start with the heroine already falling backward off the edge of a very tall skyscraper. For a brief moment, it looks like she may actually fall. Then she instantly fires one thin tether line upward, it catches, and her body snaps into a huge high-altitude swing between buildings. STYLE: Photorealistic live-action realism. Premium cinematic action quality. Bright daytime Los Angeles atmosphere with realistic haze, realistic motion blur, realistic fabric movement, realistic body weight, real inertia, and practical environmental interaction. The sequence should feel like a premium action movie shot, not animation, not a game cutscene, and not a cartoon. CAMERA: One uninterrupted drone follow shot only. No cuts. No resets. No jumpy edits. No impossible viewpoint teleporting. The drone camera must stay wide enough to show both the heroine and the environment together. It may tilt, roll, arc, climb, and dive with the motion, but it must always feel like one real flying camera tracking her. Keep the framing intense and fast, but always readable. ENVIRONMENT: Bright daytime in a dense modern city inspired by Los Angeles and Hollywood. Show: - tall glass and concrete high-rises - rooftop edges - billboards and signage - palm trees far below where visible - busy roads and traffic far beneath - bright haze and sunny atmosphere - believable large-scale urban depth The action must happen mainly high above the street between tall buildings, not low near the ground for most of the video. VILLAIN RULE: Only one villain appears in the entire video. The villain is one adult male enemy only. He wears a fitted black suit, black shirt, and black shoes. No mask, no armor, no fantasy costume. He appears only in the rooftop combat section. Do not generate multiple enemies. Do not generate background enemies. Do not clone or duplicate the villain. ACTION RULES: The heroine’s movement must feel hand-and-foot driven, not magical floating. She must visibly: - fire thin tether lines from her hands - swing with real tension and momentum - push off building surfaces - run along walls with clear foot placement - absorb landings with bent knees - sprint briefly on a rooftop - fight one villain using fast practical action - finish in control Her body mechanics must stay realistic: - core engaged during swings - arms extended or flexed according to line tension - knees bend on landing and push-off - visible transfer of momentum between swing, wall-run, leap, landing, and combat Do not make her hover weightlessly. Do not make the tether line act like magic. Do not make her float in place unnaturally. EMOTIONAL ARC: - opening: shock and immediate control - mid-swing: intense focus - rooftop approach: rising confidence - rooftop fight: sharp aggression and urgency - ending: victorious adrenaline and fearless release AUDIO: No music. Effects and ambience only: - rushing wind - tether firing and tension snaps - air pass-by - foot impacts on walls and rooftop surfaces - city ambience far below - distant traffic and horns - fabric movement - breathing - one short rooftop fight impact sequence - one powerful final shout from the heroine TIMELINE: 0:00–0:01 Start from black into a shocking rooftop-edge fall. The heroine is already dropping backward off a skyscraper. For a fraction of a second it feels dangerous and uncontrolled. She immediately flicks her wrist and fires one thin tether line upward. It catches instantly. The drone yanks back and reveals the start of a huge swing. 0:01–0:04 The heroine swings at high altitude between tall buildings. The city is far below. Her body forms a long aerodynamic arc, one arm holding tension through the line, legs trailing cleanly behind. The drone follows wide and slightly rolled, emphasizing height, speed, and scale. 0:04–0:06 At the swing’s forward rise, she releases the line and redirects toward a nearby glass-and-concrete building. She plants onto the wall and runs across it diagonally with 4 to 5 clear steps. Her feet hit the wall with visible force. Her jaw is set and focused. The drone stays close but wide enough to keep the city depth visible. 0:06–0:08 She pushes explosively off the wall, fires a new tether line, and swings again through a narrower corridor between tall buildings. The movement should feel faster and more controlled now. She threads cleanly through the urban gap and angles toward a rooftop landing zone ahead. 0:08–0:09.5 She releases the line and lands hard but controlled on a rooftop. Knees bend deeply to absorb impact. She rolls into a short forward recovery step, then rises immediately into a sprint across the rooftop surface. 0:09.5–0:12 One villain in a black suit steps in to stop her. Keep only this single enemy. The heroine engages him in a short, sharp rooftop fight. She avoids his first attack with a quick slip, grabs or redirects his arm, drives one fast body shot or elbow, then uses his off-balance momentum to throw or slam him down onto the rooftop. The fight must feel quick, practical, and decisive. Real impact reactions. No slow choreography. No extra enemies. 0:12–0:13.5 The villain is down and no longer a threat. The heroine steps past him and moves to the rooftop edge. Wind moves her hair and outfit. She looks outward over the city with intense adrenaline and triumph. 0:13.5–0:15 At the rooftop edge, she turns slightly toward the open skyline, lifts her chest, and shouts one powerful final line: “가자!” She immediately launches forward off the rooftop edge into another leap just as the clip ends. End on the feeling that the action is continuing beyond the cut. IMPORTANT RULES: - one continuous drone follow shot only - no cuts - no teleporting - no cloning - no multiple villains - only one black-suited villain - no giant web canopy - use only thin functional tether lines - no franchise logos - no copyrighted chest symbols - preserve the uploaded identity consistently - action must stay realistic and momentum-driven - rooftop fight must be short, sharp, and readable - final shout must be “가자!” NEGATIVE: no cartoon, no anime, no game-engine look, no fake CGI stiffness, no floating, no weightless hovering, no random disconnected acrobatics, no city-wide web canopy, no superhero logo, no copyrighted spider emblem, no extra enemies, no masked villain, no armored villain, no cloned villain, no empty city, no dark night setting, no rain, no slow motion, no blurred identity, no outfit drift, no face drift, no extra limbs, no broken anatomy, no unrealistic hand deformation, no collision with buildings during swings, no messy unreadable fight.show more

Sharon Riley
59,224 görüntüleme • 1 ay önce
a contractor in Shenzhen priced a ¥12,470,900 hospital contract,... about $1.7m, in one afternoon and beat firms carrying forty people he explained how he did it: the bid consultancy he used to pay took three days and ¥46,000 for the same envelope. he did this one alone, off one screen, at 11.4% margin, uploaded before the 17:00 cutoff 214 pages of tender documents read, 68 binding clauses pulled out, 9,485 building parts loaded, 14 places found where a duct and a beam sit in the same cubic metre, deepest one 38mm, all of them fixed, 3,318 lines of quantities priced and the package encrypted and uploaded before the 17:00 cutoff this is Graph Engineering: the job gets cut into small nodes, one narrow task each, wired so that one node's output is the next node's input, and any node is allowed to stop the whole run. it turns a model that answers you into a machine that finishes the job: - give every node one job and one output. a node doing two things fails at both and you cannot tell which one broke - put the cheapest rejection first. his qualification node reads clause 7.4, foreign-owned firms barred, and ends the run four seconds in, before anything expensive touches the model - what moves between nodes is a file. the model travels as a model, the quantities as a table, the price as a number - build exactly one loop: the checker finds 14 collisions, the fixer drops the duct 550mm, the checker runs again, and nothing moves on until the count is zero - cap that loop, or a graph will grind on three impossible clashes until the deadline passes - keep one node whose only job is to say no, and give it authority over everything above it - log each node's output on its own, because when the price comes out wrong you need to know which node believed the wrong thing - run the expensive nodes last, always the catch is that a graph is an extremely confident machine: point it at an outdated rate book and it prices an entire hospital off it without a single node noticing, because no node is asked to doubt the input, only to process it so the nodes that earn their keep are the ones that reject, and almost nobody builds those first bookmark this, the full build with all nine nodes and what each one hands to the next is written out in the article ↓show more

Argona
38,189 görüntüleme • 1 ay önce
Sadie Sink. Nano Banana Pro → Grok Imagine Prompt:... { "aspect_ratio": "4:5", "prompt": { "subject": { "description": "An adult woman Sadie Sink. Fair skin with a warm glow and visible light freckles across the cheeks and chest. Soft glam makeup with defined brows, subtle eyeliner/mascara, and a natural rosy-nude lip. Strawberry-auburn hair pulled into a high ponytail with a few loose face-framing strands. She wears large silver hoop earrings and oversized black rectangular eyeglasses. Expression is playful and confident, giving a flirty side-eye toward the camera.", "clothing": "A fitted deep-green short-sleeve cropped top with a front zipper that’s partially unzipped, ribbed knit texture. A red/orange plaid mini skirt with a structured fit and visible plaid pattern. Layered silver necklaces (one sits like a choker chain). Long pale manicure. She’s holding a large white starbucks coffee cup.", "pose": "She is reclined across the passenger seat of a car, leaning back into the seat with one arm raised behind her head, elbow bent. Her torso angles toward the camera, hips slightly turned. The other hand holds the coffee cup near her lap. Medium shot framed from head to mid-thigh, camera positioned close from the driver-side angle for an intimate in-car perspective." }, "environment": { "location": "Inside a car parked in an indoor parking garage.", "details": "Black leather car seats and door panel, visible stitching, dark interior trim. Through the window, a parking garage corridor is visible with brightly colored wall panels (yellow, teal, and red) and a white door in the background.", "background": "Clean indoor garage setting with geometric color blocks and soft depth; background slightly out of focus to keep attention on the subject." }, "lighting_and_quality": { "lighting": "Direct on-camera flash creating crisp highlights on skin, hair, and the ribbed green fabric; strong subject separation from the darker car interior with mild shadow falloff.", "resolution": "4K HD quality, highly detailed textures on freckles, hair strands, ribbed knit fabric, plaid skirt weave, leather seat grain, eyeglass reflections, and cup surface.", "style": "Realistic candid smartphone photo taken by someone else (not a selfie), reference-faithful composition and pose, sharp subject focus with slightly softer background depth, no text, no watermark, no extra people." } } }show more

HighkeySynth
22,098 görüntüleme • 7 ay önce
Beauty ads just changed forever. Free Claude Opus 4.8... + GPT Image 2 + Seedance 2.0 workflow to spin up 100s of video ads. No studio, no model, no macro lens, no shoot day. Here's what nobody in beauty marketing wants to say out loud. That glossy lip shot. The droplet hitting the surface in slow motion. The whip-pan into the next scene. The crystalline product splash. All the stuff that used to need a real set, a real camera op, and a full shoot day. You can generate every frame of it from a text prompt now, and stitch it into a finished ad before your coffee goes cold. The workflow is almost stupidly simple: → Tell Claude Opus 4.8 the beauty shot you want (dewy skin macro, gloss-on-lips contact, ripple transition, the works) → Claude turns it into a shot-by-shot storyboard plus a prompt for every frame → GPT Image 2 generates the photoreal stills, frame by frame → Seedance 2.0 animates each one into a clip with that buttery slow-mo glide → You drop the clips into HeyOz and assemble the full ad in one place The real unlock is volume. This isn't one hero video. Once the workflow is dialed, you spin up hundreds of variations. Different shades, different models, different hooks, different transitions. The exact creative volume Meta rewards, minus the production cost that used to make it impossible. Old way: one shoot, one look, $10k+, weeks of waiting. New way: a hundred angles, any look, a few dollars each, same afternoon. I wrote up the entire workflow. The Claude storyboard prompt, the GPT Image 2 frame prompts, the Seedance motion settings, the full assembly flow. Completely free, no email gate. Want it? Comment "GLOSS" and I'll send it straight over. (make sure you're following so it can actually reach you)show more

Ahad Shams
11,259 görüntüleme • 3 ay önce
Contact sheet prompting is the hottest AI video technique... right now 🤯 If you've seen this technique blowing up, here's why it works: You feed AI one image, and it generates a grid of consistent shots—same face, same outfit, different angles and poses. Instant storyboarding. Full creative control. No reshoots. But doing it manually is brutal: → Write the prompt from scratch → Generate the contact sheet → Crop each frame by hand → Feed frames into a video model one at a time → Repeat for every single product That's hours of work per campaign. This n8n automation handles everything: → Upload a character image + product image → AI analyzes both and writes the contact sheet prompt → Nano Banana Pro generates a 6-frame grid → System extracts each frame automatically → Kling 2.5 creates smooth transitions between frames → 5 video clips land in Airtable ready to use No manual cropping. No frame-by-frame prompting. No tedious busywork. What you get in Airtable: - AI-generated creative prompt - Hero image (model + product) - Full 6-frame contact sheet - 5 cinematic video clips - Approval gates before each step All inside n8n + Airtable. Contact sheet prompting on complete autopilot. I recorded a 20-minute Loom showing exactly how I built this. Want the walkthrough + the full n8n workflow + Airtable base? > Like this post > Comment "CONTACT" And I'll send it over (must be following so I can DM)show more

Mike Futia
24,606 görüntüleme • 7 ay önce
i just open sourced the workflow behind $2M AI... video productions... i built 7 skills that run the pipeline end to end, built for Seedance 2.5 and they work in Claude Code, Codex, Hermes or any harness (works best with 1080p using Higgsfield CLI) here's how to use them, in order: /setup writes which image and video models you run into your project, once, so every skill reads the same stack /studio-init scaffolds the whole studio as a file tree from one question, the project name /film-breakdown walks your script scene by scene and writes a 22-field card for every shot /reference-board locks your references into a visual bible, a caption on every image and a ban list for the rest /asset-passport writes the exhaustive descriptor every later prompt will quote word for word /stress-test combat-tests each asset and flips it to locked only at 10 out of 10 repeatability /shot-prompt refuses to run until everything in frame is locked, then writes the 15-block prompt and logs every attempt get access to the skills and full breakdown of the pipeline in the article below:show more

Machina
59,905 görüntüleme • 24 gün önce
EVERYONE PROMPTS THE ACTION. ALMOST NOBODY LOCKS THE IDENTITY... — WHICH IS WHY TWO-CHARACTER SCENES FALL APART. Two freerunners racing across Tokyo rooftops, eight cuts, corkscrews over a rooftop gap at the end. The parkour is the easy part. Keeping them two separate people who never blend into each other is the part that actually breaks. Here's the full prompt built that way. Attach two reference photos as image_1 and image_2, and the same structure works for any multi-character action piece: FORMAT: 15 seconds, 16:9, 1080p, 8-cut cinematic ultra-advanced parkour footage. CHARACTERS: Two realistic individuals from image_1 and image_2. Use the attached images as absolute character references, and fully maintain the facial features, hairstyles, hair colors, skin textures, body types, height differences, outfits, color schemes, and age appearances of each person across all cuts. No altering into different people, face swaps, outfit changes, hairstyle changes, or mixing of the two individuals' features. SETTING: A sunny modern Japanese city reminiscent of Tokyo, Shibuya, and Yokohama — rooftops, alleys, staircases, railings, pipes, concrete walls. The two protagonists, as equals, race through at high speed running side by side, following, crossing paths, and coordinating. CUTS: 1. (00:00–00:01.60) Low-angle rear tracking. Both accelerate side by side and simultaneously kong vault over separate obstacles. 2. (00:01.60–00:03.40) Front low-angle. One wall runs the left wall, the other the right, then tic-tac to cross in midair and land on opposite rooftops. 3. (00:03.40–00:05.20) Lateral tracking. Consecutive precision jumps, then cat leaps to grab and climb a high wall. 4. (00:05.20–00:07.20) Rooftop tracking. The leader dash vaults, the trailer websters over the gap, then they swap front and back positions. 5. (00:07.20–00:09.20) Overhead moving camera. Both dive roll, then run side by side to speed vault a long railing. 6. (00:09.20–00:11.30) Handheld retreating from the front. One underbars, the other side flips, conquering the obstacle simultaneously. 7. (00:11.30–00:13.20) Drone from diagonal rear above. Both palm spin off left and right walls, kong vault, accelerate into the final jump. 8. (00:13.20–00:15.00) Climax. Both leap a large rooftop gap, each doing a corkscrew, camera circling them in midair as they land on separate rooftop edges — then run side by side into the distance. QUALITY: Live-action film quality. World-championship-level smooth freerunning. Realistic center-of-gravity shifts, muscle movement, natural landing impacts, swaying hair and clothing. Sharp background, natural motion blur only during high-speed movement. PROHIBITED: Facial distortion, altering into different people, face or body swaps, outfit changes, hairstyle changes, body type changes, limb multiplication, duplicates, body fusion, penetration, warping, floating, unnatural landings, anime style, CG style. A few things worth noticing about why it's built this way: The character block does identity work three separate times — the reference images, the "fully maintain" list, and the prohibited list at the end. That redundancy isn't padding; each one closes a different door the model tends to walk through. The prohibited list names the exact failure modes — face swaps, body fusion, limb multiplication. Telling the model what not to do is more effective here than describing what you want, because these are the specific ways two-character scenes collapse. Every cut assigns each person a distinct action — one wall runs left, the other right; one underbars, the other side flips. Giving them separate roles keeps them functionally two people, so the model can't average them into one. And the cuts are individually timed and framed. Long continuous motion is where identity drift creeps in — breaking it into eight discrete shots gives the model less room to blend them. Made in Seedance 2.0.show more

Nexlow
114,217 görüntüleme • 1 ay önce