Wan-Animate-2 pushes character animation beyond pose-driven pipelines: reference image... + driving video in, cleaner character motion out. 🚀 🤖 🌐 🏆 Blind user study: preferred over Wan-Animate in 70%+ of overall-quality comparisons, with perceived quality comparable to Dreamina and Kling-MotionControl. 🎬 Cleaner animation: directly processes driving-video motion instead of relying on pose skeletons, reducing identity drift, artifacts, and shape distortion across different characters. 🎥 Text camera control: Viewpoint LoRA lets prompts steer the output perspective while keeping the original motion consistent. ⚡ Streaming path: Wan-Animate-2-Lite reports 24fps at 400x720, with stable long-sequence generation for live avatars and interactive virtual environments. Built with dual-branch DiT, Time-Align RoPE, Sparse-Ref Attention, 100K+ video pairs, and 50K Unreal Engine multi-view samples.show more

ModelScope
25,198 просмотров • 1 месяц назад
Wan Animate 2 is in ComfyUI with native support... he driving video goes straight into the transformer. No motion extractor in the middle, so nothing gets flattened before generation starts — subtle expression, finger position, all of it survives. Better motion. Stronger identity. And the camera is yours: viewpoint is text-driven, decoupled from whatever the driving clip was shot on. Wan Animate 2 Lite runs at real-time latency. Click the link below for the workflows and the blog👇show more

ComfyUI
26,186 просмотров • 28 дней назад
THE DEPTH MAP TRICK THAT FIXED DANCE ACCURACY IN... SEEDANCE 2.0 Feed the model a video of someone dancing and it tries to interpret everything- the person, the clothes, the lighting, the room, and somewhere in there, the movement. Feed it a depth map and there's nothing left to interpret but the motion. Most creators trying to transfer a dance to a character reference the source footage directly, then wonder why the choreography drifts. The problem isn't the model - it's that you handed it ten variables when you only wanted one. Here's the workflow 1. Lock the character reference in GPT Image 2 first -face, build, costume, so identity holds independently of whatever motion gets applied to it 2. Convert the source dance footage into a depth map instead of using the raw video -this strips out the original performer's appearance, clothing, and environment entirely 3. Feed the depth map as the motion reference and the character sheet as the identity reference- two separate inputs doing two separate jobs, not one input trying to do both 5. Let the depth map carry only spatial movement -the model receives body position and momentum with no competing information about who's moving or what they look like 6. Keep the character and motion inputs isolated throughout - the moment you mix appearance data into the motion reference, the model starts negotiating between two identities Why this works • Raw footage passes the model everything at once- performer, wardrobe, room, lighting -and the choreography competes with all of it for attention • A depth map is pure spatial information, so the only thing left to transfer is movement • Separating identity from motion means the character can stay locked while the dance stays accurate - normally you're trading one for the other • The accuracy gain isn't the model getting better, it's the model getting fewer decisions to make Use cases: ⁃ Dance and choreography transfer onto original characters ⁃ Motion capture-style workflows without motion capture ⁃ Any sequence where a specific movement needs to survive intact ⁃ Character showcase content built on existing performance footage The character sheet answers who's dancing. The depth map answers how - and keeping those two questions separate is the whole trick.show more

Nexlow
85,768 просмотров • 1 месяц назад
THE DEPTH MAP TRICK THAT FIXED DANCE ACCURACY IN... SEEDANCE 2.0 Feed the model a video of someone dancing and it tries to interpret everything- the person, the clothes, the lighting, the room, and somewhere in there, the movement. Feed it a depth map and there's nothing left to interpret but the motion. Most creators trying to transfer a dance to a character reference the source footage directly, then wonder why the choreography drifts. The problem isn't the model - it's that you handed it ten variables when you only wanted one. Here's the workflow 1. Lock the character reference in GPT Image 2 first -face, build, costume, so identity holds independently of whatever motion gets applied to it 2. Convert the source dance footage into a depth map instead of using the raw video -this strips out the original performer's appearance, clothing, and environment entirely 3. Feed the depth map as the motion reference and the character sheet as the identity reference- two separate inputs doing two separate jobs, not one input trying to do both 5. Let the depth map carry only spatial movement -the model receives body position and momentum with no competing information about who's moving or what they look like 6. Keep the character and motion inputs isolated throughout - the moment you mix appearance data into the motion reference, the model starts negotiating between two identities Why this works • Raw footage passes the model everything at once- performer, wardrobe, room, lighting -and the choreography competes with all of it for attention • A depth map is pure spatial information, so the only thing left to transfer is movement • Separating identity from motion means the character can stay locked while the dance stays accurate - normally you're trading one for the other • The accuracy gain isn't the model getting better, it's the model getting fewer decisions to make Use cases: ⁃ Dance and choreography transfer onto original characters ⁃ Motion capture-style workflows without motion capture ⁃ Any sequence where a specific movement needs to survive intact ⁃ Character showcase content built on existing performance footage The character sheet answers who's dancing. The depth map answers how - and keeping those two questions separate is the whole trick.show more

Nexlow
116,251 просмотров • 7 дней назад
I’ve used all the recent GenAI video models extensively... & here’s my 2¢: 🎬 Runway Gen3 Alpha - best image quality & motion for text-to-video & embedded words. Great at prompt travel changes over the course of 10 sec. And I’m super bullish on how gen3 will evolve, hopefully adopting the features listed below. Kling - best quality for image-to-video with prompt control, like eating food. Great clip extension that accounts for character (ie walking stride) & camera movement (speed & angle), rather than just using final frame. But it’s limited availability & Chinese native language is limiting. Used for Spider-Man video below (via Midjourney). LumaLabs - best for keyframe start & end control (it can not be overstated how important this is. other services should add it ASAP!) and their high dynamic action movements are really fun. Luma was used in my viral Multiverse of Memes video. PikaLabs - they haven’t gotten as much attention as others lately. But they did update their video model a few weeks ago and it looks great. Also, they are notable for their unique & AWESOME features, like video in-painting & out-painting. My perfect AI video platform would have the following features: 1) Gen3’s quality, prompt control & text embedding. 2) KLing’s image-to-video quality, prompt control & clip extension quality. 3) Luma’s multi-keyframe control & dynamic movement ability. 4) Pika’s inpainting & outpainting ability. And a video-to-video (aka next-gen Runway gen1) could be a game changer, too. It’s an exciting time to be alive 🫶 Who will get there first? 🔉🔉show more

Blaine Brown
26,535 просмотров • 2 лет назад
Google dropped a new AI paper called LUMIERE. It's... remarkably flexible, supporting video inpainting, image-to-video, AND stylized video generation tasks. Say hello to “space-time diffusion” for video generation! Now what the heck does that mean exactly?! 🌐⏳ → TL;DR it utilizes a “Space-Time UNet” architecture that generates the full duration of the video in one pass, rather than generating distant keyframes and interpolating between them like prior works. Because the computation is done in this “compressed space-time representation” to generate the full clip at once, it's far more temporally consistent. → Another benefit of generating the full video at once is that you can “direct” the video generation, making it easier to hand off to other models/tasks without having to stitch together partial solutions. You can condition generations on additional inputs, meaning you get the full stack of AI video capabilities – from video inpainting to image-to-video and beyond. → New SOTA for AI video generation? User study results in the paper suggest human evaluators preferred Lumiere over Runway Gen-2, Pika Labs, and Stable Video Diffusion in terms of quality, text alignment AND motion. But as always, we need to get hands-on with this tech when Google *actually* decides to ship it. → Could this end up inside YouTube? Y’all know i’m obsessed with blending reality and imagination – so it’s the video inpainting tech I'm most excited about. I really hope this model finds its way into YouTube's Generative AI efforts, and based on their prior announcements and the list of acknowledgments in the paper I think it might! 🤞🏽 Links: 🔗Paper: 🔗Project:show more

Bilawal Sidhu
44,822 просмотров • 2 лет назад
Turn Tom and Jerry in 4K reality using Seedance... 2.0 on Pollo AI Prompt: Use the uploaded reference video as the master reference. Recreate the entire scene in ultra-photorealistic live action while preserving the original video frame-by-frame. Maintain the EXACT camera movement, lens, framing, composition, timing, pacing, shot transitions, lighting direction, environment, props, object placement, character blocking, and every action from the reference video. ONLY replace the cartoon characters with realistic live-action animals while keeping everything else unchanged. ======================== CHARACTER CONSISTENCY ======================== Tom is a realistic British Shorthair cat with: • blue-gray plush fur • white chest, muzzle and paws • large amber eyes • pink nose • rounded face • thick tail • expressive eyebrows • identical appearance in every frame • identical fur pattern, facial proportions, eye color and body size throughout the video Jerry is a realistic golden Syrian hamster with: • soft golden-brown fur • cream belly • large rounded ears • black shiny eyes • tiny pink paws • small pink nose • realistic whiskers • consistent body proportions in every frame • identical appearance throughout the entire video If other Tom & Jerry characters appear, replace them with realistic animals that preserve their personality, colors, proportions and expressions while remaining identical throughout the clip. ======================== MOTION ======================== Preserve every movement exactly. The realistic animals must perform the exact same actions, walking cycle, head movement, eye movement, paw placement, facial expressions, timing and interactions as in the reference animation. No new actions. No altered timing. No changed poses. ======================== ENVIRONMENT ======================== Keep the original environment exactly the same. Do not modify: • furniture • decorations • room layout • colors • props • shadows • reflections • camera angle • camera path Everything except the characters must remain unchanged. ======================== QUALITY ======================== Hollywood-quality CGI. Photorealistic animals. Natural muscle movement. Physically accurate fur simulation. Realistic whiskers. Subsurface scattering. Realistic eye reflections. Natural breathing. Micro facial expressions. Ultra detailed textures. Soft cinematic lighting. Shallow depth of field. Global illumination. Ray-traced reflections. Macro photography realism. 4K HDR. Disney-level VFX quality. Live-action realism. Extremely stable temporal consistency. Perfect character identity consistency across all frames. Do not redesign the characters. Do not change the environment. Do not change the camera. Do not change the timing. Do not add new objects. Do not crop or zoom differently. No flickering. No morphing. No identity drift. No fur color changes. No eye color changes. No size changes. No anatomy deformation. No extra limbs. No duplicate animals. No cartoon textures. No low-quality CGI. No inconsistent lighting. No frame-to-frame variation. Maintain perfect temporal consistency and character consistency throughout the entire video.show more

Oogie
60,949 просмотров • 2 месяцев назад
Would you underestimate her just because she wears a... school uniform? GPT Image 2 + Seedance 2.0 on Sjolt Try Canvas: prompt Character Identity Lock (Highest Priority): Use the exact same young East Asian woman from the provided reference character sheet. Preserve 100% identical facial features, face shape, eye shape, nose, lips, skin tone, hairstyle, hair color, proportions, and overall identity throughout the entire video. Do not redesign, reinterpret, or substitute the character. She must remain instantly recognizable as the same person from the reference image. She has shoulder-length wavy silver-gray hair with subtle blue undertones, bright expressive eyes, fair skin, and a confident slight smile that naturally transitions into a focused, determined combat expression. She wears the identical navy blue Korean high school uniform blazer over a gray sweater vest, white collared shirt, striped tie, and matching school skirt from the reference character sheet. Video Prompt: A cinematic, hyper-realistic action sequence inside a chaotic South Korean high school classroom. The classroom is filled with overturned desks, scattered chairs, flying notebooks, broken pencils, and papers drifting through the air. Bright natural daylight streams through large classroom windows, creating realistic highlights, soft shadows, and cinematic contrast. The young female student moves with incredible speed, confidence, and precision as she expertly defends herself against multiple aggressive male students wearing matching Korean school uniforms. Every movement is fluid, athletic, and grounded in realistic martial arts choreography. The camera remains highly dynamic, featuring cinematic handheld tracking shots, fast push-ins, orbit shots, dramatic slow-motion moments, whip pans, low-angle hero shots, and close-up impact shots. Capture rapid combinations of punches, clean high kicks, evasive footwork, parries, elbow strikes, blocks, and throws. Desks slide across the floor, chairs topple over, and dust particles catch the sunlight, emphasizing the intensity of the action. Maintain a high shutter-speed action-photography aesthetic with crisp motion detail, subtle motion blur only during extremely fast movements, physically accurate body mechanics, realistic cloth simulation, natural hair physics, authentic facial expressions, and believable impact reactions. Keep the camera frequently returning to sharp close-ups of her face to reinforce character continuity and emotional intensity. Her silver-gray hair flows naturally with every movement while her determined eyes remain locked on her opponents. Photorealistic cinematic quality, 4K HDR, ultra-detailed skin textures, realistic lighting, volumetric daylight, physically based rendering, shallow depth of field during close-ups, blockbuster Korean action film aesthetic, empowering heroine energy, consistent facial identity throughout every frame, no face drift, no character variation, no animation-style exaggeration.show more

Sharon Riley
26,184 просмотров • 1 месяц назад
Household chores created using a movement sheet as reference... to animate the entire scene using ChatGPT Image 2.0 and Seedance 2.0 on Yapper GPT Image 2 Prompt; [VISUAL STYLE] Monochrome grayscale composition featuring a highly detailed 3D-rendered female character. Designed like a professional instructional guide with a technical, diagram-inspired layout. Clean white background, soft studio lighting, and strong contrast to highlight posture, actions, and object interaction. [GRID LAYOUT] Structured 4×4 panel grid (16 frames total), evenly spaced with thin black divider lines. Each panel is identical in size and clearly numbered from 1 to 16, showing a continuous sequence of household activities. [CHARACTER] Use the provided reference image for the face and overall likeness. Same facial features, skin tone, and proportions Natural makeup, soft expression Consistent identity across all 16 panels Realistic proportions and clean hairstyle (loose or tied back) [WARDROBE] Modern, modest casual outfit: Fitted crop top (not revealing, clean neckline, practical for movement) High-waisted straight or slightly wide-leg jeans (full length, neat fit) Optional minimal sneakers or barefoot indoor styling Fabric should react naturally to movement (subtle folds and tension) [SCENE APPROACH] Minimal, clean environment per panel — only essential props related to the chore. No clutter, no complex backgrounds — focus stays on the subject and action. [PANEL STRUCTURE – EACH FRAME] Top-left: Step number + task title (e.g., “Step 4 – Vacuum Floor”) Center: Full-body pose performing the chore Bottom-left: 3–4 concise instruction lines Overlay: Motion arrows and directional guides showing action flow [CHORE SEQUENCE EXAMPLES] Make the Bed Tidy Up Room Dust Surfaces Vacuum Floor Sweep Floor Mop Floor Do Laundry Hang Clothes Fold Clothes Clean Kitchen Counter Wash Dishes Take Out Trash Water Plants Clean Bathroom Organize Shelves Final Room Reset [MOTION INDICATORS] Curved arrows → fluid actions (wiping, folding) Straight arrows → directional movement Circular arrows → repetitive motions (scrubbing, mopping) [RENDER QUALITY] High-detail sculpted 3D style with smooth grayscale shading, soft shadows, and clean linework. Polished, concept-art level finish with clarity in every pose and object interaction. [RESTRICTIONS] No color, no unnecessary background detail, no extra characters, no revealing clothing, no clutter — only the subject, props, and instructional elements.show more

Johnn
73,260 просмотров • 4 месяцев назад
Made with seedance 2.0 + GPT Image 2 on... Yapper Prompt. Image mage1 fights three opponents inside a Japanese high school classroom in an intense, highly dynamic one-take action sequence. The classroom is filled with wooden desks, chairs, school bags, a chalkboard, sliding windows, curtains, fluorescent ceiling lights, posters, books, and scattered papers. The fight is fast, physical, and highly interactive with the environment. The woman moves between the desks with sharp agility, dodging attacks from all three opponents at once. She vaults over desks, slides across tabletops, kicks chairs into attackers, blocks strikes using classroom objects, grabs a backpack to deflect a hit, and uses the narrow aisles between desks to redirect momentum. Papers fly through the air, chairs scrape across the floor, desks topple, curtains whip from the movement, and sunlight cuts through the windows, catching dust particles in the air. The camera is extremely dynamic and close to the action, never a static wide shot. Use a fast-moving one-take camera that constantly follows, circles, ducks, whips, and pushes through the fight. The camera moves between desks, swings around the woman as she turns, rushes backward as opponents charge, drops low near the floor during leg sweeps, rises suddenly as she jumps over a desk, and whips around quickly to reveal the next attacker. The framing should feel urgent, handheld, immersive, and physically present inside the classroom. The action should feature strong choreography, realistic body movement, believable impact, fast reaction timing, close-range combat, and continuous motion. Make the scene feel like a high-budget martial arts action sequence captured in one uninterrupted shot. Use natural classroom lighting mixed with warm afternoon sunlight through the windows, realistic shadows, practical motion blur, grounded textures, real-world imperfections, and a raw cinematic look. No glossy AI finish, no overly polished CGI, and no static long-shot framing. Negative prompts: static camera, slow movement, shaky low-quality footage, blurry subject, distorted body, unrealistic swinging physics, cartoon style, flat lighting, dull colors, overexposed sky, broken buildings, empty streets, low detail, awkward camera cuts, poor motion continuity, AI glossy look, overly polished CGI, plastic-looking skin, waxy skin texture, hyper-smooth surfaces, artificial shine, fake cinematic bloom, excessive lens flare, unrealistic HDR, oversaturated colors, neon color grading, game-engine look, Unreal Engine render look, synthetic lighting, studio lighting, perfect clean reflections, overly sharp digital image, crispy AI detail, overprocessed image, fake depth of field, exaggerated bokeh, unnatural contrast, overly smooth motion, floating physics, rubbery body movement, distorted anatomy, warped limbs, inconsistent body proportions, blurry face, melted facial features, duplicated limbs, broken hands, unnatural pose, stiff action, low-quality motion interpolation, smeared motion blur, ghosting, frame blending artifacts, unstable subject tracking, camera jitter without purpose, awkward cuts, poor continuity, artificial city layout, empty streets, repeated cars, duplicated buildings, warped skyscrapers, fake traffic, low-detail background, superhero suit, comic-book look, stylized animation, overly dramatic VFX, unrealistic shadows, fake sun rays, unnatural haze, overexposed highlights, crushed blacks, sterile clean environments, no atmosphere, no real-world imperfections.show more

auqib
20,715 просмотров • 3 месяцев назад
It’s just a video player on localhost:8080. That’s what... I keep telling myself. Minimax H3 on Fish Creative Prompt Fixed-camera desktop screen recording of a real Windows 11 Chrome window filling the entire frame. Do not restyle the browser as a mockup, poster, or floating card. Keep the exact chrome from the reference still for every frame. Browser frame (locked, zero motion): - Tab title: Local Video Player - Address bar URL: localhost:8080 - Bookmarks bar left to right: Apps, Work, Personal, Google, GitHub, Docs, YouTube, then Other bookmarks on the right - Chrome controls: back, forward, refresh, extensions, profile avatar “M”, puzzle piece, shield, three-dot menu - Windows 11 taskbar along the bottom: search “Type here to search”, File Explorer, Google Chrome, Visual Studio Code, system tray, time 10:30 AM, date 5/20/2024 Page layout (labels, colors, and positions never change): - Left: large HTML5 video canvas - Under the canvas: play/pause triangle, timecode starting 00:00 / 00:10, purple seek bar, speaker, purple volume slider, fullscreen icon - Right sidebar on near-black: - “Video Player” with purple clapperboard icon - Controls: Space Play / Pause; ← → Seek ±5 seconds; ↑ ↓ Volume ±10%; F Toggle Fullscreen; M Mute / Unmute - Playlist: purple play arrow, “Sample Video”, 00:10 - Playback Rate: 0.5x, 1x filled purple and selected, 1.5x, 2x Character identity (must stay the same person in every pose): Photoreal young woman, long wavy blonde hair, gray school-style blazer over a white collared shirt, gray pleated mini skirt, black knee-high socks. Same face, same proportions, same outfit. Location stays the green grassy hill under a pale overcast sky. No costume change, no second character, no extra props. Motion rules: - Camera never moves. Browser chrome, sidebar, bookmarks, taskbar, and all UI text stay perfectly still. - Only the video canvas animates. Treat the canvas as a short fashion / pose test clip of the same girl on the hill. - Playback runs continuously at 1x. Do not pause the player. Do not flash a big pause icon. Do not cut away from the hill. - Timecode and the purple progress thumb advance smoothly from 00:00 to 00:10. Pose sequence inside the canvas (hold each pose ~1.5–2 seconds, then transition with natural body motion, hair, and cloth, not a hard cut): 1. Start pose (match the reference still): leaning forward, both hands on her knees, looking into camera with a pout. 2. Rise and stand tall: she straightens up, hands leave her knees, stands facing camera with arms relaxed at her sides, chin slightly lifted, wind moving her hair. 3. Hands-on-hips: she plants both hands on her hips, weight on one leg, slight hip cock, blazer and skirt settling, confident look into camera. 4. Turn / three-quarter: she twists her torso to her right, looks back over her shoulder toward camera, hair swinging, one hand lightly holding the hem of the skirt so it does not lift unnaturally. 5. Crouch / ready: she drops into a low athletic crouch, fingertips brushing the grass, still looking up at camera, same pout softening into a small smirk. 6. End pose: she stands again, steps one foot forward, both hands clasp in front of her thighs, then eases back toward the original lean-with-hands-on-knees so the last frame rhymes with the first still. Transitions must look like the same live-action take: weight shifts, knee bend, hair lag, grass and sky stay consistent. No teleport, no outfit swap, no extra readable text burned into the video. Audio: quiet outdoor hill ambience only. No music, no voiceover, no UI click spam. Style: photoreal live-action screen capture of a local demo page running in Chrome. Not animation, not a trailer, not a poster. Duration: about 10 seconds. Aspect: 16:9 desktop.show more

Sharon Riley
23,360 просмотров • 3 дней назад
A moment suspended between Saudi Arabia's football passion and... coffee tradition. GPT Image 2 + Seedance 2.0 on BudgetPixel AI prompt A highly cinematic, photorealistic single-shot sequence that preserves the exact original location, environment, architecture, objects, lighting conditions, camera perspective, and subject position from the source video. Do not replace, redesign, relocate, or alter the setting in any way. The person remains in the exact spot where they were filmed, maintaining their original pose, facial expression, body position, and interaction with the environment. The subject is wearing the official Saudi Arabia national football team uniform throughout the entire sequence: authentic green Saudi Arabia jersey with white details, official team crest, matching football shorts, athletic socks, and football boots. The uniform must appear naturally integrated into the original scene with realistic fabric folds, stitching, texture, shadows, reflections, and movement-free realism. Every background element, object, texture, structure, shadow, reflection, and environmental detail must remain identical to the original footage. The effect transforms the captured moment into a frozen-time cinematic sequence while keeping the real-world location completely unchanged. The subject is captured at the exact moment they pour a beverage from a transparent cup. Time has completely stopped. The liquid erupts from the cup in a dramatic suspended splash, forming elongated ribbons, twisting streams, intricate arcs, and hundreds of individual droplets frozen midair. Every droplet, splash fragment, and liquid strand appears perfectly suspended in space, creating the impression of a sculptural masterpiece made of liquid. The liquid spilling from the cup must be identical to the liquid inside the cup, with perfectly matching color, texture, thickness, reflections, transparency, and material properties. The beverage can be any type or color, but it must remain visually consistent throughout the scene with extreme realism. The subject remains absolutely motionless, frozen in the precise instant of action. Their posture, facial expression, fingertips, hair strands, jersey fabric folds, shorts texture, socks, football boots, accessories, and every micro-detail are perfectly preserved. Tiny condensation droplets on the cup, reflections on the surface, and subtle imperfections remain locked in place as if the entire world has been paused between two frames of time. The surrounding environment is equally frozen. Every object visible in the original footage remains completely static. Nothing moves. No wind, no shifting light, no falling droplets, no environmental motion. The entire world exists in a state of perfect suspension. The only moving element is the camera. The camera performs a slow, smooth cinematic arc movement around the subject, beginning from the original camera viewpoint and gradually orbiting to one side while maintaining focus on the frozen action. As the camera travels through three-dimensional space, it reveals changing perspectives of the suspended liquid sculpture, the Saudi Arabia football uniform, the subject, and the original environment. Strong spatial parallax is visible throughout the movement. Foreground droplets, liquid strands, the subject, nearby objects, and distant background elements shift relative to one another, creating a powerful sense of depth and dimensionality. The scene feels like moving through a perfectly preserved moment in time. Natural lighting remains consistent and unchanged throughout the shot. Shadows stay fixed, reflections remain stable, and materials such as glass, metal, stone, wood, fabric, football jersey fabric, embroidered team crest, and liquid exhibit highly detailed photorealistic textures. Captured with a premium wide-angle cinema lens, the scene emphasizes depth, scale, and immersive three-dimensional realism. The Saudi Arabia football uniform appears crisp, premium, and authentically detailed, with realistic fabric texture and professional sportswear quality. Core visual concept: The entire world is frozen in time exactly as it appeared in the original footage, like a hyper-detailed sculpture, while a subject wearing the Saudi Arabia national football team uniform pours a beverage that explodes into a suspended liquid masterpiece. The camera freely moves through the frozen moment, revealing dramatic parallax, depth, and cinematic realism from multiple angles. Style: Hyper-realistic, cinematic, ultra-detailed, 3D stop-motion illusion, frozen-time photography, volumetric depth, realistic lighting, film-quality rendering, smooth camera orbit, strong parallax, premium commercial sports production, museum-like suspended motion sculpture, exact environment preservation, original location consistency, photorealistic liquid simulation, luxury football advertisement aesthetic, FIFA World Cup promotional quality, 8K photorealism.show more

Sharon Riley
42,983 просмотров • 2 месяцев назад
Created with Seedance 2.0 on Yapper Prompt: Ultra-realistic cinematic... fantasy film scene, NOT animation, NOT cartoon, NOT CGI style, real human actress, live-action movie quality, photorealistic details, natural skin texture, realistic lighting and physics. A mysterious fog-covered forest lake during daytime. The sky is bright above, but dense ancient trees completely block direct sunlight, creating a soft, shadowy atmosphere. Thick white mist drifts slowly across the still water surface. A gigantic pure-white serpent is gracefully wrapped around a large branching tree extending above the lake. Its massive body naturally forms a living swing suspended over the water. The serpent appears calm, majestic, intelligent, and protective rather than threatening. A beautiful young woman sits peacefully on the serpent-swing about 1.5 meters above the lake surface. The serpent's head rests gently on her shoulder in a trusting and affectionate manner. She wears a luxurious flowing satin feminine dress with long elegant fabric trails. The lower edges of the dress softly touch the lake water below. She is barefoot, her feet hanging freely above the reflective water surface. A small adorable monkey stands on a nearby branch growing from the same tree. The monkey gently offers fresh fruit to the woman while she smiles and enjoys the tranquil atmosphere. The interaction feels natural, heartwarming, and magical. Very light rain drizzles through the mist, creating tiny ripples on the lake. Water droplets collect on the satin fabric, the serpent's white scales, and the tree bark. Soft wind causes the dress and fog to move naturally. Cinematic camera language: - Slow cinematic dolly movement. - Smooth orbiting camera around the woman and serpent. - Occasional close-up shots of the serpent's scales, the woman's expression, and the monkey feeding her fruit. - Slow-motion rain droplets falling through the fog. - Reflections visible on the lake surface. - Shallow depth of field. - Volumetric fog. - Atmospheric perspective. - Epic fantasy film composition. - Natural realistic motion. - Emotional, peaceful, dreamlike mood. Color palette: silver white, soft emerald green, mist gray, muted blue water reflections. Masterpiece cinematography, ultra-detailed, cinematic storytelling, realistic human proportions, live-action fantasy movie scene, 8K photorealistic quality, award-winning cinematography, magical realism, immersive atmosphere. > Maintain exact facial identity from the reference image, consistent face, consistent body proportions, no face changes, no identity drift, preserve facial features throughout the entire video. Negative Prompt: cartoon, anime, illustration, painting, CGI look, 3D render appearance, game graphics, low quality, low resolution, unrealistic human anatomy, exaggerated fantasy proportions, deformed hands, extra limbs, oversaturated colors, plastic skin, artificial motion, horror expression, aggressive serpent, violence, weapons, text, watermark, logo.16:9.show more

Calira
17,440 просмотров • 3 месяцев назад
Crafted with Seedance 2.0 + GPT Image 2 Pollo... AI Storyboard Prompt: Create a raw kung fu performance storyboard focused on extreme physical action. Use reference image for the character. 16:9 storyboard sheet, 12 cinematic panels. The actual storyboard drawings must be black and white only: rough pencil lines, minimal detail, fast gesture drawing energy, simple anatomy construction and strong silhouette readability. Keep the artwork lightweight, dynamic and unfinished like early fight choreography previs. Start directly in action. Do not begin with a calm stance, preparation shot or slow introduction. A solitary female performer executes an aggressive Tibetan kung fu master-style routine inside a vast ancient temple. The choreography is exaggerated, explosive and constantly escalating: flying diagonal kicks, monk-style low stances, rapid palm strikes, spinning cloth-like body turns, animal-form hand shapes, deep lunges, aerial twists, floor-level sweeps, sudden drops, claw-like blocks, back-arched jumps, sliding recoveries and violent sculptural impact poses. Every panel must contain visible motion and strong body momentum. Avoid static standing poses. The performer should feel like a ritual warrior moving with discipline, fury, spiritual pressure and total body control. Action progression: 1. begin mid-air with a flying diagonal kick already in motion 2. handheld close-up palm sweep cutting through air 3. orbiting wide shot of a full-body spin 4. low-angle impact palm strike with shockwave 5. long-lens side profile spinning kick 6. top-down aerial turn with body, hair and fabric flaring outward 7. hard floor stomp cracking the temple stone 8. sliding low sweep across the floor 9. aggressive close-up flurry of elbows, palms and backfist strikes 10. extreme low monk-style beast stance with energy rising 11. spinning elemental vortex around the body 12. final airborne action pose, suspended above the temple floor, body twisted in a powerful kung fu strike, all elements converging around her before impact Add selective elemental energy effects as VFX-style storyboard accents. The effects should feel spiritual, ritualistic and cinematic, not superhero-like: air bursts around spins and flying kicks, dust and stone fragments lifting from stomps, water-like floor ripples during slides, fire-like trails around explosive strikes, heat distortion around high-intensity movement, elemental vortex near the climax. Element progression: early panels: subtle wind, dust and pressure lines middle panels: stronger stone fragments, floor ripples and air shockwaves late panels: controlled fire trails and energy spirals final panel: the strongest combined elemental surge while the performer is still airborne Use cinematic arthouse action camerawork: handheld energy, whip-pan feeling, orbiting camera moves, overhead shots, side silhouettes, aggressive close-ups, long-lens compression, extreme low angles, wide negative space, strong parallax. Keep the temple environment minimal and atmospheric: towering stone columns, worn temple floor, drifting incense smoke, hanging fabric, harsh light shafts, faint dust in the air, subtle wet floor reflections. Do not overcrowd the frames. Annotation color system: red arrows = body movement blue arrows = camera movement green marks = framing / composition notes orange marks = lighting direction yellow marks = elemental VFX / energy effects black text = short lens notes and panel labels No timestamps. No dialogue. No singing. No extra characters. No enemies. No logos. No watermark. Seedance video prompt: Video Prompt Seedance 2.0 Create a 15-second cinematic kung fu performance video. Use Image1 as the fixed character sheet reference. The character must strictly match the character sheet. Use Image2 ] as the storyboard reference. Follow the storyboard shot by shot as the main source for action order, camera rhythm, body movement, framing, movement direction, camera angles and visual progression. Treat each storyboard panel as a sequential keyframe. Preserve the shot order and make the video feel like the storyboard has been translated into continuous live-action motion. The sequence must end on a frozen final frame while the performer is still airborne. Do not add text, captions, storyboard labels, arrows, UI, logos or watermarks. Do not treat the storyboard as a single image. Do not redesign the character, change the costume or alter the face. Do not begin with a calm stance, preparation pose or slow introduction. Do not make the elemental effects look like superhero powers or excessive fantasy glow. Visual style: stylized cinematic realism, high-end 3D painterly animation quality, dynamic cloth simulation, expressive silhouette design, rich cinematic lighting, controlled color palette, natural motion blur, dramatic scale, beautiful but aggressive physicality, premium feature-animation aesthetic. Environment: vast ancient temple, towering stone columns, worn temple floor, drifting incense smoke, hanging fabric, harsh light shafts, faint dust in the air, subtle wet floor reflections, high contrast shadows. The performance is a solitary female kung fu routine inside a vast ancient temple. The routine starts immediately in action, with no calm stance, no preparation pose and no slow introduction. The movement should feel aggressive, ritualistic, disciplined, physically extreme and spiritually charged. This is not a fight against an enemy. It is a solo performance of force, control, exhaustion, fury and release. Follow story board for choreography direction. Element progression: early sequence: subtle wind, dust and pressure lines responding to movement. middle sequence: stronger air shockwaves, stone fragments, floor cracks and water-like ripples across the temple floor. late sequence: controlled fire trails, heat distortion and energy spirals around explosive strikes and kicks. climax: wind, dust, stone, water ripple and fire accents combine into a stronger elemental vortex. final beat: the performer is airborne above the temple floor in a powerful kung fu strike, body twisted mid-air, hair and fabric flaring outward, with all elements converging around her before impact. Elemental VFX must feel spiritual, ritualistic and cinematic. The effects should be integrated with the choreography and motivated by physical movement. Keep the energy raw, elemental, atmospheric and grounded in the temple environment. Use Laban movement logic throughout: weight: strong, heavy, grounded during impacts, with brief lightness during jumps and aerial twists time: quick during strikes, kicks, drops and turns, sustained during suspended holds and recovery transitions space: direct during attacks, blocks and lunges, indirect during spinning turns and elemental vortex moments flow: bound during rooted stances and precise strikes, free during aerial motion, spinning fabric movement and elemental releaseSee lessshow more

Ciri
77,861 просмотров • 1 месяц назад
Can a dragon become your best friend? Watch this... legendary friendship come to life before your eyes. Made by using GPT Image 2 + Seedance 2.0 on Picsart prompt: Create a 15-second hyper-realistic live-action cinematic video in 16:9 with fast-paced, emotionally warm storytelling, spectacular action, and seamless multi-shot transitions. Absolute photorealism with feature-film quality, shot on anamorphic 35mm lenses, realistic camera physics, subtle handheld movement, natural lens breathing, cinematic motion blur, restrained film grain, and physically accurate lighting. The scene takes place entirely on a rugged alpine mountain summit with jagged gray metamorphic rocks, loose gravel, exposed cliff edges, dry golden alpine grass, distant mountain ranges, a deep blue sky with thin cirrus clouds and low white cumulus clouds, illuminated by crisp late-morning sunlight. Maintain perfect environmental continuity throughout every shot. The main character is a beautiful woman, approximately 25 years old, with long thick naturally wavy blonde hair, fair skin, bright blue eyes, and an athletic feminine build. She wears a weathered brown leather medieval explorer outfit consisting of a fitted leather tunic, dark trousers, tall leather boots, leather bracers, a travel satchel, a belt, and a medieval sword. Preserve her exact facial features, hairstyle, clothing, body proportions, and identity consistently throughout the entire video. Her companion is a gigantic biologically realistic pink-red dragon with dusty reptilian scales, amber eyes, curved horns, muscular limbs, powerful claws, a long tail, and large translucent wing membranes with visible veins. The dragon behaves like a real undiscovered animal, with subtle breathing, shifting muscles beneath its scales, moist reflective eyes, realistic weight, and physically accurate interactions with the environment. The dragon always remains vastly larger than the woman. A powerful alpine crosswind acts as a third character throughout the sequence, constantly influencing the woman's flowing blonde hair, clothing, satchel straps, grass, dust, loose gravel, and the dragon's wing membranes and neck spines. Every gust behaves naturally according to the terrain and camera angle, with believable delayed secondary motion. The sequence begins with an extreme ground-level camera hidden between dry grass and sharp rocks. Wind drives dust and gravel across the lens while the woman stands confidently on the exposed ridge. A gigantic dragon shadow sweeps rapidly across the landscape before one enormous wing passes overhead, dramatically darkening the frame and creating a violent pressure gust. Cut to a dynamic forward-moving perspective traveling low toward the woman as the dragon approaches at high speed. The mountain rocks rush past with strong parallax while her long blonde hair and leather clothing whip dramatically in the wind. The dragon's heavy breathing creates subtle camera movement. Transition into a fast lateral tracking shot racing parallel to the rocky ridge. Foreground boulders repeatedly hide and reveal the action while the dragon runs beside the woman with tremendous weight. Massive claws strike loose gravel, sending rocks toward the camera as dust trails behind. The dragon suddenly brakes beside her, carving deep tracks into the rocky ground while a sweeping cloud of dust fills the frame. Move into a close reverse circular orbit around both characters as the dragon gently lowers its enormous head. The woman smiles warmly, steps closer, and softly places one hand against the dragon's snout. Their foreheads gently touch in an intimate emotional moment as the dragon's folded wing temporarily shelters them from the wind. Focus shifts naturally from her fingers resting on the scales to the dragon's amber eye and finally to her genuine smile. Cut to an unusual snout-mounted close-up beside the dragon's muzzle. The dragon gives a playful snort, blasting a gust of wind that sends the woman's long wavy blonde hair, clothing, and satchel flying backward. Laughing naturally, she briefly loses her balance before affectionately pushing the dragon's muzzle away with both hands. The dragon playfully nudges her again while the camera receives a subtle physical bump, creating an authentic documentary feel. Transition to a perfectly vertical top-down aerial shot directly above the rocky clearing. The dragon unfolds its enormous wings around the woman, nearly filling the frame. A single powerful wingbeat creates a visible expanding pressure wave across the terrain, pushing dust, grass, gravel, and clothing outward in physically accurate concentric motion. The woman crouches, shielding her face while laughing as the dragon begins its powerful takeoff run. Finish with a dramatic cliff-edge aerial shot as the dragon launches directly over the camera. Loose stones fall past the lens while one translucent wing passes overhead, revealing veins, scars, and stretched organic membranes illuminated by sunlight. The camera dives backward along the cliff before stabilizing into a sweeping cinematic reveal of the mountain summit and expansive valley. The dragon performs one fast, low fly-by above the woman, whose hair and clothing are once again swept by the powerful wake. End with a wide composition of the woman standing alone on the exposed ridge as the dragon gracefully glides across the open sky above the vast mountain landscape. Maintain absolute live-action realism throughout with consistent lighting, geography, scale, anatomy, wind direction, environmental continuity, and character identity. Negative Prompt: CGI, animation, cartoon, stylized fantasy, magical effects, glowing eyes, fire breathing, supernatural particles, unrealistic physics, weightless movement, plastic textures, synthetic skin, morphing, duplicated characters, anatomy changes, inconsistent scale, extra limbs, deformed wings, inconsistent lighting, random landscape changes, HDR look, oversaturated colors, text, captions, subtitles, logos, watermarks, interface elements, low quality, blur, noise, artifacts.show more

Sharon Riley
67,220 просмотров • 1 месяц назад
Seedance 2.0 on FlovaAI =================== Prompt: [Reference Identity Lock]... Image 1 is ONLY the main female protagonist. Her face, hairstyle, body type, and outfit must match Image 1 exactly and stay consistent for the entire video. Image 2 is ONLY a uniform reference. All four opponents wear the school uniform shown in Image 2. Never swap, merge, duplicate, or blend identities. The protagonist's identity comes ONLY from Image 1. The four opponents have NO reference images. They are defined by the text descriptions below. The four opponents must not resemble the protagonist, and they must not resemble each other. All five characters must remain clearly distinct and recognizable until the end. [Priority Order] 1. Preserve the protagonist's identity from Image 1. 2. Keep the four opponents visually distinct from her and from each other. 3. Maintain one continuous shot with no cuts. 4. Keep the classroom layout spatially consistent. 5. Make the action fast but readable and physically connected. 6. Keep the tone as a Korean school action drama, stylish but grounded. Korean school action drama classroom fight scene — 15 seconds, ONE CONTINUOUS SHOT, NO CUTS. A single uninterrupted handheld shot. No cuts, no scene transitions, no montage. The camera should feel handheld, with micro-jitters, slight rolling shutter, and raw unstable realism. The camera must physically travel through the same classroom space. Every transition must be motivated by camera movement, not editing. Whip pans are allowed, but they must not hide a cut. Do not teleport the camera or characters. The classroom layout and character positions must remain spatially consistent. Audio: No music. Only realistic school and classroom ambient sounds: old fluorescent light hum, distant hallway noise, ceiling fan, shoes scraping the floor, desks dragging, chair legs screeching, cloth friction, dull body impacts, and breathing that gradually becomes heavier. Breathing continues throughout the scene and keeps building. Lighting: Late afternoon in a Korean high school classroom. Mixed cool fluorescent light and warm sunlight through the windows. Dust floating in the sunlight. Soft fan shadows moving across desks and school uniforms. Main character: The Korean female high school student from Image 1, age 17–18. Cold, emotionless, calm, and intimidating. She barely speaks and does not scream during the fight. She remains composed from beginning to end. Her movements are efficient, explosive, and precise. Even if her frame is not large, she dominates through speed, timing, and accuracy. Main outfit: Exactly the outfit shown in Image 1. Do not change its colors, design, or details. Her jacket or outer layer is either removed and hanging on a chair, or worn in a slightly messy way. The action must be non-sexualized and combat-focused. Fabric movement, dust, sweat, wrinkles, and impact response should feel realistic. Opponent rules: Four Korean female high school students, all wearing the Hanlim Multi Art School uniform shown in Image 2. They have no reference images. Define them strictly by these descriptions and keep each one consistent: Opponent A: short black bob with straight bangs, medium build, round face. Opponent B: long straight hair tied in a high ponytail, tall and lean, sharp jawline. Opponent C: shoulder-length hair with side-swept bangs, slim build, narrow face. Opponent D: long wavy hair worn loose, slightly stocky and broad-shouldered. A, B, C, and D must each keep clearly different faces, hairstyles, body shapes, and silhouettes. They must not resemble the protagonist, and they must not resemble each other. No face duplication, no face merging, no identity confusion. Environment: An empty classroom at Hanlim Multi Art School, a Korean performing arts high school in Seoul. Green chalkboard, chalk tray, worn wooden desks, plastic chairs, classroom clock, class schedule poster, discipline/life-guidance posters, cleaning tools, blinds or curtains, wall study materials, and a slightly scuffed floor. Desks and chairs should react naturally to impacts, sliding, shaking, and collapsing when hit. Camera framing rules: Even during kicks, framing should stay around chest-level or eye-level. No low-angle shots under the skirt. Do not focus on legs, thighs, underwear, or fetish-like details. All action framing must prioritize faces, upper-body motion, impact, and spatial choreography. Continuous action and camera choreography: From 0 to 15 seconds, the fight continues without any cuts. The action should be stylish but readable, and every movement must be physically connected. 0–3s: The camera starts behind the protagonist at a slightly low handheld angle, drifting left through the classroom aisle. Opponent A grabs the protagonist's shoulder roughly and says in Korean: "야, 너 지금 뭐 하자는 거야?" The protagonist silently turns and lands one hard straight punch to A's face. At impact, use a very brief 15% slow motion: cheek ripple, dust particles, deep thud. A falls sideways into a desk. The camera dips slightly from the shock, then whip-pans right without cutting. 3–6s: Opponent B charges in from the right. The protagonist steps forward instead of retreating. A short body shot to the stomach. Immediate uppercut to the chin. Without pausing, she drives forward into a flying knee to B's chest. B is thrown backward across or into a desk. The camera follows the forward motion low, then rebounds upward with the impact. 6–9s: Opponent D attacks with two fast punches. The protagonist deflects both strikes with her arms, then flows into a turning backfist to D's face. As D staggers, she continues the same rotation into a spinning back elbow that lands hard on D's jaw or temple. D crashes sideways into two or three desks. The camera arcs around her shoulder and jitters slightly at each impact. No cuts. 9–12s: Opponent C rushes in from the chalkboard side. The protagonist clearly grabs C's collar with her left hand. C's face must be fully visible from the front and clearly different from the protagonist. The protagonist lands one short, hard punch to C's face, then immediately throws a powerful high kick or flying high kick into C's chest. The force sends C backward into the green chalkboard. The protagonist remains in the foreground and never touches the board. The protagonist's face should be side-profile or partially obscured. C's face should be clearly visible from the front at the moment of impact. Their faces must never overlap in frame. Use a very brief 20% slow motion at the chalkboard impact: chalk dust bursts outward, and C slides down the board. The camera pushes up with the impact, then tilts down as C slides. 12–15s: Through the chalk dust, the camera hard-pans right. D makes one final charge. The protagonist sidesteps and lands a tight uppercut to D's chin, followed immediately by a cross. D crashes into a row of desks, causing a chain reaction of collapsing desks and chairs. The camera drifts forward slowly. The protagonist adjusts her loose tie or ribbon and brushes chalk dust off her shoulder. Her expression stays cold and serious. She walks past the camera and exits the frame. Dust floats in the sunlight. Natural ending. =================== Made with Flova #FlovaAI #FlovaCPPshow more

TSUBAKI
19,167 просмотров • 1 месяц назад
【Seedance2.0 4Kカリスマ - メンズダンステンプレ】 Dreamina( Dreamina AI )に4K来てたので使ってみました! 15秒で1890クレジットでした😏... #DreaminaAI #DreaminaCPP メンズを踊らせたいメンツいるよなー?🔊😎 と、いうわけでイケメンを踊らせたい民の皆さま… こちらがメンズ用のマルチカットダンスプロンプトでございます。 (ちょっと人外のエグいパワームーブしてますが、2回4K使う度胸はなかったですw ) キャラクターシート1枚参照させて15秒ドンっです! 👇 Prompt: Use the exact same character from the reference image. Preserve facial identity, hairstyle, outfit, colors, body proportions, silhouette, and overall design with absolute consistency in every frame. No redesigns. No facial drift. No outfit changes. 9:16 vertical format. The character performs an elite-level hip-hop dance performance with irresistible charisma, confidence, musicality, and effortless swagger. Dance style: premium commercial hip-hop, urban choreography, deep groove, musicality-driven movement, powerful chest isolations, sharp shoulder hits, controlled body waves, dynamic footwork, clean transitions, strong rhythm control, elite professional dancer quality. The performance feels natural, relaxed, stylish, and highly skilled. Never stiff. Never robotic. Continuous micro-groove throughout the performance. Natural breathing. Natural weight shifts. Authentic rhythm. Confident presence. The character commands attention from the very first frame. VIDEO STRUCTURE CUT 1 (0.0s–0.8s) Cowboy shot. The character is already moving when the video begins. Subtle groove active. At the first musical accent: sharp chest isolation, immediate shoulder hit, quick neck snap. The movement feels unexpected, precise, and highly skilled. Strong confidence. No posing. No stopping. CUT 2 (0.8s–3.5s) Full-body shot. The dancer explodes into groove-driven choreography. Strong bounce. Powerful rhythm. Clean footwork. Dynamic weight transfer. Large movement quality. Camera smoothly follows movement. CUT 3 (3.5s–5.5s) Low-angle hero shot. Large body roll. Strong directional movement. Powerful silhouette. Natural clothing movement. The dancer dominates the frame. CUT 4 (5.5s–7.5s) Medium cowboy shot. Focus on groove quality. Chest isolations. Shoulder grooves. Body waves. Clean musicality. Camera subtly tracks movement. Every accent feels satisfying. CUT 5 (7.5s–11.0s) Full-body performance section. Most impressive choreography. Traveling steps. Direction changes. Complex groove combinations. Elite dancer energy. Maximum visual impact. CUT 6 (11.0s–13.0s) Slow-motion sequence. Controlled body roll. Fluid body wave. Beautiful movement quality. Premium dance-film aesthetic. Elegant momentum. CUT 7 (13.0s–15.0s) Cowboy shot. The dancer advances toward the camera while continuing choreography. Confident swagger. Natural groove. Camera slowly pushes forward. Final musical accent. Strong finishing pose. Iconic ending frame. CAMERA STYLE Viral social-media dance cinematography. Fast pacing. Strong visual hooks. Character always remains dominant in frame. Alternating cowboy shots and full-body shots. Smooth tracking shots. Dynamic push-ins. Low-angle hero shots. No face close-ups. No unnecessary camera spinning. No empty establishing shots. No slow introduction. No static posing. VISUAL STYLE Luxury fashion campaign. High-end dance film. Premium music video production. Global superstar energy. Cinematic contrast. Rich blacks. Warm skin tones. Deep shadows. Subtle film grain. Beautiful skin rendering. Ultra photorealistic. 4K. Extremely detailed skin texture. Detailed fabric materials. Natural cloth simulation. Natural motion blur. Realistic inertia. Realistic weight transfer. Authentic professional dancer movement. 24fps. The final result should feel like a viral dance clip from a world-class performer, combining elite dance skill, irresistible charisma, cinematic quality, and maximum social media engagement.show more

Zeto
76,182 просмотров • 2 месяцев назад
POV: You have 15 seconds to decide what to... wear… and somehow end up with THREE completely different looks 😂✨ From casual to party-ready. which outfit would you choose? 👀💃 Created in Minimax H3 Prompt: Format: 15 seconds | 16:9 | 4K HDR | 24fps | photorealistic cinematic vlog Character: Use the girl from the uploaded reference image as the exact character reference. Preserve her facial identity, eye color, skin tone, hairstyle, body proportions, and overall appearance consistently throughout every shot. No face changes or identity drift. Scene & Story: 0–3 sec — The Mystery Package The video opens with a handheld creator-vlog shot in her cozy bedroom during late afternoon. She walks into frame holding a stylish garment bag, looking genuinely curious. She places it on the bed and looks directly into the camera with a playful expression. She says naturally: “Okay… I found something for tonight, but I have absolutely no clue if I can pull it off.” She unzips the garment bag, revealing only a glimpse of the outfit. 3–6 sec — First Reveal She quickly steps behind a nearby room divider. The camera follows with a fast handheld whip-pan. She emerges wearing a chic fitted top, tailored trousers, and elegant heels. She looks at herself in the mirror, slightly surprised by how good it looks. She turns toward the camera and says: “Wait… this might actually work.” She gives a small confident pose. 6–9 sec — The Unexpected Switch Instead of snapping her fingers, she grabs a jacket hanging beside her and dramatically throws it toward the camera. As the jacket briefly covers the lens, the outfit seamlessly transforms underneath it. When she pulls the jacket away, she is wearing a completely different fashion-forward look: a stylish mini dress with statement shoes and subtle accessories. She looks down, raises one eyebrow, then laughs. 9–12 sec — The Wild Card She walks toward the mirror while the camera circles around her. She notices a third outfit hanging on the wardrobe door. Suddenly, she steps backward and spins. During the spin, her outfit naturally morphs into a glamorous evening look with a sophisticated silhouette, elegant heels, and refined accessories. The transformation happens continuously during the movement with realistic fabric motion — no jump cut. 12–15 sec — Final Decision She walks toward the camera with confident energy, then stops very close to the lens. She looks at the camera, points toward herself, then gestures toward the three outfits visible behind her. With a playful smile, she says: “Okay… you’re choosing. Which one?” The camera quickly pushes in toward her face as she laughs, ending on a natural freeze-like moment. Visual Direction: Ultra-realistic fashion creator vlog, authentic handheld camera movement, natural body language, expressive facial reactions, realistic fabric physics, seamless outfit transitions, natural hair movement, subtle mirror reflections, cinematic depth of field, realistic skin texture, soft late-afternoon daylight mixed with warm practical bedroom lighting, detailed wardrobe and bedroom environment, premium fashion-film aesthetic while maintaining casual YouTube-vlog energy. Camera: Handheld selfie framing, medium shots, quick whip-pans, subtle camera shake, smooth tracking movements, one controlled orbit around the character, final slow push-in. Audio: Natural bedroom ambience, fabric and wardrobe sounds, footsteps, garment movement, subtle mirror-room reverb, upbeat modern fashion-vlog background music, natural spoken dialogue, soft whoosh effects during the outfit transitions. No subtitles, no logos, no watermark. Consistency & Quality: No jump cuts during transformations. No duplicated character. No distorted hands or fingers. No warped clothing. No inconsistent room layout. No facial identity changes. No unnatural skin smoothing. Keep the same character, bedroom, lighting logic, and visual realism throughout.show more

H A J R A
26,802 просмотров • 18 дней назад
Tried Minimax H3 on DomoAI official Prompt Here’s a... production-ready prompt rewritten for **the girl in your image**, with a **different graphic language** and **different typography**. Her look stays locked to the picture. Structure stays 13 cuts / 16:9 / 24fps. --- **PROMPT** Create an explosive, motion-graphics-driven character reveal trailer in 16:9, exactly 13 distinct cuts, 24fps, total 15.00s. CHARACTER (lock this design; never redesign): Anime streetwear girl from the reference still. Twin high buns of vivid mint-teal hair with long flowing tails and warm orange/gold streaks. Messy side-swept bangs. Large orange over-ear headphones with mint accents and a small logo plate. Sharp amber-orange eyes, one eye winking. Playful open-mouth grin. Oversized color-block windbreaker: navy body, vivid orange sleeves, white ribbed cuffs, silver zippers, circular teal tech buttons, a teal utility pocket on the sleeve. Orange cropped turtleneck under the open jacket. Light-wash ripped denim shorts, thick orange belt with a silver buckle. Navy thigh-high socks with orange ribbed cuffs and an orange X stitch on the left shin. Chunky white sneakers with orange details. Preserve exact face, proportions, hairstyle, outfit, materials, accessories and colors in every frame. This film is 80% bold graphic design in motion and 20% character action. Graphic language is **retro OS chrome + music-player UI**, not ice/frost/barcode-lock. Use slamming window frames, title bars, close/minimize widgets, equalizer bars, waveforms, progress ticks, folder tiles, cursor arrows, CRT scanlines, pixel shatter, vinyl-ring stamps, music-note particles, media-player transport icons. Palette: mint teal, vivid orange, navy, cream-white, silver. Style: premium AAA motion-graphics title sequence × streetwear campaign film × Windows-era media player. Every graphic element moves fast and snaps hard on the beat. CUT 01 | 0.00–1.00s — Pure graphics: a mint title-bar slams onto a cream field, an orange CLOSE widget punches the corner, two navy window borders wipe in. Tiny equalizer ticks and a progress strip flicker. The letters **P** then **LAY** punch in one after another with heavy impact shake. CUT 02 | 1.00–2.00s — The **A** becomes a headphone cup: extreme close-up of her amber eye inside the orange earcup, glancing up; RGB split flash, then the cup shatters into flat mint/orange tiles. CUT 03 | 2.00–3.10s — Navy frame with enormous cream **PLAY**. She sprints in from frame left and power-slides across the baseline of the type, speed lines and orange streaks trailing, shards of the letters kicked up like sparks. Whip-pan out. CUT 04 | 3.10–4.00s — A giant retro media-player waveform explodes across the frame as a thick mint-and-orange audio spectrum bends into a tunnel. She bursts through the center of the waveform at full speed, briefly breaking into three stroboscopic motion trails behind her. Each trail leaves behind chunky navy equalizer blocks that rise and collapse to the beat. The camera rapidly pushes through the waveform tunnel with her, while huge vertical text **TRACK 01** scrolls continuously in the background. The waveform suddenly compresses into a single horizontal line and snaps shut behind her on the final beat. CUT 05 | 4.00–5.10s — She leaps through a giant rotating ring of typography reading **MAX VOLUME**. Camera tracks her mid-air spin in slow motion as the letters scatter, then snap-zoom on her wink. CUT 06 | 5.10–6.00s — Hard cut: cream editorial card, huge navy **DROP** with an orange slash. She vaults over the word itself, palm planted on the D, legs whipping across frame; the word compresses like a spring under her hand and rebounds. CUT 07 | 6.00–7.00s — Mint field with a navy diagonal window-bar. She back-flips along the bar in three stroboscopic ghost frames, each ghost tinted mint, orange, or navy. Giant outlined text **LOOP** rotates 180 degrees in sync with her rotation. CUT 08 | 7.00–8.00s — Kinetic type barrage: the words **LOUD / WILD / TEAL / HEAT** slam in one per beat with shutter flashes and camera shake, while she slides on her knees across the foreground, jacket flaring, music-note particles bursting from her sneakers. CUT 09 | 8.00–9.00s — Navy frame, giant cream wireframe window-grid tilting in 3D. She runs up the grid like a wall, kicks off, and hangs in the air as everything freezes. An orange circular stamp locks around her pose like a media-player target, with transport icons and EQ ticks. CUT 10 | 9.00–10.10s — Freeze releases into a burst: she dives toward camera through layered flat-color window panes that shatter one by one like glass shutters, each pane revealing a bigger letter of **P-L-A-Y**. Foreground wipe with her sneaker. CUT 11 | 10.10–11.10s — Rapid-fire poster montage: four full-screen graphic posters of her in different action poses (mid-flip, sliding, landing, headphones-up wink) snap past with hard cuts. Each has a different bold layout, oversized numbers **01–04**, equalizer strips and slashes. CUT 12 | 11.10–13.00s — Hero moment on a clean cream cyclorama: she lands a final backflip dead center in slow motion, straightens with one hand on her headphones, and a shockwave of concentric mint rings, wind streaks and shattered typography blasts outward from her landing. Brief iconic freeze on her wink, overexpose to white. CUT 13 | 13.00–15.00s — Final identity card: enormous navy **PLAY** dominates a pale cream field with translucent mint rings, technical arcs, scanlines, and a rough orange circular emblem containing a ghosted headphone / waveform motif. She stands relaxed overlapping the letters, wind rippling her jacket. One last orange pulse sweeps through the typography and a window-chrome flash punctuates the end. Editing: extremely aggressive rhythm — hard cuts on every beat, graphic matches, whip pans, snap zooms, stroboscopic freezes, foreground wipes, RGB splits, shutter flashes, impact shakes and speed ramps. Every cut must feel compositionally different. Typography is always fully readable before the character overlaps it. No weapons, no combat, no fire. All energy comes from motion design, wind, glass, UI chrome and her parkour-style athleticism. BGM: hard-hitting electronic / drum-heavy future bass with aggressive drops, risers, sub hits and glitch fills locked to every cut. Sneaker impacts, whooshes, glass shatters, window-slam hits and typography slams are rhythmic sound-design elements. Peak at CUT 12 and end with a cold electronic logo stinger. Premium AAA quality, anime-inspired cinematic rendering, stylish and explosive, strong graphic-design identity, consistent character design, exactly 13 cuts. #DomoAI #DomoAICPPshow more

Sharon Riley
16,537 просмотров • 8 дней назад
Who could resist a girlfriend this sweet—and full of... surprises? 🥹 A big thanks to the original creator for sharing this delightful prompt. From adapting a reference image and matching the character to crafting the final video prompt, you can create the whole concept from scratch here: prompt: Vertical 9:16 format, 10.0 seconds, 60 fps. Realistic smartphone portrait footage with an immersive first-person interaction aesthetic, captured as one continuous uncut shot. Extreme close-up of a young adult East Asian woman, with her face occupying almost the entire frame as she rests on a soft white pillow. Her smooth, glossy black medium-length straight hair is naturally tousled across the bedding. Preserve the reference image’s distinctive wispy blunt bangs and loose face-framing strands, along with her fair, luminous complexion. 【Character Identity and Styling】 Strictly preserve the identity and facial appearance of the woman in the uploaded @ image_1: maintain the same face shape, facial-feature proportions, eye shape, eyebrows, nose, lips, hair color, wispy blunt bangs, face-framing strands, skin tone, and apparent age. She must remain the same person throughout the entire video. Her facial features must not change with movement, camera angle, facial expression, or distance from the camera. Use @ image_1 only as the reference for the woman’s identity, facial features, hairstyle, and makeup aesthetic. Do not inherit the black outfit, arm-supported pose, room background, pink heart stickers, text, watermark, or any other graphic overlays from the reference image. The woman from @ image_1 is wearing a beige-pink floral spaghetti-strap nightdress. Visual styling: preserve the delicate, translucent “tearful makeup / slightly tipsy makeup” aesthetic seen in the reference image, including naturally long upward-curled eyelashes, soft pink under-eye blush, subtly defined aegyo-sal, fine shimmering highlights at the inner corners of the eyes, glossy glass-like lip color, and watery translucent gray-brown contact lenses matching @ image_1. Her overall appearance should feel soft, innocent, affectionate, and full of fresh morning energy immediately after waking up. 【Camera and Lighting】 First-person POV from the boyfriend’s perspective. The opening shows a high-angle view looking down at the woman lying on her side in bed. During the middle and final sections, her sudden pounce causes natural mattress movement, realistic camera shaking, and an extremely close face-to-face perspective. The background is a clean, warm morning bedroom with light-colored bedding. Bright, soft morning light enters from the upper side as diffused illumination. Use high-key exposure and low contrast while preserving the subtle rise and fall of her breathing, the natural volume of her cheeks, and her moist, translucent skin texture. 【Action and Rhythm】 Within 10 seconds, complete a continuous emotional reversal: from “sleepily holding his hand and asking him to stay,” to “the boyfriend teasing her affectionately,” and finally to “suddenly waking up and energetically pouncing on him, pushing the camera back onto the bed.” Dialogue and physical actions must be precisely synchronized. The pacing should feel intensely sweet, playful, and immersive. 【Second-by-Second Timeline】 0.00–2.50 seconds | Part 1: Sleepily Reaching Out and Asking Him to Stay — “别走可以吗?” The boyfriend is about to get up and leave, causing the camera to make a subtle upward movement as if he is rising from the bed. The woman, lying on her side, sleepily opens her eyes. One fair-skinned hand instinctively reaches out from beneath the blanket and firmly holds the boyfriend’s wrist, refusing to let go. She slightly pouts her lips and gazes softly and innocently into the camera. In a gentle, sleepy voice, she says in Mandarin, with precisely synchronized lip movement: “别走可以吗?” 2.50–4.50 seconds | Part 2: Affectionate Interaction and Playful Question — “不上班你养我呀?” The boyfriend reaches out with his other hand and affectionately gives her soft cheek a gentle pinch. Her eyes remain sleepily half-open, and she rubs her cheek against his palm like a small cat. The boyfriend’s warm, smiling off-screen voice asks in Mandarin: “不上班你养我呀?” 4.50–7.50 seconds | Part 3: Instantly Waking Up and Pouncing Forward — Waking Up & Pouncing After hearing his question, her previously drooping eyelids instantly open wide, and she becomes fully alert within a second. A flash of playful cunning and excitement appears in her eyes. She plants both hands on the bed and suddenly lunges forward, directly pouncing on the boyfriend from the first-person perspective and pushing him back down onto the soft mattress. The camera experiences a realistic sense of weightlessness, strong natural shaking, and visible mattress compression as the bed sinks under their weight. 7.50–10.00 seconds | Part 4: Close-Up Dominant Gaze and Playful Declaration — “我养你呀!” The camera is now lying flat against the pillow. The woman supports herself with both hands placed on either side of the boyfriend’s body and looks down into the camera from an extremely close distance. Her wispy blunt bangs and glossy black strands of hair naturally fall around the lens. Her watery gray-brown eyes curve into smiling crescents, and her lips form a bright, proud, delighted smile. She gives the camera a playful wink and says energetically in clear Mandarin, with precisely synchronized lip movement: “我养你呀!” End on a frozen moment of intimate, heart-fluttering eye contact. 【Mandatory Continuity Rules】 Maintain throughout the entire video: 10 seconds, 60 fps, vertical 9:16 format, first-person boyfriend POV, extreme close-up facial interaction, and dynamic bed movement caused by the pounce. The pouncing action must be quick, decisive, and physically believable, with natural gravity, body momentum, mattress compression, and camera response. Her expression must shift instantly from sleepy and drowsy to bright, energetic, and playful. Maintain intense, affectionate eye contact throughout, creating a highly immersive and emotionally impactful sweet morning interaction. Strictly lock the character identity, wispy blunt bangs, face shape, and facial-feature proportions from @ image_1 throughout the video. No identity changes, face swapping, facial-feature drift, hairstyle changes, eye-color changes, age changes, excessively heavy makeup, plastic-looking skin, malformed fingers, extra limbs, body distortion, or physical clipping.show more

underwood
19,217 просмотров • 13 дней назад