Exploring diff visuals with Gen2. I'm amazed at how... it takes a single block image to build neon streets, structures, and entire cityscapes. Anybody else toying with AI tonight⁉️ AI Art Workflow: 1. Crafted starting image in #Midjourney 2. Generated a 4-sec video with #Gen2 (no text prompt) 3. Extended first 4 secs to 18 secs 4. Repeated steps 2-3 twice, each time using the final frame from the previous 18-sec video to start the next 18-secs run 5. Merged all three 18-sec videos in #FinalCut Pro 6. Adjusted speed for fluidity 7. Final version published here is a 54-sec journey distilled into a 16-sec video with 🎶 #AIart #AIArtCommunityshow more

Dave Villalva
47,946 次观看 • 2 年前
🤯 K-pop meets a full-on water obstacle challenge! 🌊... Made with Seedance in a Korean variety show style Full workflow: 1. 4:3 character sheet for the girl, swimsuit and expressions 2. 3:1 wide course image for the pool set, audience and obstacles 3. Seedance image-to-video for the full run and final splash From the confident start to the final fall, the character and scene consistency held up really well! Workflow + Prompt in the comments 👇show more

Larus Canus
45,536 次观看 • 10 天前
🎥 Comparing AI video models: Image to video •... Gen-3 • Kling AI 1.5 • Hailuo MiniMax • Luma Dream Machine I used a Midjourney image in each model 4 times with no text prompt. This type of image is difficult for the AI to separate the subject in the front from the people behind her - they tend to move the group as if they are a single unit. But the results were interesting, as you can see! I chose my favorite results, below.show more

Heather Cooper
41,354 次观看 • 1 年前
I watched so you did not have to and... here is the breakdown!!!! Nearly 2 hours. Plenty of nostalgia. Plenty of vibes. Lots of feelings. But when the net was wide open for actual clarification? Dead air. Total time Chris Albert spoke 1,501 seconds Which is 25 min and 1 sec. So on an almost 2 hour show, he spoke for about 25 minutes total. Roughly 21% of the entire show The remaining time was largely filled by the hosts themselves frequently cutting him off, reframing questions, finishing answers for him, or moving the conversation along before he ever had to sit in discomfort or explain anything fully. 0 minutes of that time addressed the major unresolved timeline and movement issues Oddly missing from this fact-based discussion No explanation for how Colin got home 20 min AFTER Chris that night, despite a four-minute head start No explanation for why Chris’s wife that went to 34 Fairview with donuts for brian Jr. pulled behind a car that was not there only an hour earlier with no explanation then, turned around, came back to get him, returned again, stayed 30 min, then left, all on the morning John was found. No mention of parents allegedly coordinating to quietly clean up a house-destroying party involving minors in which his son was one No discussion of repeated allegations of Colin being removed from bars for underage drinking. We were told this channel is about facts, not opinion. So we treated it like a data exercise. We timed everything Chris talked about, so you didn’t have to. Here’s how nearly two hours were spent 2 min saying they tried to “let the court system play out,” while explaining his New Year’s Day Twitter Space tour and condemning Sean McCabe’s comments 37 sec on moving to Canton in 1974 35 sec explaining it’s normal to work in the town you grew up in (we agree) 35 sec on Canton no longer being blue-collar and now very expensive 44 sec reminiscing about red paint stains in his childhood garage was misrepresented by the defense 1 min on town politics 15 sec explaining what a Select Board is 50 sec on handling public comment and being a “First Amendment guy… to a degree” 38 sec announcing he’s not running again 25 sec on how he keeps his cool during public comment 15 sec on pineapple pizza 52 sec explaining Colin was 17 when John died and can’t play college football with his cousin 45 sec expressing loss for how his family has been treated 10 sec referencing a blogger saying “horrible things” about Colin 26 sec blaming the media for selling and sensationalizing a story 20 sec saying he’s mind-boggled 48 sec explaining why the term “McAlberts” is dehumanizing 28 sec criticizing Yannetti’s performance was upsetting to him while stating John was his friend 42 sec explaining the fence-straddling photo was a joke 1:50 on “Nebbercracker” and a childhood yard incident 10 sec about his wife bringing donuts to a nephew for his birthday was not weird 32 sec blaming Karen for making his family look like villains 1 min on selling his house and convincing his brother to sell 24 sec saying Nebbercracker was the best part of the trial 1:12 on knowing Michael Proctor but not really knowing him 10 sec saying he was never alone with Proctor 11 sec correcting the “pin it on the girl” claim (which we agree with) 1 min praising Jillian Daniels as a great family supporter and defender 1 min explaining early lack of support made him sad 15 sec thanking the hosts 31 sec discussing starting a “Defend the Truth” media page 30 sec saying both sides were horrible and doxxed people 24 sec on feeling isolated at first 33 sec minimizing the Proctor gift as “just gratitude for solving a murder in one day” 28 sec saying he’s still friends with the O’Keefes and they eat at his pizza shop 38 sec saying Mr. O’Keefe sits with him for long periods of time 43 sec lamenting being made into villains 30 sec saying the case took years off his parents’ lives 35 sec thanking supporters and saying “the truth is on our side”show more

Dixie Normus
17,063 次观看 • 6 个月前
NVIDIA just released a very impressive text-to-video paper. Video... Latent Diffusion Models (Video LDMs) use a diffusion model in a compressed latent space to generate high-resolution videos. Here's a brief overview of how it works: 1. Pre-train image LDM on a dataset of images. 2. Turn the image LDM into a Video LDM by adding temporal layers to model video frames. 3. Fine-tune the Video LDM on encoded video sequences to create a video generator. 4. Temporally align diffusion model upsamplers to generate high-resolution videos. 5. Validate Video LDM on real driving videos of 512x1024 resolution, achieving state-of-the-art performance. 6. Apply the approach in creative content creation with text-to-video modeling. Paper: Project:show more

Lior Alexander
158,565 次观看 • 3 年前
I’ve used all the recent GenAI video models extensively... & here’s my 2¢: 🎬 Runway Gen3 Alpha - best image quality & motion for text-to-video & embedded words. Great at prompt travel changes over the course of 10 sec. And I’m super bullish on how gen3 will evolve, hopefully adopting the features listed below. Kling - best quality for image-to-video with prompt control, like eating food. Great clip extension that accounts for character (ie walking stride) & camera movement (speed & angle), rather than just using final frame. But it’s limited availability & Chinese native language is limiting. Used for Spider-Man video below (via Midjourney). LumaLabs - best for keyframe start & end control (it can not be overstated how important this is. other services should add it ASAP!) and their high dynamic action movements are really fun. Luma was used in my viral Multiverse of Memes video. PikaLabs - they haven’t gotten as much attention as others lately. But they did update their video model a few weeks ago and it looks great. Also, they are notable for their unique & AWESOME features, like video in-painting & out-painting. My perfect AI video platform would have the following features: 1) Gen3’s quality, prompt control & text embedding. 2) KLing’s image-to-video quality, prompt control & clip extension quality. 3) Luma’s multi-keyframe control & dynamic movement ability. 4) Pika’s inpainting & outpainting ability. And a video-to-video (aka next-gen Runway gen1) could be a game changer, too. It’s an exciting time to be alive 🫶 Who will get there first? 🔉🔉show more

Blaine Brown
26,535 次观看 • 2 年前
This AI UGC workflow is f*cking nuts 🤯 One... prompt -> five completely different characters + five full ad variations, all generated automatically in a single run. Perfect for DTC brands and creative agencies who need volume but don't have time to build each variation from scratch. Most people making AI UGC are doing it one video at a time. New character, new prompt, new script, new render. Repeat. It works, but it doesn't scale. By the time you've built 5 variations to test, you've burned half a day. This workflow solves it: → Enter one initial prompt with your product and angle → Auto-generates 5 unique AI characters → Builds hook scripts for each variation → Writes the bridge with your actual product image → Creates the CTA — all 5 versions in one shot No building each video manually. No copy-pasting prompts over and over. No bottleneck between idea and creative testing. What you get: > 5 unique character variations from a single prompt > Hook, bridge, and CTA scripts tailored to each character > Product image integration baked in > A repeatable system you can run every time you need fresh creative Built 100% with AI. Want a copy of the full workflow for free? > Like this post > Comment "UGC" And I'll send it over (must be following so I can DM)show more

Mike Futia
15,225 次观看 • 5 个月前
Made a fast-paced AI beauty vlog with cinematic camera... work, realistic makeup application, and influencer-level aesthetics. Honestly feels like scrolling through TikTok... except every frame was AI-generated. ✨ Thoughts? Made with Seedance 2.0 Prompt: 15-Second Fast-Paced AI Influencer Daily Makeup Tutorial Prompt Create a 15-second ultra-realistic, fast-paced cinematic beauty vlog featuring a gorgeous 22-year-old female beauty influencer in a bright, luxurious vanity room with a large LED mirror, elegant décor, and soft natural morning light. The video has energetic editing with quick 1–2 second cuts, smooth camera movement, trendy transitions, speed ramps, and close-up beauty shots. The vibe is premium, aesthetic, and perfect for TikTok or Instagram Reels. Scene 1 (0–2 sec): She appears makeup-free, smiling at the camera while holding a makeup brush. Dialogue: "Come get ready with me in under a minute!" Scene 2 (2–5 sec): Quick montage of applying primer, foundation, and concealer with seamless jump cuts. Dialogue: "Keeping it light, fresh, and glowy for today's look." Scene 3 (5–9 sec): Fast cuts of blending blush, adding bronzer, highlighter, and brushing through her brows. Close-up shots capture the smooth application. Dialogue: "A little blush and glow... because we're going for that effortless look." Scene 4 (9–12 sec): Rapid montage of neutral eyeshadow, mascara, and a glossy nude lip. She flips her hair and smiles at the mirror. Dialogue: "Mascara, lip gloss... and we're basically done!" Scene 5 (12–15 sec): Final reveal. She confidently walks toward the camera, showing off her finished makeup with a glowing smile. Dialogue: "Simple, fresh, and ready for the day. What's your favorite makeup step?" Style: Hyper-realistic 8K, cinematic beauty-commercial quality, ultra-smooth gimbal shots, macro close-ups of makeup application, soft golden lighting, flawless skin texture, shallow depth of field, vibrant color grading, realistic lip-sync, luxury influencer aesthetic, fast-paced editing with trendy transitions, no text overlays, no logos, no watermarks.show more

Maria
23,942 次观看 • 1 个月前
Create a 3D model from a single image, set... of images or a text prompt in < 1 minute 😮💨 This new AI paper called CAT3D shows us that it’ll keep getting easier to produce 3D models from 2D images — whether it’s a sparser real world 3D scan (a few photos instead of hundreds) or your favorite 2D image generator like Midjourney (just an image). How does this magic work? “This architecture is similar to video diffusion models, but with camera pose embeddings for each image instead of time embeddings. The generated views are passed into a robust 3D reconstruction pipeline to create the 3D representation (Zip-NeRF or 3DGS)”show more

Bilawal Sidhu
92,792 次观看 • 2 年前
I tried Hailuo AI (MiniMax) to see how it... handles real content creation. The workflow is simple. You just write a prompt or drop in an image, and it turns that into a dynamic video with motion, framing, and scene depth. No timeline to manage. No editing setup. No back and forth. What stood out to me: • Text to video and image to video both feel smooth. • It handles motion, camera angles, and flow on its own. • Output is fast, usually within seconds. • Works well for reels, quick ads, storytelling, and idea testing. It removes the hardest part: starting from scratch and turns your ideas into content in minutes. Instead of thinking, “How do I make this video?” You start with, “What do I want to create?” That shift alone makes it worth exploring. Try it here: #Hailuoshow more

Manish Kumar Shah
27,662 次观看 • 3 个月前
📖THE BEST SEEDANCE 2.0 WORKFLOW STARTS INSIDE CHATGPT IMAGE... 2 One creator can now go from storyboard to cinematic short film without a camera, actors, or a production crew Most creators use Seedance 2.0 as a video generator. The real power comes from combining ChatGPT Image 2 + Seedance 2.0 into a production pipeline. Here’s the workflow: 1.Create a story idea in ChatGPT. 2.Generate a shot-by-shot storyboard. 3. Build character sheets to lock consistency. 4.Generate every scene in ChatGPT Image 2. 5.Define camera movements and actions. 6.Animate each shot in Seedance 2.0. 7.Stitch the clips together into a finished short film. Why this works: • Consistent characters across scenes • Better storytelling • Precise camera control • Faster iteration • Professional-looking results • One person can do the work of an entire production team Use cases: ⁃Viral AI shorts ⁃YouTube animations ⁃Brand commercials ⁃Educational content ⁃Story-driven ads ⁃Social media content The future of AI video isn’t prompting. It’s building production pipelines. 📥Tomorrow I am sharing a workflow that almost nobody is using but absolutely should be. 🔖 The full breakdown from setup to final export is waiting in the pinned article.show more

Zentrix⌚️
49,385 次观看 • 1 个月前
Contact sheet prompting is the hottest AI video technique... right now 🤯 If you've seen this technique blowing up, here's why it works: You feed AI one image, and it generates a grid of consistent shots—same face, same outfit, different angles and poses. Instant storyboarding. Full creative control. No reshoots. But doing it manually is brutal: → Write the prompt from scratch → Generate the contact sheet → Crop each frame by hand → Feed frames into a video model one at a time → Repeat for every single product That's hours of work per campaign. This n8n automation handles everything: → Upload a character image + product image → AI analyzes both and writes the contact sheet prompt → Nano Banana Pro generates a 6-frame grid → System extracts each frame automatically → Kling 2.5 creates smooth transitions between frames → 5 video clips land in Airtable ready to use No manual cropping. No frame-by-frame prompting. No tedious busywork. What you get in Airtable: - AI-generated creative prompt - Hero image (model + product) - Full 6-frame contact sheet - 5 cinematic video clips - Approval gates before each step All inside n8n + Airtable. Contact sheet prompting on complete autopilot. I recorded a 20-minute Loom showing exactly how I built this. Want the walkthrough + the full n8n workflow + Airtable base? > Like this post > Comment "CONTACT" And I'll send it over (must be following so I can DM)show more

Mike Futia
24,559 次观看 • 6 个月前
Contact sheet prompting is the hottest AI video technique... right now 🤯 One image in → 6 consistent frames out → cinematic video ads in minutes. But everyone's doing it manually. I automated the entire workflow in n8n + Airtable. Here's why contact sheet prompting is blowing up: You give AI one reference image, and it generates a grid of consistent shots — same face, same outfit, different angles. Instant storyboarding, full creative control, no photoshoots. The problem? It's super tedious: → Write the prompt manually → Generate the contact sheet → Crop each frame by hand → Feed frames into a video model one by one → Repeat for every product This n8n automation handles all of it: → Upload character image + product image → AI analyzes both and writes the contact sheet prompt → Nano Banana Pro generates a 6-frame grid → System extracts each frame automatically → Kling 2.5 generates smooth transitions between frames → You get 5 video clips ready to stitch Approval checkpoints at every stage, no surprises. What lands in your Airtable: → AI-generated creative prompt → Core hero image (model + product) → 6-frame contact sheet → 5 cinematic video clips → Full control before each generation step Contact sheet prompting on autopilot. I filmed a 20 minute Loom video showing you exactly how I set it up. Want the Loom + the complete n8n workflow + Airtable base? > Comment "SHEET" > Like this post And I'll send it over (must be following so I can DM)show more

Mike Futia
53,632 次观看 • 7 个月前
IF I WAS FORCED to build a $20K/month AI... creative agency using nothing but Photoshop, starting from 0, here's exactly what I would do in steps: The production setup (Days 1–3) 1. Download the Higgsfield plugin inside Photoshop — takes 5 minutes 2. You now have: sketch-to-image, layer decomposer, mockup studio, relight, upscale, face swap, character swap, background removal, AI stylist — all in 1 tool 3. Old creative agency workflow: designer + photographer + editor + 3–5 day turnaround 4. New workflow: 1 person, Photoshop, 30 minutes per deliverable The offer (Days 3–7) 5. Pick 1 niche — ecom brands, real estate agents, or course creators all need visuals constantly 6. Build a simple offer: "10 ad creatives delivered in 24 hours — $500" 7. Old agencies charge $2,000–$5,000/month for the same output 8. Your cost to deliver: $0 beyond the plugin. Pure margin. 9. Create 3 sample mockups using the tool — drop a product image in, generate 9 variations, pick the best 3 10. That's your portfolio. Built in under 1 hour. Cost: $0. The client machine (Days 7–20) 11. Go on X and search "[niche] + need a designer" or "[niche] + creatives" 12. DM 50 people per day — "I'll make you 3 free ad creatives in 24 hours, no catch" 13. Deliver them in 30 minutes using the plugin 14. 50 DMs/day × 14 days = 700 outreach messages 15. Conservative 3% conversion = 21 people see the free work 16. Close 5 of them at $500 = $2,500 in week 3 The scale (Days 20–30) 17. Upsell every client to a $1,500/month retainer — 10 creatives/week, unlimited revisions 18. 1 client per day in Photoshop takes 45 minutes max 19. 10 retainer clients × $1,500 = $15,000/month 20. Add 3 one-off clients at $500/month = $1,500 21. Add a $997 "AI creative system" course teaching other people this exact workflow = $3,000+/month from 3 sales The math: 50 DMs/day × 30 days = 1,500 outreach messages 3% book a call = 45 calls 40% close at $1,500/month retainer = 18 clients 18 × $1,500 = $27,000/month recurring Time per client per day: 45 minutes Total daily work: 4–5 hours Every mockup — AI. Every restyle — AI. Every layer rebuild — AI. Every variation — AI. No photographer. No designer and no reshoot. Start it here. 👇show more

ALEX SUZUKI
20,557 次观看 • 1 个月前
Golden mornings. Fresh air. Clear mind. 🌅✨ Sometimes all... it takes is a 15-minute walk to reset your entire day. Made with Seedance 2.0 on TapNow Prompt: Create a 15-second ultra-realistic, fast-paced cinematic lifestyle vlog. A beautiful 23-year-old female influencer wearing white T shirt with black short is on her morning walk in a lush green park during golden sunrise. The editing is energetic with quick 1–2 second cuts, smooth gimbal shots, dynamic transitions, speed ramps, and cinematic close-ups. The vibe is fresh, motivating, and aesthetic, similar to viral Instagram Reels and TikTok lifestyle content. Scene 1 (0–2 sec): Close-up of her tying her white sneakers, grabbing a water bottle, and opening the park gate. Dialogue: "Morning reset... let's go!" Scene 2 (2–5 sec): Selfie shot while walking confidently through the park with sunlight behind her. Dialogue: "Nothing beats fresh air at sunrise." Scene 3 (5–8 sec): Fast montage of her walking, smiling, birds flying, trees swaying, smartwatch tracking steps, and shoes hitting the path. Dialogue: "Just fifteen minutes can completely change your mood." Scene 4 (8–12 sec): Slow-motion hair flip, laughing, stretching, and enjoying the golden sunlight. Dialogue: "Move your body... clear your mind." Scene 5 (12–15 sec): She turns toward the camera with a bright smile and gives a thumbs-up while continuing her walk. Dialogue: "See you tomorrow for another morning reset!" Style: Hyper-realistic , cinematic golden-hour lighting, vibrant colors, luxury influencer aesthetic, ultra-smooth camera movement, quick jump cuts, speed ramps, natural expressions, realistic lip sync, shallow depth of field, crisp ambient nature sounds, high-energy editing, no text overlays, no logos, no watermarks.show more

Maria
17,855 次观看 • 1 个月前
Creators are creating insane videos with AI. I created... this short ad video using GPT Image 2 + Seedance 2.0 on Creatify AI. Prompt: Create a 15-second cinematic luxury beverage advertisement featuring the same woman throughout all scenes. A beautiful young woman with long, wavy brunette hair wears a flowing light floral maxi dress on a tropical beach. She carries and drinks a chilled Summer Club Pogmosa can with visible condensation droplets. Bright tropical sunlight, premium lifestyle aesthetic, realistic photography, natural movements, luxury beverage campaign style. Scene 1 (0-2 sec) – Arrival Wide cinematic shot. A woman walks down a wooden beach pathway toward a pristine tropical beach carrying a woven beach bag. Palm trees sway gently in the breeze. The camera follows behind her. Scene 2 (2-4 sec) – Discovery Medium shot. She sits near the shoreline, opens her beach bag, and pulls out a cold Summer Club can. Condensation glistens in the sunlight. She smiles naturally. Scene 3 (4-6 sec) – Open & Refresh Close-up product shot. Her hand slowly cracks open the can. Crisp opening sound. Tiny water droplets sparkle. Slow-motion detail shot. Scene 4 (6-8 sec) – First Sip Close-up portrait. She takes a refreshing sip while looking toward the ocean. Wind gently moves her hair and dress. Warm sunlight highlights her face. Scene 5 (8-10 sec) – Beach Walk Full-body tracking shot. She walks barefoot along the shoreline holding the can. Waves roll in beside her. Dress flows beautifully in the sea breeze. Scene 6 (10-12 sec) – Summer Toast Medium shot. She raises the can toward the camera with a bright smile as if making a toast. Ocean and palm trees softly blurred in the background. Scene 7 (12-15 sec) – Sunset Finale Wide cinematic sunset shot. Woman sits on a beach blanket watching the sun set over the ocean. Summer Club can beside her. Camera slowly pulls back revealing the golden beach landscape. Visual Style Ultra-realistic commercial photography Luxury beverage advertising Tropical paradise setting Golden-hour lighting Cinematic camera movement Shallow depth of field Natural skin tones Smooth transitions between scenes Premium brand aesthetic High-end fashion and lifestyle campaign 4K HDR quality 24 fps cinematic look Soft lens flares Vibrant summer colors End Frame: Product hero shot of the chilled Summer Club Pogmosa can on the beach with ocean waves in the background. Text overlay: "Summer Club" "Taste Summer Anywhere" Premium beverage commercial ending with logo reveal and sunset glow.show more

Aaliya
11,178 次观看 • 2 个月前
🔥 VIDU Multi-Entity Consistency Give Vidu 2/3 images and... it’ll turn them into a video—it’s pure magic! ✨ Your own characters interacting with objects and in the exact environment you want! Ads, movies… endless possibilities, and this is just the beginning! Thanks @Viduforhuman The future is a carrot! 🥕 Plus, how about grabbing any frame from a Vidu-generated video "from scratch" and throwing it into another AI video or image tool to push your project even further? For now, check out the comment below: I scaled up a frame with Magnific.ai and fed it into Runway to create a dynamic shot using full camera control. But fingers crossed I can soon use #ReCapture by Bisho & team to generate new shots from the same video!show more

Hungry Donkey 🥕
37,561 次观看 • 1 年前
One trick we discovered for avoiding realistic face moderation... issues in Seedance 2.0 is using character turnaround sheets (front / side / back views). The first video is one of our experiment results — and it runs successfully. We’ve now integrated character turnarounds directly into our workflow + canvas system: 1. If your artwork was generated on our site, you can drag the image into the canvas directly from the Assets tab 2. Click the “Character Turnaround” button above the image to automatically generate a 3-view turnaround sheet 3. Create a new video node and use the turnaround sheet directly with Seedance 2.0 inside the workflow I’ve shared the workflow link in the comments if you want to explore the exact prompts, setup, and workflow details.show more

underwood
32,399 次观看 • 2 个月前
Step 1 : input your char reference, poster and... your product into GPT Image 2 and then put the prompt : Use the woman on image 1 as the main subject. create a vertical poster ad inspired by the reference poster style. show the image 1 woman holding the kimchi jar with her left hand, while the right hand eating kimchi to her mouth. with framing wide lens, low angle from the reference poster. This step is where you lock the visual direction. Step 2 : take the generated image into Seedance 2.0 and convert it into motion. Set it to 1080p if your design includes a lot of typography, this helps preserve text clarity. You can also strengthen your prompt by adding keywords like: “dynamic motion design commercial advertisement” to push the result closer to a polished ad style. Important note : If your design contains heavy typography, expect some inconsistencies in text rendering. The best approach is to generate multiple variations and select the cleanest result. For this sample, I generated it 4 times before landing on the final version.show more

DStudioproject
11,413 次观看 • 3 个月前
🚨 JUST IN: THIS FREE TOOL JUST REPLACED FOUR... AI IMAGE AND VIDEO SUBSCRIPTIONS AT ONCE. Midjourney. Krea. Higgsfield. Openart. One repo. 200+ models. Zero dollars a month. Here is what it actually does. It is a full image and video studio that runs in your browser or as a desktop app. Text to image, image to image, text to video, image to video, lip sync, cinema mode with real camera controls. All of it. 4,500 people already starred this. What you get for free: → 50+ image models including Flux, Midjourney v7, Ideogram, GPT-4o, Seedream → 60+ video models including Kling, Sora, Veo, Runway, Wan, Hailuo → lip sync studio with 9 dedicated models. upload a portrait and audio and it talks → cinema studio with real camera controls. lens, focal length, aperture, film stock → feed up to 14 reference images into one generation → self-hosted. your data never leaves your machine The crazy part is there is also a hosted version that needs zero setup. Just open the link and start generating. Now the math. Midjourney Standard: $30/month Krea AI Pro: $30/month Higgsfield Plus: $49/month Openart AI: $15/month That is $124 a month. $1,488 a year. This repo does everything all four do. With more models than any of them. For free. Forever. No subscription. No vendor lock-in. MIT licensed. Download it in one click on Mac or Windows. Someone should have told me about this sooner. I feel like an idiot. ( save this )show more

Kanika
14,749 次观看 • 3 个月前
A preview of what's next, visualized with Rerun and... PlayCanvas supersplat ✨ (Also, feel free to send me a DM 📩; I’ll be in San Francisco from July 21–29, and I'm looking to meet like-minded folks!) I'm convinced that Gaussian Splats will be an integral part of any data engine as an underlying representation. So I've started putting together a repo that: 1. Given a single image, perform image outpainting 🖼️🖌️ 2. Estimate a monocular depth map on the outpainted image 📏 3. Train a Gaussian Splat initialized from the monocular depth 🎓✨ 4. Warp to new views, perform inpainting on the missing masks -> Train new splat 🔄🎨 This is going to be integrated into exo-egoforge, but I wanted to start with the simple single-image version before moving to a multi-video implementation There's some weirdness in the final rerun visualization, but the trained splat looks great 🎉! This is all based on the very cool VistaDream paper ( .github.io/) More on this next week!show more

Pablo Vela
26,036 次观看 • 1 年前