Hailuo MiniMax Design MiniMax Design (H3) MiniMax (official) Third... Prototype - PASSED ✓ This time, I pushed it further GAUDÍ LIVING A playful furniture e-commerce UI inspired by the legendary Antoni Gaudí where every cursor hover transforms the furniture, typography, colors & motion graphics. Motion & Animations: Wiggle × Bounce × Scale Tip: Use standalone images for reference with light gray background and use hex code to lock the color. Prompt + Media below ↓show more

Shams
38,649 просмотров • 20 дней назад
GPT-5.6 Sol is unbelievably good at creating and editing... videos. It can do motion design, product demos, and animations like this one I made by simply giving it a screen recording. GPT 5.6 has the best design taste and significantly outperforms Fable, which relies heavily on repetitive design patterns. To help you experiment with video editing on it, we just launched a collection of 100 ready-to-use skills that show what’s possible and help you get started with video editing using GPT-5.6. These skills can create anything from motion graphics launch videos for your product to a 3B1B-style science explainer video. You can also use them to edit existing videos: add captions, generate motion graphics, create voiceovers, redesign visual styles, translate into new languages, and much more. If you want access to the full library, comment “VIDEO SKILLS” and I’ll share it with you. (You'll have to follow me so I can DM you.)show more

Akash Anand
517,184 просмотров • 2 месяцев назад
I tried MiniMax Design (H3) to see how it... handles real content creation. The workflow is simple. You just write a prompt or drop in an image, and it turns that into a dynamic video with motion, framing, and scene depth. No timeline to manage. No editing setup. No back and forth. What stood out to me: • Text to video and image to video both feel smooth. • It handles motion, camera angles, and flow on its own. • Output is fast, usually within seconds. • Works well for reels, quick ads, storytelling, and idea testing. It removes the hardest part: starting from scratch and turns your ideas into content in minutes. Instead of thinking, “How do I make this video?” You start with, “What do I want to create?” That shift alone makes it worth exploring. Try it here: #Hailuoshow more

Manish Kumar Shah
27,680 просмотров • 5 месяцев назад
I found 9 sites that generate Design Markdown files... for your AI builds ! here are the links: - - - - - - - - They help turn your visual direction into an actual design system your coding agent can follow. typography, colors, spacing, radius, shadows, components, states and responsive behaviour. before you build any screen, pick a visual direction and generate the design .md file aligned with your product mechanics and outcome goals add it to Claude Code or Cursor, then use /design to describe what should stay consistent across every screen this gives the agent actual constraints instead of asking it to “make the UI look premium” and hoping for a good output you'll ship much cleaner product ideations at lightning fast speed ! ( this and a folder of 5-6 visual references are the first things we prepare before starting a client build )show more

Harshil Tomar
80,784 просмотров • 11 дней назад
Claude Design + Shopify is f*cking ridiculous 🤯 You... can now publish pages from Claude Design → Claude Code → Shopify. Built 100% with Claude Design, Claude Code, and the Shopify CLI. Perfect for DTC brands and agencies who want to skip the design → dev handoff entirely. Here's how it works: → Design any landing page in Claude Design → Export as a zip and drop it into Claude Code → Install the Shopify + Shopify AI Toolkit plugins → Prompt Claude to convert the HTML into a Shopify page template + push to live theme → Claude uploads the images, deploys the files, and creates a published page No more handing designs off to a dev and waiting 2 weeks for a Shopify page. What you get: - A workflow that turns any Claude Design page into a real Shopify page template - Editable sections so your marketing team can swap copy, images, and CTAs without code - Images uploaded straight to Shopify Files automatically - A files-only deploy that only touches what's new in your live theme - A repeatable pipeline you can use every time you design a new landing page This is essentially the design-to-deploy pipeline brands have been waiting for. I put together a step-by-step playbook for going from Claude Design → published Shopify page. Every install, every plugin, every command, and the exact prompt that runs the whole thing. Want the playbook for free? > Like this post > Comment "SHOP" And I'll send it over (must be following so I can DM)show more

Mike Futia
59,057 просмотров • 4 месяцев назад
Gemini Omni's motion control is f*cking cracked i just... figured out how to turn 1 reference video into 50+ AI videos with the exact same movements... you have a video of someone eating, dancing, using a product, doing whatever complex motion you need. you feed it to Gemini Omni and it recreates that exact motion with a completely new AI character in literally one prompt i've tested this against Kling motion control and it's not even close. Kling falls apart the moment you try anything complex. eating scenes look weird, hand movements get mangled, anything multi-step breaks down completely. Gemini Omni handles all of it if you're still using kling motion control or paying creators to split test your videos, this replaces that entire workflow here's the thing though. you can't just prompt this out of the box. if you try to do motion transfer with default prompting you're going to get errors or the motion won't transfer properly. there's a specific prompting method that makes it work every time so i packaged up the whole system.. here's what you're getting: > full step by step video breakdown > how to find the best reference videos to use > the exact prompting system that allows for motion control transfer so you never get errors > the workflow for batching this out at scale (1 video → 50+) RT + reply "MOTION" and i'll send it over (must follow so i can dm)show more

Miko
61,318 просмотров • 2 месяцев назад
📖THE STEP MOST CREATORS SKIP IS WHY THEIR AI... ANIMATION LOOKS INCONSISTENT Consistency across clips doesn't come from prompting — it comes from the reference image. The pipeline, step by step: ▪ Start with ChatGPT Image 2 — generate a full character design sheet first, not just a single frame. Multiple angles, expressions, and outfit variations in one image keeps the character consistent across every scene ▪ Build a storyboard inside ChatGPT Image 2 as well — define each shot, camera angle, action, and mood before touching Seedance at all. This is the step most people skip and it's the reason clips look disconnected ▪ Define a color palette and lighting mood early — golden afternoon light, soft warm tones, dramatic shadows. Lock those values and repeat them across every prompt ▪ Take each storyboard frame into Seedance 2.0 as the reference image — one frame becomes one clip ▪ Write the Seedance prompt around the character action, not the scene description. The scene is already in the image. The prompt handles motion, camera behavior, and timing ▪ Keep clip duration between 4-6 seconds per shot — shorter clips give more control over pacing and reduce motion drift on character faces ▪ Match camera movement type across consecutive clips — if one shot dollies in, the next should hold or pull back, not dolly again The consistency across these frames comes from the character design sheet, not from luck. Seedance reads the reference image and the prompt together — if the reference is detailed enough, the output stays on-model. This video was created by ALOKXMEHTA 📥 tomorrow: the exact ChatGPT Image 2 prompt structure used to generate a multi-angle character design sheet like this one 🔖One article covers the entire workflow — it is pinned below, do not scroll past it.show more

Zentrix⌚️
14,015 просмотров • 2 месяцев назад
Impeccable 3.7 brings linting to design. Until now it... was a skill you asked for help. Now it's a design-system-aware feedback loop that runs while your agent builds, catching slop and design drift before they land. 🪝 Design hooks for Claude, Codex, and Cursor They run after every UI edit and quietly nudge your agent to fix slop and drift. The output isn't another wall of lint: it separates new findings from already-seen ones, flags clean scans, and asks the agent to use judgment. Fix real issues, leave intentional demos alone, save exceptions to config instead of littering your source. 🎨 Slop detection is now project-aware Reads your actual design system from DESIGN.md, your typography, palette, radius scale, and tokens, and flags drift from your system, not just generic AI slop: • this font isn't in your design system • this color is outside your documented palette • this radius doesn't match your rounded scale The same engine powers both the hooks and the CLI, and it's where we're investing next. 🖥️ Live Mode, ready for real projects Svelte/SvelteKit now preview variants as temporary framework components with live params, then accept cleanly back into your source component. Manual text edits got evidence / apply / discard routes, insertions preserve their anchors, and mapped lists and JSX slots clean up far more reliably. ⚡ Leaner core, sharper detector Rule-level evals across 3 providers and 4 niches cut guidance with no measurable lift and dropped examples that taught models bad patterns. The detector now skips hidden and screen-reader-only elements, understands OKLCH alpha and Sass-like inputs, and tightened checks for repeated kickers, oversized H1s, clipped overflow, and cramped padding. 🛠️ CLI caught up impeccable detect loads DESIGN.md by default, motion findings name the exact token or cubic-bezier instead of just "bounce," and impeccable ignores gives real CRUD for exceptions. Hooks and CLI share the same ignores. No split-brain config. Plus a much-improved interactive installer with hooks setup built in. Upgrade: npx impeccable install npm i -g impeccableshow more

Impeccable
233,051 просмотров • 3 месяцев назад
I built a clone of the Yeezy store with... Next.js. This was a fun challenge — the site has some smooth animations and feels very fast. But it was bothering me that I couldn't use the browser back button. Can we do better? So I rebuilt the site with v0 and Motion. Here's how it works: 1. When you click on a product, Motion is able to animate the original position of the product in the grid, to the zoomed in product detail page. 2. During this transition, we also shallow update the URL with the `/p/slug` route for the page. 3. If you press the back button in the navbar, or use the browser back button, or press escape — all options will take you back to the main product listing page. 4. If you reload the page while looking at a product, or someone sends you a link to a specific product, it still works! This is the best parts of a SPA and MPA mixed together. In the future, I can make this even better with View Transitions (I wasn't able to get the product animation just right, but if you can I'd love to see it!). I also took some creative liberties from the original design. The whole 1/2/3 size thing, where you needed to click the "?" to see SM/MD/LG was strange, so I just went directly to those sizes. Similarly, I prefered the more traditional style sheet/modal with the background color change, versus the full screen takeover. If you wanted to actually hook this up to Shopify now, you can swap the cart implementation with Next.js Commerce, which has all the APIs you need + optimistic writes 🔥 Should I make a video walking through the code?show more

Lee Robinson
148,937 просмотров • 1 год назад
How to 10x your design with Figma Make ⭐️... I spent 40+ hours testing Figma Make prompts. Most designers waste time with vague prompts and get garbage outputs. Here are the exact prompts and proven workflow that actually work: 1️⃣. Prompt formula: Bad: "Create a dashboard" Good: "Create a SaaS analytics dashboard with: → Left sidebar navigation (240px wide) → Top bar with user profile → 4 metric cards in a grid → Line chart showing revenue trend → Use blue (#2563EB) as primary color" The more you specify = higher quality. 2️⃣ Workflow: Import Your Design System First Before your first prompt: → Go to your main Figma file → Export your component library → Import it into Make → Add this to every prompt: "Use components from [Your Library Name]" Now everything matches your brand automatically. 3️⃣. Prompt for Interactive States: "Create a login form with: → Email and password inputs → Show error state when fields are empty → Disabled button state when form is incomplete → Success message after submission → Add smooth transitions between states" Gets you working prototypes, not static screens. 4️⃣. Advanced Prompts: Data States "Create a user list screen with three states: → Loading (skeleton screens) → Success (populated table with 10 users) → Empty (illustration + 'No users yet' message + 'Add User' CTA)" One prompt = complete UX coverage. 5️⃣. The "Design System Drift” Fix: Notice Make using wrong colors? → Try this Prompt: "Analyze my imported library and list all color tokens, then regenerate using only those exact values" It'll self-correct and stick to your system. 6️⃣. Responsive Design Prompt: "Create a pricing page with 3 tiers. Make it responsive: → Desktop: 3 columns side-by-side → Tablet: 2 columns with 3rd below → Mobile: Stacked vertically → Use Auto Layout for fluid scaling" This gets you mobile + desktop in one shot. 7️⃣. Magic Troubleshoot Prompts: Output looks off? → Try: "Redesign this following Material Design principles" → Or: "Make this follow iOS Human Interface Guidelines" → Or: "Apply Gestalt principles for better visual hierarchy" Give it design frameworks to follow. It works magic. Designers who master prompt engineering in 2026 will ship 10x more than everyone else. P.s. I made a Gameboy for Pokémon. (bookmark this for later)show more

Felix Lee
12,706 просмотров • 8 месяцев назад
A WEB STUDIO CHARGES $35,000 FOR AN ANIMATED SITE.... THE SAME BUILD NOW COSTS $12 - CLAUDE CODE WRITES, HIGGSFIELD RENDERS. Every agency billing $100-149/hr is just three departments. Here's each one, collapsed into a single agentic session. SYSTEM 1 - THE MOTION STUDIO (Higgsfield) Cinematic clips pulled from 30+ generative models - hero shots, transitions, ambient loops. → This used to be a motion artist on retainer. Now it's a prompt. SYSTEM 2 - THE DEV TEAM (Claude Code) Scaffolds the site, writes the GSAP ScrollTrigger timelines and Lenis smooth-scroll, extracts frames, optimizes every asset. → A full scroll-driven build with zero hand-coded keyframes. SYSTEM 3 - THE DESIGN DEPT (baked-in cinematic layer) Six effects with no config: film grain, particles, vignette, glass cards, color tints, scroll pacing. → The polish that justified the invoice - now it ships by default. Three departments. One operator. One pass. What used to take a designer, a motion artist, and a developer through weeks of handoffs now runs in a single session - for a Claude subscription and a few dollars of Higgsfield credits. The studio was never the talent. It was the overhead. And the overhead just became three systems. Reply "web-site" to this post and I will send you the step-by-step Playbook 👇show more

ZEUS⚡️
174,734 просмотров • 2 месяцев назад
Claude Design is f*cking cracked for landing pages 🤯... Find any competitor's e-com lander → feed it to Claude Design w/ your brand's design system → get back a fully rebuilt version in your brand, your copy, your design. All inside Claude Design. Perfect for DTC brands and agencies who are still paying freelancers to build advertorial pages from scratch. If you're finding landers that have been running on Meta for 6+ months and want to test the same structure for your brand —> Briefing a designer, waiting a week, getting back a flat mockup, giving notes, waiting again... Claude Design eliminates the entire loop: → Use Go Full Pag to screenshot the full competitor lander → Feed it to Claude Design with your design system → Prompt it to extract the exact section structure and rebuild it for your brand → It rewrites all copy, applies your fonts, colors, and layout → Iterate section by section — send a screenshot of what you want to fix, it fixes it → Drop in your product images and founder photos as you go No designer. No back-and-forth briefs. No starting from scratch. What you get: -> Full production-ready lander built around a proven structure that's already converting on Meta -> Live elements (countdown timers, animated sections) auto-generated -> Mobile and desktop versions you can refine with plain-English prompts -> A repeatable system — new competitor, new screenshots, same pipeline I put together a full playbook breaking down the exact process — the prompts, the section-by-section editing approach, and how to set this up for any competitor lander. Want it completely for free? > Like this post > Comment "CLONE" And I'll send it over (must be following so I can DM)show more

Mike Futia
37,935 просмотров • 3 месяцев назад
HE BUILT A $10,000-TIER ANIMATED SITE WITH CLAUDE CODE... - FOR THE COST OF A SUBSCRIPTION What's on screen isn't a landing page with a parallax background It's a fully interactive, scroll-driven site with real-time 3D rendered in the browser What's actually on the page: > 3D product models rotating and reacting to scroll in real time via WebGL > Smooth hover interactions and transitions - no hand-coded keyframes > Cinematic minimal aesthetic that got featured on Awwwards > Typography layering, editorial layouts, everything assembled in one session What it normally takes: > A 3D artist, a motion designer, and a frontend developer > Weeks of handoffs - modelling, exporting, wiring animations, layout, copy > Six separate systems integrated by hand That pipeline was the moat. It's what justified the invoice The price gap: > Studio build at this level: $5,000-10,000+ > Your cost: a Claude subscription Timeline: weeks of production -> a single session Full walkthrough in the article belowshow more

monokern
703,391 просмотров • 2 месяцев назад
today’s experiment: chimera. a little tool that let’s you... realtime morph through a design system matrix generated by 4 reference images. this one was inspired by listening to techbimbo talk about her process on Dive Club with Ridd 🤿 and How I AI with claire vo 🖤. her approach to choosing moodboards over prompts made me wonder what might be possible if we applied that same approach to generative UI. and after a few dead ends, the idea turned into chimera. the results are very generic at the moment but I might tune it up if people seem interested. the big reminder for me is how much more inspiring a tool feels when you can explore in realtime instead of waiting for results every time you make a change. lots more things to try in that direction.show more

River Marchand
202,517 просмотров • 6 месяцев назад
Hold up, here is the prompt: works with almost... any model. enjoy :) Role & Objective: Act as an Elite UI/UX Front-End Engineer specializing in Apple-tier micro-interactions and advanced CSS. Your task is to program a perfectly centered navigation bar in a strictly SINGLE HTML file containing all HTML, vanilla CSS, and vanilla JavaScript. No external libraries or frameworks (No Tailwind, React, etc.). Design Concept - "True Liquid Glass": CRITICAL INSTRUCTION: Do NOT generate standard, flat "glassmorphism" or basic frosted glass. I require a physically accurate "Liquid Glass" aesthetic. It must look like wet, poured clear resin, combining the high-gloss specular highlights of classic macOS Aqua with the volumetric spatial depth of modern Apple VisionOS. 1. The Liquid Glass Material & Lighting (CSS): - Deep Refraction: Use `backdrop-filter` with extreme blur (e.g., 50px) and over-saturation (200%). - Specular Highlight: Create a curved, semi-transparent white gradient on the top half using a pseudo-element (`::before`) to simulate a hard light reflection on a wet, rounded 3D surface. - Caustics & Volume: Use multi-layered inner and outer `box-shadow` properties to simulate light refracting at the bottom edge and casting a realistic ambient drop shadow. - Interactive Glare: Implement a soft radial-gradient spotlight inside the glass that dynamically tracks the user's mouse cursor (X/Y coordinates) using JavaScript and CSS variables (`mix-blend-mode: overlay`). 2. Navigation Layout & Elements: - Center the pill-shaped navigation bar perfectly in the middle of the viewport. - Include 3 main navigation items with minimalist, inline SVG stroke icons and text labels: "Home", "Call", and "List". - Add a subtle vertical divider line after the main buttons. - Next to the divider, add a Dark/Light Mode toggle button containing inline SVG Sun and Moon icons. 3. Animations & "Apple Magic": - Sliding Active Pill: Create a solid background "pill" that sits *behind* the active navigation item's text/icon. When a different item is clicked, this pill must dynamically recalculate its width and slide to the new position. - Spring Physics: The sliding transition MUST use an exact Apple-style bouncy spring easing curve (e.g., `transition: all 0.5s cubic-bezier(0.34, 1.2, 0.64, 1)`). - Tactile Feedback: Buttons and icons must physically press down slightly (`transform: scale(0.92)`) when clicked (`:active`). - Theme Switch: The Sun and Moon icons must smoothly rotate, scale, and cross-fade during the transition. 4. Background Environment (Crucial): - Glass needs light and color to refract! Create a full-viewport, smoothly animated mesh gradient background using 3 large, heavily blurred, floating color blobs. - Implement full Dark/Light mode logic using CSS variables (`:root` and `[data-theme="dark"]`). Toggling the theme must seamlessly transition the background blob colors, glass opacity, shadow intensity, and text colors. Output ONLY the pristine, production-ready code. Prioritize maximum visual fidelity and silky-smooth 60fps animations.show more

Leon Lin
128,501 просмотров • 6 месяцев назад
Most AI video tools are stuck in a loop,... you generate, change the prompt, and wait again. Vidu S2 is pushing things forward by making generation interactive. You can edit the video while it is being created. Change the style, swap the background, or switch outfits in real time. If you drop in a new image, the video adjusts instantly without awkward cuts. Under the hood, it combines real time 720P output with an AI agent that actively watches the video to plan what happens next. It allows you to change reference images on the fly, powered by a model optimized specifically for live streaming. This opens up massive possibilities. Imagine viewers controlling visual changes during a livestream, people transforming into game characters, or shoppers trying on different looks live. The entire S Series is available for a free trial right now so you can test it yourself. Use my code PANDAS2 with the link below to grab 300 bonus credits to explore the other features. Vidu Stream: 300 Bonus Credits: Vidu AI #ViduAI #ViduS #ViduS2show more

AI Panda
98,066 просмотров • 2 дней назад
Astra (GPT-6) is here!!! I've had early access and... tested it like crazy with things like games, code, writing, browser control, presentations and general knowledge work. This is the best model I've ever used. Period. (Incredible demos below in this thread ⬇️) Here's my take on Astra: > It's insanely capable. This feels like a massive improvement, not just an incremental change. This is especially true with zero-shot prompts. > It's all about knowledge work. Slide creation, analysis, writing, and browser control. And oh my...it's so good at browser control. GPT-5.6 was already fantastic at doing things in the browser, Astra is another level and significantly faster. > We're closer than ever (arrived?) at prompt-to-playable game. And I don't just mean only playable, these are actually fun games. I bet if someone with a great eye for games used Astra, they could create a viral game within 1-2 weeks. > Astra is better at writing but not perfect. It removes much of the "AI Smell" we're all familiar with but some stink still survived. > It has a tendency to use the same design colors and look/feel as GPT-5.6 (forrest green anyone?) but it is more steerable in design than previous models. > It's highly steerable in general. A little nudge goes a long way. When I first started using Astra, almost every task I gave it would go for ~30 minutes. I wanted it to keep working. Adding more specifics to a prompt helped greatly with it's ability to work for a long time. > Astra's 3D understanding is unmatched. 3D asset creation was consistent and easy and its spacial awareness while building complex 3D worlds blew me away. I'm still getting familiar with Astra but this will now be my go-to model for any difficult work I have. Check out the demos below: 👇show more

Matthew Berman
1,899,112 просмотров • 15 дней назад
Generated with MiniMax Design (H3) Prompt: Duration: 15 seconds... | Aspect Ratio: 16:9 | Style: Authentic UGC / iPhone selfie-vlog, handheld, natural light, TikTok/Reels aesthetic. Product Reference: Use the uploaded gourmet burger image as the only product reference. Preserve the bun shape, patty thickness, cheese melt, lettuce, tomato, sauces, and proportions exactly in every shot. Character Description Name: Hana A young Japanese woman Image 1 in her early 20s with natural beauty, long dark hair in a loose ponytail, oversized cream sweatshirt, minimal makeup, bright smile, friendly lifestyle-vlogger personality. Shot Breakdown SHOT 1 (0–2s) — Selfie showing the burger box. Dialogue: "Burger night!" SHOT 2 (2–4s) — Opens the box. SHOT 3 (4–6s) — Quick zoom on the burger. SHOT 4 (6–8s) — Hands lifting the burger with cheese stretching naturally. SHOT 5 (8–10s) — Bite reaction. Dialogue: "Okay... that's incredible." SHOT 6 (10–12s) — Casual close-up b-roll while reaching for fries. SHOT 7 (12–14s) — Toasting the burger toward the camera. Dialogue: "You need this." SHOT 8 (14–15s) — Freeze frame with overlay: "burger cravings = solved 🍔" Look & Feel Warm apartment lighting, genuine phone footage, slight grain, natural autofocus breathing, handheld imperfections, fast jump cuts. Negative Prompt cinematic grading, commercial production, CGI burger, fake cheese, distorted hands, warped food, perfect stabilization, studio lighting, text glitches, logo distortion.show more

Sairah
13,072 просмотров • 1 месяц назад
THIS GUY JUST REBUILT A $35,000 ANIMATED SITE FOR... $12. IF YOU RUN A WEB STUDIO, YOU SHOULD PROBABLY KEEP SCROLLING. Every agency billing $100-149/hr is selling you five departments wearing one invoice. Here’s each one - collapsed into a single agentic session. LAYER 1 - THE CONCEPT ROOM (Claude) Reads the brief, pulls references, and scripts the scroll: what the visitor feels at second 3, second 15, second 40. → Used to be a strategist and a wall of mood boards. Now it’s a conversation. LAYER 2 - THE MOTION STUDIO (Higgsfield) Cinematic clips from 30+ generative models - hero shots, transitions, ambient loops - all matched to the story from Layer 1. → Used to be a motion artist on retainer. Now it’s a prompt. LAYER 3 - THE DEV TEAM (Claude Code) Scaffolds the site, writes the GSAP ScrollTrigger timelines and Lenis smooth-scroll, extracts frames, optimizes every asset. → A full scroll-driven build with zero hand-coded keyframes. LAYER 4 - THE DESIGN DEPT (baked-in cinematic layer) Six effects, zero config: film grain, particles, vignette, glass cards, color tints, scroll pacing. → The polish that justified the invoice - now it ships by default. LAYER 5 - THE QA PASS (Claude) Checks load speed, mobile breakpoints, and whether the scroll actually lands - then rewrites whatever doesn’t. → Used to be a client call and a revision cycle. Now it’s one more turn in the same session. Five departments. One operator. One pass. A strategist, a motion artist, a developer, a designer, and a QA lead - weeks of handoffs - now run in a single session. For a Claude subscription and a few dollars of Higgsfield credits. The studio was never selling talent. It was selling overhead. And the overhead just became five layers. Follow me, reply “website” to this post and I will send you the step-by-step Playbook 👇show more

ZEUS⚡️
141,973 просмотров • 2 месяцев назад
Sora2 プロンプトテンプレ! だれでも簡単に凄い動画を!がコンセプトです! 参照画像ありで、実写映画風を想定しています。 参照画像なしの場合、このプロンプトの先頭に、「参照画像なしで代わりにこのテーマでお願いします」と指示してくれれば動きます! プロンプトはこちら👇 You are the... ultimate maestro of filmmaking. Create a 10-second, 2× speed live-action cinematic video for a Japanese audience. Use reference images as inference material to determine the theme, genre, style, narrative tone, and final on-screen text design. The genre must be fluid, inspired directly by the imagery — it may be action, romance, drama, comedy, documentary, or another cinematic style. Depict everything with the hyper-realistic quality of a Hollywood blockbuster, as if filmed with cutting-edge cinema technology (IMAX-grade cameras, advanced VFX, Dolby Vision lighting, cinematic slow-motion). The environment must feel like a real-world space, but elevated through exaggerated cinematic effects that make every frame spectacular. Camera work should include sweeping crane shots, extreme close-ups, POV immersion, and impossible yet cinematic movements. The narrative must be simple yet emotionally overwhelming, easy for children to understand, but delivered with adult-level cinematic depth. Use 20+ distinct cuts, averaging less than 0.5 seconds per shot. Each cut must deliver either a jaw-dropping wow moment, an exaggerated emotional beat, or a hint of narrative intrigue, combining into a condensed Hollywood-style super trailer. Transitions must be seamless, exaggerated, and cinematic, ensuring overall coherence. The final impression should be unforgettable: visually extreme, emotionally overwhelming, technologically cutting-edge, and instantly shareable. End with a blackout, then display a genre-inferred final title, exaggerated in design, motion, and style (e.g. explosive metallic typography for action, radiant handwritten glow for romance, stark monumental type for drama).show more

SHINTARO
14,889 просмотров • 11 месяцев назад
Everyone's sleeping on image-to-3D AI models. They can make... your app look incredibly unique, with just a little effort. Here's how. This is my calorie tracker, built in a week with nothing but prompting. Just Claude Code + a couple APIs. The visuals are all AI-generated. I'll be sharing the full workflow + all the crazy technical stuff Claude and I did to make this work, so nobody has to struggle through it like me. Deep dive coming soon! Till then, this is the high-level idea: 1. Get a clean image of the food (or whatever your asset is) - In my app, the user describes foods via text, or attaches images (or both) - If text, an LLM extracts the food description and formats it into a specific prompt I tuned for this design, and we generate an image using Z-Image Turbo through fal - If image, we do the same thing but with FLUX.2 [dev] to edit the user image into our reference design - Originally, both used Google Nano Banana, but switching to open models cut costs and latency a ton 2. Gaussian splatting (2D image → 3D model) - I tried various 2D-to-3D options on fal and ended up with TripoSplat as my preferred balance of speed, cost, latency; this turns an image into a 3D model that looks super high quality (link below) - The app displays the 2D image while our backend generates the 3D splat - We "groom" the splat to reduce size and load time by culling low-opacity/scale points 3. Render efficiently on device Originally, it looked great but ran at 10 FPS. Getting to 120 FPS was a crazy journey. TL;DR: - SwiftUI had to go; it forced us to render each asset in independent MTKViews, which wasn't workable - Instead, we composite every dish into one full-bleed CAMetalLayer using MetalSplatter (link below) - We had to make some optimizations within MetalSplatter's code too, to reduce the overhead of sorting points per render Then I added some finishing touches like the subtle rotation and parallax as they move around. I think it turned out pretty cool :) Overall, this took some effort, but we still got it done in less than a day. Hopefully your agent can follow in the footsteps of mine and do it much faster. Keep an eye out for the bigger writeup, which'll give your agent everything it needs. If you have any questions, drop em below!show more

Anshu
29,342 просмотров • 2 месяцев назад