🧵 Photos → Creative Code using Gemini I built... an experiment that turns photos into interactive p5xjs sketches using Gemini 2.0 Flash. Unlike UI generators, this creates code that mimics the *behavior* of what's in the image - like smoke swirling or ripples spreading. Check out this sunset + smoke simulationshow more

Trudy Painter
155,027 görüntüleme • 1 yıl önce
gemini 3 is unbelievable it can create an app... that turns image into 3D interactive The Matrix scene, adjust particle amount, color, density interact with particles using mouse or hands created with text prompts, no code neededshow more

el.cine
1,843,051 görüntüleme • 7 ay önce
I built a multimodal AI medical image diagnosis agent... using Gemini 2.0 This AI Agent can: • Analyze medical scans • Detect abnormalities • Search the web • Generate patient-friendly reports 100% Opensource Code with step-by-step tutorial.show more

Shubham Saboo
217,983 görüntüleme • 1 yıl önce
With the launch of Nano Banana Pro, we're also... rolling out the ability for Gemini users to check whether an image was generated with or edited by Google AI using SynthID, our digital watermarking technology. Now you can upload any image into the Gemini app and ask "Is this AI-generated?" Gemini then scans the image for the imperceptible SynthID watermarks that are embedded into all Google AI-generated images - including those by Nano Banana Pro! Learn more about Google’s efforts to increase AI transparency:show more

Google Gemini
190,547 görüntüleme • 8 ay önce
Create a short film like this in just 1... minute with GPT Image 2.0 + Seedance 2.0. GPT Image 2.0 can naturally combine multiple photos into one single image, while Seedance 2.0 can use that image as a reference to automatically separate the scenes, generate a coherent video sequence, and add suitable background music. This workflow greatly improves the overall creative efficiency. When using this method, simply provide the merged image as a reference for Seedance 2.0 and briefly describe each scene with a simple prompt. This can significantly increase the success rate of the final video. All of the above was created on GPT Image Prompt: Seedance Prompt:show more

Midjourney Sref and prompt Library
40,572 görüntüleme • 3 ay önce
I noticed that in Runway’s Seedance 2.0 model, we... can actually reference photos that include human faces. I have no idea how they managed to overcome this. By the way, there’s no information here about whether they’re using the fast version of Seedance 2.0 or the higher/pro model. I wish they had specified that too.show more

Ozan Sihay
22,630 görüntüleme • 3 ay önce
This is how Strike Robot turns Simulation into Reality!... One of the biggest challenges in robotics is ensuring that behaviors validated in simulation work reliably in the real world. For this experiment, we reconstructed part of a real laboratory at Eastworlds inside SR Platform. The generated layout was then deployed into MuJoCo. Using SR Agentic, the robot was tasked with finding abnormal objects in a cluttered environment and sending a Telegram notification when detected. Before deployment, everything is validated in simulation.show more

Strike Robot
15,025 görüntüleme • 1 ay önce
Character Consistency with Google Veo 3 now in Gemini... API! 🤯 Use Images as starting frame to keep character consistency! Here is an python script on how to make consistent viral videos, like you see on TikTok or Youtube shorts: 1. Based on an idea, it generates a series of scene prompts using Gemini 2.5. 2. Generates a Image based on the first scene using Imagen 3 3. For each scene prompt Veo 3 (fast) generates a video clip. 4. Uses Gemini 2.0 image editing to make sure the starting images fits the scenes 5. Combine the individual video clips into a single final video using MoviePy Veo 3 starts at $0.75 / second and Veo 3 at $0.40 / second with audio. ! 📹 🔉 Prompt: “A realistic energy drink commercial for athletes.”show more

Philipp Schmid
15,863 görüntüleme • 1 yıl önce
🚨Gemini 3.6 Flash is trash I tested it on... a 3D Golden Gate Bridge, and the results were awful. • I had to re-prompt it three times because it repeatedly ignored the instructions. • First attempt, instead of creating the requested .html file, it first tried to build the experience inside the Gemini app using simulations. • Then second attempt it started placing images from the web into the chat rather than actually producing the file. • Even after getting it to complete the task, the final output was dramatically worse than Gemini 3.1 Pro, which is 5 months old and now not even a top 10 model on leaderboards. This feels like a regression from Gemini 3.5 Flash and honestly, it is one of the weakest models I have tested in the past few months. Has anyone else tested Gemini 3.6 Flash yet, and are you seeing the same thing?show more

Lumina
71,994 görüntüleme • 10 gün önce
This trader turned $130 into $33,000 on Polymarket using... a simple script it doesn't read news, check polls or predict anything just buys contracts under 1c that everyone thinks are dead wrong 94% of the time -> one hit = 99x, one win covers 98 losses His wins: $1,191 -> $11,990 $11 -> $2,587 $133 -> $1,791 The inefficiency: markets priced at 99c/1c pull back to 50c far more often than the price implies I found 8 wallets doing this -> $550K combined fed 400M trades into Claude, built the bot full code inside copy his trades with ARES: bookmark this before the code gets taken downshow more

Paone
13,390 görüntüleme • 3 ay önce
Claude Code + Gemini Omni + GPT Image 2... is f*cking cracked i just built an AUTOMATED AI UGC content system that generates full ad creatives end to end drop in your product photos and a one line pitch, and the system analyzes your brand, generates a realistic AI avatar, writes the script, renders the video, adds captions, and stitches everything together. in literally one click. if you're still wasting hours generating videos or hiring expensive editors with long turnaround times, this is for you.. here's how it works: > drop in your product photos and brand info > the system finds a reference image that matches your target audience it generates a detailed JSON prompt and creates a hyper realistic starting frame > feeds everything into Gemini Omni and renders the full video with voice outputs and captions, all ready to post and all through Kie AI MCP or Higgsfield MCP (you have full control) the whole process runs inside Claude Code. so you're not jumping between 6 different tabs trying to piece things together. one system handles everything its only a 3 minute setup and 4-12 min per ad. each 30 second ad costs under $3 in API credits which is way cheaper than any editor and 10x faster.. RT + reply "UGC" and i'll send it over (must follow so i can dm)show more

Miko
55,550 görüntüleme • 13 gün önce
Claude Code can now watch & analyze ANY video... 🤯 I built a skill that gives Claude the ability to watch any video file you drop in — UGC ads, competitor Meta ads, organic TikToks, screen recordings, anything. All inside Claude Code. Perfect for DTC brands and agencies who study competitor creative every week to figure out what's working and what to test next. Here's the problem: If you're studying competitor ads on Meta or hooks on TikTok, you're scrubbing through videos manually, pausing to write down hooks, screenshotting on-screen text, and trying to remember what made the ad land by the time you've watched 10 of them. This skill solves it: → Drop any video file into Claude Code → Skill routes it through the Gemini API for native video understanding → Returns a full creative teardown — hook breakdown, target audience, angle, beat-by-beat, on-screen text verbatim → Surfaces the steal-worthy patterns you can apply to your own creative → Same skill works on UGC ads, produced video ads, organic TikToks, and Loom recordings No manual scrubbing. No pausing every 5 seconds. No $200/mo ad intelligence platform. What you get: - Native video understanding via Gemini (not just transcripts) - Structured analysis — hook, angle, audience, pain point, CTA - Verbatim on-screen text and dialogue with timestamps - Hook variations generated directly from competitor ads - About 27 cents per 30-minute video Built 100% in Claude Code with the Gemini API. I recorded a full breakdown showing exactly how I built this and I'm giving away the skill for free. Want the skill? > Comment "CLAUDE" + > Like this post And I'll send it over (must be following so I can DM)show more

Mike Futia
35,592 görüntüleme • 2 ay önce
ECOMMERCE CREATORS NEED TO SEE THIS. Creating product video... ads used to be the slowest part of launching a product. I tested Creative Studio inside Pollo AI using just one product photo. Minutes later, it turned that single image into a ready-to-publish product video ad, without touching any editing software. No filming. No expensive gear. No production team. Creative Studio honestly feels like having an AI creative team in your pocket.show more

Aria
27,457 görüntüleme • 28 gün önce
I just built a $10K/month creative strategist inside Claude... Code 🤯 Give it your competitor Facebook page URLs → it scrapes their ads, watches every video with AI, and delivers a data-backed creative brief with 10 ad concepts in your brand voice. All inside Claude Code. Perfect for DTC brands and agencies who are still manually scrolling the Meta Ad Library, screenshotting ads into Google Docs, and guessing at what's working. If you're spending hours every week pulling competitor ads one by one, watching videos to figure out the hook, copying notes into a brief, and rewriting concepts from scratch every time... This system eliminates the entire loop: → Apify scrapes your competitors' active ads from Meta Ad Library (video + image) → Downloads every creative asset locally → Gemini watches each video and analyzes the hook, angle, visual format, copy framework, CTA, and emotional trigger → Runs the full batch and finds the patterns that repeat across 3+ ads → Claude generates 10 ad concepts using the proven mechanics, matched to your brand voice No manually scrolling the Ad Library. No screenshotting ads into docs. No guessing which hooks are actually working. What you get: → Individual creative breakdowns for every competitor ad (7 dimensions each) → A pattern report showing which hooks, formats, and triggers keep repeating → 10 ready-to-brief ad concepts traced back to real competitor data → A reusable system — new competitors, new brief, same pipeline The research that takes your team a full day now runs in 15 minutes for ~$3 in API costs. Built 100% in Claude Code with Apify + Gemini. I put together a full playbook showing you can build the entire thing step-by-step from scratch. Want the playbook for free? > Like this post > Comment "ADS" And I'll send it over (must be following so I can DM)show more

Mike Futia
54,684 görüntüleme • 4 ay önce
I just built a skill that lets Claude Code... watch & analyze ANY video 🤯 Drop in any video file — UGC ads, competitor Meta ads, organic TikToks, screen recordings — and Claude hands you back a full creative teardown. All inside Claude Code. Perfect for media buyers and creative strategists who reverse-engineer competitor ads every week — and lose half a day doing it by hand. If your creative process starts with studying what's already working, you're scrubbing through competitor ads frame by frame, pausing to write down every hook, screenshotting the on-screen text, and by the tenth video you can't remember what made the first one land... This skill solves it: → Drop any video file into Claude Code → Skill routes it through the Gemini API for native video understanding → Returns a full creative teardown — hook breakdown, target audience, angle, beat-by-beat, on-screen text verbatim → Surfaces the steal-worthy patterns you can apply to your own creative → Same skill works on UGC ads, produced video ads, organic TikToks, and Loom recordings No manual scrubbing. No pausing every 5 seconds. No $200/mo ad intelligence platform. What you get: → Native video understanding via Gemini (not just transcripts) → Structured analysis — hook, angle, audience, pain point, CTA → Verbatim on-screen text and dialogue with timestamps → Hook variations generated directly from competitor ads → About 27 cents per 30-minute video Built 100% in Claude Code with the Gemini API. I recorded a full breakdown showing exactly how I built this, and I'm giving away the skill for free. Want the skill? > Like this post > Comment "CLAUDE" And I'll send it over (must be following so I can DM)show more

Mike Futia
41,315 görüntüleme • 1 ay önce
found a tool on product hunt that turns claude... code into an ad agency.. it’s called Goose Ad Remixer. The setup was surprisingly simple. First, I installed the Gooseworks skill library: npx gooseworks install --all Then I asked Claude Code to generate high-performing ad creatives with: /goose-ads make ads for my brand From there, Goose researched the brand, studied the website, messaging, visual identity, positioning, and existing assets before creating multiple ad concepts. What stood out is that it did not just generate random designs. It studied proven ad patterns, including the hooks, offers, copy, layouts, and creative structures, then remixed them for the brand. It felt less like prompting an image generator and more like giving Claude Code an ad strategist, copywriter, and designer in one workflow. Here are a few of the creatives it generated:show more

Robin Delta
16,309 görüntüleme • 17 gün önce
Loving how this turned out! IronSight turns Meta Ray-Ban... clips (from the range) into 4D reconstructions you can replay from any angle -- including an AR view that sees targets straight through walls. It 3D tracks both runs, auto locates every target, and scores hits vs misses using audio cues + Gemini for multimodal reasoning. Full breakdown coming to the channel. The test below is where this started, and then Fable showed up and I blitzed through my whole roadmap in a few days.show more

Bilawal Sidhu
26,387 görüntüleme • 24 gün önce
🤩 I just learned a super cool trick! Do... you know how sometimes you want to record while you scroll on a website, like for a demo or something... But by using the mouse to scroll it turns out to be kinda janky and it doesn't look very good? No? Just me? 🤭 Well, I found out that if you paste this code in the console of your browser: ` setInterval(() => window.scrollBy(0, 4), 16); ` You can make the browser automatically scroll for you! And it looks super cool in the recording! 😎show more

Florin Pop 👨🏻💻
29,839 görüntüleme • 10 ay önce
Create a 3D model from a single image, set... of images or a text prompt in < 1 minute 😮💨 This new AI paper called CAT3D shows us that it’ll keep getting easier to produce 3D models from 2D images — whether it’s a sparser real world 3D scan (a few photos instead of hundreds) or your favorite 2D image generator like Midjourney (just an image). How does this magic work? “This architecture is similar to video diffusion models, but with camera pose embeddings for each image instead of time embeddings. The generated views are passed into a robust 3D reconstruction pipeline to create the 3D representation (Zip-NeRF or 3DGS)”show more

Bilawal Sidhu
92,792 görüntüleme • 2 yıl önce