Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

I made a sprite animator using Google's Gemini 2.0 flash experimental. Here's how it works: 1. Gemini 2.0 exp generates the images 2. They are uploaded each in an S3 bucket/folder 3. Then you can use WebGPU /w canvas or just img Let's watch some examples and generate some:

353,002 görüntüleme • 1 yıl önce •via X (Twitter)

12 Yorum

wavefnx profil fotoğrafı
wavefnx1 yıl önce

This is actually one out of many iterations, in the other ones you could crop/decide and add the frame, auto-generated dungeons and so on. e.g. I could also imagine another vision model segmenting the characters out of each frame or sprite-sheet

wavefnx profil fotoğrafı
wavefnx1 yıl önce

gg @GoogleDeepMind

Rainmaker profil fotoğrafı
Rainmaker2 yıl önce

Can Machine Learning beat the market? Check out this post on my free Substack where I share code and commentary for an XGBoost model and a Random Forest model that both deliver powerful performances.

wavefnx profil fotoğrafı
wavefnx1 yıl önce

If you have any fun features in mind, drop them below, the more crazy and "unachievable" the better

atharva profil fotoğrafı
atharva1 yıl önce

so clean

wavefnx profil fotoğrafı
wavefnx1 yıl önce

thanks, the goblin dancing and the flower ones are crazy

Ash ▵ profil fotoğrafı
Ash ▵1 yıl önce

I wrote a script to do the same thing . UI is much better , Really cool

wavefnx profil fotoğrafı
wavefnx1 yıl önce

amazing; there's much to do to it, caching CDN etc and the UI can be even better too but very easy to develop I bet it can have great results with realistic images also, mini-veo 2.0

dan ⚡️ profil fotoğrafı
dan ⚡️1 yıl önce

so cool holy shit

wavefnx profil fotoğrafı
wavefnx1 yıl önce

early access?

Ryan Lucchese profil fotoğrafı
Ryan Lucchese1 yıl önce

whoa this is so cool

wavefnx profil fotoğrafı
wavefnx1 yıl önce

thanks a lot, can early-access you too, I think the sprite/8bit stuff it's fun but imagine: 1. being able to control the sprites e.g. assign keys to each one 2. realistic images (easy) 3. many more video is just images after all and with the edit feature you could do a lot

Benzer Videolar

GEMINI 3 LAUNCH IS HERE I got a SNEAK PEEK at Gemini 3 with Logan Kilpatrick (Google Deepmind), and it might be the most POWERFUL vibe-coding tool on the planet. A little breakdown: 1. Anyone can build 3D and casual games now You can vibecode full, playable 3D video games generated in minutes. Actual games with physics, characters, controls, and loops you can remix instantly. Pure insanity. I can see founders and brands spinning up games on the fly to ride trends and drive growth. 2. Intelligent apps are becoming the default We built apps where reasoning, memory, and multi-step planning were baked in from the start. Once you’re building apps with ACTUAL intelligence baked in, there’s a whole wave of new opportunities that weren’t possible before. 3. Gemini acts like a creative partner You describe the idea, Gemini fills in the gaps, challenges decisions, proposes alternatives, and iterates in real time. 4. Vibe coding hits a new level Gemini 3 can generate assets, code, game logic, UI, and narrative in one flow. Tools like Claude and Cursor feel fast. This feels like the next layer, the one where a single builder can compete with full teams. Logan Kilpatrick and I pushed Google Gemini 3 hard, and the outputs were solid. A few times we had to give it a few extra prompts but it took feedback really well. I think 1 year ago, a lot of people discounted Google in the AI arms race. Can you discount them anymore? Doubt it. After this, it feels like they at best leading, at worst leading. What do you think of Google's AI efforts/Gemini 3 My biggest takeaway was how it just felt like Gemini 3 had a little more vibe coding horsepower than anything I’ve used.

GREG ISENBERG

73,838 görüntüleme • 10 ay önce