正在加载视频...
视频加载失败
The model sometimes enters a self-critique loop by itself, but you can trigger this manually, and the model tunes the prompt for itself through self-conversation. [Add e.g., "Verify the image, if it's incorrect, write your own prompt, try again, and repeat the process." ]
11 条评论

I've been having a lot of fun seeing all the examples from folks trying Gemini’s native image generation! So many impressive edits (yes, Gemini is the best at this), and people love the magical consistency & pixel-perfect tweaks.

But what really sets native image generation apart from stand-alone models is that it’s your multimodal inspiration engine. Here are a few fun ways to push interleaved text-image generation (which isn’t perfect yet but already very impressive) even further:

Use it as your "multimodal brainstorming partner." Ask for creative ideas & let the model generate its own prompt for the images. At least for me, they’re always better than mine!

If the model makes a mistake, don’t reset. Don’t even tell it how to fix it. Just tell the model that it's wrong and ask to fix it. It's fun to watch the model figuring it out. (I’ve found Gemini to be very polite in such interactions😀)

I picked up this trick for force triggering self-critique from @a7b2_3 🧠🧠🧠

I’ve helped 800+ agencies generate over $200,000,000 in revenue. But I’m often asked this question: “How did you even get 800+ clients?” The answer is pretty simple. A clear offer and diversified acquisition system. Let me explain.

neat! reminds me of this from the old times

What am I doing wrong?

Looks like the image might have triggered a safety filter that prevents editing certain types of content. You can try a different image... we're actively working on improving safety filters to reduce false positives!

This is beautiful. Please take it further with such a verifier-based post-training approach. These abilities might emerge naturally later. Till then why not change the system prompt to include this self-critique ability?

Tried to generate tilesets from our base ones for our games. It's almost there. I wish we could control more the output. Any advices?
