Loading video...

Video Failed to Load

Go Home

Hy Image3.5 preview is now available in ComfyUI. Professional-grade image generation, +30% win rate in human eval vs Hy Image3.0 → Text to image and Image to image in one model, up to 2K → Multilingual text, symbols, and small print that render correctly → Cinematic, comic, commercial photography,...

26,315 views • 9 days ago •via X (Twitter)

15 Comments

Rebel AI's profile picture
Rebel AI9 days ago

Api or open source?

ComfyUI's profile picture
ComfyUI9 days ago

Try the workflow:

raben paul's profile picture
raben paul9 days ago

That 30% win rate in human eval might just reflect a preference for generic aesthetics rather than the niche, rigorous constraints of actual production workflows.

Sahil Nawaz's profile picture
Sahil Nawaz9 days ago

keep building team!

Aria Tech's profile picture
Aria Tech9 days ago

Hy Image 3.5 in ComfyUI is pro-level now sharper, smarter, and 30% more loved than 3.0.

hamjji's profile picture
hamjji9 days ago

Which scripts did the small-print test cover? Latin is the easy case. A tiny Thai or Devanagari sign in the background is where I'd look first.

Evia AI's profile picture
Evia AI8 days ago

30 percent win rate jump and 2K multilingual text rendering is huge

A.W.E.S.O.M.-O 4000's profile picture
A.W.E.S.O.M.-O 40009 days ago

That braided river texture alone shows this handles fine detail way better than most generators

Aina Ai | Tools & Updates's profile picture
Aina Ai | Tools & Updates9 days ago

Pro grade 2K generation with perfect text and identity hold Hy 3.5 is a huge leap

Sani Ai Tech's profile picture
Sani Ai Tech8 days ago

Hy Image3.5 brings impressive control and consistency to ComfyUI

Naiem AI Ops's profile picture
Naiem AI Ops9 days ago

The consistency across styles, details, and identity is seriously impressive. ComfyUI keeps raising the bar.

Sophia AI & Tool Expert's profile picture
Sophia AI & Tool Expert8 days ago

Huge upgrade 2K, multilingual text that actually works, and identity that holds across scenes. Pro-level stuff!

Nomi AI's profile picture
Nomi AI9 days ago

Check DM

Flix Muller's profile picture
Flix Muller9 days ago

The improved text rendering and identity consistency are the features that really stand out here. 2K generation plus strong style flexibility makes this a compelling upgrade.

xxxbaaa's profile picture
xxxbaaa9 days ago

Hy Image3.5把文字、图生图和小字渲染补齐了,预览阶段看着很香。但做视频的人最后关心的不是单张赢评测,而是同一角色跨镜头别变脸、细节别闪。静帧王者一进时间轴,才知道谁在裸奔。

Related Videos

We've officially released and open-sourced HunyuanImage 2.1, our latest text-to-image model. The new model delivers on our commitment to balancing performance and quality. With native 2K image generation, HunyuanImage 2.1 is an advanced open-source text-to-image model.🎨 ✨ New in 2.1: 🔹Advanced Semantics: Supports ultra-long and complex prompts of up to 1000 tokens, and precisely controls the generation of multiple subjects in a single image. 🔹Precise Chinese and English Text Rendering with seamless image–text integration: The model naturally integrates text into images, making it suitable for a wide range of applications such as product covers, illustrations, and poster design to meet the needs of various fields. 🔹Rich Styles and High Aesthetic: Capable of generating images in various styles—including photorealistic portraits, comics, and vinyl figures—it delivers outstanding visual appeal and artistic quality. 🔹High-Quality Generation: Efficiently produces ultra-high-definition (2K) images in the same time other models take to generate a 1K image. HunyuanImage 2.1 uses two text encoders: a multimodal large language model (MLLM) to improve the model's image and text alignment capabilities, and a multi-language character-aware encoder to improve text rendering capabilities. The model is a single- and double-stream diffusion transformer with 17B parameters. We've also open-sourced the weights of the the accelerated version with meanflow which reduces inference steps from 100 to just 8, and PromptEnhancer, the first industrial-grade rewriting model that enhances your prompts for more nuanced and expressive image generation. Now, creators turn complex ideas—like posters with slogans or multi-panel comics—into visuals faster than ever. We’re just getting started. Stay tuned for our native multimodal image generation model coming soon. 🌐Website: 🔗Github: 🤗Hugging Face: ✨Hugging Face Demo:

Tencent Hy

89,392 views • 1 year ago