Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

Hy Image3.5 preview is now available in ComfyUI. Professional-grade image generation, +30% win rate in human eval vs Hy Image3.0 → Text to image and Image to image in one model, up to 2K → Multilingual text, symbols, and small print that render correctly → Cinematic, comic, commercial photography,...

26,315 görüntüleme • 9 gün önce •via X (Twitter)

15 Yorum

Rebel AI profil fotoğrafı
Rebel AI9 gün önce

Api or open source?

ComfyUI profil fotoğrafı
ComfyUI9 gün önce

Try the workflow:

raben paul profil fotoğrafı
raben paul9 gün önce

That 30% win rate in human eval might just reflect a preference for generic aesthetics rather than the niche, rigorous constraints of actual production workflows.

Sahil Nawaz profil fotoğrafı
Sahil Nawaz9 gün önce

keep building team!

Aria Tech profil fotoğrafı
Aria Tech9 gün önce

Hy Image 3.5 in ComfyUI is pro-level now sharper, smarter, and 30% more loved than 3.0.

hamjji profil fotoğrafı
hamjji8 gün önce

Which scripts did the small-print test cover? Latin is the easy case. A tiny Thai or Devanagari sign in the background is where I'd look first.

Evia AI profil fotoğrafı
Evia AI8 gün önce

30 percent win rate jump and 2K multilingual text rendering is huge

A.W.E.S.O.M.-O 4000 profil fotoğrafı
A.W.E.S.O.M.-O 40009 gün önce

That braided river texture alone shows this handles fine detail way better than most generators

Aina Ai | Tools & Updates profil fotoğrafı
Aina Ai | Tools & Updates9 gün önce

Pro grade 2K generation with perfect text and identity hold Hy 3.5 is a huge leap

Sani Ai Tech profil fotoğrafı
Sani Ai Tech8 gün önce

Hy Image3.5 brings impressive control and consistency to ComfyUI

Naiem AI Ops profil fotoğrafı
Naiem AI Ops9 gün önce

The consistency across styles, details, and identity is seriously impressive. ComfyUI keeps raising the bar.

Sophia AI & Tool Expert profil fotoğrafı
Sophia AI & Tool Expert8 gün önce

Huge upgrade 2K, multilingual text that actually works, and identity that holds across scenes. Pro-level stuff!

Nomi AI profil fotoğrafı
Nomi AI9 gün önce

Check DM

Flix Muller profil fotoğrafı
Flix Muller9 gün önce

The improved text rendering and identity consistency are the features that really stand out here. 2K generation plus strong style flexibility makes this a compelling upgrade.

xxxbaaa profil fotoğrafı
xxxbaaa9 gün önce

Hy Image3.5把文字、图生图和小字渲染补齐了,预览阶段看着很香。但做视频的人最后关心的不是单张赢评测,而是同一角色跨镜头别变脸、细节别闪。静帧王者一进时间轴,才知道谁在裸奔。

Benzer Videolar

We've officially released and open-sourced HunyuanImage 2.1, our latest text-to-image model. The new model delivers on our commitment to balancing performance and quality. With native 2K image generation, HunyuanImage 2.1 is an advanced open-source text-to-image model.🎨 ✨ New in 2.1: 🔹Advanced Semantics: Supports ultra-long and complex prompts of up to 1000 tokens, and precisely controls the generation of multiple subjects in a single image. 🔹Precise Chinese and English Text Rendering with seamless image–text integration: The model naturally integrates text into images, making it suitable for a wide range of applications such as product covers, illustrations, and poster design to meet the needs of various fields. 🔹Rich Styles and High Aesthetic: Capable of generating images in various styles—including photorealistic portraits, comics, and vinyl figures—it delivers outstanding visual appeal and artistic quality. 🔹High-Quality Generation: Efficiently produces ultra-high-definition (2K) images in the same time other models take to generate a 1K image. HunyuanImage 2.1 uses two text encoders: a multimodal large language model (MLLM) to improve the model's image and text alignment capabilities, and a multi-language character-aware encoder to improve text rendering capabilities. The model is a single- and double-stream diffusion transformer with 17B parameters. We've also open-sourced the weights of the the accelerated version with meanflow which reduces inference steps from 100 to just 8, and PromptEnhancer, the first industrial-grade rewriting model that enhances your prompts for more nuanced and expressive image generation. Now, creators turn complex ideas—like posters with slogans or multi-panel comics—into visuals faster than ever. We’re just getting started. Stay tuned for our native multimodal image generation model coming soon. 🌐Website: 🔗Github: 🤗Hugging Face: ✨Hugging Face Demo:

Tencent Hy

89,392 görüntüleme • 1 yıl önce