Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

Hy Image3.5 preview is now available in ComfyUI. Professional-grade image generation, +30% win rate in human eval vs Hy Image3.0 → Text to image and Image to image in one model, up to 2K → Multilingual text, symbols, and small print that render correctly → Cinematic, comic, commercial photography,...

26,315 Aufrufe • vor 9 Tagen •via X (Twitter)

15 Kommentare

Profilbild von Rebel AI
Rebel AIvor 9 Tagen

Api or open source?

Profilbild von ComfyUI
ComfyUIvor 9 Tagen

Try the workflow:

Profilbild von raben paul
raben paulvor 9 Tagen

That 30% win rate in human eval might just reflect a preference for generic aesthetics rather than the niche, rigorous constraints of actual production workflows.

Profilbild von Sahil Nawaz
Sahil Nawazvor 9 Tagen

keep building team!

Profilbild von Aria Tech
Aria Techvor 9 Tagen

Hy Image 3.5 in ComfyUI is pro-level now sharper, smarter, and 30% more loved than 3.0.

Profilbild von hamjji
hamjjivor 8 Tagen

Which scripts did the small-print test cover? Latin is the easy case. A tiny Thai or Devanagari sign in the background is where I'd look first.

Profilbild von Evia AI
Evia AIvor 7 Tagen

30 percent win rate jump and 2K multilingual text rendering is huge

Profilbild von A.W.E.S.O.M.-O 4000
A.W.E.S.O.M.-O 4000vor 9 Tagen

That braided river texture alone shows this handles fine detail way better than most generators

Profilbild von Aina Ai | Tools & Updates
Aina Ai | Tools & Updatesvor 9 Tagen

Pro grade 2K generation with perfect text and identity hold Hy 3.5 is a huge leap

Profilbild von Sani Ai Tech
Sani Ai Techvor 8 Tagen

Hy Image3.5 brings impressive control and consistency to ComfyUI

Profilbild von Naiem AI Ops
Naiem AI Opsvor 9 Tagen

The consistency across styles, details, and identity is seriously impressive. ComfyUI keeps raising the bar.

Profilbild von Sophia AI & Tool Expert
Sophia AI & Tool Expertvor 8 Tagen

Huge upgrade 2K, multilingual text that actually works, and identity that holds across scenes. Pro-level stuff!

Profilbild von Nomi AI
Nomi AIvor 9 Tagen

Check DM

Profilbild von Flix Muller
Flix Mullervor 9 Tagen

The improved text rendering and identity consistency are the features that really stand out here. 2K generation plus strong style flexibility makes this a compelling upgrade.

Profilbild von xxxbaaa
xxxbaaavor 9 Tagen

Hy Image3.5把文字、图生图和小字渲染补齐了,预览阶段看着很香。但做视频的人最后关心的不是单张赢评测,而是同一角色跨镜头别变脸、细节别闪。静帧王者一进时间轴,才知道谁在裸奔。

Ähnliche Videos

We've officially released and open-sourced HunyuanImage 2.1, our latest text-to-image model. The new model delivers on our commitment to balancing performance and quality. With native 2K image generation, HunyuanImage 2.1 is an advanced open-source text-to-image model.🎨 ✨ New in 2.1: 🔹Advanced Semantics: Supports ultra-long and complex prompts of up to 1000 tokens, and precisely controls the generation of multiple subjects in a single image. 🔹Precise Chinese and English Text Rendering with seamless image–text integration: The model naturally integrates text into images, making it suitable for a wide range of applications such as product covers, illustrations, and poster design to meet the needs of various fields. 🔹Rich Styles and High Aesthetic: Capable of generating images in various styles—including photorealistic portraits, comics, and vinyl figures—it delivers outstanding visual appeal and artistic quality. 🔹High-Quality Generation: Efficiently produces ultra-high-definition (2K) images in the same time other models take to generate a 1K image. HunyuanImage 2.1 uses two text encoders: a multimodal large language model (MLLM) to improve the model's image and text alignment capabilities, and a multi-language character-aware encoder to improve text rendering capabilities. The model is a single- and double-stream diffusion transformer with 17B parameters. We've also open-sourced the weights of the the accelerated version with meanflow which reduces inference steps from 100 to just 8, and PromptEnhancer, the first industrial-grade rewriting model that enhances your prompts for more nuanced and expressive image generation. Now, creators turn complex ideas—like posters with slogans or multi-panel comics—into visuals faster than ever. We’re just getting started. Stay tuned for our native multimodal image generation model coming soon. 🌐Website: 🔗Github: 🤗Hugging Face: ✨Hugging Face Demo:

Tencent Hy

89,392 Aufrufe • vor 1 Jahr