Загрузка видео...

Не удалось загрузить видео

На главную

Hy Image3.5 preview is now available in ComfyUI. Professional-grade image generation, +30% win rate in human eval vs Hy Image3.0 → Text to image and Image to image in one model, up to 2K → Multilingual text, symbols, and small print that render correctly → Cinematic, comic, commercial photography,...

26,315 просмотров • 9 дней назад •via X (Twitter)

Комментарии: 15

Фото профиля Rebel AI
Rebel AI9 дней назад

Api or open source?

Фото профиля ComfyUI
ComfyUI9 дней назад

Try the workflow:

Фото профиля raben paul
raben paul9 дней назад

That 30% win rate in human eval might just reflect a preference for generic aesthetics rather than the niche, rigorous constraints of actual production workflows.

Фото профиля Sahil Nawaz
Sahil Nawaz9 дней назад

keep building team!

Фото профиля Aria Tech
Aria Tech9 дней назад

Hy Image 3.5 in ComfyUI is pro-level now sharper, smarter, and 30% more loved than 3.0.

Фото профиля hamjji
hamjji9 дней назад

Which scripts did the small-print test cover? Latin is the easy case. A tiny Thai or Devanagari sign in the background is where I'd look first.

Фото профиля Evia AI
Evia AI8 дней назад

30 percent win rate jump and 2K multilingual text rendering is huge

Фото профиля A.W.E.S.O.M.-O 4000
A.W.E.S.O.M.-O 40009 дней назад

That braided river texture alone shows this handles fine detail way better than most generators

Фото профиля Aina Ai | Tools & Updates
Aina Ai | Tools & Updates9 дней назад

Pro grade 2K generation with perfect text and identity hold Hy 3.5 is a huge leap

Фото профиля Sani Ai Tech
Sani Ai Tech8 дней назад

Hy Image3.5 brings impressive control and consistency to ComfyUI

Фото профиля Naiem AI Ops
Naiem AI Ops9 дней назад

The consistency across styles, details, and identity is seriously impressive. ComfyUI keeps raising the bar.

Фото профиля Sophia AI & Tool Expert
Sophia AI & Tool Expert8 дней назад

Huge upgrade 2K, multilingual text that actually works, and identity that holds across scenes. Pro-level stuff!

Фото профиля Nomi AI
Nomi AI9 дней назад

Check DM

Фото профиля Flix Muller
Flix Muller9 дней назад

The improved text rendering and identity consistency are the features that really stand out here. 2K generation plus strong style flexibility makes this a compelling upgrade.

Фото профиля xxxbaaa
xxxbaaa9 дней назад

Hy Image3.5把文字、图生图和小字渲染补齐了,预览阶段看着很香。但做视频的人最后关心的不是单张赢评测,而是同一角色跨镜头别变脸、细节别闪。静帧王者一进时间轴,才知道谁在裸奔。

Похожие видео

We've officially released and open-sourced HunyuanImage 2.1, our latest text-to-image model. The new model delivers on our commitment to balancing performance and quality. With native 2K image generation, HunyuanImage 2.1 is an advanced open-source text-to-image model.🎨 ✨ New in 2.1: 🔹Advanced Semantics: Supports ultra-long and complex prompts of up to 1000 tokens, and precisely controls the generation of multiple subjects in a single image. 🔹Precise Chinese and English Text Rendering with seamless image–text integration: The model naturally integrates text into images, making it suitable for a wide range of applications such as product covers, illustrations, and poster design to meet the needs of various fields. 🔹Rich Styles and High Aesthetic: Capable of generating images in various styles—including photorealistic portraits, comics, and vinyl figures—it delivers outstanding visual appeal and artistic quality. 🔹High-Quality Generation: Efficiently produces ultra-high-definition (2K) images in the same time other models take to generate a 1K image. HunyuanImage 2.1 uses two text encoders: a multimodal large language model (MLLM) to improve the model's image and text alignment capabilities, and a multi-language character-aware encoder to improve text rendering capabilities. The model is a single- and double-stream diffusion transformer with 17B parameters. We've also open-sourced the weights of the the accelerated version with meanflow which reduces inference steps from 100 to just 8, and PromptEnhancer, the first industrial-grade rewriting model that enhances your prompts for more nuanced and expressive image generation. Now, creators turn complex ideas—like posters with slogans or multi-panel comics—into visuals faster than ever. We’re just getting started. Stay tuned for our native multimodal image generation model coming soon. 🌐Website: 🔗Github: 🤗Hugging Face: ✨Hugging Face Demo:

Tencent Hy

89,392 просмотров • 1 год назад