Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

🚀 Excited to introduce Qwen-Image-Edit! Built on 20B Qwen-Image, it brings precise bilingual text editing (Chinese & English) while preserving style, and supports both semantic and appearance-level editing. ✨ Key Features ✅ Accurate text editing with bilingual support ✅ High-level semantic editing (e.g. object rotation, IP creation) ✅ Low-level...

659,536 Aufrufe • vor 1 Jahr •via X (Twitter)

21 Kommentare

Profilbild von Qwen
Qwenvor 1 Jahr

HF demo:

Profilbild von batuhan the fal guy
batuhan the fal guyvor 1 Jahr

Available day 0 at @fal as the launch partner

Profilbild von GosuCoder
GosuCodervor 1 Jahr

Is Qwen legit one of the best AI companies now?

Profilbild von Capx AI
Capx AIvor 1 Jahr

time to experiment

Profilbild von Rayan
Rayanvor 1 Jahr

who wants an über fast version of the model? Available with our inference partners :)

Profilbild von AK
AKvor 1 Jahr

now available in anycoder:

Profilbild von Chong-U
Chong-Uvor 1 Jahr

With this there’s almost no reason to keep Photoshop

Profilbild von apolinario (poli)
apolinario (poli)vor 1 Jahr

try it out on the official demo out on @huggingface spaces!

Profilbild von Petri Kuittinen
Petri Kuittinenvor 1 Jahr

Cool! Here is a young female and anime version of me, James Bond poster and demonic being. I used my profile photo as source for edits. Prompts for the latter two: "turn into James Bond in a Bond 007 movie poster" "turn into demonic creature with fangs and glowing red eyes."

Profilbild von sbstndbs
sbstndbsvor 1 Jahr

So this is the nano banana? 🍌

Profilbild von Shridhar
Shridharvor 1 Jahr

this eatss

Profilbild von Rizki
Rizkivor 1 Jahr

Qwen keep giving free tools, its crazy

Profilbild von Anes Valentic
Anes Valenticvor 1 Jahr

Amazing news! Been waiting for this one! 🚀 A huge thank you to the Qwen for open sourcing SOTA models! Keep up the amazing work!

Profilbild von FluidVoice
FluidVoicevor 1 Jahr

Oh wow. If this works as intended out of the box, freaking game changer!! Let me make a quick test run!! Thanks @Alibaba_Qwen for changing the open source game.

Profilbild von William Lamkin
William Lamkinvor 1 Jahr

nano banana, is that you? 🥺

Profilbild von 0 ツ
0 ツvor 1 Jahr

Bro wtf is this

Profilbild von Himanshu Kumar
Himanshu Kumarvor 1 Jahr

Impressive. Bridging language barriers in visual content creation opens exciting possibilities for global collaboration.

Profilbild von Techgoalz
Techgoalzvor 1 Jahr

This is getting ridiculous. How is Qwen able to keep churning out state of the art model update every day?

Profilbild von Respect
Respectvor 1 Jahr

Whoa, kinda impressed how well you keep context of the rest of the image after a change, we see a lot of models change other parts of the image even when not asked

Profilbild von bruce
brucevor 1 Jahr

cool

Profilbild von Gadgetify
Gadgetifyvor 1 Jahr

very cool. Just tested it with a cat :)

Ähnliche Videos

We’re excited to announce the release and open-source of HunyuanImage 3.0 — the largest and most powerful open-source text-to-image model to date, with over 80 billion total parameters, of which 13 billion are activated per token during inference.The effect is completely comparable to the industry’s flagship closed-source model.🚀🚀🚀 HunyuanImage 3.0 originates from our internally developed native multimodal large language model, with fine-tuning and post-training focused on text-to-image generation. This unique foundation gives the model a powerful set of capabilities: ✅Reason with world knowledge ✅Understand complex, thousand-word prompts ✅Generate precise text within images Different from traditional DiT architecture image generation models, HunyuanImage 3.0’s MoE architecture uses a Transfusion-based approach to deeply couple Diffusion and LLM training for a single, powerful system. Built on Hunyuan-A13B, HunyuanImage 3.0 was trained on a massive dataset: 5 billion image-text pairs, video frames, interleaved image-text data, and 6 trillion tokens of text corpora. This hybrid training across multimodal generation, understanding, and LLM capabilities allows the model to seamlessly integrate multiple tasks. Whether you're an illustrator, designer, or creator, this is built to slash your workflow from hours to minutes. HunyuanImage 3.0 can generate intricate text, detailed comics, expressive emojis, and lively, engaging illustrations for educational content. The current release focuses solely on text-to-image generation and future updates will include image-to-image, image editing, multi-turn interaction, and more. 👉🏻Try it now: 🔗GitHub: 🤗Hugging Face:

Tencent Hy

413,175 Aufrufe • vor 1 Jahr

Blended-NeRF: Zero-Shot Object Generation and Blending in Existing Neural Radiance Fields paper page: Editing a local region or a specific object in a 3D scene represented by a NeRF is challenging, mainly due to the implicit nature of the scene representation. Consistently blending a new realistic object into the scene adds an additional level of difficulty. We present Blended-NeRF, a robust and flexible framework for editing a specific region of interest in an existing NeRF scene, based on text prompts or image patches, along with a 3D ROI box. Our method leverages a pretrained language-image model to steer the synthesis towards a user-provided text prompt or image patch, along with a 3D MLP model initialized on an existing NeRF scene to generate the object and blend it into a specified region in the original scene. We allow local editing by localizing a 3D ROI box in the input scene, and seamlessly blend the content synthesized inside the ROI with the existing scene using a novel volumetric blending technique. To obtain natural looking and view-consistent results, we leverage existing and new geometric priors and 3D augmentations for improving the visual fidelity of the final result. We test our framework both qualitatively and quantitatively on a variety of real 3D scenes and text prompts, demonstrating realistic multi-view consistent results with much flexibility and diversity compared to the baselines. Finally, we show the applicability of our framework for several 3D editing applications, including adding new objects to a scene, removing/replacing/altering existing objects, and texture conversion.

AK

62,768 Aufrufe • vor 3 Jahren

We've officially released and open-sourced HunyuanImage 2.1, our latest text-to-image model. The new model delivers on our commitment to balancing performance and quality. With native 2K image generation, HunyuanImage 2.1 is an advanced open-source text-to-image model.🎨 ✨ New in 2.1: 🔹Advanced Semantics: Supports ultra-long and complex prompts of up to 1000 tokens, and precisely controls the generation of multiple subjects in a single image. 🔹Precise Chinese and English Text Rendering with seamless image–text integration: The model naturally integrates text into images, making it suitable for a wide range of applications such as product covers, illustrations, and poster design to meet the needs of various fields. 🔹Rich Styles and High Aesthetic: Capable of generating images in various styles—including photorealistic portraits, comics, and vinyl figures—it delivers outstanding visual appeal and artistic quality. 🔹High-Quality Generation: Efficiently produces ultra-high-definition (2K) images in the same time other models take to generate a 1K image. HunyuanImage 2.1 uses two text encoders: a multimodal large language model (MLLM) to improve the model's image and text alignment capabilities, and a multi-language character-aware encoder to improve text rendering capabilities. The model is a single- and double-stream diffusion transformer with 17B parameters. We've also open-sourced the weights of the the accelerated version with meanflow which reduces inference steps from 100 to just 8, and PromptEnhancer, the first industrial-grade rewriting model that enhances your prompts for more nuanced and expressive image generation. Now, creators turn complex ideas—like posters with slogans or multi-panel comics—into visuals faster than ever. We’re just getting started. Stay tuned for our native multimodal image generation model coming soon. 🌐Website: 🔗Github: 🤗Hugging Face: ✨Hugging Face Demo:

Tencent Hy

89,392 Aufrufe • vor 1 Jahr

🧬 BREAKING: Our CRISPR-GPT paper is out TODAY in Nature Biomedical Engineering Nature Biomedical Engineering ! 🤯 We built an AI agent that turns ANYONE into a gene-editing expert in 1 DAY instead of months. An undergrad with ZERO experience achieved 90%+ editing efficiency on their FIRST attempt. 🧵 Here's how we're building expert AI agents for cutting-edge biotechnology: 🎯 The Problem: CRISPR is revolutionary but requires PhD-level expertise, it can take weeks to learn, adopt, and design, analyze a CRISPR experiment for R&D or making life-saving medicine. Even Pro scientists can make small mistakes (e.g. typos in guideRNA or cloning design) that cost months to find out, slowing us down. 💡 Our Solution: CRISPR-GPT - an AI co-pilot from Stanford University Princeton University Google DeepMind that guides you through EVERY step via simple conversation 🔬 Real Results: -Novice researcher: ~90% editing on 1st go -Training time: Months → 1 day -100% success rate in our trials -Even experts save days/weeks on data analysis & troubleshooting 🤖 How it works: Our multi-agent system handles: CRISPR system and delivery method selection, guideRNA design, Protocol generation, Real-time troubleshooting, Data analysis, and beyond. All through natural language! No coding, no complex software. 📊 We benchmarked it extensively: -288 evaluation scenarios/cases -Outperformed GPT-4o on ALL gene editing tasks -Trained on 11 years of expert discussions -Covers knockout, base-editing, prime-editing & epigenetic editing 🌍 Why this matters: -Every lab can now use CRISPR with an AI system distilling expert knowledge and skills. -Every student can learn faster. -Every researcher can tackle bigger challenges without worrying about small mistakes. -Customized CRISPR design can be automated based on your need and the context of R&D workflow. -Agentic AI ensure safety, privacy, and responsibility -We're not just automating gene editing - we're using AI to power scientists to cure diseases. 🚀 Try it yourself! Beta access available at: Paper: Code: Benchmark (companion work, Genome-bench): Co-first and key authors: Yuanhao Qu Kaixuan Huang Ming Yin PIs: Le Cong@Stanford, AI+Bio+Gene-Editing Mengdi Wang Key collaborators: Russ Altman Denny Zhou The future of biology and science is conversational. The future is now. Nature Biomedical Engineering Nature Portfolio #CRISPR #AI #GeneEditing #Biotech #Science #AISafety

Le Cong@Stanford, AI+Bio+Gene-Editing

111,938 Aufrufe • vor 1 Jahr