Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

We're reintroducing and open-sourcing project "See-through". Given a single anime illustration, it automatically decomposes the character into fully-inpainted semantic layers with depth ordering. One image in, layered PSD out. (1/n) Repo:

580,679 görüntüleme • 5 ay önce •via X (Twitter)

70 Yorum

柔情猫娘 profil fotoğrafı
柔情猫娘5 ay önce

Under the hood: a diffusion-based body part consistency module (built on SDXL), paired with pixel-level pseudo-depth fine-tuned on Marigold. It automatically resolves interleaving structures like overlapping hair strands, accessories behind hair, clothing layers, etc. (2/n)

柔情猫娘 profil fotoğrafı
柔情猫娘5 ay önce

Is this Image-to-Live2D? Not really. Live2D needs artistic decomposition choices + rigging. We only automate segmentation and occlusion inpainting. But it's a solid starting point that may save plenty of manual work. (3/n)

柔情猫娘 profil fotoğrafı
柔情猫娘5 ay önce

Looking for community contributions — ComfyUI nodes, Colab notebooks, workflow integrations. PRs welcome, we'll feature your work. Enjoy! Finally, props to Lvmin Zhang. He is a legend. Honestly, without his work I might have lost all motivation to pursue this direction. (n/n)

夏落荷 希/Live2D委託開放中 profil fotoğrafı
夏落荷 希/Live2D委託開放中5 ay önce

It only took about 10 minutes, that's amazing!

柔情猫娘 profil fotoğrafı
柔情猫娘5 ay önce

Thanks for your report! It should be faster like 2mins on H200, or 3-4 mins on 4090. Maybe we will consider some optimizations in the future...

Even profil fotoğrafı
Even5 ay önce

Can you make it so it only redraws/create what’s missing and hidden and not the whole illustration bcs AI-ing over the input doesn’t look good and kinda ruins the point It’d be way more hidden if it could keep the input and add the output to it, not overwrite input yk

柔情猫娘 profil fotoğrafı
柔情猫娘5 ay önce

Potentially there will be seams near the inpainting boundary. Poisson image editing may ultimately solve this ( but please understand this is a research project and all codes are more PoC-like. We encourage all kinds of improvements over our work.

もみじ profil fotoğrafı
もみじ5 ay önce

@3rrabundus この素晴らしいモデルの分割情報をMetaのSAMのような分離モデルに渡して「分離」し、パーツごとの描き足し部分をPhotoshopの生成塗りつぶしのように元のレイヤーと綺麗に馴染むように「生成」するというようなことができたら、業界に大きなゲームチェンジが起こせると思いました💭

柔情猫娘 profil fotoğrafı
柔情猫娘5 ay önce

@3rrabundus I guess you can try adapting it with SAM3+LaMa, it should work but I am not sure about the final quality.

Sin(Game in Developing) profil fotoğrafı
Sin(Game in Developing)5 ay önce

The resolution of the test image is 3649x2770 pixels. The tested graphics card is 4090Ti 24G. The problems encountered are as follows: 1. The resolution has decreased significantly. 2. The details of the generated image have become chaotic. 3. The recognition of non-standing postures is not very good.😳

Mingfei Guo profil fotoğrafı
Mingfei Guo5 ay önce

弹弹

Reigen profil fotoğrafı
Reigen5 ay önce

woah! this can be super useful to artists in some scenarios. a question, why choose marigold instead of depth anything2? is just because marigold is also diffusion based.

柔情猫娘 profil fotoğrafı
柔情猫娘5 ay önce

Yes. Quite the reason. We have also tested with depth anything but find the results quite unstable. Guess this is related to the internal depth representation (but not investigate thoroughly).

Yesith Thomas profil fotoğrafı
Yesith Thomas5 ay önce

Does it break it down into layers of image or in vectors? but nonetheless super cool! gonna check and test it out myself! Good Stuff @ljsabc

柔情猫娘 profil fotoğrafı
柔情猫娘5 ay önce

It’s in pixels. There are also researches for vector decomposition but the motivations are quite different there.

Yesith Thomas profil fotoğrafı
Yesith Thomas5 ay önce

Interesting... gives me more reason to try it out now though! Thanks!

SagaSu profil fotoğrafı
SagaSu5 ay önce

非常棒的项目!!!期待更多更新!!

柔情猫娘 profil fotoğrafı
柔情猫娘5 ay önce

还缺几个训练脚本我们会尽快补上,也欢迎大家多出意见多给建议,看看这东西还能咋用…

SagaSu profil fotoğrafı
SagaSu5 ay önce

如果能拆的很自然就很棒

Cloud profil fotoğrafı
Cloud5 ay önce

喜报!!! 以后大价钱约的有PSD的图也是AI图了.jpg

鳥遊 profil fotoğrafı
鳥遊5 ay önce

言語翻訳のおかげであなたのポストが即読めた、ありがたい。 ものすごいツールだ

HOShura profil fotoğrafı
HOShura5 ay önce

我只想拿來抅走背景

柔情猫娘 profil fotoğrafı
柔情猫娘5 ay önce

那这东西就够了

HOShura profil fotoğrafı
HOShura5 ay önce

看情況 有些第一身不太行 想試你這個

CosmicCorsair profil fotoğrafı
CosmicCorsair5 ay önce

I think this is a match made in heaven with @MohoAnimation

LucidPlay profil fotoğrafı
LucidPlay5 ay önce

Hi, thanks for your work on see-through! 1. Can it handle game screenshots with UI/effects for layer separation, or mainly anime-style images? 2. How does it differ from end-to-end layered models like Qwen-Image-Layered? Thanks!

柔情猫娘 profil fotoğrafı
柔情猫娘5 ay önce

@NexusUI_Offical 1. We trained our model mainly based on anime characters so we cannot really promise the generalization on out of distribution samples (humanoids, furry, dragons, etc., but it’s worth trying).

柔情猫娘 profil fotoğrafı
柔情猫娘5 ay önce

@NexusUI_Offical 2. The Qwen model requires prompting and may not always follow the desired semantics ( as in Figure 6, It’s hard to isolate a specific part apart from the composition.

LucidPlay profil fotoğrafı
LucidPlay5 ay önce

Thanks a lot for the detailed reply!

Aderek profil fotoğrafı
Aderek5 ay önce

This, if it were for pixel art, would speed up my process by about 30%. xD

Yuja✦ | PROPIC profil fotoğrafı
Yuja✦ | PROPIC5 ay önce

Is there any problem if I change the size from 1280 to 1920? It takes time for the 1280, but I confirmed that it works well I haven't confirmed that it's working yet with 1920

柔情猫娘 profil fotoğrafı
柔情猫娘5 ay önce

I don’t see these will be obvious problems except the VRAM usage. I think you may have a try.

Yuja✦ | PROPIC profil fotoğrafı
Yuja✦ | PROPIC5 ay önce

May I ask you something When I set the resolution value of the file 'inference_psd.py' to 1920 to output the size to 1920, png is output normally, but psd is not generated Is there any other way? It's okay to combine png files, but I'm asking if there's a way

柔情猫娘 profil fotoğrafı
柔情猫娘5 ay önce

Yes. We see this issue and the update is on its way. Before that, you can modify: on line 50 of inference_psd.py, change: apply_marigold() by adding a new parameter: resolution=args.resolution, and it should temporarily work.

Yuja✦ | PROPIC profil fotoğrafı
Yuja✦ | PROPIC5 ay önce

Thank you for your consideration. I'll try it!

柔情猫娘 profil fotoğrafı
柔情猫娘5 ay önce

No problem. We hard-coded our depth part for a reason, and we will notice you when the update arrives.

Yuja✦ | PROPIC profil fotoğrafı
Yuja✦ | PROPIC5 ay önce

Thank you so much !

柔情猫娘 profil fotoğrafı
柔情猫娘5 ay önce

now we have updated the codebase to separate diffusion/marigold resolution, git pull and rerun should produce PSD correctly. Thanks again for your feedback (w)

Sebastian Buzdugan profil fotoğrafı
Sebastian Buzdugan5 ay önce

cool demo, next step is integrating rigging so animators can actually ship it

DenomRS (OPEN comms) profil fotoğrafı
DenomRS (OPEN comms)5 ay önce

I may be doing something wrong. But is it normal for it to take up to 24 GB of disk space after 1 use?

柔情猫娘 profil fotoğrafı
柔情猫娘5 ay önce

You need to download the model and deps. But this will only happy once and no additional spaces will be consumed.

DenomRS (OPEN comms) profil fotoğrafı
DenomRS (OPEN comms)5 ay önce

Got it, I will give it a try. Thank you

Megaaziib profil fotoğrafı
Megaaziib5 ay önce

this insane!, thanks for sharing this.

elecpure profil fotoğrafı
elecpure5 ay önce

Very useful!

Uipalent transport profil fotoğrafı
Uipalent transport5 ay önce

Really cool!

Jacky Fan profil fotoğrafı
Jacky Fan5 ay önce

how to use this to make animation?

LucidPlay profil fotoğrafı
LucidPlay1 ay önce

Why doesn’t my generated output have ears?

Lorelai White profil fotoğrafı
Lorelai White5 ay önce

This is why technology exists... Absolutely incredible work! 🤔 I need to tinker with this...

Youshou Xi profil fotoğrafı
Youshou Xi5 ay önce

这个需要多少显存才能跑的动?

Farr profil fotoğrafı
Farr5 ay önce

holyshit

nuru profil fotoğrafı
nuru5 ay önce

Hello I love you ok bye

Jase Hazard 💭 profil fotoğrafı
Jase Hazard 💭5 ay önce

Ah so will this be working on Dragonbones Pro?

woctordho profil fotoğrafı
woctordho5 ay önce

Do you think it can be used to extract lineart, or decompose base color/highlight/shadow? There are some classic works like Illyasviel's Erasing Appearance Preservation, but it seems lineart extraction is still an unsolved problem.

柔情猫娘 profil fotoğrafı
柔情猫娘5 ay önce

Check this out (but I don’t know when they will open source their models) By the way the vanilla Qwen-Image-Layered has already done a good job in separating the line arts (if prompted carefully). So I believe high-quality data is sticky.

柔情猫娘 profil fotoğrafı
柔情猫娘5 ay önce

* is the key.

Mash profil fotoğrafı
Mash5 ay önce

Damn! Fr 🤯

A.I.Warper profil fotoğrafı
A.I.Warper5 ay önce

@sin_ceriously

Court Reinland profil fotoğrafı
Court Reinland5 ay önce

Amazing

Kazuki profil fotoğrafı
Kazuki5 ay önce

Hello ! Are you planning to do a version compatible with Apple silicon chip? Because I see that we need PyTorch with Cuda 12.8 so on my Mac not useable. Thanks

柔情猫娘 profil fotoğrafı
柔情猫娘5 ay önce

I did test with my mps device but find some ops are mps-incompatible. Running everything on CPU is incredibly slow. We will consider this in the future, but I think potentially a better solution could be a Colab solution so it can balance cost and execution time.

Gary W 🌐 profil fotoğrafı
Gary W 🌐5 ay önce

What kind of hardware would you suggest is the minimum? I saw a reply saying a h200 should take around 2 minutes, 3 - 4 on a 4090. I've a 4060ti, so 16GB of VRAM, would that be enough or are we looking at 24GB - 32GB? I don't mind it taking 10-20, it still saves hours of work!

柔情猫娘 profil fotoğrafı
柔情猫娘5 ay önce

loading in fp16 should fit, but you may expect long running like 10mins or something.

Gary W 🌐 profil fotoğrafı
Gary W 🌐5 ay önce

That's perfectly fine, considering the time saved over all. Will definitely be fun to experiment with! Great work and thank you for sharing :)

柔情猫娘 profil fotoğrafı
柔情猫娘5 ay önce

Happy to know that. Always let me know if you meet any issues later.

bluebear profil fotoğrafı
bluebear5 ay önce

这玩意应该受限于sam的性能?如果jpeg退化的图或者由北京的情况会不会糟糕一些?

柔情猫娘 profil fotoğrafı
柔情猫娘5 ay önce

这是个diffusion model,可能会受VAE的影响退点画质,但是跟SAM没啥关系

bluebear profil fotoğrafı
bluebear5 ay önce

训练数据是不是没有画师授权没法分发?

柔情猫娘 profil fotoğrafı
柔情猫娘5 ay önce

是的,这是我们最大的遗憾。不过我觉得大公司很容易就有这些数据,至于小团队可以先用我们的模型infer然后bootstrap一个差不多的数据出来,再用我们的UI洗一遍

穆泽 profil fotoğrafı
穆泽5 ay önce

很厉害的项目,这样应该会更加的节省独立游戏的制作时间...总之很强,过几天试一下插件部署到云端..

柔情猫娘 profil fotoğrafı
柔情猫娘5 ay önce

好的!未来如果遇到什么问题或者有什么意见建议都欢迎和我们说!

Benzer Videolar