正在加载视频...

视频加载失败

(1/n) Time to unify your favorite visual generative models, VLMs, and simulators for controllable visual generation—Introducing a Product of Experts (PoE) framework for inference-time knowledge composition from heterogeneous models.

49,081 次观看 • 1 年前 •via X (Twitter)

6 条评论

Yunzhi Zhang 的头像
Yunzhi Zhang1 年前

(2/n) The composition yields better controllability and provides flexible user interfaces for specifying visual synthesis goals, enabling applications such as composing physics simulation into generated videos…

Yunzhi Zhang 的头像
Yunzhi Zhang1 年前

(3/n) …inserting graphics engine rendering into images, and more.

Yunzhi Zhang 的头像
Yunzhi Zhang1 年前

(4/n) PoE sampling is non-trivial in high dimensions. We adopt Annealed Importance Sampling, where particles are initially drawn from a simple base distribution and steered towards the target, with transition kernels computed from expert models.. Two possible annealing paths:

Yunzhi Zhang 的头像
Yunzhi Zhang1 年前

(5/5) Page: More details in paper: Team work with the incredible Carson Murtuza-Lanier, @zizhang_li, @du_yilun, and @jiajunwu_cs!

Rainmaker 的头像
Rainmaker2 年前

Join me as I put several Machine Learning models head-to-head to see which one can beat the market and deliver strong returns. In this free Substack post I share several models that deliver better returns with much lower drawdown compared to Buy-and-Hold approach.

Aisha 的头像
Aisha1 年前

Whats the best visual video generative models in your experience ?

相关视频