Loading video...

Video Failed to Load

Go Home

๐ŸšจHow do LLMs acquire human values?๐Ÿค” We often point to preference optimization. However, in our new work, we trace how and when model values shift during post-training and uncover surprising dynamics. We ask: How do data, algorithms, and their interaction shape model values?๐Ÿงต

41,593 views โ€ข 10 months ago โ€ขvia X (Twitter)

0 Comments

No comments available

Comments from the original post will appear here

Related Videos