Loading video...
Video Failed to Load
I’m very excited for Silico to accelerate alignment research! One example, done in Silico over just a couple days: RL erodes guardrails, but we can use reward shaping to prevent that. (1/6)
12,957 views • 1 month ago •via X (Twitter)
0 Comments
No comments available
Comments from the original post will appear here


