Loading video...
Video Failed to Load
🚨 RL for LLMs is finally accessible. Introducing OpenTinker: The first community-driven, open-source framework designed to democratize Reinforcement Learning for LLMs. Inspired by Thinking Machines's amazing Tinker, we realize the biggest bottleneck in agentic LLM research isn’t the math—it’s the setup. Current RL pipelines are messy. Configuring VeRL for... show more
58,326 views • 9 months ago •via X (Twitter)
15 Comments

@thinkymachines How does this compare to berkeley sky RL tinker library?

The '10x Efficiency' claim holds up because the bottleneck in RL isn't training time; it's Setup Time. If I have to spend 4 hours configuring a distributed environment just to test a new reward function, I only run one experiment a day. OpenTinker effectively commoditizes the infrastructure layer, allowing researchers to treat the training loop like a simple function call. This tightens the feedback loop from 'Idea' to 'Loss Curve' drastically.

@thinkymachines Whats the difference between this and tinker

@thinkymachines looks promising! thanks for sharing, would love to check it out

@thinkymachines Accessible frameworks are great until you need to explain to your team why the RL training failed overnight. What's the debugging experience look like for non-engineers who just want the model to improve?

@thinkymachines important work

@thinkymachines Great work Jiaxuan

@thinkymachines Looks great!

@thinkymachines

@thinkymachines This hits the real bottleneck RL for LLMs isn’t “hard math,” it’s plumbing + reproducibility. Turning idle GPUs into an internal RL/SFT/inference service is 🔥 too. Curious: does OpenTinker bake in eval/trace + reward/RM versioning + multi-tenant isolation out of the box?

@thinkymachines sounds cool cant wait to check it out

@thinkymachines open sourcing this is a really nice move honestly

@thinkymachines Whats wrong with verifiers from PI? What’s your differentiator?

@thinkymachines I've observed the bottleneck; OpenTinker offers a practical, accessible solution.

@thinkymachines Sounds very useful! Plan to try it out.

