Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

1/ Introducing RL Swarm’s new backend: GenRL. A modular reinforcement learning library built for distributed, fault-tolerant training - now powering RL Swarm from the ground up. 🧵

82,191 Aufrufe • vor 1 Jahr •via X (Twitter)

10 Kommentare

Profilbild von gensyn
gensynvor 1 Jahr

2/ Each worker runs its own environment instance, contributes asynchronously to a shared rollout buffer, and updates its model weights independently, so no central controller is required.

Profilbild von gensyn
gensynvor 1 Jahr

3/ GenRL allows RL Swarm to work with any environment, described intuitively through code. This launch incorporates Reasoning Gym out-of-the-box, giving access to >100 community-created environments with no extra configuration required.

Profilbild von gensyn
gensynvor 1 Jahr

4/ What’s new: – Modular GenRL backend – Expanded configuration surface – Prebuilt Docker image for easy deployment – Reasoning Gym environment to enhance model reasoning capabilities – New multi-task swarm

Profilbild von gensyn
gensynvor 1 Jahr

5/ Now live on the Gensyn testnet. You can run RL-Swarm with GenRL today. Full code + setup:

Profilbild von gensyn
gensynvor 1 Jahr

6/ A node update is required for GenRL. Please visit ⁠support-discussion in the Discord if you have any questions.

Profilbild von Gautamgg 🕵
Gautamgg 🕵vor 1 Jahr

I want to ask 1 que What about previous trained model data rewards & participants bec it's not showing Is that data saved in your database? @fenbielding @_jamico @_grieve waiting for ans 💙

Profilbild von Mintair | One Click Node🪄
Mintair | One Click Node🪄vor 1 Jahr

Looks really interesting, we gotta setup our own custom environment.

Profilbild von AJDominic (🐱,🐐)
AJDominic (🐱,🐐)vor 1 Jahr

What gensyn cooking is unmatched!

Profilbild von Bitduke
Bitdukevor 1 Jahr

Cool, cool - more modularity

Profilbild von lior.eth (Lior Messika)
lior.eth (Lior Messika)vor 1 Jahr

These retro vibes are everything I ever wanted from an AI lab

Ähnliche Videos

🚨 RL for LLMs is finally accessible. Introducing OpenTinker: The first community-driven, open-source framework designed to democratize Reinforcement Learning for LLMs. Inspired by Thinking Machines's amazing Tinker, we realize the biggest bottleneck in agentic LLM research isn’t the math—it’s the setup. Current RL pipelines are messy. Configuring VeRL for every single experiment is a productivity killer. OpenTinker fixed it. 🛠 How OpenTinker Works: Decoupled Design of Server and Client - Setup Once, Run Forever: Configure the OpenTinker backend on your GPU cluster once. - Develop Locally: Define your RL environments directly on your laptop. - Train on the Cloud: Simply point your local client to the backend. The cluster handles the compute; you handle the science. 📉 The 10x Development Efficiency Thanks to our elegant architectural decomposition, OpenTinker reduces the time to develop a new RL training pipeline by at least an order of magnitude. ⚡ Turn Idle GPU Compute into Gold Small labs often have underutilized hardware. OpenTinker turns your idle GPUs into an internal/external API service for - RL Training - SFT - Inference 🎯 Who needs OpenTinker? - Researchers tired of infrastructure hell. - Labs needing to standardize workflows. - Teams wanting to maximize hardware ROI. Thanks my amazing PhD student Siqi Zhu for leading the project. We are building the future of open RL infra. Be the first to build with us. 👇 Start Building with OpenTinker Now 🚀 Repo: 🌐 Blog: If you believe RL should be accessible to everyone, give us a star, repost this 🔄 post, and let us know what agents you plan to build!

Jiaxuan You

58,258 Aufrufe • vor 8 Monaten