Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

Today we’re launching INTELLECT-2: The first decentralized 32B-parameter RL training run open to join for anyone with compute — fully permissionless. Scaling towards frontier reasoning across coding, math and science.

353,791 Aufrufe • vor 1 Jahr •via X (Twitter)

11 Kommentare

Profilbild von Prime Intellect
Prime Intellectvor 1 Jahr

INTELLECT-2 brings decentralized training into the inference-time compute era: • Fully async, decentralized reinforcement learning • Eliminating communication overhead • Scalable across heterogeneous GPUs worldwide

Profilbild von Prime Intellect
Prime Intellectvor 1 Jahr

Over the past months, we’ve built the full open-source stack to enable INTELLECT-2: • PRIME-RL: fully async decentralized RL • GENESYS & SYNTHETIC-1: crowdsourced tasks & verifiers for RL • TOPLOC validation: verifiable inference with low overhead • Protocol Testnet: global AI coordination infrastructure

Profilbild von Prime Intellect
Prime Intellectvor 1 Jahr

How INTELLECT-2 operates: • Inference Rollout Workers: Decentralized swarm collects rollouts using the latest policy model • TOPLOC Validators: Verify the inference computations • GRPO Training Workers: Train on new data and broadcast weights via our shardcast library

Profilbild von Prime Intellect
Prime Intellectvor 1 Jahr

Asynchronous RL completely eliminates communication bottlenecks. Our ablation studies confirm we maintain performance even with 4-step delays, making decentralized training viable with weak global interconnects.

Profilbild von Prime Intellect
Prime Intellectvor 1 Jahr

With INTELLECT-2 we aim for frontier reasoning performance with a controllable thinking budget. By incorporating length rewards into our training run, users can specify how long the model should reason for a given task.

Profilbild von Prime Intellect
Prime Intellectvor 1 Jahr

Contribute compute and watch training live on our dashboard: • Participate in our testnet powering an open reasoning model • Slashing and validation keep contributions honest Join us in building towards open and decentralized AGI

Profilbild von Lab4crypto
Lab4cryptovor 1 Jahr

🚀 Don't gamble with your portfolio! Use our advanced hybrid quant risk tool using on/off-chain data and make informed decisions. 📈 Acess to 1000+ charts for your crypto journey. 📚Join our Premium Telegram for daily alerts. 📊+21 projects supported. 🏗️ Beginners and experts.

Profilbild von Michael Luo
Michael Luovor 1 Jahr

This is amazing work! Looks like asynchronous RL for LLMs is pretty stable 🤩, a paradigm shift away from traditional Deep RL where synchronous PPO won, as opposed to A3C and IMPALA

Profilbild von Tom Bennet
Tom Bennetvor 1 Jahr

Decentralized RL training? Sounds like we're breeding digital superminds in a worldwide compute farm. Count me in for the chaos.

Profilbild von jesse.base.eth
jesse.base.ethvor 1 Jahr

ok this is really sick

Profilbild von Burny — Effective Omni
Burny — Effective Omnivor 1 Jahr

Decentralized open source is climbing

Ähnliche Videos