正在加载视频...
视频加载失败
New Course: Reinforcement Fine-Tuning LLMs with GRPO! Learn to use reinforcement learning to improve your LLM performance in this short course, built in collaboration with Predibase by Rubrik, and taught by Travis Addair, its Co-Founder and CTO, and Arnav Garg, its Senior Engineer and Machine Learning Lead. Reasoning models... show more
9 条评论

@predibase @TravisAddair @grg_arnav I am gonna binge learn tonight!!!

💡 Learn how Reinforcement Learning can boost your trading performance! In this free Substack article I share full code of a trading algorithm based on Reinforcement Learning that beats other Machine Learning models as well as simply buying and holding the stock.

@predibase @TravisAddair @grg_arnav exciting news. this course sounds like a fantastic opportunity to deepen our understanding of reinforcement learning in llms.

@predibase @TravisAddair @grg_arnav The intersection of reinforcement learning and LLMs presents fascinating possibilities for performance enhancement. How impactful is this for your work? 🤔 #LLMs

@predibase @TravisAddair @grg_arnav @AndrewYNg, how can reinforcement learning reshape our approach to LLM performance? Exciting developments ahead. #AIInnovation

@predibase @TravisAddair @grg_arnav This course could vastly improve strategic abilities with LLM models.

@TravisAddair @grg_arnav Thank you @AndrewYNg! It was an honor working with you and the team to bring this course to life. Now anyone can take a small open-source #LLM and turn it into a reasoning powerhouse tailored to their use case with as little as 10 labeled data examples!

@predibase @TravisAddair @grg_arnav Their "effect" got me ☠️

@predibase @TravisAddair @grg_arnav Dear Professor, I kindly request you to consider reinstating 100% financial aid on Coursera. This would be immensely beneficial for learners like me who rely on such support to access quality education. Thank you for your time and consideration.
