Media Summary: November 20 session where we are diving into the paper " Authors: Zichen Liu, Changyu Chen, Wenjun Li, Penghui Qi, Tianyu Pang, Chao Du, Wee Sun Lee, Min Lin DeepSeek- Want to ask live questions and join a community of over 1200 AI researchers, engineers, and nerds who LOVE AI? Join Arxiv ...
Understanding R1 Zero Like Training - Detailed Analysis & Overview
November 20 session where we are diving into the paper " Authors: Zichen Liu, Changyu Chen, Wenjun Li, Penghui Qi, Tianyu Pang, Chao Du, Wee Sun Lee, Min Lin DeepSeek- Want to ask live questions and join a community of over 1200 AI researchers, engineers, and nerds who LOVE AI? Join Arxiv ... 0:00 Intro 0:55 Chain of thought 1:31 Reinforcement Learning 2:23 Bonus GRPO 3:02 Model Distillation 3:41 Outro Curious about ... 0:00 - 2:24 Paper Overview 2:24 - 7:41 Code Walkthrough 1 7:41 - 15:33 GRPO Full Can a large language model learn to reason — not just guess — using reinforcement learning alone? In this video, Rina walks ...
Frustrated your company isn't maximizing AI? Get your AI score out of 10 (free, 2 min): ... DeepSeek AI has released multiple models, each designed for different tasks. In this video, I explain the differences between ...