Media Summary: November 20 session where we are diving into the paper " Authors: Zichen Liu, Changyu Chen, Wenjun Li, Penghui Qi, Tianyu Pang, Chao Du, Wee Sun Lee, Min Lin DeepSeek- Want to ask live questions and join a community of over 1200 AI researchers, engineers, and nerds who LOVE AI? Join Arxiv ...

Understanding R1 Zero Like Training - Detailed Analysis & Overview

November 20 session where we are diving into the paper " Authors: Zichen Liu, Changyu Chen, Wenjun Li, Penghui Qi, Tianyu Pang, Chao Du, Wee Sun Lee, Min Lin DeepSeek- Want to ask live questions and join a community of over 1200 AI researchers, engineers, and nerds who LOVE AI? Join Arxiv ... 0:00 Intro 0:55 Chain of thought 1:31 Reinforcement Learning 2:23 Bonus GRPO 3:02 Model Distillation 3:41 Outro Curious about ... 0:00 - 2:24 Paper Overview 2:24 - 7:41 Code Walkthrough 1 7:41 - 15:33 GRPO Full Can a large language model learn to reason — not just guess — using reinforcement learning alone? In this video, Rina walks ...

Frustrated your company isn't maximizing AI? Get your AI score out of 10 (free, 2 min): ... DeepSeek AI has released multiple models, each designed for different tasks. In this video, I explain the differences between ...

Photo Gallery

Exploring "Understanding R1-Zero-Like Training (Dr. GRPO)" | Deep Learning Study Session
Understanding R1-Zero-Like Training: A Critical Perspective
Dr. GRPO: Understanding R1-Zero-Like Training with Zichen Liu
2503.20783 - Understanding R1 Zero Like Training: A Critical Perspective
GitHub - sail-sg/understand-r1-zero: Understanding R1-Zero-Like Training: A Critical Perspective
DeepSeek R1 Explained by AI Expert: How R1-Zero Led to an AI Breakthrough
DeepSeek R1 Theory Overview | GRPO + RL + SFT
How R1 and GRPO Work (Deep Technical Dive into DeepSeeks Models)
DeepSeek R1 Explained like you're 5
DeepSeek R1 TRAINING SECRETS You Need to Know! (With Code)
DeepSeek-R1 Explained: How Reinforcement Learning Teaches LLMs to Reason (Open-Source AI
How to Train LLMs to "Think" (o1 & DeepSeek-R1)
View Detailed Profile
Exploring "Understanding R1-Zero-Like Training (Dr. GRPO)" | Deep Learning Study Session

Exploring "Understanding R1-Zero-Like Training (Dr. GRPO)" | Deep Learning Study Session

November 20 session where we are diving into the paper "

Understanding R1-Zero-Like Training: A Critical Perspective

Understanding R1-Zero-Like Training: A Critical Perspective

Authors: Zichen Liu, Changyu Chen, Wenjun Li, Penghui Qi, Tianyu Pang, Chao Du, Wee Sun Lee, Min Lin DeepSeek-

Dr. GRPO: Understanding R1-Zero-Like Training with Zichen Liu

Dr. GRPO: Understanding R1-Zero-Like Training with Zichen Liu

R1

2503.20783 - Understanding R1 Zero Like Training: A Critical Perspective

2503.20783 - Understanding R1 Zero Like Training: A Critical Perspective

title:

GitHub - sail-sg/understand-r1-zero: Understanding R1-Zero-Like Training: A Critical Perspective

GitHub - sail-sg/understand-r1-zero: Understanding R1-Zero-Like Training: A Critical Perspective

https://github.com/sail-sg/understand-r1-zero

DeepSeek R1 Explained by AI Expert: How R1-Zero Led to an AI Breakthrough

DeepSeek R1 Explained by AI Expert: How R1-Zero Led to an AI Breakthrough

Watch the full video: ...

DeepSeek R1 Theory Overview | GRPO + RL + SFT

DeepSeek R1 Theory Overview | GRPO + RL + SFT

Here's an overview of the DeepSeek

How R1 and GRPO Work (Deep Technical Dive into DeepSeeks Models)

How R1 and GRPO Work (Deep Technical Dive into DeepSeeks Models)

Want to ask live questions and join a community of over 1200 AI researchers, engineers, and nerds who LOVE AI? Join Arxiv ...

DeepSeek R1 Explained like you're 5

DeepSeek R1 Explained like you're 5

0:00 Intro 0:55 Chain of thought 1:31 Reinforcement Learning 2:23 Bonus GRPO 3:02 Model Distillation 3:41 Outro Curious about ...

DeepSeek R1 TRAINING SECRETS You Need to Know! (With Code)

DeepSeek R1 TRAINING SECRETS You Need to Know! (With Code)

0:00 - 2:24 Paper Overview 2:24 - 7:41 Code Walkthrough 1 7:41 - 15:33 GRPO Full

DeepSeek-R1 Explained: How Reinforcement Learning Teaches LLMs to Reason (Open-Source AI

DeepSeek-R1 Explained: How Reinforcement Learning Teaches LLMs to Reason (Open-Source AI

Can a large language model learn to reason — not just guess — using reinforcement learning alone? In this video, Rina walks ...

How to Train LLMs to "Think" (o1 & DeepSeek-R1)

How to Train LLMs to "Think" (o1 & DeepSeek-R1)

Frustrated your company isn't maximizing AI? Get your AI score out of 10 (free, 2 min): ...

DeepSeek R1 vs DeepSeek R1 Zero [Architecture Explained] | Run DeepSeek R1 Locally with Ollama

DeepSeek R1 vs DeepSeek R1 Zero [Architecture Explained] | Run DeepSeek R1 Locally with Ollama

DeepSeek AI has released multiple models, each designed for different tasks. In this video, I explain the differences between ...