Media Summary: [PoD] Self-Distilled Reasoner : On-Policy Self-Distillation for Large Language Models I recently met Sasha Rush and he started giving me an impromptu lecture on how targeted on- How do we train a small yet powerful AI model? This video explains knowledge
Self Distilled Reasoner On Policy - Detailed Analysis & Overview
[PoD] Self-Distilled Reasoner : On-Policy Self-Distillation for Large Language Models I recently met Sasha Rush and he started giving me an impromptu lecture on how targeted on- How do we train a small yet powerful AI model? This video explains knowledge Disclaimer: This video is generated with Google's NotebookLM. Rethinking On- In this AI Research Roundup episode, Alex discusses the paper: ' In this video, we sit down with Jonas Hübotter (ETH Zurich) and Idan Shenfeld (MIT) to break down
00:00 - Intro 01:39 - Motivation: Why On-