Hosted by Jinoo B. and Ngoc Long T.
Join us in discussing the paper — Reinforcement Learning via Self-Distillation https://arxiv.org/abs/2601.20802v2