Joint On-Policy Learning and Teaching for Reinforcement Learning
This work explores joint on-policy learning and self-distillation to mitigate sparse supervision issues in reinforcement learning tasks.
This work explores joint on-policy learning and self-distillation to mitigate sparse supervision issues in reinforcement learning tasks.