World Wires · story 28623 · corroborated · 1 source(s)

Joint On-Policy Learning and Teaching for Reinforcement Learning

This work explores joint on-policy learning and self-distillation to mitigate sparse supervision issues in reinforcement learning tasks.

Open in the desk

Coverage

What this site indexes