World Wires · story 2569 · corroborated · 1 source(s)

Distillation Defenses Easily Break After Reinforcement Learning

A study shows that existing defenses against model distillation attacks fail after reinforcement learning is applied to the distilled models.

Open in the desk

Coverage

What this site indexes