Switch language한국어
Back to the list

The Distillation Game: Adaptive Attacks & Efficient Defenses

TL;DR AI

Key summary

2 min read
  1. A machine-learning paper reframes model distillation as a minimax game between a utility-constrained teacher and an adaptive student.

  2. It introduces adaptive evaluation methods and a forward-pass-only Product-of-Experts defense against model extraction and imitation.

  3. The study finds adaptive students recover much more capability than passive tests suggest, weakening confidence in standard evaluations.

  4. On GSM8K and MATH, the cheaper Product-of-Experts defense performs nearly as well as more expensive defenses under stronger testing.

Read the original