Anthropic just announced it uncovered distillation attacks by Chinese AI labs. But what even *is* model distillation? L…
From @harpercarrollai on Instagram — 619.9K followers · See full profile →
Anthropic just announced it uncovered distillation attacks by Chinese AI labs. But what even *is* model distillation? Let’s break it down. (I just taught this last weekend in my small, private, live course, so it’s fresh on my (& my students’!) mind(s)!) Anthropic added, “Distillation can be legitimate: Al labs use it to create smaller, cheaper models for their customers… But foreign labs that illicitly distill American models can remove safeguards, feeding model capabilities into their own…