Anthropic says Alibaba illicitly extracted Claude AI model capabilities (opens in new tab)
There's two basic kinds of distillation: 1) the massive [and dumb] method where you ask a question and use the answer as reinforcement (Black Box), and 2) more targeted distillation where you use one model to directly inform/train/guide another model (RLAIF).
Read the original article