Anthropic latest "misuse of AI" report claims distillation can improve general reasoning enough to increase dangerous capabilities beyond the subjects covered in the training conversations.
It also says Claude’s safeguards do not transfer through unauthorized distillation, but provides no quantified evaluation for these claims.