
The article explains that "distillation attacks" refer to the use of outputs from stronger AI models to train smaller models, a legitimate technique in AI development. However, its misuse by Chinese labs is being misinterpreted as an illicit practice, leading to confusion and potential regulatory challenges. The text warns that such misuse could harm U.S. companies by giving them a competitive disadvantage, while advocating for a balanced approach to prevent a negative connotation of the term and ensure the proper use of APIs.
Reflections on AI's writing ability and how AI models get more capable.

After a few long years of finding time to document my lessons from training open models, my post-training book is done!

Musings on model alignment, what determines safety, and where we go from here.
