
The episode discusses how distillation allows researchers to train smaller and cheaper AI models using larger teacher models.
Fundamental technique lets researchers use a big, expensive “teacher” model to train a “student” model for less. The story How Distillation Makes AI Models Smaller and Cheaper first appeared on Quanta Magazine.
Explore listener stats, chart rankings, contacts and more on the The Quanta Podcast podcast page.