Firehose

Filtered to tagged “model distillation” · clear filters

All PeopleCompaniesPapersPodcastsHacker News

Browse by tag

12 SEP 2026 · Paper

This paper explores how specialist models, trained without explicit reasoning supervision, can still effectively transfer domain expertise to student models through implicit trajectory selection. Practitioners may care because this finding has implications for efficient and effective model distillation.

1 MAY 2026 · Podcast · "The Cognitive Revolution" | AI Builders, Researchers, and Live Player Analysis

Kyle Corbitt, founder of OpenPipe and leader of CoreWeave's serverless training team, provides a master class on reinforcement learning (RL) and custom fine-tuning for AI models. He explains how RL differs from supervised fine-tuning (SFT) …