This paper explores how specialist models, trained without explicit reasoning supervision, can still effectively transfer domain expertise to student models through implicit trajectory selection. Practitioners may care because this finding has implications for efficient and effective model distillation.
Firehose
Filtered to Papers, tagged “model distillation” · clear filters
Browse: People · Companies · Papers · Podcasts · Hacker News · Deep dives