This paper explores how specialist models, trained without explicit reasoning supervision, can still effectively transfer domain expertise to student models through implicit trajectory selection. Practitioners may care because this finding has implications for efficient and effective model distillation.
Firehose
Filtered to Papers, tagged “implicit trajectory selection” · clear filters
Browse: People · Companies · Papers · Podcasts · Hacker News · Deep dives