This paper explores how transformer representations change over time and how these changes can be understood and manipulated. Practitioners might care about this research because it could lead to more robust and efficient transformer models, especially in applications where model updates need to be edited or compressed.
Firehose
Filtered to Papers, tagged “neural compression” · clear filters
Browse: People · Companies · Papers · Podcasts · Hacker News · Deep dives