This paper helps developers make stronger backdoor attacks on large language models by learning to select the most effective set of poisoned examples. Practitioners might care about this because it can be used to improve the security of these models in real-world applications.
Firehose
Filtered to Papers, tagged “poisoning attacks” · clear filters
Browse: People · Companies · Papers · Podcasts · Hacker News · Deep dives