This paper develops a technique to identify and mitigate demographic bias in large language models by selectively pruning specific neurons in neural networks, without significantly impacting the model's overall performance. Practitioners in AI and NLP may care about this method as it could help create more fair and transparent language models.
Firehose
Filtered to Papers, tagged “bias” · clear filters
Browse: People · Companies · Papers · Podcasts · Hacker News · Deep dives