This paper shows that large language models struggle with commonsense reasoning due to a bias towards explicit conditions, which can be misled by irrelevant information, and that this issue can be improved by adjusting the task framing or using lightweight prompting. Practitioners caring about the reliability of language models in real-world applications might want to consider this when using them for tasks that require critical thinking.
Firehose
Filtered to Papers, tagged “inference-time prompting” · clear filters
Browse: People · Companies · Papers · Podcasts · Hacker News · Deep dives