
Finds only a weak -memory conflict effect when LLMs are given counterfactual context, suggesting faithfulness issues in systems may stem less from belief conflict than assumed.
“Contrary to our expectations, we observe only a weak context-memory conflict on the human-annotated sample.”
“We also find that a suboptimal choice of LLM judge would lead to overestimating the strength of the context-memory conflict.”
articleLarge Language Models and Language Server Protocol: a match made in contextAlessandro Schena, Ilgiz Mustafin, Julia Kotovich
articleSWORD: Wikidata-based Distortions Reveal Hidden Cross-Lingual Inconsistencies in LLM Factual Error RejectionSanghyeok Park, Minji Kang, Hosung Kwak, Jinhyuk Yun
articleA Removal Based Approach to Improve LLM Faithfulness at Test-TimeQinglan Luo, S M A Nahian, John Guttag, S. Mazdak Abulnaga, Katie MattonChecking sign-in…
Loading comments…