A removal based approach to improve LLM faithfulness at test time
Read the original at arxiv.org→arXiv:2609.04343v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for consequential decisions, making their explanations an important tool for auditing model behavior. Unfortunately,...
Original headline: "A Removal Based Approach to Improve LLM Faithfulness at Test-Time"
Coverage timeline
- Sep 7, 04:00 UTC arXiv cs.AI lead source A Removal Based Approach to Improve LLM Faithfulness at Test-Time