Research1 min read
Test-time removal approach improves LLM explanation faithfulness
A new test-time method enhances LLM explanation faithfulness by removing uncredited concepts from input, applicable without model modifications, and tested across datasets and models.
From arXiv cs.AI