Necessary or Sufficient? Evaluating LLM Explanations With Behavioural Evidence

Administrator 0 阅读

AI Digest - ArXiv AI

Necessary or Sufficient? Evaluating LLM Explanations With Behavioural Evidence

LLM decision components that can operate within agent workflows often produce action-relevant recommendations or judgements together with explanations. Operators may use the named factors to monitor a system, diagnose errors, or decide when to escalate an output. Such use assumes that the explanations agree with the component's observable decision behaviour. We test two interpretations of the named factors: necessity, meaning that changing a factor would change the output, and sufficiency, meani


Source: ArXiv AI