Model Explanations under Calibration
Explaining and interpreting the decisions of recommender systems are becoming extremely relevant both, for improving predictive performance, and providing valid explanations to users. While most of the recent interest has focused on providing local explanations, there has been a much lower emphasis on studying the effects of model dynamics and its impact on explanation. In this paper, we perform a focused study on the impact of model interpretability in the context of calibration. Specifically, we address the challenges of both over-confident and under-confident predictions with interpretability using attention distribution. Our results indicate that the means of using attention distributions for interpretability are highly unstable for un-calibrated models. Our empirical analysis on the stability of attention distribution raises questions on the utility of attention for explainability.
Code (1)
Tasks
modelRecommendation SystemsvalidMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Improving Perturbation-based Explanations by Understanding the Role of Uncertainty Calibration
Perturbation-based explanations are widely utilized to enhance the transparency of machine-learning models in practice. However, their reliability is often compromised by the unknown model behavior under the specific per…
Calibration Meets Explanation: A Simple and Effective Approach for Model Confidence Estimates
Calibration strengthens the trustworthiness of black-box models by producing better accurate confidence estimates on given examples. However, little is known about if model explanations can help confidence calibration. I…
A Study on the Calibration of In-context Learning
Accurate uncertainty quantification is crucial for the safe deployment of machine learning models, and prior research has demonstrated improvements in the calibration of modern language models (LMs). We study in-context …
In-Context LearningNatural Language UnderstandingUncertainty QuantificationExplain then Rank: Scale Calibration of Neural Rankers Using Natural Language Explanations from LLMs
In search settings, calibrating the scores during the ranking process to quantities such as click-through rates or relevance levels enhances a system's usefulness and trustworthiness for downstream users. While previous …
Document RankingLearning-To-RankExplanation-based Counterfactual Retraining(XCR): A Calibration Method for Black-box Models
With the rapid development of eXplainable Artificial Intelligence (XAI), a long line of past work has shown concerns about the Out-of-Distribution (OOD) problem in perturbation-based post-hoc XAI models and explanations …
counterfactualExplainable artificial intelligenceExplainable Artificial Intelligence (XAI)Feature Importance