Quantus: An Explainable AI Toolkit for Responsible Evaluation of Neural Network Explanations and Beyond
The evaluation of explanation methods is a research topic that has not yet been explored deeply, however, since explainability is supposed to strengthen trust in artificial intelligence, it is necessary to systematically review and compare explanation methods in order to confirm their correctness. Until now, no tool with focus on XAI evaluation exists that exhaustively and speedily allows researchers to evaluate the performance of explanations of neural network predictions. To increase transparency and reproducibility in the field, we therefore built Quantus -- a comprehensive, evaluation toolkit in Python that includes a growing, well-organised collection of evaluation metrics and tutorials for evaluating explainable methods. The toolkit has been thoroughly tested and is available under an open-source license on PyPi (or on https://github.com/understandable-machine-intelligence-lab/Quantus/).
Code (1)
Tasks
Explainable Artificial Intelligence (XAI)Similar Papers 제목 키워드 기반
CNN-based explanation ensembling for dataset, representation and explanations evaluation
Explainable Artificial Intelligence has gained significant attention due to the widespread use of complex deep learning models in high-stake domains such as medicine, finance, and autonomous cars. However, different expl…
Explainable artificial intelligenceThe Meta-Evaluation Problem in Explainable AI: Identifying Reliable Estimators with MetaQuantus
One of the unsolved challenges in the field of Explainable AI (XAI) is determining how to most reliably estimate the quality of an explanation method in the absence of ground truth explanation labels. Resolving this issu…
Explainable Artificial Intelligence (XAI)The Explabox: Model-Agnostic Machine Learning Transparency & Analysis
We present the Explabox: an open-source toolkit for transparent and responsible machine learning (ML) model development and usage. Explabox aids in achieving explainable, fair and robust models by employing a four-step s…
DescriptiveFairnessDeepSight: An All-in-One LM Safety Toolkit
As the development of Large Models (LMs) progresses rapidly, their safety is also a priority. In current Large Language Models (LLMs) and Multimodal Large Language Models (MLLMs) safety workflow, evaluation, diagnosis, a…
Nuquantus: Machine learning software for the characterization and quantification of cell nuclei in complex immunofluorescent tissue images
Determination of fundamental mechanisms of disease often hinges on histopathology visualization and quantitative image analysis. Currently, the analysis of multi-channel fluorescence tissue images is primarily achieved b…