paper-with-me

홈 › Papers

Unifying Corroborative and Contributive Attributions in Large Language Models

2023-11-20 · Theodora Worledge, Judy Hanwen Shen, Nicole Meister, Caleb Winston, Carlos Guestrin

As businesses, products, and services spring up around large language models, the trustworthiness of these models hinges on the verifiability of their outputs. However, methods for explaining language model outputs largely fall across two distinct fields of study which both use the term "attribution" to refer to entirely separate techniques: citation generation and training data attribution. In many modern applications, such as legal document generation and medical question answering, both types of attributions are important. In this work, we argue for and present a unified framework of large language model attributions. We show how existing methods of different types of attribution fall under the unified framework. We also use the framework to discuss real-world use cases where one or both types of attributions are required. We believe that this unified framework will guide the use case driven development of systems that leverage both types of attribution, as well as the standardization of their evaluation.

📄 PDF Abstract BibTeX arXiv:2311.12233

Code (0)

등록된 구현이 없습니다.

Tasks

Language ModelingLanguage ModellingLarge Language ModelMedical Question AnsweringQuestion Answering

Similar Papers 제목 키워드 기반

EventMapper: Detecting Real-World Physical Events Using Corroborative and Probabilistic Sources

2020-01-23 · Abhijit Suprem, Calton Pu

The ubiquity of social media makes it a rich source for physical event detection, such as disasters, and as a potential resource for crisis management resource allocation. There have been some recent works on leveraging …

BIG-bench Machine LearningEvent DetectionManagement

Autonomous Robotic Swarms: A Corroborative Approach for Verification and Validation

2024-07-22 · Dhaminda B. Abeywickrama, Suet Lee, Chris Bennett, Razanne Abu-Aisheh 외

The emergent behaviour of autonomous robotic swarms poses a significant challenge to their safety assurance. Assurance tasks encompass adherence to standards, certification processes, and the execution of verification an…

Distributing Synergy Functions: Unifying Game-Theoretic Interaction Methods for Machine-Learning Explainability

2023-05-04 · Daniel Lundstrom, Meisam Razaviyayn

Deep learning has revolutionized many areas of machine learning, from computer vision to natural language processing, but these high-performance models are generally "black box." Explaining such models would improve tran…

Decision MakingFairness

Dbnary: Wiktionary as a LMF based Multilingual RDF network

2012-05-01 · LREC 2012 5 · Gilles S{\'e}rasset

Contributive resources, such as wikipedia, have proved to be valuable in Natural Language Processing or Multilingual Information Retrieval applications.This article focusses on Wiktionary, the dictionary part of the coll…

Information RetrievalRetrieval

Unifying Perplexing Behaviors in Modified BP Attributions through Alignment Perspective

2025-03-14 · Guanhua Zheng, Jitao Sang, Changsheng Xu

Attributions aim to identify input pixels that are relevant to the decision-making process. A popular approach involves using modified backpropagation (BP) rules to reverse decisions, which improves interpretability comp…

Decision MakingSensitivity