paper-with-me

Papers

TEDI: Trustworthy and Ethical Dataset Indicators to Analyze and Compare Dataset Documentation

2025-05-23 · Wiebke Hutiri, Mircea Cimpoi, Morgan Scheuerman, Victoria Matthews, Alice Xiang

Dataset transparency is a key enabler of responsible AI, but insights into multimodal dataset attributes that impact trustworthy and ethical aspects of AI applications remain scarce and are difficult to compare across datasets. To address this challenge, we introduce Trustworthy and Ethical Dataset Indicators (TEDI) that facilitate the systematic, empirical analysis of dataset documentation. TEDI encompasses 143 fine-grained indicators that characterize trustworthy and ethical attributes of multimodal datasets and their collection processes. The indicators are framed to extract verifiable information from dataset documentation. Using TEDI, we manually annotated and analyzed over 100 multimodal datasets that include human voices. We further annotated data sourcing, size, and modality details to gain insights into the factors that shape trustworthy and ethical dimensions across datasets. We find that only a select few datasets have documented attributes and practices pertaining to consent, privacy, and harmful content indicators. The extent to which these and other ethical indicators are addressed varies based on the data collection method, with documentation of datasets collected via crowdsourced and direct collection approaches being more likely to mention them. Scraping dominates scale at the cost of ethical indicators, but is not the only viable collection method. Our approach and empirical insights contribute to increasing dataset transparency along trustworthy and ethical dimensions and pave the way for automating the tedious task of extracting information from dataset documentation in future.

📄 PDF Abstract BibTeX arXiv:2505.17841

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Optimizing Ethical Risk Reduction for Medical Intelligent Systems with Constraint Programming

2025-10-08 · Clotilde Brayé, Aurélien Bricout, Arnaud Gotlieb, Nadjib Lazaar 외 arxiv

Medical Intelligent Systems (MIS) are increasingly integrated into healthcare workflows, offering significant benefits but also raising critical safety and ethical concerns. According to the European Union AI Act, most M…

Towards Transparent Ethical AI: A Roadmap for Trustworthy Robotic Systems

2025-08-07 · Ahmad Farooq, Kamran Iqbal arxiv

As artificial intelligence (AI) and robotics increasingly permeate society, ensuring the ethical behavior of these systems has become paramount. This paper contends that transparency in AI decision-making processes is fu…

RE-centric Recommendations for the Development of Trustworthy(er) Autonomous Systems

2023-05-29 · Krishna Ronanki, Beatriz Cabrero-Daniel, Jennifer Horkoff, Christian Berger

Complying with the EU AI Act (AIA) guidelines while developing and implementing AI systems will soon be mandatory within the EU. However, practitioners lack actionable instructions to operationalise ethics during AI syst…

Ethics

"Model Cards for Model Reporting" in 2024: Reclassifying Category of Ethical Considerations in Terms of Trustworthiness and Risk Management

2024-02-15 · DeBrae Kennedy-Mayo, Jake Gord

In 2019, the paper entitled "Model Cards for Model Reporting" introduced a new tool for documenting model performance and encouraged the practice of transparent reporting for a defined list of categories. One of the cate…

FairnessManagementmodel

Finding differences in perspectives between designers and engineers to develop trustworthy AI for autonomous cars

2023-07-01 · Gustav Jonelid, K. R. Larsson

In the context of designing and implementing ethical Artificial Intelligence (AI), varying perspectives exist regarding developing trustworthy AI for autonomous cars. This study sheds light on the differences in perspect…