paper-with-me

Papers

Closing the AI Trust Gap: The Case for Independent Certification for Trustworthy AI

2026-07-17 · Trisevgeni Papakonstantinou, Cansu Canca, Farah Nanji, Waheedullah Pardess, Jen Weedon, Jasmijn Remmers, Eliza Krigman, Matthew Ball, Yalda Daryani, Kiran Iqbal, Francielle Vargas, María Llorente Sánchez, Joe Humphreys, Fendi Tsim, Kelly Fitzpatrick, Jeff Dunn, Catherine Feldman arxiv

Over the past decade, responsible AI (RAI) has produced a substantial body of practice for identifying and mitigating the risks AI poses in high-stakes settings. Yet this work has not produced a market that rewards trustworthiness. Firms that invest seriously in safety, fairness, and oversight cannot consistently prove to consumers, regulators, and shareholders that their systems go beyond the bare minimum of compliance. What is missing is a way for society to recognize or compare the difference. The result is a trust gap: a structural condition in which responsible development efforts happen inside organizations but produce no external, independently recognized and verifiable signal of trustworthy outcomes. We argue this gap is sustained in part because of a focus on responsible AI (a matter of internal process) as opposed to trustworthy AI (a matter of independently verifiable real-world outcomes), and that it persists because of three compounding failures: (1) the market cannot distinguish trustworthy systems from their imitations; (2) evaluation targets models and outputs rather than deployed sociotechnical systems and their outcomes; (3) the measurement ecosystem is oriented toward avoiding harm rather than demonstrating benefit. Reviewing existing AI governance instruments and comparing them to certification regimes in healthcare, sustainability, and security, we show that none integrate a governance baseline, independently verified positive-outcome evidence, and market signaling in a single framework. We propose independent, outcome-oriented certification as the connective layer that can close the trust gap, complementing regulation and internal governance by making trustworthiness measurable, comparable, and commercially rewarded.

📄 PDF Abstract BibTeX arXiv:2607.15992

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Certification Labels for Trustworthy AI: Insights From an Empirical Mixed-Method Study

2023-05-15 · Nicolas Scharowski, Michaela Benk, Swen J. Kühne, Léane Wettstein 외

Auditing plays a pivotal role in the development of trustworthy AI. However, current research primarily focuses on creating auditable AI documentation, which is intended for regulators and experts rather than end-users a…

Survey

Trustworthy Artificial Intelligence in the Context of Metrology

2024-06-14 · Tameem Adel, Sam Bilson, Mark Levene, Andrew Thompson

We review research at the National Physical Laboratory (NPL) in the area of trustworthy artificial intelligence (TAI), and more specifically trustworthy machine learning (TML), in the context of metrology, the science of…

Uncertainty Quantification

Rethinking Certification for Trustworthy Machine Learning-Based Applications

2023-05-26 · Marco Anisetti, Claudio A. Ardagna, Nicola Bena, Ernesto Damiani

Machine Learning (ML) is increasingly used to implement advanced applications with non-deterministic behavior, which operate on the cloud-edge continuum. The pervasive adoption of ML is urgently calling for assurance sol…

Fairness

Ensuring trustworthy and ethical behaviour in intelligent logical agents

2024-02-12 · Stefania Costantini

Autonomous Intelligent Agents are employed in many applications upon which the life and welfare of living beings and vital social functions may depend. Therefore, agents should be trustworthy. A priori certification tech…

Trustworthy Machine Learning through the Lens of Combinatorial Optimization: Survey and Research Perspectives

2026-07-08 · Thibaut Vidal, Julien Ferry arxiv

Modern machine learning (ML) increasingly relies on complex models whose behavior is difficult to characterize beyond empirical performance metrics. Across a wide range of tasks, including prediction, generation, and dec…

Explanation GenerationModel Compression