paper-with-me

Papers

Using AI Agents to Automate Black-Box Audits of Personalization Algorithms at Scale

2026-06-29 · Alessandro Morosini, Sarah H. Cen, Andrew Ilyas, Hedi Driss, Aleksander Mądry, Chara Podimata arxiv

Personalization algorithms determine what content users encounter on online platforms. Auditing these systems is difficult because independent auditors have only black-box access to the algorithms, while personalization depends on users' attributes, behavior, and evolving interaction histories. Existing auditing methods face a tradeoff: studies with real users capture realistic behavior but are costly and hard to control, whereas sock-puppet audits scale more easily but often rely on scripted behavior that limits realism. Beyond this, both approaches struggle to decouple user attributes from user behavior, limiting our ability to causally understand personalization. To address this gap, we introduce a framework for black-box audits of personalization algorithms using generative AI agents as behavioral engines for synthetic accounts. Each agent is instantiated with a fixed persona, grounded in demographic and political survey data, and interacts with a platform's content by reasoning about it and choosing actions. Because behavior is fixed within each persona while platform-visible signals such as age, gender, or location can be experimentally perturbed, our design enables counterfactual auditing of how platforms respond to user attributes. As a case study, we deploy 1,120 agents on X shortly after the 2024 U.S. election, spanning 14 personas and three counterfactual conditions, collecting over 200,000 content exposures. We find that X's algorithmic feed amplifies toxic, polarizing, political, and right-leaning content relative to the chronological feed, with amplification varying sharply by user ideology. Counterfactual analyses show that demographic signals affect content delivery in persona-dependent ways: pooled effects are largely null, while subgroup-level effects vary in direction and magnitude. Our work establishes GenAI-based agents as a new tool for algorithmic auditing.

📄 PDF Abstract BibTeX arXiv:2606.30801

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Black-Box Access is Insufficient for Rigorous AI Audits

2024-01-25 · Stephen Casper, Carson Ezell, Charlotte Siegmann, Noam Kolt 외

External audits of AI systems are increasingly recognized as a key mechanism for AI governance. The effectiveness of an audit, however, depends on the degree of access granted to auditors. Recent audits of state-of-the-a…

Narrow Secret Loyalty Dodges Black-Box Audits

2026-05-07 · Alfie Lamerton, Fabien Roger arxiv

Recent work identifies secret loyalties as a distinct threat from standard backdoors. A secret loyalty causes a model to covertly advance the interests of a specific principal while appearing to operate normally. We cons…

Deployment-Time Memorization in Foundation-Model Agents

2026-06-08 · Lei, Chen, Guilin Zhang, Kai Zhao 외 arxiv

Foundation-model agents are increasingly long-lived systems that remember users across interactions, making memorization an explicit deployment-time function rather than solely a property of model weights. Existing work …

Algorithmic audits of algorithms, and the law

2022-02-15 · Erwan Le Merrer, Ronan Pons, Gilles Trédan

Algorithmic decision making is now widespread, ranging from health care allocation to more common actions such as recommendation or information ranking. The aim to audit these algorithms has grown alongside. In this pape…

Decision Making

System Cards for AI-Based Decision-Making for Public Policy

2022-03-01 · Furkan Gursoy, Ioannis A. Kakadiaris

Decisions impacting human lives are increasingly being made or assisted by automated decision-making algorithms. Many of these algorithms process personal data for predicting recidivism, credit risk analysis, identifying…

Decision MakingFace Recognition