paper-with-me

Papers

De-anonymization of authors through arXiv submissions during double-blind review

2020-07-01 · Homanga Bharadhwaj, Dylan Turpin, Animesh Garg, Ashton Anderson

In this paper, we investigate the effects of releasing arXiv preprints of papers that are undergoing a double-blind review process. In particular, we ask the following research question: What is the relation between de-anonymization of authors through arXiv preprints and acceptance of a research paper at a (nominally) double-blind venue? Under two conditions: papers that are released on arXiv before the review phase and papers that are not, we examine the correlation between the reputation of their authors with the review scores and acceptance decisions. By analyzing a dataset of ICLR 2020 and ICLR 2019 submissions (n=5050), we find statistically significant evidence of positive correlation between percentage acceptance and papers with high reputation released on arXiv. In order to understand this observed association better, we perform additional analyses based on self-specified confidence scores of reviewers and observe that less confident reviewers are more likely to assign high review scores to papers with well known authors and low review scores to papers with less known authors, where reputation is quantified in terms of number of Google Scholar citations. We emphasize upfront that our results are purely correlational and we neither can nor intend to make any causal claims. A blog post accompanying the paper and our scraping code will be linked in the project website https://sites.google.com/view/deanon-arxiv/home

📄 PDF Abstract BibTeX arXiv:2007.00177

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Topics, Authors, and Institutions in Large Language Model Research: Trends from 17K arXiv Papers

2023-07-20 · Rajiv Movva, Sidhika Balachandar, Kenny Peng, Gabriel Agostini 외

Large language models (LLMs) are dramatically influencing AI research, spurring discussions on what has changed so far and how to shape the field's future. To clarify such questions, we analyze a new dataset of 16,979 LL…

Language ModelingLanguage ModellingLarge Language Model

Can AI Put Gamma-Ray Astrophysicists Out of a Job?

2023-03-31 · Samuel T. Spencer, Vikas Joshi, Alison M. W. Mitchell

In what will likely be a litany of generative-model-themed arXiv submissions celebrating April the 1st, we evaluate the capacity of state-of-the-art transformer models to create a paper detailing the detection of a Pulsa…

ER-AE: Differentially Private Text Generation for Authorship Anonymization

2019-07-20 · NAACL 2021 4 · Haohan Bo, Steven H. H. Ding, Benjamin C. M. Fung, Farkhund Iqbal

Most of privacy protection studies for textual data focus on removing explicit sensitive identifiers. However, personal writing style, as a strong indicator of the authorship, is often neglected. Recent studies, such as …

Privacy PreservingText Generation

Protecting Anonymous Speech: A Generative Adversarial Network Methodology for Removing Stylistic Indicators in Text

2021-10-18 · Rishi Balakrishnan, Stephen Sloan, Anil Aswani

With Internet users constantly leaving a trail of text, whether through blogs, emails, or social media posts, the ability to write and protest anonymously is being eroded because artificial intelligence, when given a sam…

Generative Adversarial NetworkSentence

Assessing Deanonymization Risks with Stylometry-Assisted LLM Agent

2026-02-26 · Boyang Zhang, Yang Zhang arxiv

The rapid advancement of large language models (LLMs) has enabled powerful authorship inference capabilities, raising growing concerns about unintended deanonymization risks in textual data such as news articles. In this…