paper-with-me

홈 › Papers

AI evaluation may bias perceptions: The importance of context in interpreting academic writing

2026-05-26 · Shang Wu, Randol Yao arxiv

This paper examines how estimates of AI use in scientific writing can be biased when evaluation methods ignore contextual differences across countries and fields. Using large-scale data on journal publications from Dimensions, we construct AI-likeness benchmarks based on differences between human-written and LLM-rephrased abstracts. We show that a pooled benchmark may confound pre-existing stylistic variation with AI-generated text, producing substantial distortions across country-field groups even in pre-LLM publications. In contrast, country-field-specific benchmarks attenuate such distortions and provide a more credible baseline for comparison. Applying these methods to publications in 2025 reveals that the pooled benchmark systematically overestimates AI use in certain countries and fields while underestimating it in others. These findings highlight the importance of context-aware measurement for accurate and equitable evaluation of AI use in science.

📄 PDF Abstract BibTeX arXiv:2605.26662

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Bridging Human and Model Perspectives: A Comparative Analysis of Political Bias Detection in News Media Using Large Language Models

2025-11-18 · Shreya Adrita Banik, Niaz Nafi Rahman, Tahsina Moiukh, Farig Sadeque arxiv

Detecting political bias in news media is a complex task that requires interpreting subtle linguistic and contextual cues. Although recent advances in Natural Language Processing (NLP) have enabled automatic bias classif…

Bias Detection

In-Depth Analysis of Emotion Recognition through Knowledge-Based Large Language Models

2024-07-17 · Bin Han, Cleo Yau, Su Lei, Jonathan Gratch

Emotion recognition in social situations is a complex task that requires integrating information from both facial expressions and the situational context. While traditional approaches to automatic emotion recognition hav…

Emotion Recognition

Interpreting Text Classifiers by Learning Context-sensitive Influence of Words

2021-06-01 · NAACL (TrustNLP) 2021 6 · Sawan Kumar, Kalpit Dixit, Kashif Shah

Many existing approaches for interpreting text classification models focus on providing importance scores for parts of the input text, such as words, but without a way to test or improve the interpretation method itself.…

Sentiment Analysistext-classificationText Classification

Public Perceptions of Gender Bias in Large Language Models: Cases of ChatGPT and Ernie

2023-09-17 · Kyrie Zhixuan Zhou, Madelyn Rose Sanfilippo

Large language models are quickly gaining momentum, yet are found to demonstrate gender bias in their responses. In this paper, we conducted a content analysis of social media discussions to gauge public perceptions of g…

Algorithmic Inheritance: Surname Bias in AI Decisions Reinforces Intergenerational Inequality

2025-01-23 · Pat Pataranutaporn, Nattavudh Powdthavee, Pattie Maes

Surnames often convey implicit markers of social status, wealth, and lineage, shaping perceptions in ways that can perpetuate systemic biases and intergenerational inequality. This study is the first of its kind to inves…

Decision MakingFairness