paper-with-me

홈 › Papers

Decoding Demographic un-fairness from Indian Names

2022-09-07 · Medidoddi Vahini, Jalend Bantupalli, Souvic Chakraborty, Animesh Mukherjee

Demographic classification is essential in fairness assessment in recommender systems or in measuring unintended bias in online networks and voting systems. Important fields like education and politics, which often lay a foundation for the future of equality in society, need scrutiny to design policies that can better foster equality in resource distribution constrained by the unbalanced demographic distribution of people in the country. We collect three publicly available datasets to train state-of-the-art classifiers in the domain of gender and caste classification. We train the models in the Indian context, where the same name can have different styling conventions (Jolly Abraham/Kumar Abhishikta in one state may be written as Abraham Jolly/Abishikta Kumar in the other). Finally, we also perform cross-testing (training and testing on different datasets) to understand the efficacy of the above models. We also perform an error analysis of the prediction models. Finally, we attempt to assess the bias in the existing Indian system as case studies and find some intriguing patterns manifesting in the complex demographic layout of the sub-continent across the dimensions of gender and caste.

📄 PDF Abstract BibTeX arXiv:2209.03089

Code (1)

vahini01/indiandemographics 공식 구현

Tasks

FairnessRecommendation Systems

Similar Papers 제목 키워드 기반

SamaVaani: Auditing and Debiasing Multilingual Clinical ASR for Indian Languages

2026-06-25 · Subham Kumar, Prakrithi Shivaprakash, Abhishek Manoharan, Astut Kurariya 외 arxiv

Automatic Speech Recognition (ASR) is increasingly used to document clinical encounters, yet its reliability in multilingual and demographically diverse Indian healthcare context remains largely unknown. In this study, w…

Speech Recognition

Are Models Trained on Indian Legal Data Fair?

2023-03-13 · Sahil Girhepuje, Anmol Goel, Gokul S Krishnan, Shreya Goyal 외

Recent advances and applications of language technology and artificial intelligence have enabled much success across multiple domains like law, medical and mental health. AI-based Language Models, like Judgement Predicti…

FairnessPrediction

ASR Under the Stethoscope: Evaluating Biases in Clinical Speech Recognition across Indian Languages

2025-11-30 · Subham Kumar, Prakrithi Shivaprakash, Abhishek Manoharan, Astut Kurariya 외 arxiv

Automatic Speech Recognition (ASR) is increasingly used to document clinical encounters, yet its reliability in multilingual and demographically diverse Indian healthcare contexts remains largely unknown. In this study, …

Speech Recognition

An Analysis of the Effects of Decoding Algorithms on Fairness in Open-Ended Language Generation

2022-10-07 · Jwala Dhamala, Varun Kumar, Rahul Gupta, Kai-Wei Chang 외

Several prior works have shown that language models (LMs) can generate text containing harmful social biases and stereotypes. While decoding algorithms play a central role in determining properties of LM generated text, …

DiversityFairnessText Generation

Linear socio-demographic representations emerge in Large Language Models from indirect cues

2025-12-10 · Paul Bouchaud, Pedro Ramaciotti arxiv

We investigate how LLMs encode sociodemographic attributes of human conversational partners inferred from indirect cues such as names and occupations. We show that LLMs develop linear representations of user demographics…