paper-with-me

Papers

Towards Algorithmic Fidelity: Mental Health Representation across Demographics in Synthetic vs. Human-generated Data

2024-03-25 · Shinka Mori, Oana Ignat, Andrew Lee, Rada Mihalcea

Synthetic data generation has the potential to impact applications and domains with scarce data. However, before such data is used for sensitive tasks such as mental health, we need an understanding of how different demographics are represented in it. In our paper, we analyze the potential of producing synthetic data using GPT-3 by exploring the various stressors it attributes to different race and gender combinations, to provide insight for future researchers looking into using LLMs for data generation. Using GPT-3, we develop HEADROOM, a synthetic dataset of 3,120 posts about depression-triggering stressors, by controlling for race, gender, and time frame (before and after COVID-19). Using this dataset, we conduct semantic and lexical analyses to (1) identify the predominant stressors for each demographic group; and (2) compare our synthetic data to a human-generated dataset. We present the procedures to generate queries to develop depression data using GPT-3, and conduct analyzes to uncover the types of stressors it assigns to demographic groups, which could be used to test the limitations of LLMs for synthetic data generation for depression data. Our findings show that synthetic data mimics some of the human-generated data distribution for the predominant depression stressors across diverse demographics.

📄 PDF Abstract BibTeX arXiv:2403.16909

Code (1)

michigannlp/depression_synthetic_data 공식 구현

Tasks

Synthetic Data Generation

Methods 이 논문이 사용한 방법론

Refunds@Expedia|||How do I get a full refund from Expedia? “How do I get a full refund from Expedia? How do I get a full refund from Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Quick Help &…
15 Ways to Contact How can i speak to someone at Delta Airlines 설명 없음
Attention 설명 없음
Linear Layer A Linear Layer is a projection $\mathbf{XW + b}$.
Attention Dropout Attention Dropout is a type of dropout used in attention-based architectures, where elements are randomly dropped out of the…
Residual Connection 설명 없음
Cosine Annealing Cosine Annealing is a type of learning rate schedule that has the effect of starting with a large learning rate that is relatively rapidly decreased to a minimum value before…
Multi-Head Attention 설명 없음

Similar Papers 제목 키워드 기반

PsychBench: Auditing Epidemiological Fidelity in Large Language Model Mental Health Simulations

2026-04-19 · Patrick Keough arxiv

Large language models are increasingly deployed to simulate patients for clinical training, research, and mental health tools, yet population-level validity remains largely untested. We introduce PsychBench, the first ep…

Representation Fidelity:Auditing Algorithmic Decisions About Humans Using Self-Descriptions

2026-03-05 · Theresa Elstner, Martin Potthast arxiv

This paper introduces a new dimension for validating algorithmic decisions about humans by measuring the fidelity of their representations. Representation Fidelity measures if decisions about a person rest on reasonable …

LLM Generated Distribution-Based Prediction of US Electoral Results, Part I

2024-11-05 · Caleb Bradshaw, Caelen Miller, Sean Warnick

This paper introduces distribution-based prediction, a novel approach to using Large Language Models (LLMs) as predictive tools by interpreting output token probabilities as distributions representing the models' learned…

Prediction

Representational Ethical Model Calibration

2022-07-25 · Robert Carruthers, Isabel Straw, James K Ruffle, Daniel Herron 외

Equity is widely held to be fundamental to the ethics of healthcare. In the context of clinical decision-making, it rests on the comparative fidelity of the intelligence -- evidence-based or intuitive -- guiding the mana…

Decision MakingDiversityEthicsManagement+1

An Empirical Characterization of Fair Machine Learning For Clinical Risk Prediction

2020-07-20 · Stephen R. Pfohl, Agata Foryciarz, Nigam H. Shah

The use of machine learning to guide clinical decision making has the potential to worsen existing health disparities. Several recent works frame the problem as that of algorithmic fairness, a framework that has attracte…

BIG-bench Machine LearningDecision MakingFairness