paper-with-me

홈 › Papers

Language Models Surface the Unwritten Code of Science and Society

2025-05-25 · Honglin Bao, Siyang Wu, Jiwoong Choi, Yingrong Mao, James A. Evans

This paper calls on the research community not only to investigate how human biases are inherited by large language models (LLMs) but also to explore how these biases in LLMs can be leveraged to make society's "unwritten code" - such as implicit stereotypes and heuristics - visible and accessible for critique. We introduce a conceptual framework through a case study in science: uncovering hidden rules in peer review - the factors that reviewers care about but rarely state explicitly due to normative scientific expectations. The idea of the framework is to push LLMs to speak out their heuristics through generating self-consistent hypotheses - why one paper appeared stronger in reviewer scoring - among paired papers submitted to 45 computer science conferences, while iteratively searching deeper hypotheses from remaining pairs where existing hypotheses cannot explain. We observed that LLMs' normative priors about the internal characteristics of good science extracted from their self-talk, e.g. theoretical rigor, were systematically updated toward posteriors that emphasize storytelling about external connections, such as how the work is positioned and connected within and across literatures. This shift reveals the primacy of scientific myths about intrinsic properties driving scientific excellence rather than extrinsic contextualization and storytelling that influence conceptions of relevance and significance. Human reviewers tend to explicitly reward aspects that moderately align with LLMs' normative priors (correlation = 0.49) but avoid articulating contextualization and storytelling posteriors in their review comments (correlation = -0.14), despite giving implicit reward to them with positive scores. We discuss the broad applicability of the framework, leveraging LLMs as diagnostic tools to surface the tacit codes underlying human society, enabling more precisely targeted responsible AI.

📄 PDF Abstract BibTeX arXiv:2505.18942

Code (0)

등록된 구현이 없습니다.

Tasks

Diagnostic

Methods 이 논문이 사용한 방법론

ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…

Similar Papers 제목 키워드 기반

UWSpeech: Speech to Speech Translation for Unwritten Languages

2020-06-14 · Chen Zhang, Xu Tan, Yi Ren, Tao Qin 외

Existing speech to speech translation systems heavily rely on the text of target language: they usually translate source language either to target text and then synthesize target speech from text, or directly to target s…

speech-recognitionSpeech RecognitionSpeech-to-Speech TranslationTranslation

Acquisition of Translation Lexicons for Historically Unwritten Languages via Bridging Loanwords

2017-06-06 · WS 2017 8 · Michael Bloodgood, Benjamin Strauss

With the advent of informal electronic communications such as social media, colloquial languages that were historically unwritten are being written for the first time in heavily code-switched environments. We present a m…

Machine TranslationTranslation

HAVA: Hybrid Approach to Value-Alignment through Reward Weighing for Reinforcement Learning

2025-05-21 · Kryspin Varys, Federico Cerutti, Adam Sobey, Timothy J. Norman

Our society is governed by a set of norms which together bring about the values we cherish such as safety, fairness or trustworthiness. The goal of value-alignment is to create agents that not only do their tasks but thr…

Fairness

Unwritten Languages Demand Attention Too! Word Discovery with Encoder-Decoder Models

2017-09-17 · Marcely Zanon Boito, Alexandre Berard, Aline Villavicencio, Laurent Besacier

Word discovery is the task of extracting words from unsegmented text. In this paper we examine to what extent neural networks can be applied to this task in a realistic unwritten language scenario, where only small corpo…

DecoderMachine TranslationTranslation

Large-Scale Text Collection for Unwritten Languages

2013-10-01 · IJCNLP 2013 10 · Florian R. Hanke, Steven Bird