paper-with-me

Papers

A Japanese Benchmark for Evaluating Social Bias in Reasoning Based on Attribution Theory

2026-04-01 · Taihei Shiotani, Masahiro Kaneko, Naoaki Okazaki arxiv

In enhancing the fairness of Large Language Models (LLMs), evaluating social biases rooted in the cultural contexts of specific linguistic regions is essential. However, most existing Japanese benchmarks heavily rely on translating English data, which does not necessarily provide an evaluation suitable for Japanese culture. Furthermore, they only evaluate bias in the conclusion, failing to capture biases lurking in the reasoning. In this study, based on attribution theory in social psychology, we constructed a new dataset, ``JUBAKU-v2,'' which evaluates the bias in attributing behaviors to in-groups and out-groups within reasoning while fixing the conclusion. This dataset consists of 216 examples reflecting cultural biases specific to Japan. Experimental results verified that it can detect performance differences across models more sensitively than existing benchmarks.

📄 PDF Abstract BibTeX arXiv:2604.00568

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Analyzing Social Biases in Japanese Large Language Models

2024-06-04 · Hitomi Yanaka, Namgi Han, Ryoma Kumon, Jie Lu 외

With the development of Large Language Models (LLMs), social biases in the LLMs have become a crucial issue. While various benchmarks for social biases have been provided across languages, the extent to which Japanese LL…

Question Answering

Bias Mitigation or Cultural Commonsense? Evaluating LLMs with a Japanese Dataset

2025-09-29 · Taisei Yamamoto, Ryoma Kumon, Danushka Bollegala, Hitomi Yanaka arxiv

Large language models (LLMs) exhibit social biases, prompting the development of various debiasing methods. However, debiasing methods may degrade the capabilities of LLMs. Previous research has evaluated the impact of b…

JUBAKU: An Adversarial Benchmark for Exposing Culturally Grounded Stereotypes in Japanese LLMs

2026-03-21 · Taihei Shiotani, Masahiro Kaneko, Ayana Niwa, Yuki Maruyama 외 arxiv

Social biases reflected in language are inherently shaped by cultural norms, which vary significantly across regions and lead to diverse manifestations of stereotypes. Existing evaluations of social bias in large languag…

Evaluating the Effect of Retrieval Augmentation on Social Biases

2025-02-24 · Tianhui Zhang, Yi Zhou, Danushka Bollegala

Retrieval Augmented Generation (RAG) has gained popularity as a method for conveniently incorporating novel facts that were not seen during the pre-training stage in Large Language Model (LLM)-based Natural Language Gene…

Large Language ModelQuestion AnsweringRAGRetrieval+2

Can Language Models Handle a Non-Gregorian Calendar? The Case of the Japanese wareki

2025-09-04 · Mutsumi Sasaki, Go Kamoda, Ryosuke Takahashi, Kosuke Sato 외 arxiv

Temporal reasoning and knowledge are essential capabilities for language models (LMs). While much prior work has analyzed and improved temporal reasoning in LMs, most studies have focused solely on the Gregorian calendar…