Modeling the Sacred: Considerations when Using Religious Texts in Natural Language Processing
This position paper concerns the use of religious texts in Natural Language Processing (NLP), which is of special interest to the Ethics of NLP. Religious texts are expressions of culturally important values, and machine learned models have a propensity to reproduce cultural values encoded in their training data. Furthermore, translations of religious texts are frequently used by NLP researchers when language data is scarce. This repurposes the translations from their original uses and motivations, which often involve attracting new followers. This paper argues that NLP's use of such texts raises considerations that go beyond model biases, including data provenance, cultural contexts, and their use in proselytism. We argue for more consideration of researcher positionality, and of the perspectives of marginalized linguistic and religious communities.
Code (0)
등록된 구현이 없습니다.
Tasks
EthicsPositionSimilar Papers 제목 키워드 기반
Generative AI Alignment with Hinduism's Theological Plurality and Sacred Representation
Generative AI systems are increasingly used to answer personal questions and mediate everyday practices, including religion. However, existing discussions around AI alignment and ethics have largely centered secular, Wes…
What do Asian Religions Have in Common? An Unsupervised Text Analytics Exploration
The main source of various religious teachings is their sacred texts which vary from religion to religion based on different factors like the geographical location or time of the birth of a particular religion. Despite t…
ClusteringDetecting Religious Language in Climate Discourse
Religious language continues to permeate contemporary discourse, even in ostensibly secular domains such as environmental activism and climate change debates. This paper investigates how explicit and implicit forms of re…
Sacred or Secular? Religious Bias in AI-Generated Financial Advice
This study examines religious biases in AI-generated financial advice, focusing on ChatGPT's responses to financial queries. Using a prompt-based methodology and content analysis, we find that 50% of the financial emails…
FrictionSacred or Synthetic? Evaluating LLM Reliability and Abstention for Religious Questions
Despite the increasing usage of Large Language Models (LLMs) in answering questions in a variety of domains, their reliability and accuracy remain unexamined for a plethora of domains including the religious domains. In …