paper-with-me

홈 › Papers

Why Do Self-Harm Prediction Models Struggle to Generalise? Lexical and Semantic Variations in Emergency Department Triage Notes

2026-06-01 · Liuliu Chen, Mike Conway, Jo Robinson, Vlada Rozova arxiv

Self-harm presentations to emergency departments (EDs) are strongly associated with higher suicide risk. NLP models have shown robust performance in detecting self-harm from triage notes within single hospitals, yet performance often declines across institutions. To examine potential causes, we compare ED triage notes from two hospitals by analyzing lexical characteristics, highly associated predictive features, and salient topics. Our results reveal variation in lexical expression and feature importance related to self-harm across hospitals, despite consistent core themes such as self-poisoning and self-injury. These documentation differences are associated with reduced cross-site performance. Our findings provide insight into how institutional variation affects the identification of self-harm in clinical text and highlight potential methods to improve model generalisability.

📄 PDF Abstract BibTeX arXiv:2606.01678

Code (0)

등록된 구현이 없습니다.

Tasks

Feature Importance

Similar Papers 제목 키워드 기반

An Empirical Study of Multi-Generation Sampling for Jailbreak Detection in Large Language Models

2026-04-20 · Hanrui Luo, Shreyank N Gowda arxiv

Detecting jailbreak behaviour in large language models remains challenging, particularly when strongly aligned models produce harmful outputs only rarely. In this work, we present an empirical study of output based jailb…

Towards generalisable hate speech detection: a review on obstacles and solutions

2021-02-17 · Wenjie Yin, Arkaitz Zubiaga

Hate speech is one type of harmful online content which directly attacks or promotes hate towards a group or an individual member based on their actual or perceived aspects of identity, such as ethnicity, religion, and s…

Hate Speech Detection

Just a Scratch: Enhancing LLM Capabilities for Self-harm Detection through Intent Differentiation and Emoji Interpretation

2025-06-05 · Soumitra Ghosh, Gopendra Vikram Singh, Shambhavi, Sabarna Choudhury 외

Self-harm detection on social media is critical for early intervention and mental health support, yet remains challenging due to the subtle, context-dependent nature of such expressions. Identifying self-harm intent aids…

Multi-Task LearningSensitivity

Take and Took, Gaggle and Goose, Book and Read: Evaluating the Utility of Vector Differences for Lexical Relation Learning

2015-09-05 · ACL 2016 8 · Ekaterina Vylomova, Laura Rimell, Trevor Cohn, Timothy Baldwin

Recent work on word embeddings has shown that simple vector subtraction over pre-trained embeddings is surprisingly effective at capturing different lexical relations, despite lacking explicit supervision. Prior work has…

ClusteringRelationWord Embeddings

Harmonization of German Lexical Resources for Opinion Mining

2014-05-01 · LREC 2014 5 · Thierry Declerck, Hans-Ulrich Krieger

We present on-going work on the harmonization of existing German lexical resources in the field of opinion and sentiment mining. The input of our harmonization effort consisted in four distinct lexicons of German word fo…

Opinion MiningSentiment Analysis