paper-with-me

Papers

Geometric Deviation as an Unsupervised Pre-Generation Reliability Signal: Probing LLM Representations for Answerability

2026-05-04 · Yucheng Du arxiv

A reliable language model should be able to signal, prior to generation, when a query falls outside its knowledge. We investigate whether representation geometry can provide such a pre-generation signal by measuring the deviation of hidden states from an answerable reference set, requiring no labeled failure data and no access to model outputs. Across three instruction-tuned models (Llama 3.1-8B, Qwen 2.5-7B, and Mistral-7B-Instruct) and three prompt forms (Math, Fact, Code), we find that geometry primarily encodes task form. Within mathematical prompts, unanswerable inputs consistently deviate from the answerable centroid, yielding strong separation (ROC-AUC 0.78-0.84). This single-pass pre-generation signal outperforms a simple refusal baseline and compares favorably to self-consistency. It also captures cases where models do not explicitly refuse. In contrast, no reliable geometric signal emerges for factual prompts, indicating that the effect is form-conditional rather than universal. Code prompts show large effect sizes with higher variance, suggesting partial generalization beyond mathematical form. A layer-wise analysis reveals that the signal arises in early layers and gradually attenuates toward the output. These results suggest that answerability-related geometry is established before the final stages of generation. Together, these findings indicate that geometric deviation can serve as a lightweight pre-generation signal that is reliable in structured domains with formal answerability constraints, with clear boundaries on where it generalizes.

📄 PDF Abstract BibTeX arXiv:2605.03196

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Reliability-Aware Prototype Calibration for Frozen Pose-Flow Video Anomaly Detection

2026-06-18 · Ning Dong, Yingna Su, Xin Dong, Ziyun Jiao 외 arxiv

Pose-flow video anomaly detectors are attractive for one-class surveillance because they provide likelihood-based rankings for tracked skeleton windows. However, a single likelihood score may hide multimodal normal behav…

Video Anomaly Detection

Integrating Local and Global Entropy for Uncertainty Quantification in LLMs

2026-06-02 · Johanne Medina, Tianyi Zhou, Keivin Isufaj, Aristides Gionis 외 arxiv

Large language models hallucinate confidently, making uncertainty quantification (UQ) essential for reliable deployment. Existing methods rely predominantly on token-level signals, leaving the geometric structure of inte…

REMA: A Unified Reasoning Manifold Framework for Interpreting Large Language Model

2025-09-26 · Bo Li, Guanzhi Deng, Ronghao Chen, Junrong Yue 외 arxiv

Understanding how Large Language Models (LLMs) perform complex reasoning and their failure mechanisms is a challenge in interpretability research. To provide a measurable geometric analysis perspective, we define the con…

GeoAvatar: Adaptive Geometrical Gaussian Splatting for 3D Head Avatar

2025-07-24 · SeungJun Moon, Hah Min Lew, Seungeun Lee, Ji-Su Kang 외 arxiv

Despite recent progress in 3D head avatar generation, balancing identity preservation, i.e., reconstruction, with novel poses and expressions, i.e., animation, remains a challenge. Existing methods struggle to adapt Gaus…

Discriminately Treating Motion Components Evolves Joint Depth and Ego-Motion Learning

2025-11-03 · Mengtan Zhang, Zizhan Guo, Hongbo Zhao, Yi Feng 외 arxiv

Unsupervised learning of depth and ego-motion, two fundamental 3D perception tasks, has made significant strides in recent years. However, most methods treat ego-motion as an auxiliary task, either mixing all motion type…