paper-with-me

Papers

Microsaccade-Inspired Probing: Positional Encoding Perturbations Reveal LLM Misbehaviours

2025-10-01 · Rui Melo, Rui Abreu, Corina S. Pasareanu arxiv

We draw inspiration from microsaccades, tiny involuntary eye movements that reveal hidden dynamics of human perception, to propose an analogous probing method for large language models (LLMs). Just as microsaccades expose subtle but informative shifts in vision, we show that lightweight position encoding perturbations elicit latent signals that indicate model misbehaviour. Our method requires no fine-tuning or task-specific supervision, yet detects failures across diverse settings including factuality, safety, toxicity, and backdoor attacks. Experiments on multiple state-of-the-art LLMs demonstrate that these perturbation-based probes surface misbehaviours while remaining computationally efficient. These findings suggest that pretrained LLMs already encode the internal evidence needed to flag their own failures, and that microsaccade-inspired interventions provide a pathway for detecting and mitigating undesirable behaviours.

📄 PDF Abstract BibTeX arXiv:2510.01288

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Microsaccade-inspired Event Camera for Robotics

2024-05-28 · Botao He, Ze Wang, Yuan Zhou, Jingxi Chen 외

Neuromorphic vision sensors or event cameras have made the visual perception of extremely low reaction time possible, opening new avenues for high-dynamic robotics applications. These event cameras' output is dependent o…

Artificial Microsaccade Compensation: Stable Vision for an Ornithopter

2025-12-03 · Levi Burner, Guido de Croon, Yiannis Aloimonos arxiv

Animals with foveated vision, including humans, experience microsaccades, small, rapid eye movements that they are not aware of. Inspired by this phenomenon, we develop a method for "Artificial Microsaccade Compensation"…

What DINO saw: ALiBi positional encoding reduces positional bias in Vision Transformers

2026-03-17 · Moritz Pawlowsky, Antonis Vamvakeros, Alexander Weiss, Anja Bielefeld 외 arxiv

Vision transformers (ViTs) - especially feature foundation models like DINOv2 - learn rich representations useful for many downstream tasks. However, architectural choices (such as positional encoding) can lead to these …

Benchmarking Microsaccade Recognition with Event Cameras: A Novel Dataset and Evaluation

2025-10-28 · Waseem Shariff, Timothy Hanley, Maciej Stec, Hossein Javidnia 외 arxiv

Microsaccades are small, involuntary eye movements vital for visual perception and neural processing. Traditional microsaccade studies typically use eye trackers or frame-based analysis, which, while precise, are costly …

Event-based vision

Transformer Language Models without Positional Encodings Still Learn Positional Information

2022-03-30 · Adi Haviv, Ori Ram, Ofir Press, Peter Izsak 외

Causal transformer language models (LMs), such as GPT-3, typically require some form of positional encoding, such as positional embeddings. However, we show that LMs without any explicit positional encoding are still com…

Position