paper-with-me

홈 › Papers

Correctness-Optimized Residual Activation Lens (CORAL): Transferrable and Calibration-Aware Inference-Time Steering

2026-02-05 · Miranda Muqing Miao, Young-Min Cho, Lyle Ungar arxiv

Large language models (LLMs) exhibit persistent miscalibration, especially after instruction tuning and preference alignment. Modified training objectives can improve calibration, but retraining is expensive. Inference-time steering offers a lightweight alternative, yet most existing methods optimize proxies for correctness rather than correctness itself. We introduce CORAL (Correctness-Optimized Residual Activation Lens), a regularized inference-time steering method that captures distributed correctness signals from model internal activations using weight-decay MLP probes. We evaluate CORAL across three 7B-parameter models and find that it consistently improves accuracy by 10\% and expected calibration error (ECE) by 50\% on average. We additionally demonstrate that these gains transfer without retraining to the complete published test sets of four held-out benchmarks (ARC-Challenge, HellaSwag, Math-MC, OpenBookQA), averaging 14\% accuracy improvements and 49\% ECE improvements. Our results support the hypothesis that distributed information in model internals can be extracted using regularized probes when individual neurons are insufficient. CORAL thus provides a compute-efficient, transferable, and calibration-aware approach to improve MCQA performance during inference.

📄 PDF Abstract BibTeX arXiv:2602.06022

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Deep Learning Models for Coral Bleaching Classification in Multi-Condition Underwater Image Datasets

2025-10-24 · Julio Jerison E. Macrohon, Gordon Hung arxiv

Coral reefs support numerous marine organisms and are an important source of coastal protection from storms and floods, representing a major part of marine ecosystems. However coral reefs face increasing threats from pol…

CoralStyleCLIP: Co-optimized Region and Layer Selection for Image Editing

2023-03-09 · CVPR 2023 1 · Ambareesh Revanur, Debraj Basu, Shradha Agrawal, Dhwanit Agarwal 외

Edit fidelity is a significant issue in open-world controllable generative image editing. Recently, CLIP-based approaches have traded off simplicity to alleviate these problems by introducing spatial attention in a handp…

Correlation Alignment for Unsupervised Domain Adaptation

2016-12-06 · Baochen Sun, Jiashi Feng, Kate Saenko

In this chapter, we present CORrelation ALignment (CORAL), a simple yet effective method for unsupervised domain adaptation. CORAL minimizes domain shift by aligning the second-order statistics of source and target distr…

Domain AdaptationUnsupervised Domain Adaptation

Nanopore Base Calling on the Edge

2020-11-09 · Peter Perešíni, Vladimír Boža, Broňa Brejová, Tomáš Vinař

We developed a new base caller DeepNano-coral for nanopore sequencing, which is optimized to run on the Coral Edge Tensor Processing Unit, a small USB-attached hardware accelerator. To achieve this goal, we have designed…

speech-recognitionSpeech Recognition

Deep CORAL: Correlation Alignment for Deep Domain Adaptation

2016-07-06 · Baochen Sun, Kate Saenko

Deep neural networks are able to learn powerful representations from large quantities of labeled input data, however they cannot always generalize well across changes in input distributions. Domain adaptation algorithms …

Domain AdaptationDomain GeneralizationImage ClassificationUnsupervised Domain Adaptation