paper-with-me

홈 › Papers

Neuron-Aware Active Few-Shot Learning for LLMs

2026-07-02 · Zhuowei Chen, Liwei Chen, Christian Schunn, Raquel Coelho, Xiang Lorraine Li arxiv

Active Few-Shot Learning (AFSL) adapts LLMs to specialized domains by identifying the most valuable unlabeled samples for annotation and use as few-shot demonstrations, effectively reducing human annotation costs while promoting high performance. However, existing methods typically rely on output-level signals for sample identification, such as predictive entropy or semantic similarities with test-time data based on external embeddings, which often overlook models' internal dynamics, which could pinpoint specific knowledge gaps. To bridge this gap, we propose NeuFS, a Neuron-Aware Active Few-Shot Learning framework that shifts the selection paradigm from output-level proxies to models' internal dynamics. NeuFS utilizes neuron activation patterns to represent sample directly, and includes a dual-criteria selection strategy that: (1) ensures few-shot sample diversity with neuron patterns for broader example coverage, while (2) prioritizing on identifying informative and challenging few-shot samples LLMs tend to hallucinate by quantifying neuron consensus. Experiments on three datasets demonstrate that NeuFS excels in both reasoning and text classification tasks, outperforming existing AFSL baselines. Ablation studies further highlight that internal neuron activations provide a more principled and effective selection signal than external embeddings, validating the superiority of the proposed NeuFS.

📄 PDF Abstract BibTeX arXiv:2607.02423

Code (0)

등록된 구현이 없습니다.

Tasks

Text ClassificationFew-Shot Learning

Similar Papers 제목 키워드 기반

StrucSum: Graph-Structured Reasoning for Long Document Extractive Summarization with LLMs

2025-05-29 · Haohan Yuan, Sukhwa Hong, Haopeng Zhang

Large language models (LLMs) have shown strong performance in zero-shot summarization, but often struggle to model document structure and identify salient information in long texts. In this work, we introduce StrucSum, a…

Extractive SummarizationSentence

Improving Generalization in LLM Structured Pruning via Function-Aware Neuron Grouping

2025-12-28 · Tao Yu, Yongqi An, Kuan Zhu, Guibo Zhu 외 arxiv

Large Language Models (LLMs) demonstrate impressive performance across natural language tasks but incur substantial computational and storage costs due to their scale. Post-training structured pruning offers an efficient…

Investigating Neurons and Heads in Transformer-based LLMs for Typographical Errors

2025-02-27 · Kohei Tsuji, Tatsuya Hiraoka, Yuchang Cheng, Eiji Aramaki 외

This paper investigates how LLMs encode inputs with typos. We hypothesize that specific neurons and attention heads recognize typos and fix them internally using local and global contexts. We introduce a method to identi…

Modality-Aware Neuron Pruning for Unlearning in Multimodal Large Language Models

2025-02-21 · Zheyuan Liu, Guangyao Dou, Xiangchi Yuan, Chunhui Zhang 외

Generative models such as Large Language Models (LLMs) and Multimodal Large Language Models (MLLMs) trained on massive datasets can lead them to memorize and inadvertently reveal sensitive information, raising ethical an…

NeuPAT: Neuron-aware Plasticity Allocation Tuning for Language-Preserving MLLMs

2026-08-08 · Jiayue Jin, Jingwei Zhang, Chen Wang, Jing Liu 외 hf

Multimodal expansion of large language models (LLMs) enables new perceptual capabilities but often compromises the language intelligence acquired during pretraining. In this work, we investigate this phenomenon from the …