paper-with-me

홈 › Papers

Householder Pseudo-Rotation: A Novel Approach to Activation Editing in LLMs with Direction-Magnitude Perspective

2024-09-16 · Van-Cuong Pham, Thien Huu Nguyen

Activation Editing, which involves directly editting the internal representations of large language models (LLMs) to alter their behaviors and achieve desired properties, has emerged as a promising area of research. Existing works primarily treat LLMs' activations as points in space and modify them by adding steering vectors. However, this approach is limited in its ability to achieve greater performance improvement while maintaining the necessary consistency of activation magnitudes. To overcome these issues, we propose a novel editing method that views activations in terms of their directions and magnitudes. Our method, named Householder Pseudo-Rotation (HPR), mimics the rotation transformation, thus preserving activation norms and resulting in an improved performance on various safety benchmarks.

📄 PDF Abstract BibTeX arXiv:2409.10053

Code (1)

VinAIResearch/HPR 공식 구현

Similar Papers 제목 키워드 기반

RUQuant: Towards Refining Uniform Quantization for Large Language Models

2026-04-05 · Han Liu, Haotian Gao, Changya Li, Feng Zhang 외 arxiv

The increasing size and complexity of large language models (LLMs) have raised significant challenges in deployment efficiency, particularly under resource constraints. Post-training quantization (PTQ) has emerged as a p…

HousE: Knowledge Graph Embedding with Householder Parameterization

2022-02-16 · Rui Li, Jianan Zhao, Chaozhuo Li, Di He 외

The effectiveness of knowledge graph embedding (KGE) largely depends on the ability to model intrinsic relation patterns and mapping properties. However, existing approaches can only capture some of them with insufficien…

Graph EmbeddingKnowledge Graph EmbeddingRelationRelation Mapping

Rotation Invariant Householder Parameterization for Bayesian PCA

2019-05-12 · Rajbir S. Nirwan, Nils Bertschinger

We consider probabilistic PCA and related factor models from a Bayesian perspective. These models are in general not identifiable as the likelihood has a rotational symmetry. This gives rise to complicated posterior dist…

Probabilistic Programming

Novel Quadratic Constraints for Extending LipSDP beyond Slope-Restricted Activations

2024-01-25 · Patricia Pauli, Aaron Havens, Alexandre Araujo, Siddharth Garg 외

Recently, semidefinite programming (SDP) techniques have shown great promise in providing accurate Lipschitz bounds for neural networks. Specifically, the LipSDP approach (Fazlyab et al., 2019) has received much attentio…

Why Do Accumulated Transformations Extrapolate?

2026-06-23 · Mahesh Godavarti arxiv

PaTH Attention showed that replacing RoPE's position-indexed rotations with accumulated data-dependent Householder reflections yields strong length extrapolation, though performance degrades at extreme context lengths. W…