paper-with-me

Patching

Activation Patching

2000년 도입 · 논문 102편에서 사용

Activation patching studies the model's computation by altering its latent representations, the token embeddings in transformer-based language models, during the inference process

출처: Patchscopes: A Unifying Framework for Inspecting Hidden Representations of Language Models

소개 논문: Patchscopes: A Unifying Framework for Inspecting Hidden Representations of Language Models

Inference Extrapolation · Natural Language Processing