paper-with-me

Papers

Getting aligned on representational alignment

2023-10-18 · Ilia Sucholutsky, Lukas Muttenthaler, Adrian Weller, Andi Peng, Andreea Bobu, Been Kim, Bradley C. Love, Christopher J. Cueva, Erin Grant, Iris Groen, Jascha Achterberg, Joshua B. Tenenbaum, Katherine M. Collins, Katherine L. Hermann, Kerem Oktar, Klaus Greff, Martin N. Hebart, Nathan Cloos, Nikolaus Kriegeskorte, Nori Jacoby, Qiuyi Zhang, Raja Marjieh, Robert Geirhos, Sherol Chen, Simon Kornblith, Sunayana Rane, Talia Konkle, Thomas P. O'Connell, Thomas Unterthiner, Andrew K. Lampinen, Klaus-Robert Müller, Mariya Toneva, Thomas L. Griffiths

Biological and artificial information processing systems form representations of the world that they can use to categorize, reason, plan, navigate, and make decisions. How can we measure the similarity between the representations formed by these diverse systems? Do similarities in representations then translate into similar behavior? If so, then how can a system's representations be modified to better match those of another system? These questions pertaining to the study of representational alignment are at the heart of some of the most promising research areas in contemporary cognitive science, neuroscience, and machine learning. In this Perspective, we survey the exciting recent developments in representational alignment research in the fields of cognitive science, neuroscience, and machine learning. Despite their overlapping interests, there is limited knowledge transfer between these fields, so work in one field ends up duplicated in another, and useful innovations are not shared effectively. To improve communication, we propose a unifying framework that can serve as a common language for research on representational alignment, and map several streams of existing work across fields within our framework. We also lay out open problems in representational alignment where progress can benefit all three of these fields. We hope that this paper will catalyze cross-disciplinary collaboration and accelerate progress for all communities studying and developing information processing systems.

📄 PDF Abstract BibTeX arXiv:2310.13018

Code (1)

qglht/repal pytorch

Tasks

NavigateTransfer Learning

Methods 이 논문이 사용한 방법론

ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…

Similar Papers 제목 키워드 기반

Diagnosing Catastrophe: Large parts of accuracy loss in continual learning can be accounted for by readout misalignment

2023-10-09 · Daniel Anthes, Sushrut Thorat, Peter König, Tim C. Kietzmann

Unlike primates, training artificial neural networks on changing data distributions leads to a rapid decrease in performance on old tasks. This phenomenon is commonly referred to as catastrophic forgetting. In this paper…

Continual Learning

STAER: Temporal Aligned Rehearsal for Continual Spiking Neural Network

2026-01-16 · Matteo Gianferrari, Omayma Moussadek, Riccardo Salami, Cosimo Fiorini 외 arxiv

Spiking Neural Networks (SNNs) are inherently suited for continuous learning due to their event-driven temporal dynamics; however, their application to Class-Incremental Learning (CIL) has been hindered by catastrophic f…

class-incremental learningRepresentation Learning

Visualizing Representational Dynamics with Multidimensional Scaling Alignment

2019-06-21 · Baihan Lin, Marieke Mur, Tim Kietzmann, Nikolaus Kriegeskorte

Representational similarity analysis (RSA) has been shown to be an effective framework to characterize brain-activity profiles and deep neural network activations as representational geometry by computing the pairwise di…

Object Categorization

Diagnosing the Performance Trade-off in Moral Alignment: A Case Study on Gender Stereotypes

2025-09-25 · Guangliang Liu, Bocheng Chen, Han Zi, Xitong Zhang 외 arxiv

Moral alignment has emerged as a widely adopted approach for regulating the behavior of pretrained language models (PLMs), typically through fine-tuning on curated datasets. Gender stereotype mitigation is a representati…

Towards Aligned Data Forgetting via Twin Machine Unlearning

2025-01-15 · Zhenxing Niu, Haoxuan Ji, Yuyao Sun, Zheng Lin 외

Modern privacy regulations have spurred the evolution of machine unlearning, a technique enabling a trained model to efficiently forget specific training data. In prior unlearning methods, the concept of "data forgetting…

Machine Unlearning