paper-with-me

홈 › Papers

Projective Latent Interventions for Understanding and Fine-tuning Classifiers

2020-06-23 · Andreas Hinterreiter, Marc Streit, Bernhard Kainz

High-dimensional latent representations learned by neural network classifiers are notoriously hard to interpret. Especially in medical applications, model developers and domain experts desire a better understanding of how these latent representations relate to the resulting classification performance. We present Projective Latent Interventions (PLIs), a technique for retraining classifiers by back-propagating manual changes made to low-dimensional embeddings of the latent space. The back-propagation is based on parametric approximations of t-distributed stochastic neighbourhood embeddings. PLIs allow domain experts to control the latent decision space in an intuitive way in order to better match their expectations. For instance, the performance for specific pairs of classes can be enhanced by manually separating the class clusters in the embedding. We evaluate our technique on a real-world scenario in fetal ultrasound imaging.

📄 PDF Abstract BibTeX arXiv:2006.12902

Code (1)

einbandi/latent-projective-interventions pytorch

Tasks

General Classification

Similar Papers 제목 키워드 기반

Beyond Parameter Finetuning: Test-Time Representation Refinement for Node Classification

2026-01-29 · Jiaxin Zhang, Yiqi Wang, Siwei Wang, Xihong Yang 외 arxiv

Graph Neural Networks frequently exhibit significant performance degradation in the out-of-distribution test scenario. While test-time training (TTT) offers a promising solution, existing Parameter Finetuning (PaFT) para…

Node Classification

IPPRO: Importance-based Pruning with PRojective Offset for Magnitude-indifferent Structural Pruning

2025-07-10 · Jaeheun Jung, Jaehyuk Lee, Yeajin Lee, Donghun Lee arxiv

Importance-based structured pruning overwhelmingly relies on filter magnitude. This proxy is fundamentally flawed: due to scale invariance, functionally identical filters can receive arbitrarily different importance scor…

Neural Network Compression

Learning Structured Twin-Incoherent Twin-Projective Latent Dictionary Pairs for Classification

2019-08-21 · Zhao Zhang, Yulin Sun, Zheng Zhang, Yang Wang 외

In this paper, we extend the popular dictionary pair learning (DPL) into the scenario of twin-projective latent flexible DPL under a structured twin-incoherence. Technically, a novel framework called Twin-Projective Late…

General Classification

Building Comparative Motivation Profiles with Instrumental Interventions

2026-06-06 · David Vella Zarb, Rustem Turtayev, Taywon Min, Jinghua Ou 외 arxiv

Safety evaluations often infer latent motivations from behavioral patterns, but the construct validity of these inferences is unclear. We study this problem in alignment faking, where models comply with training objectiv…

Activation Steering of Video Generation Models via Reduced-Order Linear Optimal Control

2026-06-03 · Jihoon Hong, Alice Chan, Qiyue Dai, Julian Skifstad 외 arxiv

Text-to-video (T2V) models trained on large-scale web data can generate undesired content, motivating interventions that reduce harmful outputs without sacrificing visual quality. Activation steering offers an attractive…

Video Generation