Matrix-Driven Identification and Reconstruction of LLM Weight Homology
We propose Matrix-Driven Identification and Reconstruction (MDIR), a SOTA large language model homology method that accurately detects weight correspondences between models and provides rigorous $p$-value estimation of the statistical significance of these correspondences. Our method does not require model inference, and allows the detection of unattributed reuse or replication of model weights even on low-resource devices as it compares only a single pair of matrices at a time. We leverage matrix analysis, polar decomposition, and Large Deviation Theory (LDT) to achieve accurate reconstruction of weight relationships between models. Notably, MDIR is the first method to achieve perfect scores on both Area-Under-Curve (AUC) and accuracy metrics across different source models on LeaFBench.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
Topological Invariant-Based Iris Identification via Digital Homology and Machine Learning
Objective - This study presents a biometric identification method based on topological invariants from 2D iris images, representing iris texture via formally defined digital homology and evaluating classification perform…
A Topology Layer for Machine Learning
Topology applied to real world data using persistent homology has started to find applications within machine learning, including deep learning. We present a differentiable topology layer that computes persistent homolog…
BIG-bench Machine LearningDeep LearningAnalyzing Brain Tumor Connectomics using Graphs and Persistent Homology
Recent advances in molecular and genetic research have identified a diverse range of brain tumor sub-types, shedding light on differences in their molecular mechanisms, heterogeneity, and origins. The present study perfo…
Topological Data AnalysisCubical Ripser: Software for computing persistent homology of image and volume data
We introduce Cubical Ripser for computing persistent homology of image and volume data (more precisely, weighted cubical complexes). To our best knowledge, Cubical Ripser is currently the fastest and the most memory-effi…
Selecting Interpretable Circular Coordinates from Data
Circular coordinates obtained from persistent cohomology reveal loop structure in data, but they usually remain abstract: A detected circle does not tell us which measured angle, phase, torsion, or decoder explains it. W…
Point Clouds