paper-with-me

홈 › Papers

Towards a unified and verified understanding of group-operation networks

2024-10-09 · Wilson Wu, Louis Jaburi, Jacob Drori, Jason Gross

A recent line of work in mechanistic interpretability has focused on reverse-engineering the computation performed by neural networks trained on the binary operation of finite groups. We investigate the internals of one-hidden-layer neural networks trained on this task, revealing previously unidentified structure and producing a more complete description of such models in a step towards unifying the explanations of previous works (Chughtai et al., 2023; Stander et al., 2024). Notably, these models approximate equivariance in each input argument. We verify that our explanation applies to a large fraction of networks trained on this task by translating it into a compact proof of model performance, a quantitative evaluation of the extent to which we faithfully and concisely explain model internals. In the main text, we focus on the symmetric group S5. For models trained on this group, our explanation yields a guarantee of model accuracy that runs 3x faster than brute force and gives a >=95% accuracy bound for 45% of the models we trained. We were unable to obtain nontrivial non-vacuous accuracy bounds using only explanations from previous works.

📄 PDF Abstract BibTeX arXiv:2410.07476

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

ReasonOps: A Unified Operational Paradigm for Trustworthy Verified LLM Reasoning

2026-05-26 · Adnan Rashid arxiv

Large Language Models (LLMs) have transformed artificial intelligence from primarily generative systems into increasingly capable reasoning agents. Recent advances in theorem proving, autoformalization, symbolic reasonin…

Status hierarchy and group cooperation: A generalized model

2020-08-17

In a refreshing mathematical investigation, Mark (2018) shows that status hierarchy may facilitate the emergence of cooperation in groups. Despite the contribution, the present paper notes that there are limitations in M…

Crab: A Unified Audio-Visual Scene Understanding Model with Explicit Cooperation

2025-03-17 · CVPR 2025 1 · Henghui Du, Guangyao Li, Chang Zhou, Chunjie Zhang 외

In recent years, numerous tasks have been proposed to encourage model to develop specified capability in understanding audio-visual scene, primarily categorized into temporal localization, spatial localization, spatio-te…

Data InteractionScene UnderstandingTemporal LocalizationUIE

Differentiable Learning-to-Group Channels via Groupable Convolutional Neural Networks

2019-08-16 · ICCV 2019 10 · Zhaoyang Zhang, Jingyu Li, Wenqi Shao, Zhanglin Peng 외

Group convolution, which divides the channels of ConvNets into groups, has achieved impressive improvement over the regular convolution operation. However, existing models, eg. ResNeXt, still suffers from the sub-optimal…

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation

2026-05-26 · Jianzhe Gao, Rui Liu, Wenguan Wang arxiv

Vision-language navigation (VLN) requires an agent to traverse complex 3D environments based on natural language instructions, necessitating a thorough scene understanding. While existing works equip agents with various …

Vision-Language NavigationScene UnderstandingPoint Clouds