paper-with-me

홈 › Papers

Stimulative Training of Residual Networks: A Social Psychology Perspective of Loafing

2022-10-09 · Peng Ye, Shengji Tang, Baopu Li, Tao Chen, Wanli Ouyang

Residual networks have shown great success and become indispensable in today's deep models. In this work, we aim to re-investigate the training process of residual networks from a novel social psychology perspective of loafing, and further propose a new training strategy to strengthen the performance of residual networks. As residual networks can be viewed as ensembles of relatively shallow networks (i.e., \textit{unraveled view}) in prior works, we also start from such view and consider that the final performance of a residual network is co-determined by a group of sub-networks. Inspired by the social loafing problem of social psychology, we find that residual networks invariably suffer from similar problem, where sub-networks in a residual network are prone to exert less effort when working as part of the group compared to working alone. We define this previously overlooked problem as \textit{network loafing}. As social loafing will ultimately cause the low individual productivity and the reduced overall performance, network loafing will also hinder the performance of a given residual network and its sub-networks. Referring to the solutions of social psychology, we propose \textit{stimulative training}, which randomly samples a residual sub-network and calculates the KL-divergence loss between the sampled sub-network and the given residual network, to act as extra supervision for sub-networks and make the overall goal consistent. Comprehensive empirical results and theoretical analyses verify that stimulative training can well handle the loafing problem, and improve the performance of a residual network by improving the performance of its sub-networks. The code is available at https://github.com/Sunshine-Ye/NIPS22-ST .

📄 PDF Abstract BibTeX arXiv:2210.04153

Code (1)

sunshine-ye/nips22-st 공식 구현 pytorch

Similar Papers 제목 키워드 기반

Stimulative Training++: Go Beyond The Performance Limits of Residual Networks

2023-05-04 · Peng Ye, Tong He, Shengji Tang, Baopu Li 외

Residual networks have shown great success and become indispensable in recent deep neural network models. In this work, we aim to re-investigate the training process of residual networks from a novel social psychology pe…

Boosting Residual Networks with Group Knowledge

2023-08-26 · Shengji Tang, Peng Ye, Baopu Li, Weihao Lin 외

Recent research understands the residual networks from a new perspective of the implicit ensemble model. From this view, previous methods such as stochastic depth and stimulative training have further improved the perfor…

Knowledge Distillation

Social Skill Training with Large Language Models

2024-04-05 · Diyi Yang, Caleb Ziems, William Held, Omar Shaikh 외

People rely on social skills like conflict resolution to communicate effectively and to thrive in both work and personal life. However, practice environments for social skills are typically out of reach for most people. …

Survey and Perspective on Social Emotions in Robotics

2021-05-20 · Chie Hieida, Takayuki Nagai

This study reviews research on social emotions in robotics. In robotics, the study of emotions has been pursued for a long time, including the study of their recognition, expression, and computational modeling of the bas…

Survey

Examining psychology of science as a potential contributor to science policy

2023-09-17 · Arash Mousavi, Reza Hafezi, Hasan Ahmadi

The psychology of science is the least developed member of the family of science studies. It is growing, however, increasingly into a promising discipline. After a very brief review of this emerging sub-field of psycholo…