Beyond Preserved Accuracy: Evaluating Loyalty and Robustness of BERT Compression
Recent studies on compression of pretrained language models (e.g., BERT) usually use preserved accuracy as the metric for evaluation. In this paper, we propose two new metrics, label loyalty and probability loyalty that measure how closely a compressed model (i.e., student) mimics the original model (i.e., teacher). We also explore the effect of compression with regard to robustness under adversarial attacks. We benchmark quantization, pruning, knowledge distillation and progressive module replacing with loyalty and robustness. By combining multiple compression techniques, we provide a practical strategy to achieve better accuracy, loyalty and robustness.
Code (1)
Tasks
Knowledge DistillationQuantizationMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Attitudinal Loyalty Manifestation in Banking CSR: Cross-Buying Behavior and Customer Advocacy
This study in the banking industry examines the influence of attitudinal loyalty on customer advocacy and cross buying behavior, alongside the moderating roles of Quality of Life and Corporate Social Responsibility suppo…
Price Discrimination in the Presence of Customer Loyalty and Differing Firm Costs
We study how loyalty behavior of customers and differing costs to produce undifferentiated products by firms can influence market outcomes. In prior works that study such markets, firm costs have generally been assumed n…
Loyalty in Online Communities
Loyalty is an essential component of multi-community engagement. When users have the choice to engage with a variety of different communities, they often become loyal to just one, focusing on that community at the expens…
The Green Advantage: Analyzing the Effects of Eco-Friendly Marketing on Consumer Loyalty
The idea that marketing, in addition to profitability and sales, should also consider the consumer's health is not and has not been a far-fetched concept. It can be stated that there is no longer a way back to producing …
MarketingPADBen: A Comprehensive Benchmark for Evaluating AI Text Detectors Against Paraphrase Attacks
While AI-generated text (AIGT) detectors achieve over 90\% accuracy on direct LLM outputs, they fail catastrophically against iteratively-paraphrased content. We investigate why iteratively-paraphrased text -- itself AI-…