paper-with-me

홈 › Papers

Why GANs are overkill for NLP

2022-05-19 · David Alvarez-Melis, Vikas Garg, Adam Tauman Kalai

This work offers a novel theoretical perspective on why, despite numerous attempts, adversarial approaches to generative modeling (e.g., GANs) have not been as popular for certain generation tasks, particularly sequential tasks such as Natural Language Generation, as they have in others, such as Computer Vision. In particular, on sequential data such as text, maximum-likelihood approaches are significantly more utilized than GANs. We show that, while it may seem that maximizing likelihood is inherently different than minimizing distinguishability, this distinction is largely artificial and only holds for limited models. We argue that minimizing KL-divergence (i.e., maximizing likelihood) is a more efficient approach to effectively minimizing the same distinguishability criteria that adversarial models seek to optimize. Reductions show that minimizing distinguishability can be seen as simply boosting likelihood for certain families of models including n-gram models and neural networks with a softmax output layer. To achieve a full polynomial-time reduction, a novel next-token distinguishability model is considered.

📄 PDF Abstract BibTeX arXiv:2205.09838

Code (0)

등록된 구현이 없습니다.

Tasks

Text Generation

Methods 이 논문이 사용한 방법론

Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…

Similar Papers 제목 키워드 기반

Navigating the OverKill in Large Language Models

2024-01-31 · Chenyu Shi, Xiao Wang, Qiming Ge, Songyang Gao 외

Large language models are meticulously aligned to be both helpful and harmless. However, recent research points to a potential overkill which means models may refuse to answer benign queries. In this paper, we investigat…

Gaining Free or Low-Cost Transparency with Interpretable Partial Substitute

2018-02-12 · Tong Wang

This work addresses the situation where a black-box model with good predictive performance is chosen over its interpretable competitors, and we show interpretability is still achievable in this case. Our solution is to f…

Decision MakingInterpretable Machine Learning

Neural Networks for Local Search and Crossover in Vehicle Routing: A Possible Overkill?

2022-09-09 · Ítalo Santana, Andrea Lodi, Thibaut Vidal

Extensive research has been conducted, over recent years, on various ways of enhancing heuristic search for combinatorial optimization problems with machine learning algorithms. In this study, we investigate the use of p…

Combinatorial OptimizationHeuristic Search

Hybrid Predictive Model: When an Interpretable Model Collaborates with a Black-box Model

2019-05-10 · Tong Wang, Qihang Lin

Interpretable machine learning has become a strong competitor for traditional black-box models. However, the possible loss of the predictive performance for gaining interpretability is often inevitable, putting practitio…

Interpretable Machine Learningmodel

Stop overkilling simple tasks with black-box models and use transparent models instead

2023-02-06 · Matteo Rizzo, Matteo Marcuzzo, Alessandro Zangari, Andrea Gasparetto 외

In recent years, the employment of deep learning methods has led to several significant breakthroughs in artificial intelligence. Different from traditional machine learning models, deep learning-based approaches are abl…

Deep LearningFeature Engineering