paper-with-me

홈 › Papers

Classical Structured Prediction Losses for Sequence to Sequence Learning

2017-11-14 · NAACL 2018 6 · Sergey Edunov, Myle Ott, Michael Auli, David Grangier, Marc'Aurelio Ranzato

There has been much recent work on training neural attention models at the sequence-level using either reinforcement learning-style methods or by optimizing the beam. In this paper, we survey a range of classical objective functions that have been widely used to train linear models for structured prediction and apply them to neural sequence to sequence models. Our experiments show that these losses can perform surprisingly well by slightly outperforming beam search optimization in a like for like setup. We also report new state of the art results on both IWSLT'14 German-English translation as well as Gigaword abstractive summarization. On the larger WMT'14 English-French translation task, sequence-level training achieves 41.5 BLEU which is on par with the state of the art.

📄 PDF Abstract BibTeX arXiv:1711.04956

Code (1)

pytorch/fairseq 공식 구현 pytorch

Tasks

Abstractive Text SummarizationMachine TranslationPredictionReinforcement LearningReinforcement Learning (RL)Structured PredictionTranslation

Similar Papers 제목 키워드 기반

On Structured Prediction Theory with Calibrated Convex Surrogate Losses

2017-03-07 · NeurIPS 2017 12 · Anton Osokin, Francis Bach, Simon Lacoste-Julien

We provide novel theoretical insights on structured prediction in the context of efficient convex surrogate loss minimization with consistency guarantees. For any task loss, we construct a convex surrogate that can be op…

PredictionStructured Prediction

Second Order Regret Bounds Against Generalized Expert Sequences under Partial Bandit Feedback

2022-04-13 · Kaan Gokcesu, Hakan Gokcesu

We study the problem of expert advice under partial bandit feedback setting and create a sequential minimax optimal algorithm. Our algorithm works with a more general partial monitoring setting, where, in contrast to the…

Compositional Generalization for Neural Semantic Parsing via Span-level Supervised Attention

2021-06-01 · NAACL 2021 4 · Pengcheng Yin, Hao Fang, Graham Neubig, Adam Pauls 외

We describe a span-level supervised attention loss that improves compositional generalization in semantic parsers. Our approach builds on existing losses that encourage attention maps in neural sequence-to-sequence model…

Machine TranslationSemantic ParsingTranslationWord Alignment

Structured Recommendation

2017-06-27 · Chen Dawei, Xie Lexing, Menon Aditya Krishna, Ong Cheng Soon

Current recommender systems largely focus on static, unstructured content. In many scenarios, we would like to recommend content that has structure, such as a trajectory of points-of-interests in a city, or a playlist of…

Recommendation SystemsStructured Predictionvalid

Efficient Gradient Computation for Structured Output Learning with Rational and Tropical Losses

2018-12-01 · NeurIPS 2018 12 · Corinna Cortes, Vitaly Kuznetsov, Mehryar Mohri, Dmitry Storcheus 외

Many structured prediction problems admit a natural loss function for evaluation such as the edit-distance or $n$-gram loss. However, existing learning algorithms are typically designed to optimize alternative objectives…

Structured Prediction