paper-with-me

홈 › Papers

Leveraging Sentence-level Information with Encoder LSTM for Semantic Slot Filling

2016-01-07 · EMNLP 2016 11 · Gakuto Kurata, Bing Xiang, Bo-Wen Zhou, Mo Yu

Recurrent Neural Network (RNN) and one of its specific architectures, Long Short-Term Memory (LSTM), have been widely used for sequence labeling. In this paper, we first enhance LSTM-based sequence labeling to explicitly model label dependencies. Then we propose another enhancement to incorporate the global information spanning over the whole input sequence. The latter proposed method, encoder-labeler LSTM, first encodes the whole input sequence into a fixed length vector with the encoder LSTM, and then uses this encoded vector as the initial state of another LSTM for sequence labeling. Combining these methods, we can predict the label sequence with considering label dependencies and information of whole input sequence. In the experiments of a slot filling task, which is an essential component of natural language understanding, with using the standard ATIS corpus, we achieved the state-of-the-art F1-score of 95.66%.

📄 PDF Abstract BibTeX arXiv:1601.01530

Code (0)

등록된 구현이 없습니다.

Tasks

Natural Language UnderstandingSentenceslot-fillingSlot Filling

Methods 이 논문이 사용한 방법론

Sigmoid Activation 설명 없음
Tanh Activation 설명 없음
LSTM An LSTM is a type of recurrent neural network that addresses the vanishing gradient problem in vanilla…

Similar Papers 제목 키워드 기반

A Sequential Neural Encoder with Latent Structured Description for Modeling Sentences

2017-11-15 · Yu-Ping Ruan, Qian Chen, Zhen-Hua Ling

In this paper, we propose a sequential neural encoder with latent structured description (SNELSD) for modeling sentences. This model introduces latent chunk-level representations into conventional sequential neural encod…

ChunkingNatural Language InferenceSentenceSentence Embeddings+1

Fake Sentence Detection as a Training Task for Sentence Encoding

2018-08-11 · ICLR 2019 5 · Viresh Ranjan, Heeyoung Kwon, Niranjan Balasubramanian, Minh Hoai

Sentence encoders are typically trained on language modeling tasks with large unlabeled datasets. While these encoders achieve state-of-the-art results on many sentence-level tasks, they are difficult to train with long …

Binary ClassificationLanguage ModelingLanguage ModellingSentence

Dual Convolutional LSTM Network for Referring Image Segmentation

2020-01-30 · Linwei Ye, Zhi Liu, Yang Wang

We consider referring image segmentation. It is a problem at the intersection of computer vision and natural language understanding. Given an input image and a referring expression in the form of a natural language sente…

DecoderImage Segmentationmultimodal interactionNatural Language Understanding+4

Interpretable Structure-aware Document Encoders with Hierarchical Attention

2019-02-26 · Khalil Mrini, Claudiu Musat, Michael Baeriswyl, Martin Jaggi

We propose a method to create document representations that reflect their internal structure. We modify Tree-LSTMs to hierarchically merge basic elements such as words and sentences into blocks of increasing complexity. …

Document ClassificationSentenceWord Embeddings

Exploring Overall Contextual Information for Image Captioning in Human-Like Cognitive Style

2019-10-15 · ICCV 2019 10 · Hongwei Ge, Zehang Yan, Kai Zhang, Mingde Zhao 외

Image captioning is a research hotspot where encoder-decoder models combining convolutional neural network (CNN) and long short-term memory (LSTM) achieve promising results. Despite significant progress, these models gen…

DecoderImage CaptioningSentence