paper-with-me

홈 › Papers

Improving Text Auto-Completion with Next Phrase Prediction

2021-09-15 · Findings (EMNLP) 2021 11 · Dong-Ho Lee, Zhiqiang Hu, Roy Ka-Wei Lee

Language models such as GPT-2 have performed well on constructing syntactically sound sentences for text auto-completion task. However, such models often require considerable training effort to adapt to specific writing domains (e.g., medical). In this paper, we propose an intermediate training strategy to enhance pre-trained language models' performance in the text auto-completion task and fastly adapt them to specific domains. Our strategy includes a novel self-supervised training objective called Next Phrase Prediction (NPP), which encourages a language model to complete the partial query with enriched phrases and eventually improve the model's text auto-completion performance. Preliminary experiments have shown that our approach is able to outperform the baselines in auto-completion for email and academic writing domains.

📄 PDF Abstract BibTeX arXiv:2109.07067

Code (0)

등록된 구현이 없습니다.

Tasks

Language ModelingLanguage ModellingPrediction

Methods 이 논문이 사용한 방법론

Refunds@Expedia|||How do I get a full refund from Expedia? “How do I get a full refund from Expedia? How do I get a full refund from Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Quick Help &…
Multi-Head Attention 설명 없음
Attention 설명 없음
Linear Layer A Linear Layer is a projection $\mathbf{XW + b}$.
Adam 설명 없음
Dense Connections Dense Connections, or Fully Connected Connections, are a type of layer in a deep neural network that use a linear operation where every input is connected to every output…
Cosine Annealing Cosine Annealing is a type of learning rate schedule that has the effect of starting with a large learning rate that is relatively rapidly decreased to a minimum value before…
Dropout Dropout is a regularization technique for neural networks that drops a unit (along with connections) at training time with a specified probability $p$ (a common value is…

Similar Papers 제목 키워드 기반

Deep Keyphrase Completion

2021-10-29 · Yu Zhao, Jia Song, Huali Feng, Fuzhen Zhuang 외

Keyphrase provides accurate information of document content that is highly compact, concise, full of meanings, and widely used for discourse comprehension, organization, and text retrieval. Though previous studies have m…

DecoderKeyphrase ExtractionKeyphrase GenerationRetrieval+1

Contextual LSTM (CLSTM) models for Large scale NLP tasks

2016-02-19 · Shalini Ghosh, Oriol Vinyals, Brian Strope, Scott Roy 외

Documents exhibit sequential structure at multiple levels of abstraction (e.g., sentences, paragraphs, sections). These abstractions constitute a natural hierarchy for representing the context in which to infer the meani…

ArticlesParaphrase GenerationPredictionQuestion Answering+2

RepoBench: Benchmarking Repository-Level Code Auto-Completion Systems

2023-06-05 · Tianyang Liu, Canwen Xu, Julian McAuley

Large Language Models (LLMs) have greatly advanced code auto-completion systems, with a potential for substantial productivity enhancements for developers. However, current benchmarks mainly focus on single-file tasks, l…

BenchmarkingC++ codeCode CompletionRetrieval

Effidit: Your AI Writing Assistant

2022-08-03 · Shuming Shi, Enbo Zhao, Duyu Tang, Yan Wang 외

In this technical report, we introduce Effidit (Efficient and Intelligent Editing), a digital writing assistant that facilitates users to write higher-quality text more efficiently by using artificial intelligence (AI) t…

Keywords to SentencesRetrievalSentenceSentence Completion+1

mOKB6: A Multilingual Open Knowledge Base Completion Benchmark

2022-11-13 · Shubham Mittal, Keshav Kolluru, Soumen Chakrabarti, Mausam

Automated completion of open knowledge bases (Open KBs), which are constructed from triples of the form (subject phrase, relation phrase, object phrase), obtained via open information extraction (Open IE) system, are use…

coreference-resolutionCoreference ResolutionKnowledge Base CompletionOpen Information Extraction