paper-with-me

홈 › Papers

Empowering NLG: Offline Reinforcement Learning for Informal Summarization in Online Domains

2023-06-17 · Zhi-Xuan Tai, Po-Chuan Chen

Our research introduces an innovative Natural Language Generation (NLG) approach that aims to optimize user experience and alleviate the workload of human customer support agents. Our primary objective is to generate informal summaries for online articles and posts using an offline reinforcement learning technique. In our study, we compare our proposed method with existing approaches to text generation and provide a comprehensive overview of our architectural design, which incorporates crawling, reinforcement learning, and text generation modules. By presenting this original approach, our paper makes a valuable contribution to the field of NLG by offering a fresh perspective on generating natural language summaries for online content. Through the implementation of Empowering NLG, we are able to generate higher-quality replies in the online domain. The experimental results demonstrate a significant improvement in the average "like" score, increasing from 0.09954378 to 0.5000152. This advancement has the potential to enhance the efficiency and effectiveness of customer support services and elevate the overall user experience when consuming online content.

📄 PDF Abstract BibTeX arXiv:2306.17174

Code (1)

jacksonchen1998/Empowering-NLG 공식 구현 pytorch

Tasks

Articlesreinforcement-learningReinforcement LearningText Generation

Methods 이 논문이 사용한 방법론

customer support 설명 없음

Similar Papers 제목 키워드 기반

Streetwise Agents: Empowering Offline RL Policies to Outsmart Exogenous Stochastic Disturbances in RTC

2024-11-11 · Aditya Soni, Mayukh Das, Anjaly Parayil, Supriyo Ghosh 외

The difficulty of exploring and training online on real production systems limits the scope of real-time online data/feedback-driven decision making. The most feasible approach is to adopt offline reinforcement learning …

Offline RL

Value-Incentivized Preference Optimization: A Unified Approach to Online and Offline RLHF

2024-05-29 · Shicong Cen, Jincheng Mei, Katayoon Goshvadi, Hanjun Dai 외

Reinforcement learning from human feedback (RLHF) has demonstrated great promise in aligning large language models (LLMs) with human preference. Depending on the availability of preference data, both online and offline R…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)Text Summarization

Empowering Embodied Visual Tracking with Visual Foundation Models and Offline RL

2024-04-15 · Fangwei Zhong, Kui Wu, Hai Ci, Churan Wang 외

Embodied visual tracking is to follow a target object in dynamic 3D environments using an agent's egocentric vision. This is a vital and challenging skill for embodied agents. However, existing methods suffer from ineffi…

GPUOffline RLQ-LearningSemantic Segmentation+1

Abstractive Summarization of Reddit Posts with Multi-level Memory Networks

2018-11-02 · NAACL 2019 6 · Byeongchang Kim, Hyunwoo Kim, Gunhee Kim

We address the problem of abstractive summarization in two directions: proposing a novel dataset and a new model. First, we collect Reddit TIFU dataset, consisting of 120K posts from the online discussion forum Reddit. W…

Abstractive Text SummarizationArticles

MOORe: Model-based Offline-to-Online Reinforcement Learning

2022-01-25 · Yihuan Mao, Chao Wang, Bin Wang, Chongjie Zhang

With the success of offline reinforcement learning (RL), offline trained RL policies have the potential to be further improved when deployed online. A smooth transfer of the policy matters in safe real-world deployment. …

D4RLmodelreinforcement-learningReinforcement Learning+1