paper-with-me

Papers

LlaMaVAE: Guiding Large Language Model Generation via Continuous Latent Sentence Spaces

2023-12-20 · Yingji Zhang, Danilo S. Carvalho, Ian Pratt-Hartmann, André Freitas

Deep generative neural networks, such as Variational AutoEncoders (VAEs), offer an opportunity to better understand and control language models from the perspective of sentence-level latent spaces. To combine the controllability of VAE latent spaces with the state-of-the-art performance of recent large language models (LLMs), we present in this work LlaMaVAE, which combines expressive encoder and decoder models (sentenceT5 and LlaMA) with a VAE architecture, aiming to provide better text generation control to LLMs. In addition, to conditionally guide the VAE generation, we investigate a new approach based on flow-based invertible neural networks (INNs) named Invertible CVAE. Experimental results reveal that LlaMaVAE can outperform the previous state-of-the-art VAE language model, Optimus, across various tasks, including language modelling, semantic textual similarity and definition modelling. Qualitative analysis on interpolation and traversal experiments also indicates an increased degree of semantic clustering and geometric consistency, which enables better generation control.

📄 PDF Abstract BibTeX arXiv:2312.13208

Code (0)

등록된 구현이 없습니다.

Tasks

DecoderDefinition ModellingLanguage ModelingLanguage ModellingLarge Language ModelSemantic Textual SimilaritySentenceText Generation

Similar Papers 제목 키워드 기반

EmoFeedback$^2$: Reinforcement of Continuous Emotional Image Generation via LVLM-based Reward and Textual Feedback

2025-11-25 · Jingyang Jia, Kai Shu, Gang Yang, Long Xing 외 arxiv

Continuous emotional image content generation (C-EICG) is emerging rapidly due to its ability to produce images aligned with both user descriptions and continuous emotional values. However, existing approaches lack emoti…

Image Generation

FlowCoMotion: Text-to-Motion Generation via Token-Latent Flow Modeling

2026-04-13 · Dawei Guan, Di Yang, Chengjie Jin, Jiangtao Wang arxiv

Text-to-motion generation is driven by learning motion representations for semantic alignment with language. Existing methods rely on either continuous or discrete motion representations. However, continuous representati…

Local Action-Guided Motion Diffusion Model for Text-to-Motion Generation

2024-07-15 · Peng Jin, Hao Li, Zesen Cheng, Kehan Li 외

Text-to-motion generation requires not only grounding local actions in language but also seamlessly blending these individual actions to synthesize diverse and realistic global motions. However, existing motion generatio…

Graph AttentionMotion GenerationMotion Synthesis

Augmented Large Language Models with Parametric Knowledge Guiding

2023-05-08 · Ziyang Luo, Can Xu, Pu Zhao, Xiubo Geng 외

Large Language Models (LLMs) have significantly advanced natural language processing (NLP) with their impressive language understanding and generation capabilities. However, their performance may be suboptimal for domain…

StreamUni: Achieving Streaming Speech Translation with a Unified Large Speech-Language Model

2025-07-10 · Shoutao Guo, Xiang Li, Mengge Liu, Wei Chen 외 arxiv

Streaming speech translation (StreamST) requires determining appropriate timing, known as policy, to generate translations while continuously receiving source speech inputs, balancing low latency with high translation qu…