paper-with-me

Papers

Let Your Graph Do the Talking: Encoding Structured Data for LLMs

2024-02-08 · Bryan Perozzi, Bahare Fatemi, Dustin Zelle, Anton Tsitsulin, Mehran Kazemi, Rami Al-Rfou, Jonathan Halcrow

How can we best encode structured data into sequential form for use in large language models (LLMs)? In this work, we introduce a parameter-efficient method to explicitly represent structured data for LLMs. Our method, GraphToken, learns an encoding function to extend prompts with explicit structured information. Unlike other work which focuses on limited domains (e.g. knowledge graph representation), our work is the first effort focused on the general encoding of structured data to be used for various reasoning tasks. We show that explicitly representing the graph structure allows significant improvements to graph reasoning tasks. Specifically, we see across the board improvements - up to 73% points - on node, edge and, graph-level tasks from the GraphQA benchmark.

📄 PDF Abstract BibTeX arXiv:2402.05862

Code (2)

google-research/talk-like-a-graph 공식 구현
wxxshirley/gnn4taskplan pytorch

Similar Papers 제목 키워드 기반

EditYourself: Audio-Driven Generation and Manipulation of Talking Head Videos with Diffusion Transformers

2026-01-29 · John Flynn, Wolfgang Paier, Dimitar Dinev, Sam Nhut Nguyen 외 arxiv

Current generative video models excel at producing novel content from text and image prompts, but leave a critical gap in editing existing pre-recorded videos, where minor alterations to the spoken script require preserv…

Audio-Plane: Audio Factorization Plane Gaussian Splatting for Real-Time Talking Head Synthesis

2025-03-28 · Shuai Shen, Wanhua Li, Yunpeng Zhang, Weipeng Hu 외

Talking head synthesis has become a key research area in computer graphics and multimedia, yet most existing methods often struggle to balance generation quality with computational efficiency. In this paper, we present a…

Computational EfficiencyTalking Head Generation

Audio-Driven Talking Face Generation with Blink Embedding and Hash Grid Landmarks Encoding

2026-01-26 · Yuhui Zhang, Hui Yu, Wei Liang, Sunjie Zhang arxiv

Dynamic Neural Radiance Fields (NeRF) have demonstrated considerable success in generating high-fidelity 3D models of talking portraits. Despite significant advancements in the rendering speed and generation quality, cha…

Talking Face Generation

VAST: Vivify Your Talking Avatar via Zero-Shot Expressive Facial Style Transfer

2023-08-09 · Liyang Chen, Zhiyong Wu, Runnan Li, Weihong Bao 외

Current talking face generation methods mainly focus on speech-lip synchronization. However, insufficient investigation on the facial talking style leads to a lifeless and monotonous avatar. Most previous works fail to i…

DecoderFace GenerationStyle TransferTalking Face Generation

Real-time Neural Radiance Talking Portrait Synthesis via Audio-spatial Decomposition

2022-11-22 · Jiaxiang Tang, Kaisiyuan Wang, Hang Zhou, Xiaokang Chen 외

While dynamic Neural Radiance Fields (NeRF) have shown success in high-fidelity 3D modeling of talking portraits, the slow training and inference speed severely obstruct their potential usage. In this paper, we propose a…

NeRFTalking Face Generation