paper-with-me

홈 › Papers

PosterGen: Aesthetic-Aware Multi-Modal Paper-to-Poster Generation via Multi-Agent LLMs

2025-08-24 · Zhilin Zhang, Xiang Zhang, Jiaqi Wei, Yiwei Xu, Chenyu You arxiv

Multi-agent systems built upon large language models (LLMs) have demonstrated remarkable capabilities in tackling complex compositional tasks. In this work, we apply this paradigm to the paper-to-poster generation problem, a practical yet time-consuming process faced by researchers preparing for conferences. While recent approaches have attempted to automate this task, most neglect core design and aesthetic principles, resulting in posters that require substantial manual refinement. To address these design limitations, we propose PosterGen, a multi-agent framework that mirrors the workflow of professional poster designers. It consists of four collaborative specialized agents: (1) Parser and Curator agents extract content from the paper and organize storyboard; (2) Layout agent maps the content into a coherent spatial layout; (3) Stylist agents apply visual design elements such as color and typography; and (4) Renderer composes the final poster. Together, these agents produce posters that are both semantically grounded and visually appealing. To evaluate design quality, we introduce a vision-language model (VLM)-based rubric that measures layout balance, readability, and aesthetic coherence. Experimental results show that PosterGen consistently matches in content fidelity, and significantly outperforms existing methods in visual designs, generating posters that are presentation-ready with minimal human refinements.

📄 PDF Abstract BibTeX arXiv:2508.17188

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

EfficientPosterGen: Semantic-aware Efficient Poster Generation via Token Compression and Accurate Violation Detection

2026-02-25 · Wenxin Tang, Jingyu Xiao, Yanpei Gong, Fengyuan Ran 외 arxiv

Automated academic poster generation aims to distill lengthy research papers into concise, visually coherent presentations. Existing Multimodal Large Language Models (MLLMs) based approaches, however, suffer from three c…

Information Retrievalmultimodal generation

AesthetiQ: Enhancing Graphic Layout Design via Aesthetic-Aware Preference Alignment of Multi-modal Large Language Models

2025-03-01 · CVPR 2025 1 · Sohan Patnaik, Rishabh Jain, Balaji Krishnamurthy, Mausoom Sarkar

Visual layouts are essential in graphic design fields such as advertising, posters, and web interfaces. The application of generative models for content-aware layout generation has recently gained traction. However, thes…

Large Language ModelLayout DesignLayout Generation

Presenting a Paper is an Art: Self-Improvement Aesthetic Agents for Academic Presentations

2025-10-07 · Chengzhi Liu, Yuzhe Yang, Kaiwen Zhou, Zhen Zhang 외 arxiv

The promotion of academic papers has become an important means of enhancing research visibility. However, existing automated methods struggle limited storytelling, insufficient aesthetic quality, and constrained self-adj…

Reinforcement Learning

TATTOO: Training-free AesTheTic-aware Outfit recOmmendation

2025-09-27 · Yuntian Wu, Xiaonan Hu, Ziqi Zhou, Hao Lu arxiv

The global fashion e-commerce market relies significantly on intelligent and aesthetic-aware outfit-completion tools to promote sales. While previous studies have approached the problem of fashion outfit-completion and c…

Enhancing Zero-shot Personalized Image Aesthetics Assessment with Profile-aware Multimodal LLM

2026-04-19 · Chun Wang, Chenfeng Wei, Chenyang Liu, Weihong Deng arxiv

Personalized image aesthetics assessment (PIAA) aims to predict an individual user's subjective rating of an image, which requires modeling user-specific aesthetic preferences. Existing methods rely on historical user ra…