paper-with-me

홈 › Papers

WebGen-V Bench: Structured Representation for Enhancing Visual Design in LLM-based Web Generation and Evaluation

2025-10-17 · Kuang-Da Wang, Zhao Wang, Yotaro Shimose, Wei-Yao Wang, Shingo Takamatsu arxiv

Witnessed by the recent advancements on leveraging LLM for coding and multimodal understanding, we present WebGen-V, a new benchmark and framework for instruction-to-HTML generation that enhances both data quality and evaluation granularity. WebGen-V contributes three key innovations: (1) an unbounded and extensible agentic crawling framework that continuously collects real-world webpages and can leveraged to augment existing benchmarks; (2) a structured, section-wise data representation that integrates metadata, localized UI screenshots, and JSON-formatted text and image assets, explicit alignment between content, layout, and visual components for detailed multimodal supervision; and (3) a section-level multimodal evaluation protocol aligning text, layout, and visuals for high-granularity assessment. Experiments with state-of-the-art LLMs and ablation studies validate the effectiveness of our structured data and section-wise evaluation, as well as the contribution of each component. To the best of our knowledge, WebGen-V is the first work to enable high-granularity agentic crawling and evaluation for instruction-to-HTML generation, providing a unified pipeline from real-world data acquisition and webpage generation to structured multimodal assessment.

📄 PDF Abstract BibTeX arXiv:2510.15306

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

WebGen-Agent: Enhancing Interactive Website Generation with Multi-Level Feedback and Step-Level Reinforcement Learning

2025-09-26 · Zimu Lu, Houxing Ren, Yunqiao Yang, Ke Wang 외 arxiv

Agent systems powered by large language models (LLMs) have demonstrated impressive performance on repository-level code-generation tasks. However, for tasks such as website codebase generation, which depend heavily on vi…

Reinforcement Learning

ProductWebGen: Benchmarking Multimodal Product Webpage Generation

2026-05-31 · Zhihong Liu, Siqi Kou, Zheng Li, Ye Ma 외 arxiv

Crafting a product display webpage from a source product image, along with layout and visual content instructions, holds significant practical value for domains such as marketing, advertising, and E-commerce. Intuitively…

Instruction FollowingImage GenerationImage Editing

WebGen-R1: Incentivizing Large Language Models to Generate Functional and Aesthetic Websites with Reinforcement Learning

2026-04-22 · Juyong Jiang, Chenglin Cai, Chansung Park, Jiasi Shen 외 arxiv

While Large Language Models (LLMs) excel at function-level code generation, project-level tasks such as generating functional and visually aesthetic multi-page websites remain highly challenging. Existing works are often…

Reinforcement LearningCode Generation

I-WebGenBench : Evaluating Interactivity in LLM-Generated Scientific Web Applications

2026-05-30 · Dasen Dai, Biao Wu, Meng Fang, Shuoqi Li 외 arxiv

Recent advances in visual language models have enabled autonomous agents for complex reasoning, tool use, and document understanding. However, existing document agents mainly transform papers into static artifacts such a…

WebGen-Bench: Evaluating LLMs on Generating Interactive and Functional Websites from Scratch

2025-05-06 · Zimu Lu, Yunqiao Yang, Houxing Ren, Haotian Hou 외

LLM-based agents have demonstrated great potential in generating and managing code within complex codebases. In this paper, we introduce WebGen-Bench, a novel benchmark designed to measure an LLM-based agent's ability to…