paper-with-me

홈 › Papers

How do Visual Attributes Influence Web Agents? A Comprehensive Evaluation of User Interface Design Factors

2026-01-29 · Kuai Yu, Naicheng Yu, Han Wang, Rui Yang, Huan Zhang arxiv

Web agents have demonstrated strong performance on a wide range of web-based tasks. However, existing research on the effect of environmental variation has mostly focused on robustness to adversarial attacks, with less attention to agents' preferences in benign scenarios. Although early studies have examined how textual attributes influence agent behavior, a systematic understanding of how visual attributes shape agent decision-making remains limited. To address this, we introduce VAF, a controlled evaluation pipeline for quantifying how webpage Visual Attribute Factors influence web-agent decision-making. Specifically, VAF consists of three stages: (i) variant generation, which ensures the variants share identical semantics as the original item while only differ in visual attributes; (ii) browsing interaction, where agents navigate the page via scrolling and clicking the interested item, mirroring how human users browse online; (iii) validating through both click action and reasoning from agents, which we use the Target Click Rate and Target Mention Rate to jointly evaluate the effect of visual attributes. By quantitatively measuring the decision-making difference between the original and variant, we identify which visual attributes influence agents' behavior most. Extensive experiments, across 8 variant families (48 variants total), 5 real-world websites (including shopping, travel, and news browsing), and 4 representative web agents, show that background color contrast, item size, position, and card clarity have a strong influence on agents' actions, whereas font styling, text color, and item image clarity exhibit minor effects.

📄 PDF Abstract BibTeX arXiv:2601.21961

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

SocialBench: Sociality Evaluation of Role-Playing Conversational Agents

2024-03-20 · Hongzhan Chen, Hehong Chen, Ming Yan, Wenshen Xu 외

Large language models (LLMs) have advanced the development of various AI conversational agents, including role-playing conversational agents that mimic diverse characters and human behaviors. While prior research has pre…

Credit Assignment and Efficient Exploration based on Influence Scope in Multi-agent Reinforcement Learning

2025-05-13 · Shuai Han, Mehdi Dastani, Shihan Wang

Training cooperative agents in sparse-reward scenarios poses significant challenges for multi-agent reinforcement learning (MARL). Without clear feedback on actions at each step in sparse-reward setting, previous methods…

Efficient ExplorationMulti-agent Reinforcement Learning

Mitigating Hallucinations on Object Attributes using Multiview Images and Negative Instructions

2025-01-17 · Zhijie Tan, Yuzhi Li, Shengwei Meng, Xiang Yuan 외

Current popular Large Vision-Language Models (LVLMs) are suffering from Hallucinations on Object Attributes (HoOA), leading to incorrect determination of fine-grained attributes in the input images. Leveraging significan…

3D Generation

TRACE: Trajectory-Aware Comprehensive Evaluation for Deep Research Agents

2026-02-05 · Yanyu Chen, Jiyue Jiang, Jiahong Liu, Yifei Zhang 외 arxiv

The evaluation of Deep Research Agents is a critical challenge, as conventional outcome-based metrics fail to capture the nuances of their complex reasoning. Current evaluation faces two primary challenges: 1) a reliance…

Training on Art Composition Attributes to Influence CycleGAN Art Generation

2018-12-19 · Holly Grimm

I consider how to influence CycleGAN, image-to-image translation, by using additional constraints from a neural network trained on art composition attributes. I show how I trained the the Art Composition Attributes Netwo…

AttributeImage-to-Image TranslationTranslation