paper-with-me

Papers

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents

2025-04-24 · Chaoran Chen, Zhiping Zhang, Ibrahim Khalilov, Bingcan Guo, Simret A Gebreegziabher, Yanfang Ye, Ziang Xiao, Yaxing Yao, Tianshi Li, Toby Jia-Jun Li

The rise of Large Language Models (LLMs) has revolutionized Graphical User Interface (GUI) automation through LLM-powered GUI agents, yet their ability to process sensitive data with limited human oversight raises significant privacy and security risks. This position paper identifies three key risks of GUI agents and examines how they differ from traditional GUI automation and general autonomous agents. Despite these risks, existing evaluations focus primarily on performance, leaving privacy and security assessments largely unexplored. We review current evaluation metrics for both GUI and general LLM agents and outline five key challenges in integrating human evaluators for GUI agent assessments. To address these gaps, we advocate for a human-centered evaluation framework that incorporates risk assessments, enhances user awareness through in-context consent, and embeds privacy and security considerations into GUI agent design and evaluation.

📄 PDF Abstract BibTeX arXiv:2504.17934

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

HELM: A Human-Centered Evaluation Framework for LLM-Powered Recommender Systems

2026-01-27 · Sushant Mehta arxiv

The integration of Large Language Models (LLMs) into recommendation systems has introduced unprecedented capabilities for natural language understanding, explanation generation, and conversational interactions. However, …

Natural Language UnderstandingCollaborative FilteringExplanation GenerationRecommendation Systems

Toward Human-Centered Readability Evaluation

2025-10-12 · Bahar İlgen, Georges Hattab arxiv

Text simplification is essential for making public health information accessible to diverse populations, including those with limited health literacy. However, commonly used evaluation metrics in Natural Language Process…

Text Simplification

The Challenges and Opportunities of Human-Centered AI for Trustworthy Robots and Autonomous Systems

2021-05-07 · Hongmei He, John Gray, Angelo Cangelosi, Qinggang Meng 외

The trustworthiness of Robots and Autonomous Systems (RAS) has gained a prominent position on many research agendas towards fully autonomous systems. This research systematically explores, for the first time, the key fac…

Ethics

Human-Centered Artificial Intelligence: Reliable, Safe & Trustworthy

2020-02-10 · Ben Shneiderman

Well-designed technologies that offer high levels of human control and high levels of computer automation can increase human performance, leading to wider adoption. The Human-Centered Artificial Intelligence (HCAI) frame…

Explanation User Interfaces: A Systematic Literature Review

2025-05-26 · Eleonora Cappuccio, Andrea Esposito, Francesco Greco, Giuseppe Desolda 외

Artificial Intelligence (AI) is one of the major technological advancements of this century, bearing incredible potential for users through AI-powered applications and tools in numerous domains. Being often black-box (i.…

Decision MakingExplainable artificial intelligenceExplainable Artificial Intelligence (XAI)Systematic Literature Review