paper-with-me

Papers

How Different AI Chatbots Behave? Benchmarking Large Language Models in Behavioral Economics Games

2024-12-16 · Yutong Xie, Yiyao Liu, Zhuang Ma, Lin Shi, Xiyuan Wang, Walter Yuan, Matthew O. Jackson, Qiaozhu Mei

The deployment of large language models (LLMs) in diverse applications requires a thorough understanding of their decision-making strategies and behavioral patterns. As a supplement to a recent study on the behavioral Turing test, this paper presents a comprehensive analysis of five leading LLM-based chatbot families as they navigate a series of behavioral economics games. By benchmarking these AI chatbots, we aim to uncover and document both common and distinct behavioral patterns across a range of scenarios. The findings provide valuable insights into the strategic preferences of each LLM, highlighting potential implications for their deployment in critical decision-making roles.

📄 PDF Abstract BibTeX arXiv:2412.12362

Code (0)

등록된 구현이 없습니다.

Tasks

BenchmarkingChatbotDecision MakingNavigate

Similar Papers 제목 키워드 기반

Investigating person-specific errors in chat-oriented dialogue systems

2022-05-01 · ACL 2022 5 · Koh Mitsuda, Ryuichiro Higashinaka, Tingxuan Li, Sen Yoshida

Creating chatbots to behave like real people is important in terms of believability. Errors in general chatbots and chatbots that follow a rough persona have been studied, but those in chatbots that behave like real peop…

Chatbot

Benchmarking LLM powered Chatbots: Methods and Metrics

2023-08-08 · Debarag Banerjee, Pooja Singh, Arjun Avadhanam, Saksham Srivastava

Autonomous conversational agents, i.e. chatbots, are becoming an increasingly common mechanism for enterprises to provide support to customers and partners. In order to rate chatbots, especially ones powered by Generativ…

BenchmarkingChatbot

A Turing Test: Are AI Chatbots Behaviorally Similar to Humans?

2023-11-19 · Qiaozhu Mei, Yutong Xie, Walter Yuan, Matthew O. Jackson

We administer a Turing Test to AI Chatbots. We examine how Chatbots behave in a suite of classic behavioral games that are designed to elicit characteristics such as trust, fairness, risk-aversion, cooperation, \textit{e…

Fairness

Towards Ethical Machines Via Logic Programming

2019-09-18 · Abeer Dyoub, Stefania Costantini, Francesca A. Lisi

Autonomous intelligent agents are playing increasingly important roles in our lives. They contain information about us and start to perform tasks on our behalves. Chatbots are an example of such agents that need to engag…

One Chatbot Per Person: Creating Personalized Chatbots based on Implicit User Profiles

2021-08-20 · Zhengyi Ma, Zhicheng Dou, Yutao Zhu, Hanxun Zhong 외

Personalized chatbots focus on endowing chatbots with a consistent personality to behave like real users, give more informative responses, and further act as personal assistants. Existing personalized approaches tried to…

ChatbotLanguage Modelling