paper-with-me

홈 › Papers

A Categorical Archive of ChatGPT Failures

2023-02-06 · Ali Borji

Large language models have been demonstrated to be valuable in different fields. ChatGPT, developed by OpenAI, has been trained using massive amounts of data and simulates human conversation by comprehending context and generating appropriate responses. It has garnered significant attention due to its ability to effectively answer a broad range of human inquiries, with fluent and comprehensive answers surpassing prior public chatbots in both security and usefulness. However, a comprehensive analysis of ChatGPT's failures is lacking, which is the focus of this study. Eleven categories of failures, including reasoning, factual errors, math, coding, and bias, are presented and discussed. The risks, limitations, and societal implications of ChatGPT are also highlighted. The goal of this study is to assist researchers and developers in enhancing future language models and chatbots.

📄 PDF Abstract BibTeX arXiv:2302.03494

Code (1)

aliborji/chatgpt_failures 공식 구현

Tasks

Math

Similar Papers 제목 키워드 기반

Artificial Intelligence in archival and historical scholarship workflow: HTS and ChatGPT

2023-07-05 · Salvatore Spina

This article examines the impact of Artificial Intelligence on the archival heritage digitization processes, specifically regarding the manuscripts' automatic transcription, their correction, and normalization. It highli…

Dynamics of Human-AI Collective Knowledge on the Web: A Scalable Model and Insights for Sustainable Growth

2026-01-27 · Buddhika Nettasinghe, Kang Zhao arxiv

Humans and large language models (LLMs) now co-produce and co-consume the web's shared knowledge archives. Such human-AI collective knowledge ecosystems contain feedback loops with both benefits (e.g., faster growth, eas…

ChatGPT (Feb 13 Version) is a Chinese Room

2023-02-19 · Maurice HT Ling

ChatGPT has gained both positive and negative publicity after reports suggesting that it is able to pass various professional and licensing examinations. This suggests that ChatGPT may pass Turing Test in the near future…

Hallucination

Prompting Implicit Discourse Relation Annotation

2024-02-07 · Frances Yung, Mansoor Ahmad, Merel Scholman, Vera Demberg

Pre-trained large language models, such as ChatGPT, archive outstanding performance in various reasoning tasks without supervised training and were found to have outperformed crowdsourcing workers. Nonetheless, ChatGPT's…

ClassificationImplicit Discourse Relation ClassificationMultiple-choicePrompt Engineering+2

ChatGPT for Programming Numerical Methods

2023-03-21 · Ali Kashefi, Tapan Mukerji

ChatGPT is a large language model recently released by the OpenAI company. In this technical report, we explore for the first time the capability of ChatGPT for programming numerical algorithms. Specifically, we examine …

Language ModellingLarge Language Model