paper-with-me

홈 › Papers

Evaluating the Capabilities of LLMs for Supporting Anticipatory Impact Assessment

2024-01-31 · Mowafak Allaham, Nicholas Diakopoulos

Gaining insight into the potential negative impacts of emerging Artificial Intelligence (AI) technologies in society is a challenge for implementing anticipatory governance approaches. One approach to produce such insight is to use Large Language Models (LLMs) to support and guide experts in the process of ideating and exploring the range of undesirable consequences of emerging technologies. However, performance evaluations of LLMs for such tasks are still needed, including examining the general quality of generated impacts but also the range of types of impacts produced and resulting biases. In this paper, we demonstrate the potential for generating high-quality and diverse impacts of AI in society by fine-tuning completion models (GPT-3 and Mistral-7B) on a diverse sample of articles from news media and comparing those outputs to the impacts generated by instruction-based (GPT-4 and Mistral-7B-Instruct) models. We examine the generated impacts for coherence, structure, relevance, and plausibility and find that the generated impacts using Mistral-7B, a small open-source model fine-tuned on impacts from the news media, tend to be qualitatively on par with impacts generated using a more capable and larger scale model such as GPT-4. Moreover, we find that impacts produced by instruction-based models had gaps in the production of certain categories of impacts in comparison to fine-tuned models. This research highlights a potential bias in the range of impacts generated by state-of-the-art LLMs and the potential of aligning smaller LLMs on news media as a scalable alternative to generate high quality and more diverse impacts in support of anticipatory governance approaches.

📄 PDF Abstract BibTeX arXiv:2401.18028

Code (0)

등록된 구현이 없습니다.

Tasks

Articles

Methods 이 논문이 사용한 방법론

Attention 설명 없음
Linear Layer A Linear Layer is a projection $\mathbf{XW + b}$.
BPE Byte Pair Encoding, or BPE, is a subword segmentation algorithm that encodes rare and unknown words as sequences of subword units. The intuition is that various word…
Residual Connection 설명 없음
Dropout Dropout is a regularization technique for neural networks that drops a unit (along with connections) at training time with a specified probability $p$ (a common value is…
Layer Normalization Unlike batch normalization, Layer Normalization directly estimates the normalization statistics from the summed inputs…
Multi-Head Attention 설명 없음
Adam 설명 없음

Similar Papers 제목 키워드 기반

AdaptEval: A Benchmark for Evaluating Large Language Models on Code Snippet Adaptation

2026-01-08 · Tanghaoran Zhang, Xinjun Mao, Shangwen Wang, Yuxin Zhao 외 arxiv

Recent advancements in large language models (LLMs) have automated various software engineering tasks, with benchmarks emerging to evaluate their capabilities. However, for adaptation, a critical activity during code reu…

SayNext-Bench: Why Do LLMs Struggle with Next-Utterance Anticipation?

2026-01-30 · Yueyi Yang, Haotian Liu, Fang Kang, Mengqi Zhang 외 arxiv

We explore the use of large language models (LLMs) for next-utterance anticipation in human dialogue. Despite recent advances in LLMs demonstrating their ability to engage in natural conversations with users, we show tha…

AI for Anticipatory Action: Moving Beyond Climate Forecasting

2023-07-28 · Benjamin Q. Huynh, Mathew V. Kiang

Disaster response agencies have been shifting from a paradigm of climate forecasting towards one of anticipatory action: assessing not just what the climate will be, but how it will impact specific populations, thereby e…

Disaster Response

Investigating the Factual Knowledge Boundary of Large Language Models with Retrieval Augmentation

2023-07-20 · Ruiyang Ren, Yuhao Wang, Yingqi Qu, Wayne Xin Zhao 외

Large language models (LLMs) have shown impressive prowess in solving a wide range of tasks with world knowledge. However, it remains unclear how well LLMs are able to perceive their factual knowledge boundaries, particu…

Open-Domain Question AnsweringQuestion AnsweringRetrievalWorld Knowledge

Anticipatory Thinking Challenges in Open Worlds: Risk Management

2023-06-22 · Adam Amos-Binks, Dustin Dannenhauer, Leilani H. Gilpin

Anticipatory thinking drives our ability to manage risk - identification and mitigation - in everyday life, from bringing an umbrella when it might rain to buying car insurance. As AI systems become part of everyday life…

Adversarial RobustnessAutonomous VehiclesManagementModel Poisoning+1