paper-with-me

홈 › Papers

Stable LLM Ensemble: Interaction between Example Representativeness and Diversity

2025-10-15 · Junichiro Niimi arxiv

Large language models (LLMs) have achieved remarkable results in wide range of domains. However, the accuracy and robustness of one-shot LLM predictions remain highly sensitive to the examples and the diversity among ensemble members. This study systematically investigates the effects of example representativeness (one-shot strategy) and output diversity (sampling temperature) on LLM ensemble performance. Two one-shot strategies are compared: centroid-based representative examples (proposed) and randomly sampled examples (baseline) and sampling temperature also is varied. The proposed approach with higher temperature setting significantly outperforms random selection by +7.6% (macro-F1) and -10.5% (RMSE). Furthermore, the proposed model exceeds 5-shot prompting by +21.1% (macro-F1) and -24.0% (RMSE). Our findings demonstrate that combining representative example selection with increased temperature provides the appropriate level of diversity to the ensemble. This work highlights the practical importance of both example selection and controlled diversity in designing effective one-shot LLM ensembles.

📄 PDF Abstract BibTeX arXiv:2510.13143

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Uncertainty-Aware Measurement of Scenario Suite Representativeness for Autonomous Systems

2025-11-18 · Robab Aghazadeh Chakherlou, Siddartha Khastgir, Xingyu Zhao, Jerein Jeyachandran 외 arxiv

Assuring the trustworthiness and safety of AI systems, e.g., autonomous vehicles (AV), depends critically on the data-related safety properties, e.g., representativeness, completeness, etc., of the datasets used for thei…

Autonomous Vehicles

Inspecting the Geographical Representativeness of Images from Text-to-Image Models

2023-05-18 · ICCV 2023 1 · Abhipsa Basu, R. Venkatesh Babu, Danish Pruthi

Recent progress in generative models has resulted in models that produce both realistic as well as relevant images for most textual inputs. These models are being used to generate millions of images everyday, and hold th…

Data AugmentationMarketing

Phrasing for UX: Enhancing Information Engagement through Computational Linguistics and Creative Analytics

2024-08-23 · Nimrod Dvir

This study explores the relationship between textual features and Information Engagement (IE) on digital platforms. It highlights the impact of computational linguistics and analytics on user interaction. The READ model …

Will the Real Linda Please Stand up...to Large Language Models? Examining the Representativeness Heuristic in LLMs

2024-04-01 · Pengda Wang, Zilin Xiao, Hanjie Chen, Frederick L. Oswald

Although large language models (LLMs) have demonstrated remarkable proficiency in modeling text and generating human-like text, they may exhibit biases acquired from training data in doing so. Specifically, LLMs may be s…

Decision Making

Beyond Marginal Distributions: A Framework to Evaluate the Representativeness of Demographic-Aligned LLMs

2026-01-22 · Tristan Williams, Franziska Weeber, Sebastian Padó, Alan Akbik arxiv

Large language models are increasingly used to represent human opinions, values, or beliefs, and their steerability towards these ideals is an active area of research. Existing work focuses predominantly on aligning marg…