paper-with-me

홈 › Papers

Is Temperature the Creativity Parameter of Large Language Models?

2024-05-01 · Max Peeperkorn, Tom Kouwenhoven, Dan Brown, Anna Jordanous

Large language models (LLMs) are applied to all sorts of creative tasks, and their outputs vary from beautiful, to peculiar, to pastiche, into plain plagiarism. The temperature parameter of an LLM regulates the amount of randomness, leading to more diverse outputs; therefore, it is often claimed to be the creativity parameter. Here, we investigate this claim using a narrative generation task with a predetermined fixed context, model and prompt. Specifically, we present an empirical analysis of the LLM output for different temperature values using four necessary conditions for creativity in narrative generation: novelty, typicality, cohesion, and coherence. We find that temperature is weakly correlated with novelty, and unsurprisingly, moderately correlated with incoherence, but there is no relationship with either cohesion or typicality. However, the influence of temperature on creativity is far more nuanced and weak than suggested by the "creativity parameter" claim; overall results suggest that the LLM generates slightly more novel outputs as temperatures get higher. Finally, we discuss ideas to allow more controlled LLM creativity, rather than relying on chance via changing the temperature parameter.

📄 PDF Abstract BibTeX arXiv:2405.00492

Code (1)

maxpeeperkorn/creativity-parameter 공식 구현

Similar Papers 제목 키워드 기반

Probing the Creativity of Large Language Models: Can models produce divergent semantic association?

2023-10-17 · Honghua Chen, Nai Ding

Large language models possess remarkable capacity for processing language, but it remains unclear whether these models can further generate creative content. The present study aims to investigate the creative thinking of…

The Paradox of Stochasticity: Limited Creativity and Computational Decoupling in Temperature-Varied LLM Outputs of Structured Fictional Data

2025-02-12 · Evgenii Evstafev

This study examines how temperature settings and model architectures affect the generation of structured fictional data (names, birthdates) across three large language models (LLMs): llama3.1:8b, deepseek-r1:8b, and mist…

Computational EfficiencyDiversityModel Selection

Automated Creativity Evaluation of Language Models Across Open-Ended Tasks

2026-06-10 · Min Sen Tan, Zachary Kit Chun Choy, Syed Ali Redha Alsagoff, Nadya Yuki Wangsajaya 외 arxiv

Large language models (LLMs) have achieved remarkable progress in language understanding, reasoning, and generation, sparking growing interest in their creative potential. Realizing this potential requires systematic and…

Before and After Temperature: A Distributional View of Creative LLM Generation

2026-05-31 · V. S. Raghu Parupudi, Harsha Ponnada, Aditi Kaushal, S. Shria Parupudi 외 arxiv

Reference-free evaluation of large language model (LLM) creativity relies on perplexity, entropy, and top-1 margin. We show that a much stronger signal lives one step earlier in the pipeline: in how sampling temperature …

Benchmarking Large Language Model Volatility

2023-11-26 · Boyang Yu

The impact of non-deterministic outputs from Large Language Models (LLMs) is not well examined for financial text understanding tasks. Through a compelling case study on investing in the US equity market via news sentime…

BenchmarkingDecision MakingDecoderLanguage Modeling+6