paper-with-me

홈 › Papers

Compositional Literary Primitives in Instruction-Tuned LLMs: Cross-Architectural SAE Features for Self, Style, and Affect

2026-05-11 · Joao Paulo Cavalcante Presa, Savio Salvarino Teles de Oliveira arxiv

We characterize a compositional architecture of literary primitives in two instruction-tuned large language models (Llama 3.1 8B-Instruct and Gemma 2 9B-IT) via sparse autoencoders on mid-depth residual streams. Four feature classes emerge: naming-gates that promote lexical tokens of a target affect, an eleven-self cluster of first-person register features, stylistic register modulators (show-don't-tell and defamiliarization), and compositional emotions that arise only from multi-feature steering. Under a forced-choice 5-LLM judge panel applied to a 27-category emotion taxonomy (Cowen-Keltner), Llama reaches full 27/27 coverage by combining naming-gates, multi-feature recipes, and single self-feature steering; Gemma reaches 23/27 with adoration as the single residual strict-fail. Under random judging, the per-cell pass probability is on the order of $10^{-3}$ and the expected number of two-seed false-positive cells across the catalog is negligible, so the observed coverage is not consistent with chance. A cross-architectural asymmetry sits in the strict-versus-soft judge contrast: on the same generations, judges agree more often on Llama outputs than on Gemma outputs because Llama outputs name the target affect more directly while Gemma outputs evoke it through scene and imagery. Both architectures contain self-features that serve simultaneously as register markers and as emotion emitters, including a single most-RLHF-loaded self-feature per architecture that intensifies the institutional Helper-AI persona at one operating regime and produces affect-categorizable output at the same calibrated coefficient. Methodologically, the paper presents a three-stage validation pipeline (logit-lens, LLM-rate, 5-LLM judge) with documented anti-patterns; the total compute is single-GPU and about 15 minutes per emotion-feature discovery cycle.

📄 PDF Abstract BibTeX arXiv:2605.18808

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Evaluating Morphological Compositional Generalization in Large Language Models

2024-10-16 · Mete Ismayilzada, Defne Circi, Jonne Sälevä, Hale Sirin 외

Large language models (LLMs) have demonstrated significant progress in various natural language generation and understanding tasks. However, their linguistic generalization capabilities remain questionable, raising doubt…

Text Generation

Algorithmic Primitives and Compositional Geometry of Reasoning in Language Models

2025-10-13 · Samuel Lippl, Thomas McGee, Kimberly Lopez, Ziwen Pan 외 arxiv

How do latent and inference time computations enable large language models (LLMs) to solve multi-step reasoning? We introduce a framework for tracing and steering algorithmic primitives that underlie model reasoning. Our…

Named Entity Recognition in Historical Italian: The Case of Giacomo Leopardi's Zibaldone

2025-05-26 · Cristian Santini, Laura Melosi, Emanuele Frontoni

The increased digitization of world's textual heritage poses significant challenges for both computer science and literary studies. Overall, there is an urgent need of computational techniques able to adapt to the challe…

named-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)NER

LLMs Exhibit Significantly Lower Uncertainty in Creative Writing Than Professional Writers

2026-02-18 · Peiqi Sui arxiv

We argue that uncertainty is a key and understudied limitation of LLMs' performance in creative writing, which is often characterized as trite and cliché-ridden. Literary theory identifies uncertainty as a necessary cond…

A Modern Turkish Poet: Fine-Tuned GPT-2

2023-09-15 · 8th International Conference on Computer Science and Engineering (UBMK) 2023 9 · Uygar Kurt, Aykut Çayır

Generative tasks are getting more realistic thanks to the improvements in deep learning. Text generation is getting increasingly important as LLMs (large language models) get more advanced. Even though ChatGPT gave rise …

Text Generation