paper-with-me

홈 › Papers

How Controllable Are Large Language Models? A Unified Evaluation across Behavioral Granularities

2026-03-03 · Ziwen Xu, Kewei Xu, Haoming Xu, Haiwen Hong, Longtao Huang, Hui Xue, Ningyu Zhang, Yongliang Shen, Guozhou Zheng, Huajun Chen, Shumin Deng arxiv

Large Language Models (LLMs) are increasingly deployed in socially sensitive domains, yet their unpredictable behaviors, ranging from misaligned intent to inconsistent personality, pose significant risks. We introduce SteerEval, a hierarchical benchmark for evaluating LLM controllability across three domains: language features, sentiment, and personality. Each domain is structured into three specification levels: L1 (what to express), L2 (how to express), and L3 (how to instantiate), connecting high-level behavioral intent to concrete textual output. Using SteerEval, we systematically evaluate contemporary steering methods, revealing that control often degrades at finer-grained levels. Our benchmark offers a principled and interpretable framework for safe and controllable LLM behavior, serving as a foundation for future research.

📄 PDF Abstract BibTeX arXiv:2603.02578

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

FlexCAD: Unified and Versatile Controllable CAD Generation with Fine-tuned Large Language Models

2024-11-05 · Zhanwei Zhang, Shizhao Sun, Wenxiao Wang, Deng Cai 외

Recently, there is a growing interest in creating computer-aided design (CAD) models based on user intent, known as controllable CAD generation. Existing work offers limited controllability and needs separate models for …

REACT: Representation Extraction And Controllable Tuning to Overcome Overfitting in LLM Knowledge Editing

2025-05-25 · Haitian Zhong, Yuhuan Liu, Ziyang Xu, Guofan Liu 외

Large language model editing methods frequently suffer from overfitting, wherein factual updates can propagate beyond their intended scope, overemphasizing the edited target even when it's contextually inappropriate. To …

knowledge editingLanguage ModelingLanguage ModellingLarge Language Model+1

ALM2Vec: Learning Audio Embeddings for Universal Audio Retrieval with Large Audio-Language Models

2026-06-27 · Fengjie Lu, Chenang Jiang, Jiarui Hai, Helin Wang 외 arxiv

Recent advances in language--audio retrieval have been largely driven by contrastive dual-encoder architectures that align audio and text in a shared embedding space. While effective, existing retrieval embeddings are pr…

Question Answering

JoyAI-Image: Awaking Spatial Intelligence in Unified Multimodal Understanding and Generation

2026-05-05 · Lin Song, Wenbo Li, Guoqing Ma, Wei Tang 외 arxiv

We present JoyAI-Image, a unified multimodal foundation model for visual understanding, text-to-image generation, and instruction-guided image editing. JoyAI-Image couples a spatially enhanced Multimodal Large Language M…

Text-to-Image GenerationImage Editing

Unified Personalized Understanding, Generating and Editing

2026-01-11 · Yu Zhong, Tianwei Lin, Ruike Zhu, Yuqian Yuan 외 arxiv

Unified large multimodal models (LMMs) have achieved remarkable progress in general-purpose multimodal understanding and generation. However, they still operate under a ``one-size-fits-all'' paradigm and struggle to mode…

Image Editing