paper-with-me

홈 › Papers

Spontaneous Emerging Preference in Two-tower Language Model

2022-10-13 · Zhengqi He, Taro Toyoizumi

The ever-growing size of the foundation language model has brought significant performance gains in various types of downstream tasks. With the existence of side-effects brought about by the large size of the foundation language model such as deployment cost, availability issues, and environmental cost, there is some interest in exploring other possible directions, such as a divide-and-conquer scheme. In this paper, we are asking a basic question: are language processes naturally dividable? We study this problem with a simple two-tower language model setting, where two language models with identical configurations are trained side-by-side cooperatively. With this setting, we discover the spontaneous emerging preference phenomenon, where some of the tokens are consistently better predicted by one tower while others by another tower. This phenomenon is qualitatively stable, regardless of model configuration and type, suggesting this as an intrinsic property of natural language. This study suggests that interesting properties of natural language are still waiting to be discovered, which may aid the future development of natural language processing techniques.

📄 PDF Abstract BibTeX arXiv:2210.07041

Code (0)

등록된 구현이 없습니다.

Tasks

Language ModelingLanguage ModellingVocal Bursts Valence Prediction

Similar Papers 제목 키워드 기반

SLIDE: Integrating Speech Language Model with LLM for Spontaneous Spoken Dialogue Generation

2025-01-01 · Haitian Lu, Gaofeng Cheng, Liuping Luo, Leying Zhang 외

Recently, ``textless" speech language models (SLMs) based on speech units have made huge progress in generating naturalistic speech, including non-verbal vocalizations. However, the generated speech samples often lack se…

Dialogue GenerationLanguage ModelingLanguage Modelling

TowerMind: A Tower Defence Game Learning Environment and Benchmark for LLM as Agents

2026-01-09 · Dawei Wang, Chengming Zhou, Di Zhao, Xinyuan Liu 외 arxiv

Recent breakthroughs in Large Language Models (LLMs) have positioned them as a promising paradigm for agents, with long-term planning and decision-making emerging as core general-purpose capabilities for adapting to dive…

Reinforcement Learning

Accuracy meets Diversity in a News Recommender System

2022-10-01 · COLING 2022 10 · Shaina Raza, Syed Raza Bashir, Usman Naseem

News recommender systems face certain challenges. These challenges arise due to evolving users’ preferences over dynamically created news articles. The diversity is necessary for a news recommender system to expose users…

ArticlesDiversityRecommendation Systems

Zero Shot on the Cold-Start Problem: Model-Agnostic Interest Learning for Recommender Systems

2021-08-31 · Philip J. Feng, Pingjun Pan, Tingting Zhou, Hongxiang Chen 외

User behavior has been validated to be effective in revealing personalized preferences for commercial recommendations. However, few user-item interactions can be collected for new users, which results in a null space for…

Recommendation Systems

Spontaneous Reward Hacking in Iterative Self-Refinement

2024-07-05 · Jane Pan, He He, Samuel R. Bowman, Shi Feng

Language models are capable of iteratively improving their outputs based on natural language feedback, thus enabling in-context optimization of user preference. In place of human users, a second language model can be use…

Language ModelingLanguage Modelling