paper-with-me

Papers

Actions Speak Louder than Words: Trillion-Parameter Sequential Transducers for Generative Recommendations

2024-02-27 · Jiaqi Zhai, Lucy Liao, Xing Liu, Yueming Wang, Rui Li, Xuan Cao, Leon Gao, Zhaojie Gong, Fangda Gu, Michael He, Yinghai Lu, Yu Shi

Large-scale recommendation systems are characterized by their reliance on high cardinality, heterogeneous features and the need to handle tens of billions of user actions on a daily basis. Despite being trained on huge volume of data with thousands of features, most Deep Learning Recommendation Models (DLRMs) in industry fail to scale with compute. Inspired by success achieved by Transformers in language and vision domains, we revisit fundamental design choices in recommendation systems. We reformulate recommendation problems as sequential transduction tasks within a generative modeling framework ("Generative Recommenders"), and propose a new architecture, HSTU, designed for high cardinality, non-stationary streaming recommendation data. HSTU outperforms baselines over synthetic and public datasets by up to 65.8% in NDCG, and is 5.3x to 15.2x faster than FlashAttention2-based Transformers on 8192 length sequences. HSTU-based Generative Recommenders, with 1.5 trillion parameters, improve metrics in online A/B tests by 12.4% and have been deployed on multiple surfaces of a large internet platform with billions of users. More importantly, the model quality of Generative Recommenders empirically scales as a power-law of training compute across three orders of magnitude, up to GPT-3/LLaMa-2 scale, which reduces carbon footprint needed for future model developments, and further paves the way for the first foundational models in recommendations.

📄 PDF Abstract BibTeX arXiv:2402.17152

Code (12)

facebookresearch/generative-recommenders 공식 구현 pytorch
1hb6s7t/Gr-HSTU mindspore
MindSpore-scientific-2/code-9/tree/main/GR-HSTU mindspore
MindSpore-scientific/code-12/tree/main/GR-HSTU mindspore
MindSpore-scientific/code-4/tree/main/GR-HSTU mindspore
alibaba/TorchEasyRec pytorch
bailuding/rails pytorch
foreverYoungGitHub/generative-recommenders-pl pytorch
glb400/Toy-RecLM pytorch
pwc-1/Paper-9/tree/main/4/GR-HSTU mindspore
roman-dusek/GR-HSTU pytorch
snapfinger/hstu-blair pytorch

Tasks

Recommendation Systems

Similar Papers 제목 키워드 기반

My Actions Speak Louder Than Your Words: When User Behavior Predicts Their Beliefs about Agents' Attributes

2023-01-21 · Nikolos Gurney, David Pynadath, Ning Wang

An implicit expectation of asking users to rate agents, such as an AI decision-aid, is that they will use only relevant information -- ask them about an agent's benevolence, and they should consider whether or not it was…

The Effect of eHealth Training on Dysarthric Speech

2022-06-01 · RaPID (LREC) 2022 6 · Chiara Pesenti, Loes Van Bemmel, Roeland van Hout, Helmer Strik

In the current study on dysarthric speech, we investigate the effect of web-based treatment, and whether there is a difference between content and function words. Since the goal of the treatment is to speak louder, witho…

Riemannian Geometry Speaks Louder Than Words: From Graph Foundation Model to Next-Generation Graph Intelligence

2026-03-23 · Philip S. Yu, Li Sun arxiv

Graphs provide a natural description of the complex relationships among objects, and play a pivotal role in communications, transportation, social computing, the life sciences, etc. Currently, there is strong agreement t…

Graph Learning

Actions Speak Louder Than (Pass)words: Passive Authentication of Smartphone Users via Deep Temporal Features

2019-01-16 · Debayan Deb, Arun Ross, Anil K. Jain, Kwaku Prakah-Asante 외

Prevailing user authentication schemes on smartphones rely on explicit user interaction, where a user types in a passcode or presents a biometric cue such as face, fingerprint, or iris. In addition to being cumbersome an…

Actions Speak Louder than Words: Agent Decisions Reveal Implicit Biases in Language Models

2025-01-29 · YuXuan Li, Hirokazu Shirado, Sauvik Das

While advances in fairness and alignment have helped mitigate overt biases exhibited by large language models (LLMs) when explicitly prompted, we hypothesize that these models may still exhibit implicit biases when simul…

Decision MakingFairness