paper-with-me

Papers

SPECTRA: Sparse Entity-centric Transitions

2019-09-25 · Rim Assouel, Yoshua Bengio

Learning an agent that interacts with objects is ubiquituous in many RL tasks. In most of them the agent's actions have sparse effects : only a small subset of objects in the visual scene will be affected by the action taken. We introduce SPECTRA, a model for learning slot-structured transitions from raw visual observations that embodies this sparsity assumption. Our model is composed of a perception module that decomposes the visual scene into a set of latent objects representations (i.e. slot-structured) and a transition module that predicts the next latent set slot-wise and in a sparse way. We show that learning a perception module jointly with a sparse slot-structured transition model not only biases the model towards more entity-centric perceptual groupings but also enables intrinsic exploration strategy that aims at maximizing the number of objects changed in the agent’s trajectory.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Entity Retrieval for Answering Entity-Centric Questions

2024-08-05 · Hassan S. Shavarani, Anoop Sarkar

The similarity between the question and indexed documents is a crucial factor in document retrieval for retrieval-augmented question answering. Although this is typically the only method for obtaining the relevant docume…

Entity RetrievalQuestion AnsweringRetrieval

Sparse Joint Transmission for Cell-Free Massive MIMO: A Sparse PCA Approach

2019-12-11 · Deokhwan Han, Jeonghun Park, Namyoon Lee

Cell-free massive multiple-input multiple-output (MIMO) is a promising cellular network. In this network, a large number of distributed and multi-antenna access points (APs) jointly serve many single antenna users using …

CineOrchestra: Unified Entity-Centric Conditioning for Cinematic Video Generation

2026-06-11 · Sharath Girish, Tsai-Shien Chen, Zhikang Dong, Mukesh Singhal 외 arxiv

Cinematic video depicts multiple subjects acting or interacting at specific moments, captured with deliberate camera movement, and stitched together by shot transitions. Together, these elements demand a level of fine-gr…

Video Generation

Unveiling the Significance of Toddler-Inspired Reward Transition in Goal-Oriented Reinforcement Learning

2024-03-11 · Junseok Park, Yoonsung Kim, Hee Bin Yoo, Min Whoo Lee 외

Toddlers evolve from free exploration with sparse feedback to exploiting prior experiences for goal-directed learning with denser rewards. Drawing inspiration from this Toddler-Inspired Reward Transition, we set out to e…

Reinforcement Learning (RL)

NarrativeTrack: Evaluating Entity-Centric Reasoning for Narrative Understanding

2026-01-03 · Hyeonjeong Ha, Jinjin Ge, Bo Feng, Kaixin Ma 외 arxiv

Multimodal large language models (MLLMs) have achieved impressive progress in vision-language reasoning, yet their ability to understand temporally unfolding narratives in videos remains underexplored. True narrative und…