paper-with-me

홈 › Papers

Extracting Latent Attributes from Video Scenes Using Text as Background Knowledge

2014-08-01 · SEMEVAL 2014 8 · Anh Tran, Mihai Surdeanu, Paul Cohen
📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Coreference ResolutionInformation Retrieval

Similar Papers 제목 키워드 기반

Complex Event Detection via Multi-source Video Attributes

2013-06-01 · CVPR 2013 6 · Zhigang Ma, Yi Yang, Zhongwen Xu, Shuicheng Yan 외

Complex events essentially include human, scenes, objects and actions that can be summarized by visual attributes, so leveraging relevant attributes properly could be helpful for event detection. Many works have exploite…

Event Detection

CoNeRF: Controllable Neural Radiance Fields

2021-12-03 · CVPR 2022 1 · Kacper Kania, Kwang Moo Yi, Marek Kowalski, Tomasz Trzciński 외

We extend neural 3D representations to allow for intuitive and interpretable user control beyond novel view rendering (i.e. camera control). We allow the user to annotate which part of the scene one wishes to control wit…

3D Face Modelling3D ReconstructionAttributeFew-Shot Learning

Exploring MLLM-Diffusion Information Transfer with MetaCanvas

2025-12-12 · Han Lin, Xichen Pan, Ziqi Huang, Ji Hou 외 arxiv

Multimodal learning has rapidly advanced visual understanding, largely via multimodal large language models (MLLMs) that use powerful LLMs as cognitive cores. In visual generation, however, these powerful core models are…

Text-to-Image GenerationVideo Generation

Pre-training Contextualized World Models with In-the-wild Videos for Reinforcement Learning

2023-05-29 · NeurIPS 2023 11 · Jialong Wu, Haoyu Ma, Chaoyi Deng, Mingsheng Long

Unsupervised pre-training methods utilizing large and diverse datasets have achieved tremendous success across a range of domains. Recent work has investigated such unsupervised pre-training methods for model-based reinf…

Autonomous DrivingDecoderModel-based Reinforcement LearningTransfer Learning+2

SIMONe: View-Invariant, Temporally-Abstracted Object Representations via Unsupervised Video Decomposition

2021-06-07 · NeurIPS 2021 12 · Rishabh Kabra, Daniel Zoran, Goker Erdogan, Loic Matthey 외

To help agents reason about scenes in terms of their building blocks, we wish to extract the compositional structure of any given scene (in particular, the configuration and characteristics of objects comprising the scen…

Instance SegmentationObjectSemantic Segmentation