paper-with-me

Papers

Tree-Structured Reinforcement Learning for Sequential Object Localization

2017-03-08 · NeurIPS 2016 12 · Zequn Jie, Xiaodan Liang, Jiashi Feng, Xiaojie Jin, Wen Feng Lu, Shuicheng Yan

Existing object proposal algorithms usually search for possible object regions over multiple locations and scales separately, which ignore the interdependency among different objects and deviate from the human perception procedure. To incorporate global interdependency between objects into object localization, we propose an effective Tree-structured Reinforcement Learning (Tree-RL) approach to sequentially search for objects by fully exploiting both the current observation and historical search paths. The Tree-RL approach learns multiple searching policies through maximizing the long-term reward that reflects localization accuracies over all the objects. Starting with taking the entire image as a proposal, the Tree-RL approach allows the agent to sequentially discover multiple objects via a tree-structured traversing scheme. Allowing multiple near-optimal policies, Tree-RL offers more diversity in search paths and is able to find multiple objects with a single feed-forward pass. Therefore, Tree-RL can better cover different objects with various scales which is quite appealing in the context of object proposal. Experiments on PASCAL VOC 2007 and 2012 validate the effectiveness of the Tree-RL, which can achieve comparable recalls with current object proposal algorithms via much fewer candidate windows.

📄 PDF Abstract BibTeX arXiv:1703.02710

Code (0)

등록된 구현이 없습니다.

Tasks

DiversityObjectObject Localizationreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Learning to Compose Words into Sentences with Reinforcement Learning

2016-11-28 · Dani Yogatama, Phil Blunsom, Chris Dyer, Edward Grefenstette 외

We use reinforcement learning to learn tree-structured neural networks for computing representations of natural language sentences. In contrast with prior work on tree-structured models in which the trees are either prov…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Tree-Structured Policy based Progressive Reinforcement Learning for Temporally Language Grounding in Video

2020-01-18 · Jie Wu, Guanbin Li, Si Liu, Liang Lin

Temporally language grounding in untrimmed videos is a newly-raised task in video understanding. Most of the existing methods suffer from inferior efficiency, lacking interpretability, and deviating from the human percep…

Decision Makingreinforcement-learningReinforcement LearningReinforcement Learning (RL)+1

Temporal Attention for Cross-View Sequential Image Localization

2024-08-28 · Dong Yuan, Frederic Maire, Feras Dayoub

This paper introduces a novel approach to enhancing cross-view localization, focusing on the fine-grained, sequential localization of street-view images within a single known satellite image patch, a significant departur…

Image RetrievalRetrieval

Improving Object Detection with Deep Convolutional Networks via Bayesian Optimization and Structured Prediction

2015-04-13 · CVPR 2015 6 · Yuting Zhang, Kihyuk Sohn, Ruben Villegas, Gang Pan 외

Object detection systems based on the deep convolutional neural network (CNN) have recently made ground- breaking advances on several object detection benchmarks. While the features learned by these high-capacity neural …

Bayesian OptimizationObjectobject-detectionObject Detection+1

Attention-driven Tree-structured Convolutional LSTM for High Dimensional Data Understanding

2019-01-29 · Bin Kong, Xin Wang, Junjie Bai, Yi Lu 외

Modeling the sequential information of image sequences has been a vital step of various vision tasks and convolutional long short-term memory (ConvLSTM) has demonstrated its superb performance in such spatiotemporal prob…

Vocal Bursts Intensity Prediction