paper-with-me

Papers

Video Description using Bidirectional Recurrent Neural Networks

2016-04-12 · Álvaro Peris, Marc Bolaños, Petia Radeva, Francisco Casacuberta

Although traditionally used in the machine translation field, the encoder-decoder framework has been recently applied for the generation of video and image descriptions. The combination of Convolutional and Recurrent Neural Networks in these models has proven to outperform the previous state of the art, obtaining more accurate video descriptions. In this work we propose pushing further this model by introducing two contributions into the encoding stage. First, producing richer image representations by combining object and location information from Convolutional Neural Networks and second, introducing Bidirectional Recurrent Neural Networks for capturing both forward and backward temporal relationships in the input frames.

📄 PDF Abstract BibTeX arXiv:1604.03390

Code (1)

lvapeab/ABiViRNet 공식 구현 tf

Tasks

DecoderText GenerationTranslationVideo CaptioningVideo Description

Similar Papers 제목 키워드 기반

WMT 2016 Multimodal Translation System Description based on Bidirectional Recurrent Neural Networks with Double-Embeddings

2016-08-01 · WS 2016 8 · Sergio Rodr{\'\i}guez Guasch, Marta R. Costa-juss{\`a}
Image CaptioningLanguage ModelingLanguage ModellingMachine Translation+2

BIRNAT: Bidirectional Recurrent Neural Networks with Adversarial Training for Video Snapshot Compressive Imaging

2020-08-01 · ECCV 2020 8 · Ziheng Cheng, Ruiying Lu, Zhengjue Wang, Hao Zhang 외

We consider the problem of video snapshot compressive imaging (SCI), where multiple high-speed frames are coded by different masks and then summed to a single measurement. This measurement and the modulation masks are fe…

Hierarchical Boundary-Aware Neural Encoder for Video Captioning

2016-11-28 · CVPR 2017 7 · Lorenzo Baraldi, Costantino Grana, Rita Cucchiara

The use of Recurrent Neural Networks for video captioning has recently gained a lot of attention, since they can be used both to encode the input video and to generate the corresponding description. In this paper, we pre…

DecoderVideo CaptioningVideo Description

Bidirectional Long-Short Term Memory for Video Description

2016-06-15 · Yi Bin, Yang Yang, Zi Huang, Fumin Shen 외

Video captioning has been attracting broad research attention in multimedia community. However, most existing approaches either ignore temporal information among video frames or just employ local contextual temporal know…

Language ModelingLanguage ModellingVideo CaptioningVideo Description

Bidirectional Recurrent Convolutional Networks for Multi-Frame Super-Resolution

2015-12-01 · NeurIPS 2015 12 · Yan Huang, Wei Wang, Liang Wang

Super resolving a low-resolution video is usually handled by either single-image super-resolution (SR) or multi-frame SR. Single-Image SR deals with each video frame independently, and ignores intrinsic temporal dependen…

Image Super-ResolutionMulti-Frame Super-ResolutionOptical Flow EstimationSuper-Resolution+2