paper-with-me

Papers

An Improved Video Analysis using Context based Extension of LSH

2017-05-10 · Angana Chakraborty, Sanghamitra Bandyopadhyay

Locality Sensitive Hashing (LSH) based algorithms have already shown their promise in finding approximate nearest neighbors in high dimen- sional data space. However, there are certain scenarios, as in sequential data, where the proximity of a pair of points cannot be captured without considering their surroundings or context. In videos, as for example, a particular frame is meaningful only when it is seen in the context of its preceding and following frames. LSH has no mechanism to handle the con- texts of the data points. In this article, a novel scheme of Context based Locality Sensitive Hashing (conLSH) has been introduced, in which points are hashed together not only based on their closeness, but also because of similar context. The contribution made in this article is three fold. First, conLSH is integrated with a recently proposed fast optimal sequence alignment algorithm (FOGSAA) using a layered approach. The resultant method is applied to video retrieval for extracting similar sequences. The pro- posed algorithm yields more than 80% accuracy on an average in different datasets. It has been found to save 36.3% of the total time, consumed by the exhaustive search. conLSH reduces the search space to approximately 42% of the entire dataset, when compared with an exhaustive search by the aforementioned FOGSAA, Bag of Words method and the standard LSH implementations. Secondly, the effectiveness of conLSH is demon- strated in action recognition of the video clips, which yields an average gain of 12.83% in terms of classification accuracy over the state of the art methods using STIP descriptors. The last but of great significance is that this article provides a way of automatically annotating long and composite real life videos. The source code of conLSH is made available at http://www.isical.ac.in/~bioinfo_miu/conLSH/conLSH.html

📄 PDF Abstract BibTeX arXiv:1705.03933

Code (0)

등록된 구현이 없습니다.

Tasks

Action RecognitionRetrievalTemporal Action LocalizationVideo Retrieval

Similar Papers 제목 키워드 기반

VideoRun2D Demo: Markerless Body Tracking for Biomechanical Analysis of Running

2026-08-19 · Luis F. Gomez, Julian Fierrez, Roberto Daza, Ruben Tolosana 외 arxiv

Human pose estimation has advanced significantly due to the development of deep learning models, increased data availability, and improved computing resources. These developments have led to highly accurate body tracking…

Outlier DetectionPose Estimation

Analysis of displacement compensation methods for wavelet lifting of medical 3-D thorax CT volume data

2023-01-11 · Wolfgang Schnurrer, Jürgen Seiler, Eugen Wige, André Kaup

A huge advantage of the wavelet transform in image and video compression is its scalability. Wavelet-based coding of medical computed tomography (CT) data becomes more and more popular. While much effort has been spent o…

Computed Tomography (CT)Video Compression

High Fidelity Interactive Video Segmentation Using Tensor Decomposition Boundary Loss Convolutional Tessellations and Context Aware Skip Connections

2020-11-23 · Anthony D. Rhodes, Manan Goel

We provide a high fidelity deep learning algorithm (HyperSeg) for interactive video segmentation tasks using a convolutional network with context-aware skip connections, and compressed, hypercolumn image features combine…

Interactive SegmentationSegmentationTensor DecompositionVideo Segmentation+1

ConTra: (Con)text (Tra)nsformer for Cross-Modal Video Retrieval

2022-10-09 · Adriano Fragomeni, Michael Wray, Dima Damen

In this paper, we re-examine the task of cross-modal clip-sentence retrieval, where the clip is part of a longer untrimmed video. When the clip is short or visually ambiguous, knowledge of its local temporal context (i.e…

RetrievalSentenceSentence RetrievalVideo Retrieval

Improving Video Generation for Multi-functional Applications

2017-11-30 · Bernhard Kratzwald, Zhiwu Huang, Danda Pani Paudel, Acharya Dinesh 외

In this paper, we aim to improve the state-of-the-art video generative adversarial networks (GANs) with a view towards multi-functional applications. Our improved video GAN model does not separate foreground from backgro…

ColorizationFuture predictionVideo GenerationVideo Inpainting