Video Text Localization using Wavelet and Shearlet Transforms
Text in video is useful and important in indexing and retrieving the video documents efficiently and accurately. In this paper, we present a new method of text detection using a combined dictionary consisting of wavelets and a recently introduced transform called shearlets. Wavelets provide optimally sparse expansion for point-like structures and shearlets provide optimally sparse expansions for curve-like structures. By combining these two features we have computed a high frequency sub-band to brighten the text part. Then K-means clustering is used for obtaining text pixels from the Standard Deviation (SD) of combined coefficient of wavelets and shearlets as well as the union of wavelets and shearlets features. Text parts are obtained by grouping neighboring regions based on geometric properties of the classified output frame of unsupervised K-means classification. The proposed method tested on a standard as well as newly collected database shows to be superior to some existing methods.
Code (0)
등록된 구현이 없습니다.
Tasks
ClusteringText DetectionSimilar Papers 제목 키워드 기반
CoShNet: A Hybrid Complex Valued Neural Network using Shearlets
In a hybrid neural network, the expensive convolutional layers are replaced by a non-trainable fixed transform with a great reduction in parameters. In previous works, good results were obtained by replacing the convolut…
Superresolution of Noisy Remotely Sensed Images Through Directional Representations
We develop an algorithm for single-image superresolution of remotely sensed data, based on the discrete shearlet transform. The shearlet transform extracts directional features of signals, and is known to provide near-op…
DenoisingEdge DetectionSSIMShearlet-Based Detection of Flame Fronts
Identifying and characterizing flame fronts is the most common task in the computer-assisted analysis of data obtained from imaging techniques such as planar laser-induced fluorescence (PLIF), laser Rayleigh scattering (…
Line DetectionShearlet-based compressed sensing for fast 3D cardiac MR imaging using iterative reweighting
High-resolution three-dimensional (3D) cardiovascular magnetic resonance (CMR) is a valuable medical imaging technique, but its widespread application in clinical practice is hampered by long acquisition times. Here we p…
Anatomycompressed sensingImage ReconstructionDiscrete Wavelet Transform and Gradient Difference based approach for text localization in videos
The text detection and localization is important for video analysis and understanding. The scene text in video contains semantic information and thus can contribute significantly to video retrieval and understanding. How…
RetrievalText DetectionVideo Retrieval