paper-with-me

홈 › Papers

RoadText-1K: Text Detection & Recognition Dataset for Driving Videos

2020-05-19 · Sangeeth Reddy, Minesh Mathew, Lluis Gomez, Marcal Rusinol, Dimosthenis Karatzas., C. V. Jawahar

Perceiving text is crucial to understand semantics of outdoor scenes and hence is a critical requirement to build intelligent systems for driver assistance and self-driving. Most of the existing datasets for text detection and recognition comprise still images and are mostly compiled keeping text in mind. This paper introduces a new "RoadText-1K" dataset for text in driving videos. The dataset is 20 times larger than the existing largest dataset for text in videos. Our dataset comprises 1000 video clips of driving without any bias towards text and with annotations for text bounding boxes and transcriptions in every frame. State of the art methods for text detection, recognition and tracking are evaluated on the new dataset and the results signify the challenges in unconstrained driving videos compared to existing datasets. This suggests that RoadText-1K is suited for research and development of reading systems, robust enough to be incorporated into more complex downstream tasks like driver assistance and self-driving. The dataset can be found at http://cvit.iiit.ac.in/research/projects/cvit-projects/roadtext-1k

📄 PDF Abstract BibTeX arXiv:2005.09496

Code (0)

등록된 구현이 없습니다.

Tasks

Text Detection

Similar Papers 제목 키워드 기반

Reading Between the Lanes: Text VideoQA on the Road

2023-07-08 · George Tom, Minesh Mathew, Sergi Garcia, Dimosthenis Karatzas 외

Text and signs around roads provide crucial information for drivers, vital for safe navigation and situational awareness. Scene text recognition in motion is a challenging problem, while textual cues typically appear for…

Question AnsweringScene Text RecognitionVideo Question Answering

TraRA: Trajectory-level Recognition Aggregation for Video Text Spotting in Urban Surveillance

2026-06-05 · Duc Tri Tran, Trung Thanh Nguyen, Vijay John, Phi Le Nguyen 외 arxiv

Video Text Spotting (VTS) is essential for urban surveillance and intelligent transportation systems, enabling automated reading of street signs, vehicle markings, and scene text in video streams. However, reliable recog…

Text Spotting

TraffSign: Multilingual Traffic Signboard Text Detection and Recognition for Urdu and English

2022-05-18 · Document Analysis Systems 2022 5 · Muhammad Atif Butt, Adnan Ul-Hasan, and Faisal Shafait

Scene-text detection and recognition methods have demonstrated remarkable performance on standard benchmark datasets. These methods can be utilized in human-driven/self-driving cars to perform navigation assistance throu…

Scene Text DetectionSelf-Driving CarsText Detection

The First Swahili Language Scene Text Detection and Recognition Dataset

2024-05-19 · Fadila Wendigoundi Douamba, Jianjun Song, Ling Fu, Yuliang Liu 외

Scene text recognition is essential in many applications, including automated translation, information retrieval, driving assistance, and enhancing accessibility for individuals with visual impairments. Much research has…

Information RetrievalScene Text DetectionScene Text RecognitionText Detection

Video-based Traffic Light Recognition by Rockchip RV1126 for Autonomous Driving

2025-03-31 · Miao Fan, Xuxu Kong, Shengtong Xu, Haoyi Xiong 외

Real-time traffic light recognition is fundamental for autonomous driving safety and navigation in urban environments. While existing approaches rely on single-frame analysis from onboard cameras, they struggle with comp…

Autonomous Driving