paper-with-me

홈 › Papers

1st Place Solution to ECCV 2022 Challenge on Out of Vocabulary Scene Text Understanding: End-to-End Recognition of Out of Vocabulary Words

2022-09-01 · Zhangzi Zhu, Chuhui Xue, Yu Hao, Wenqing Zhang, Song Bai

Scene text recognition has attracted increasing interest in recent years due to its wide range of applications in multilingual translation, autonomous driving, etc. In this report, we describe our solution to the Out of Vocabulary Scene Text Understanding (OOV-ST) Challenge, which aims to extract out-of-vocabulary (OOV) words from natural scene images. Our oCLIP-based model achieves 28.59\% in h-mean which ranks 1st in end-to-end OOV word recognition track of OOV Challenge in ECCV2022 TiE Workshop.

📄 PDF Abstract BibTeX arXiv:2209.00224

Code (0)

등록된 구현이 없습니다.

Tasks

Autonomous DrivingScene Text RecognitionTranslation

Similar Papers 제목 키워드 기반

Runner-Up Solution to ECCV 2022 Challenge on Out of Vocabulary Scene Text Understanding: Cropped Word Recognition

2022-08-04 · Zhangzi Zhu, Yu Hao, Wenqing Zhang, Chuhui Xue 외

This report presents our 2nd place solution to ECCV 2022 challenge on Out-of-Vocabulary Scene Text Understanding (OOV-ST) : Cropped Word Recognition. This challenge is held in the context of ECCV 2022 workshop on Text in…

First Place Solution to the ECCV 2024 BRAVO Challenge: Evaluating Robustness of Vision Foundation Models for Semantic Segmentation

2024-09-25 · Tommie Kerssies, Daan de Geus, Gijs Dubbelman

In this report, we present the first place solution to the ECCV 2024 BRAVO Challenge, where a model is trained on Cityscapes and its robustness is evaluated on several out-of-distribution datasets. Our solution leverages…

DecoderSemantic Segmentation

2nd Place Solution to ECCV 2020 VIPriors Object Detection Challenge

2020-07-17 · Yinzheng Gu, Yihan Pan, Shi-Zhe Chen

In this report, we descibe our approach to the ECCV 2020 VIPriors Object Detection Challenge which took place from March to July in 2020. We show that by using state-of-the-art data augmentation strategies, model designs…

Data Augmentationobject-detectionObject DetectionTransfer Learning

First Place Solution to the ECCV 2024 ROAD++ Challenge @ ROAD++ Spatiotemporal Agent Detection 2024

2024-10-30 · Tengfei Zhang, Heng Zhang, Ruyang Li, Qi Deng 외

This report presents our team's solutions for the Track 1 of the 2024 ECCV ROAD++ Challenge. The task of Track 1 is spatiotemporal agent detection, which aims to construct an "agent tube" for road agents in consecutive v…

Data Augmentationobject-detectionObject Detection

Vision-Language Adaptive Mutual Decoder for OOV-STR

2022-09-02 · Jinshui Hu, Chenyu Liu, Qiandong Yan, Xuyang Zhu 외

Recent works have shown huge success of deep learning models for common in vocabulary (IV) scene text recognition. However, in real-world scenarios, out-of-vocabulary (OOV) words are of great importance and SOTA recognit…

DecoderLanguage ModelingLanguage ModellingRepresentation Learning+1