paper-with-me

Papers

The 8th AI City Challenge

2024-04-15 · Shuo Wang, David C. Anastasiu, Zheng Tang, Ming-Ching Chang, Yue Yao, Liang Zheng, Mohammed Shaiqur Rahman, Meenakshi S. Arya, Anuj Sharma, Pranamesh Chakraborty, Sanjita Prajapati, Quan Kong, Norimasa Kobori, Munkhjargal Gochoo, Munkh-Erdene Otgonbold, Fady Alnajjar, Ganzorig Batnasan, Ping-Yang Chen, Jun-Wei Hsieh, Xunlei Wu, Sameer Satish Pusegaonkar, Yizhou Wang, Sujit Biswas, Rama Chellappa

The eighth AI City Challenge highlighted the convergence of computer vision and artificial intelligence in areas like retail, warehouse settings, and Intelligent Traffic Systems (ITS), presenting significant research opportunities. The 2024 edition featured five tracks, attracting unprecedented interest from 726 teams in 47 countries and regions. Track 1 dealt with multi-target multi-camera (MTMC) people tracking, highlighting significant enhancements in camera count, character number, 3D annotation, and camera matrices, alongside new rules for 3D tracking and online tracking algorithm encouragement. Track 2 introduced dense video captioning for traffic safety, focusing on pedestrian accidents using multi-camera feeds to improve insights for insurance and prevention. Track 3 required teams to classify driver actions in a naturalistic driving analysis. Track 4 explored fish-eye camera analytics using the FishEye8K dataset. Track 5 focused on motorcycle helmet rule violation detection. The challenge utilized two leaderboards to showcase methods, with participants setting new benchmarks, some surpassing existing state-of-the-art achievements.

📄 PDF Abstract BibTeX arXiv:2404.09432

Code (0)

등록된 구현이 없습니다.

Tasks

Dense Video CaptioningVideo Captioning

Similar Papers 제목 키워드 기반

CityTrack: Improving City-Scale Multi-Camera Multi-Target Tracking by Location-Aware Tracking and Box-Grained Matching

2023-07-06 · Jincheng Lu, Xipeng Yang, Jin Ye, Yifu Zhang 외

Multi-Camera Multi-Target Tracking (MCMT) is a computer vision technique that involves tracking multiple targets simultaneously across multiple cameras. MCMT in urban traffic visual analysis faces great challenges due to…

Improving Acoustic Scene Classification with City Features

2025-03-21 · Yiqiang Cai, Yizhou Tan, Shengchen Li, Xi Shao 외

Acoustic scene recordings are often collected from a diverse range of cities. Most existing acoustic scene classification (ASC) approaches focus on identifying common acoustic scene patterns across cities to enhance gene…

Acoustic Scene ClassificationClassificationKnowledge DistillationScene Classification

Toxicity Prediction using Deep Learning

2015-03-04 · Thomas Unterthiner, Andreas Mayr, Günter Klambauer, Sepp Hochreiter

Everyday we are exposed to various chemicals via food additives, cleaning and cosmetic products and medicines -- and some of them might be toxic. However testing the toxicity of all existing compounds by biological exper…

Deep LearningPrediction

WildCity: A Real-World City-Scale Testbed for Rendering, Simulation, and Spatial Intelligence

2026-07-07 · Xiangyu Han, Mengyu Yang, Jiaqi Li, Bowen Chang 외 arxiv

Humans can navigate an unfamiliar city and gradually form a coherent spatial mental map spanning tens of square kilometers. Can AI build spatial representations at a comparable scale? Although recent foundation models ha…

Camera-based vehicle velocity estimation from monocular video

2018-02-20 · Moritz Kampelmühler, Michael G. Müller, Christoph Feichtenhofer

This paper documents the winning entry at the CVPR2017 vehicle velocity estimation challenge. Velocity estimation is an emerging task in autonomous driving which has not yet been thoroughly explored. The goal is to estim…

Autonomous DrivingCPUOptical Flow Estimation