Situation Recognition
1개 벤치마크 · 논문 15편 · 이 태스크의 논문 보기 →
Benchmarks
imSitu
Most implemented
Collaborative Transformers for Grounded Situation Recognition
Commonly Uncommon: Semantic Sparsity in Situation Recognition
ClipSitu: Effectively Leveraging CLIP for Conditional Predictions in Situation Recognition
Rethinking the Two-Stage Framework for Grounded Situation Recognition
Grounded Situation Recognition with Transformers
Attention-Based Context Aware Reasoning for Situation Recognition
Papers
One Identity, Many Roles: Multimodal Entity Coreference for Enhanced Video Situation Recognition
Video Situation Recognition (VidSitu) addresses the challenging problem of "who did what to whom, with what, how, and where" in a video. It tests thorough video understanding by requiring identification of salient action…
Situation RecognitionVisual GroundingKIRETT -- A wearable device to support rescue operations using artificial intelligence to improve first aid
This short paper presents first steps in the scientific part of the KIRETT project, which aims to improve first aid during rescue operations using a wearable device. The wearable is used for computer-aided situation reco…
Situation RecognitionThe Demon is in Ambiguity: Revisiting Situation Recognition with Single Positive Multi-Label Learning
Context recognition (SR) is a fundamental task in computer vision that aims to extract structured semantic summaries from images by identifying key events and their associated entities. Specifically, given an input image…
Semantic Role LabelingSituation RecognitionMulti-Label LearningDynamic Scene Understanding from Vision-Language Representations
Images depicting complex, dynamic scenes are challenging to parse automatically, requiring both high-level comprehension of the overall situation and fine-grained identification of participating entities and their intera…
Grounded Situation RecognitionHuman-Human Interaction RecognitionHuman Interaction RecognitionHuman-Object Interaction Detection+2ClipSitu: Effectively Leveraging CLIP for Conditional Predictions in Situation Recognition
Situation Recognition is the task of generating a structured summary of what is happening in an image using an activity verb and the semantic roles played by actors and objects. In this task, the same activity verb can d…
Grounded Situation RecognitionSituation RecognitionCollaborative Transformers for Grounded Situation Recognition
Grounded situation recognition is the task of predicting the main activity, entities playing certain roles within the activity, and bounding-box groundings of the entities in the given image. To effectively deal with thi…
Grounded Situation RecognitionImage ClassificationObject DetectionScene Understanding+3