paper-with-me

홈 › Papers

POLO -- Point-based, multi-class animal detection

2024-10-15 · Giacomo May, Emanuele Dalsasso, Benjamin Kellenberger, Devis Tuia

Automated wildlife surveys based on drone imagery and object detection technology are a powerful and increasingly popular tool in conservation biology. Most detectors require training images with annotated bounding boxes, which are tedious, expensive, and not always unambiguous to create. To reduce the annotation load associated with this practice, we develop POLO, a multi-class object detection model that can be trained entirely on point labels. POLO is based on simple, yet effective modifications to the YOLOv8 architecture, including alterations to the prediction process, training losses, and post-processing. We test POLO on drone recordings of waterfowl containing up to multiple thousands of individual birds in one image and compare it to a regular YOLOv8. Our experiments show that at the same annotation cost, POLO achieves improved accuracy in counting animals in aerial imagery.

📄 PDF Abstract BibTeX arXiv:2410.11741

Code (0)

등록된 구현이 없습니다.

Tasks

object-detectionObject Detection

Methods 이 논문이 사용한 방법론

YOLOv8 설명 없음

Similar Papers 제목 키워드 기반

A Novel Dataset for Keypoint Detection of quadruped Animals from Images

2021-08-31 · Prianka Banik, Lin Li, Xishuang Dong

In this paper, we studied the problem of localizing a generic set of keypoints across multiple quadruped or four-legged animal species from images. Due to the lack of large scale animal keypoint dataset with ground truth…

Keypoint Detection

Interspecies Knowledge Transfer for Facial Keypoint Detection

2017-04-13 · CVPR 2017 7 · Maheen Rashid, Xiuye Gu, Yong Jae Lee

We present a method for localizing facial keypoints on animals by transferring knowledge gained from human faces. Instead of directly finetuning a network trained to detect keypoints on human faces to animal faces (which…

Human DetectionKeypoint DetectionTransfer Learning

OmniFaceRig: Fully Automatic Inner-Mouth-Aware Face Rigging Across Diverse 3D Character Topologies

2026-06-06 · Chao Wang, Guangyao Ma, John Doublestein, Junming Chen 외 arxiv

Facial rigging - creating FACS-based blendshapes together with inner-mouth geometry (teeth, gums, and tongue) - remains a major bottleneck in 3D character production. Existing pipelines still require substantial designer…

Face DetectionFace Parsing

Open-Vocabulary Animal Keypoint Detection with Semantic-feature Matching

2023-10-08 · Hao Zhang, Lumin Xu, Shenqi Lai, Wenqi Shao 외

Current image-based keypoint detection methods for animal (including human) bodies and faces are generally divided into full-supervised and few-shot class-agnostic approaches. The former typically relies on laborious and…

Keypoint DetectionOpen Vocabulary Keypoint Detection

Actor-agnostic Multi-label Action Recognition with Multi-modal Query

2023-07-20 · Anindya Mondal, Sauradip Nag, Joaquin M Prada, Xiatian Zhu 외

Existing action recognition methods are typically actor-specific due to the intrinsic topological and apparent differences among the actors. This requires actor-specific pose estimation (e.g., humans vs. animals), leadin…

Action ClassificationAction RecognitionAction Recognition In VideosAction Recognition on HMDB-51+2