paper-with-me

홈 › Papers

High Accurate and Explainable Multi-Pill Detection Framework with Graph Neural Network-Assisted Multimodal Data Fusion

2023-03-17 · Anh Duy Nguyen, Huy Hieu Pham, Huynh Thanh Trung, Quoc Viet Hung Nguyen, Thao Nguyen Truong, Phi Le Nguyen

Due to the significant resemblance in visual appearance, pill misuse is prevalent and has become a critical issue, responsible for one-third of all deaths worldwide. Pill identification, thus, is a crucial concern needed to be investigated thoroughly. Recently, several attempts have been made to exploit deep learning to tackle the pill identification problem. However, most published works consider only single-pill identification and fail to distinguish hard samples with identical appearances. Also, most existing pill image datasets only feature single pill images captured in carefully controlled environments under ideal lighting conditions and clean backgrounds. In this work, we are the first to tackle the multi-pill detection problem in real-world settings, aiming at localizing and identifying pills captured by users in a pill intake. Moreover, we also introduce a multi-pill image dataset taken in unconstrained conditions. To handle hard samples, we propose a novel method for constructing heterogeneous a priori graphs incorporating three forms of inter-pill relationships, including co-occurrence likelihood, relative size, and visual semantic correlation. We then offer a framework for integrating a priori with pills' visual features to enhance detection accuracy. Our experimental results have proved the robustness, reliability, and explainability of the proposed framework. Experimentally, it outperforms all detection benchmarks in terms of all evaluation metrics. Specifically, our proposed framework improves COCO mAP metrics by 9.4% over Faster R-CNN and 12.0% compared to vanilla YOLOv5. Our study opens up new opportunities for protecting patients from medication errors using an AI-based pill identification solution.

📄 PDF Abstract BibTeX arXiv:2303.09782

Code (0)

등록된 구현이 없습니다.

Tasks

Graph Neural Network

Methods 이 논문이 사용한 방법론

fail 설명 없음
Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Convolution A convolution is a type of matrix operation, consisting of a kernel, a small matrix of weights, that slides over input data performing element-wise multiplication with the…
RPN A Region Proposal Network, or RPN, is a fully convolutional network that simultaneously predicts object bounds and objectness scores at each position. The RPN is trained…
RoIPool 설명 없음
Faster R-CNN Faster R-CNN is an object detection model that improves on Fast R-CNN by utilising a region proposal network…

Similar Papers 제목 키워드 기반

Automating Detection of Papilledema in Pediatric Fundus Images with Explainable Machine Learning

2022-07-10 · Kleanthis Avramidis, Mohammad Rostami, Melinda Chang, Shrikanth Narayanan

Papilledema is an ophthalmic neurologic disorder in which increased intracranial pressure leads to swelling of the optic nerves. Undiagnosed papilledema in children may lead to blindness and may be a sign of life-threate…

BIG-bench Machine LearningData AugmentationDeep LearningDiagnostic

PillarDETR: YOLO-Backbone and RT-DETR Head for Real-Time 3D Object Detection

2026-06-01 · Smit Kadvani, Shriya Gumber, Kriti Faujdar, Harsh Dave arxiv

Real-time 3D object detection is a critical component for the safe operation of autonomous driving systems and robotics. While LiDAR point clouds provide accurate spatial information, processing them efficiently remains …

3D Object DetectionAutonomous DrivingPoint Clouds

PointPillars Backbone Type Selection For Fast and Accurate LiDAR Object Detection

2022-09-30 · Konrad Lis, Tomasz Kryjak

3D object detection from LiDAR sensor data is an important topic in the context of autonomous cars and drones. In this paper, we present the results of experiments on the impact of backbone selection of a deep convolutio…

3D Object Detectionobject-detectionObject DetectionVocal Bursts Type Prediction

TransPillars: Coarse-to-Fine Aggregation for Multi-Frame 3D Object Detection

2022-08-04 · Zhipeng Luo, Gongjie Zhang, Changqing Zhou, Tianrui Liu 외

3D object detection using point clouds has attracted increasing attention due to its wide applications in autonomous driving and robotics. However, most existing studies focus on single point cloud frames without harness…

3D Object DetectionAutonomous DrivingObjectobject-detection+2

Accurate and Real-time 3D Pedestrian Detection Using an Efficient Attentive Pillar Network

2021-12-31 · Duy-Tho Le, Hengcan Shi, Hamid Rezatofighi, Jianfei Cai

Efficiently and accurately detecting people from 3D point cloud data is of great importance in many robotic and autonomous driving applications. This fundamental perception task is still very challenging due to (i) signi…

3D Object DetectionAutonomous DrivingBirds Eye View Object Detectionobject-detection+2