OrthographicNet: A Deep Transfer Learning Approach for 3D Object Recognition in Open-Ended Domains
Nowadays, service robots are appearing more and more in our daily life. For this type of robot, open-ended object category learning and recognition is necessary since no matter how extensive the training data used for batch learning, the robot might be faced with a new object when operating in a real-world environment. In this work, we present OrthographicNet, a Convolutional Neural Network (CNN)-based model, for 3D object recognition in open-ended domains. In particular, OrthographicNet generates a global rotation- and scale-invariant representation for a given 3D object, enabling robots to recognize the same or similar objects seen from different perspectives. Experimental results show that our approach yields significant improvements over the previous state-of-the-art approaches concerning object recognition performance and scalability in open-ended scenarios. Moreover, OrthographicNet demonstrates the capability of learning new categories from very few examples on-site. Regarding real-time performance, three real-world demonstrations validate the promising performance of the proposed architecture.
Code (1)
Tasks
3D Object RecognitionObjectObject RecognitionTransfer LearningSimilar Papers 제목 키워드 기반
Bridging Coarse and Fine Recognition: A Hybrid Approach for Open-Ended Multi-Granularity Object Recognition in Interactive Educational Games
Recent advances in Multimodal Large Language Models (MLLMs) have enabled open-ended object recognition, yet they struggle with fine-grained tasks. In contrast, CLIP-style models excel at fine-grained recognition but lack…
Object RecognitionSimultaneous Multi-View Object Recognition and Grasping in Open-Ended Domains
To aid humans in everyday tasks, robots need to know which objects exist in the scene, where they are, and how to grasp and manipulate them in different situations. Therefore, object recognition and grasping are two key …
Active LearningObjectObject RecognitionExplain What You See: Open-Ended Segmentation and Recognition of Occluded 3D Objects
Local-HDP (for Local Hierarchical Dirichlet Process) is a hierarchical Bayesian method that has recently been used for open-ended 3D object category recognition. This method has been proven to be efficient in real-time r…
Incremental LearningObject3D_DEN: Open-ended 3D Object Recognition using Dynamically Expandable Networks
Service robots, in general, have to work independently and adapt to the dynamic changes happening in the environment in real-time. One important aspect in such scenarios is to continually learn to recognize newer object …
3D Object RecognitionContinual LearningObjectObject Recognition+1Lifelong Ensemble Learning based on Multiple Representations for Few-Shot Object Recognition
Service robots are integrating more and more into our daily lives to help us with various tasks. In such environments, robots frequently face new objects while working in the environment and need to learn them in an open…
3D Object RecognitionEnsemble LearningFew-Shot LearningLifelong learning+2