paper-with-me

Papers

Enhancing Computer Vision with Knowledge: a Rummikub Case Study

2024-11-27 · Simon Vandevelde, Laurent Mertens, Sverre Lauwers, Joost Vennekens

Artificial Neural Networks excel at identifying individual components in an image. However, out-of-the-box, they do not manage to correctly integrate and interpret these components as a whole. One way to alleviate this weakness is to expand the network with explicit knowledge and a separate reasoning component. In this paper, we evaluate an approach to this end, applied to the solving of the popular board game Rummikub. We demonstrate that, for this particular example, the added background knowledge is equally valuable as two-thirds of the data set, and allows to bring down the training time to half the original time.

📄 PDF Abstract BibTeX arXiv:2411.18172

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Commonsense Reasoning in Computer Vision: Foundations, Recent Advancements, and Future Directions

2026-09-04 · Bahar Uddin Mahmud, Sumit Barua, Guan Yue Hong, Ajay Gupta 외 arxiv

Commonsense reasoning in computer vision encompasses integrating visual data and contextual knowledge, crucial for enhancing AI's understanding of everyday scenarios. This understanding not only improves machine learning…

Object RecognitionKnowledge Graphs

Challenges and Practices of Deep Learning Model Reengineering: A Case Study on Computer Vision

2023-03-13 · Wenxin Jiang, Vishnu Banna, Naveen Vivek, Abhinav Goel 외

Many engineering organizations are reimplementing and extending deep neural networks from the research community. We describe this process as deep learning model reengineering. Deep learning model reengineering - reusing…

Deep Learning

VILA-M3: Enhancing Vision-Language Models with Medical Expert Knowledge

2024-11-19 · CVPR 2025 1 · Vishwesh Nath, Wenqi Li, Dong Yang, Andriy Myronenko 외

Generalist vision language models (VLMs) have made significant strides in computer vision, but they fall short in specialized fields like healthcare, where expert knowledge is essential. In traditional computer vision ta…

Using Computer Vision for Skin Disease Diagnosis in Bangladesh Enhancing Interpretability and Transparency in Deep Learning Models for Skin Cancer Classification

2025-01-30 · Rafiul Islam, Jihad Khan Dipu, Mehedi Hasan Tusar

With over 2 million new cases identified annually, skin cancer is the most prevalent type of cancer globally and the second most common in Bangladesh, following breast cancer. Early detection and treatment are crucial fo…

Cancer ClassificationDecision MakingDeep LearningSkin Cancer Classification

Leveraging Large-Scale Pretrained Vision Foundation Models for Label-Efficient 3D Point Cloud Segmentation

2023-11-03 · Shichao Dong, Fayao Liu, Guosheng Lin

Recently, large-scale pre-trained models such as Segment-Anything Model (SAM) and Contrastive Language-Image Pre-training (CLIP) have demonstrated remarkable success and revolutionized the field of computer vision. These…

3D Semantic SegmentationPoint Cloud SegmentationScene UnderstandingSegmentation+3