paper-with-me

Papers

Bridging the Gap between Semantics and Multimedia Processing

2019-11-25 · Marcio Ferreira Moreno, Guilherme Lima, Rodrigo Costa Mesquita Santos, Roberto Azevedo, Markus Endler

In this paper, we give an overview of the semantic gap problem in multimedia and discuss how machine learning and symbolic AI can be combined to narrow this gap. We describe the gap in terms of a classical architecture for multimedia processing and discuss a structured approach to bridge it. This approach combines machine learning (for mapping signals to objects) and symbolic AI (for linking objects to meanings). Our main goal is to raise awareness and discuss the challenges involved in this structured approach to multimedia understanding, especially in the view of the latest developments in machine learning and symbolic AI.

📄 PDF Abstract BibTeX arXiv:1911.11631

Code (0)

등록된 구현이 없습니다.

Tasks

BIG-bench Machine Learning

Similar Papers 제목 키워드 기반

Towards Bridging the Cross-modal Semantic Gap for Multi-modal Recommendation

2024-07-07 · Xinglong Wu, Anfeng Huang, HongWei Yang, Hui He 외

Multi-modal recommendation greatly enhances the performance of recommender systems by modeling the auxiliary information from multi-modality contents. Most existing multi-modal recommendation models primarily exploit mul…

cross-modal alignmentMulti-modal RecommendationRecommendation SystemsSemantic Similarity+1

STIMONT: A core ontology for multimedia stimuli description

2014-01-10 · Marko Horvat, Nikola Bogunović, Krešimir Ćosić

Affective multimedia documents such as images, sounds or videos elicit emotional responses in exposed human subjects. These stimuli are stored in affective multimedia databases and successfully used for a wide variety of…

Retrieval

A Survey of Multimedia Technologies and Robust Algorithms

2021-03-24 · Zijian Kuang, Xinran Tie

Multimedia technologies are now more practical and deployable in real life, and the algorithms are widely used in various researching areas such as deep learning, signal processing, haptics, computer vision, robotics, an…

Survey

LecEval: An Automated Metric for Multimodal Knowledge Acquisition in Multimedia Learning

2025-05-04 · Joy Lim Jia Yin, Daniel Zhang-li, Jifan Yu, Haoxuan Li 외

Evaluating the quality of slide-based multimedia instruction is challenging. Existing methods like manual assessment, reference-based metrics, and large language model evaluators face limitations in scalability, context …

Language ModelingLanguage ModellingLarge Language Model

Semantics for Large-Scale Multimedia: New Challenges for NLP

2014-06-01 · ACL 2014 6 · Florian Metze, Koichi Shinoda
Active LearningInformation RetrievalSpeech RecognitionVideo Summarization