paper-with-me

Papers

Face-voice Association in Multilingual Environments (FAME) Challenge 2024 Evaluation Plan

2024-04-14 · Muhammad Saad Saeed, Shah Nawaz, Muhammad Salman Tahir, Rohan Kumar Das, Muhammad Zaigham Zaheer, Marta Moscati, Markus Schedl, Muhammad Haris Khan, Karthik Nandakumar, Muhammad Haroon Yousaf

The advancements of technology have led to the use of multimodal systems in various real-world applications. Among them, the audio-visual systems are one of the widely used multimodal systems. In the recent years, associating face and voice of a person has gained attention due to presence of unique correlation between them. The Face-voice Association in Multilingual Environments (FAME) Challenge 2024 focuses on exploring face-voice association under a unique condition of multilingual scenario. This condition is inspired from the fact that half of the world's population is bilingual and most often people communicate under multilingual scenario. The challenge uses a dataset namely, Multilingual Audio-Visual (MAV-Celeb) for exploring face-voice association in multilingual environments. This report provides the details of the challenge, dataset, baselines and task details for the FAME Challenge.

📄 PDF Abstract BibTeX arXiv:2404.09342

Code (1)

mavceleb/mavceleb_baseline 공식 구현 pytorch

Similar Papers 제목 키워드 기반

Face-voice Association in Multilingual Environments (FAME) 2026 Challenge Evaluation Plan

2025-08-06 · Marta Moscati, Ahmed Abdullah, Muhammad Saad Saeed, Shah Nawaz 외 arxiv

The advancements of technology have led to the use of multimodal systems in various real-world applications. Among them, audio-visual systems are among the most widely used multimodal systems. In the recent years, associ…

Linking Faces and Voices Across Languages: Insights from the FAME 2026 Challenge

2025-12-23 · Marta Moscati, Ahmed Abdullah, Muhammad Saad Saeed, Shah Nawaz 외 arxiv

Over half of the world's population is bilingual and people often communicate under multilingual scenarios. The Face-Voice Association in Multilingual Environments (FAME) 2026 Challenge, held at ICASSP 2026, focuses on d…

Contrastive Learning-based Chaining-Cluster for Multilingual Voice-Face Association

2024-08-04 · Wuyang Chen, Yanjie Sun, Kele Xu, Yong Dou

The innate correlation between a person's face and voice has recently emerged as a compelling area of study, especially within the context of multilingual environments. This paper introduces our novel solution to the Fac…

Contrastive Learning

RFOP: Rethinking Fusion and Orthogonal Projection for Face-Voice Association

2025-12-02 · Abdul Hannan, Furqan Malik, Hina Jabbar, Syed Suleman Sadiq 외 arxiv

Face-voice association in multilingual environment challenge 2026 aims to investigate the face-voice association task in multilingual scenario. The challenge introduces English-German face-voice pairs to be utilized in t…

Shared Multi-modal Embedding Space for Face-Voice Association

2025-12-04 · Christopher Simic, Korbinian Riedhammer, Tobias Bocklet arxiv

The FAME 2026 challenge comprises two demanding tasks: training face-voice associations combined with a multilingual setting that includes testing on languages on which the model was not trained. Our approach consists of…