Match without a Referee: Evaluating MT Adequacy without Reference Translations
Code (0)
등록된 구현이 없습니다.
Tasks
Machine TranslationNatural Language InferenceSimilar Papers 제목 키워드 기반
FAIEr: Fidelity and Adequacy Ensured Image Caption Evaluation
Image caption evaluation is a crucial task, which involves the semantic perception and matching of image and text. Good evaluation metrics aim to be fair, comprehensive, and consistent with human judge intentions. Wh…
Image CaptioningReFEree: Reference-Free and Fine-Grained Method for Evaluating Factual Consistency in Real-World Code Summarization
As Large Language Models (LLMs) have become capable of generating long and descriptive code summaries, accurate and reliable evaluation of factual consistency has become a critical challenge. However, previous evaluation…
RefereeBench: Are Video MLLMs Ready to be Multi-Sport Referees
While Multimodal Large Language Models (MLLMs) excel at generic video understanding, their ability to support specialized, rule-grounded decision-making remains insufficiently explored. In this paper, we introduce Refere…
Automated Tackle Injury Risk Assessment in Contact-Based Sports -- A Rugby Union Example
Video analysis in tackle-collision based sports is highly subjective and exposed to bias, which is inherent in human observation, especially under time constraints. This limitation of match analysis in tackle-collision b…
ManagementReferee: Reference-aware Audiovisual Deepfake Detection
Deepfakes generated by advanced generative models have rapidly posed serious threats, yet existing audiovisual deepfake detection approaches struggle to generalize to unseen manipulation methods. To address this, we prop…
DeepFake Detection