paper-with-me

홈 › Papers

A New Evaluation Protocol and Benchmarking Results for Extendable Cross-media Retrieval

2017-03-10 · Ruoyu Liu, Yao Zhao, Liang Zheng, Shikui Wei, Yi Yang

This paper proposes a new evaluation protocol for cross-media retrieval which better fits the real-word applications. Both image-text and text-image retrieval modes are considered. Traditionally, class labels in the training and testing sets are identical. That is, it is usually assumed that the query falls into some pre-defined classes. However, in practice, the content of a query image/text may vary extensively, and the retrieval system does not necessarily know in advance the class label of a query. Considering the inconsistency between the real-world applications and laboratory assumptions, we think that the existing protocol that works under identical train/test classes can be modified and improved. This work is dedicated to addressing this problem by considering the protocol under an extendable scenario, \ie, the training and testing classes do not overlap. We provide extensive benchmarking results obtained by the existing protocol and the proposed new protocol on several commonly used datasets. We demonstrate a noticeable performance drop when the testing classes are unseen during training. Additionally, a trivial solution, \ie, directly using the predicted class label for cross-media retrieval, is tested. We show that the trivial solution is very competitive in traditional non-extendable retrieval, but becomes less so under the new settings. The train/test split, evaluation code, and benchmarking results are publicly available on our website.

📄 PDF Abstract BibTeX arXiv:1703.03567

Code (0)

등록된 구현이 없습니다.

Tasks

BenchmarkingImage RetrievalRetrieval

Similar Papers 제목 키워드 기반

How much progress have we made in neural network training? A New Evaluation Protocol for Benchmarking Optimizers

2020-10-19 · Yuanhao Xiong, Xuanqing Liu, Li-Cheng Lan, Yang You 외

Many optimizers have been proposed for training deep neural networks, and they often have multiple hyperparameters, which make it tricky to benchmark their performance. In this work, we propose a new benchmarking protoco…

BenchmarkingGraph Mining

A Transparent Fairness Evaluation Protocol for Open-Source Language Model Benchmarking on the Blockchain

2025-07-29 · Hugo Massaroli, Leonardo Iara, Emmanuel Iarussi, Viviana Siless arxiv

Large language models (LLMs) are increasingly deployed in realworld applications, yet concerns about their fairness persist especially in highstakes domains like criminal justice, education, healthcare, and finance. This…

Quantifying Ranking Instability Across Evaluation Protocol Axes in Gene Regulatory Network Benchmarking

2026-03-03 · Ihor Kendiukhov arxiv

Benchmark rankings are routinely used to justify scientific claims about method quality in gene regulatory network (GRN) inference, yet the stability of these rankings under plausible evaluation protocol choices is rarel…

GraphBench: Next-generation graph learning benchmarking

2025-12-04 · Timo Stoll, Chendi Qian, Ben Finkelshtein, Ali Parviz 외 arxiv

Machine learning on graphs has made substantial progress across domains such as molecular property prediction and chip design. Yet benchmarking practices remain fragmented, often relying on narrow, task-specific datasets…

Molecular Property PredictionGraph Learning

EBES: Easy Benchmarking for Event Sequences

2024-10-04 · Dmitry Osin, Igor Udovichenko, Viktor Moskvoretskii, Egor Shvetsov 외

Event sequences, characterized by irregular sampling intervals and a mix of categorical and numerical features, are common data structures in various real-world domains such as healthcare, finance, and user interaction l…

Benchmarking