paper-with-me

Benchmarking

2개 벤치마크 · 논문 5,548편 · 이 태스크의 논문 보기 →

Benchmarks

CloudEval-YAML

결과 3개

Wiki-40B

결과 3개

Most implemented

The StarCraft Multi-Agent Challenge

2019-02-11 · 구현 23개

Benchmarking Graph Neural Networks

2020-03-02 · 구현 15개

Papers

Visual Place Recognition for Large-Scale UAV Applications

2025-07-20 · Ioannis Tsampikos Papapetros, Ioannis Kansizoglou, Antonios Gasteratos

Visual Place Recognition (vPR) plays a crucial role in Unmanned Aerial Vehicle (UAV) navigation, enabling robust localization across diverse environments. Despite significant advancements, aerial vPR faces unique challen…

BenchmarkingDiversityVisual Place Recognition

Training Transformers with Enforced Lipschitz Constants

2025-07-17 · Laker Newhouse, R. Preston Hess, Franz Cesista, Andrii Zahorodnii 외

Neural networks are often highly sensitive to input and weight perturbations. This sensitivity has been linked to pathologies such as vulnerability to adversarial examples, divergent training, and overfitting. To combat …

Benchmarking

Disentangling coincident cell events using deep transfer learning and compressive sensing

2025-07-17 · Moritz Leuthner, Rafael Vorländer, Oliver Hayden

Accurate single-cell analysis is critical for diagnostics, immunomonitoring, and cell therapy, but coincident events - where multiple cells overlap in a sensing zone - can severely compromise signal fidelity. We present …

BenchmarkingCompressive SensingTransfer Learning

MUPAX: Multidimensional Problem Agnostic eXplainable AI

2025-07-17 · Vincenzo Dentamaro, Felice Franchini, Giuseppe Pirlo, Irina Voiculescu

Robust XAI techniques should ideally be simultaneously deterministic, model agnostic, and guaranteed to converge. We propose MULTIDIMENSIONAL PROBLEM AGNOSTIC EXPLAINABLE AI (MUPAX), a deterministic, model agnostic expla…

Anatomical Landmark DetectionAudio ClassificationBenchmarkingFeature Importance+3

DVFL-Net: A Lightweight Distilled Video Focal Modulation Network for Spatio-Temporal Action Recognition

2025-07-16 · Hayat Ullah, Muhammad Ali Shafique, Abbas Khan, Arslan Munir

The landscape of video recognition has evolved significantly, shifting from traditional Convolutional Neural Networks (CNNs) to Transformer-based architectures for improved accuracy. While 3D CNNs have been effective at …

BenchmarkingKnowledge DistillationSpatio-temporal Action RecognitionTemporal Action Localization+2

DCR: Quantifying Data Contamination in LLMs Evaluation

2025-07-15 · Cheng Xu, Nan Yan, Shuhao Guan, Changhong Jin 외

The rapid advancement of large language models (LLMs) has heightened concerns about benchmark data contamination (BDC), where models inadvertently memorize evaluation data, inflating performance metrics and undermining g…

Arithmetic ReasoningBenchmarkingComputational EfficiencyFake News Detection+1

전체 5,548편 보기 →