paper-with-me

홈 › Papers

BioBench: A Blueprint to Move Beyond ImageNet for Scientific ML Benchmarks

2025-11-20 · Samuel Stevens arxiv

ImageNet-1K linear-probe transfer accuracy remains the default proxy for visual representation quality, yet it no longer predicts performance on scientific imagery. Across 46 modern vision model checkpoints, ImageNet top-1 accuracy explains only 34% of variance on ecology tasks and mis-ranks 30% of models above 75% accuracy. We present BioBench, an open ecology vision benchmark that captures what ImageNet misses. BioBench unifies 9 publicly released, application-driven tasks, 4 taxonomic kingdoms, and 6 acquisition modalities (drone RGB, web video, micrographs, in-situ and specimen photos, camera-trap frames), totaling 3.1M images. A single Python API downloads data, fits lightweight classifiers to frozen backbones, and reports class-balanced macro-F1 (plus domain metrics for FishNet and FungiCLEF); ViT-L models evaluate in 6 hours on an A6000 GPU. BioBench provides new signal for computer vision in ecology and a template recipe for building reliable AI-for-science benchmarks in any domain. Code and predictions are available at https://github.com/samuelstevens/biobench and results at https://samuelstevens.me/biobench.

📄 PDF Abstract BibTeX arXiv:2511.16315

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Computational Charisma -- A Brick by Brick Blueprint for Building Charismatic Artificial Intelligence

2022-12-31 · Björn W. Schuller, Shahin Amiriparian, Anton Batliner, Alexander Gebhard 외

Charisma is considered as one's ability to attract and potentially also influence others. Clearly, there can be considerable interest from an artificial intelligence's (AI) perspective to provide it with such skill. Beyo…

Meta-Designing Quantum Experiments with Language Models

2024-06-04 · Sören Arlt, Haonan Duan, Felix Li, Sang Michael Xie 외

Artificial Intelligence (AI) has the potential to significantly advance scientific discovery by finding solutions beyond human capabilities. However, these super-human solutions are often unintuitive and require consider…

Language ModelingLanguage Modellingscientific discovery

Beyond Shapley Values: Cooperative Games for the Interpretation of Machine Learning Models

2025-06-16 · Marouane Il Idrissi, Agathe Fernandes Machado, Arthur Charpentier

Cooperative game theory has become a cornerstone of post-hoc interpretability in machine learning, largely through the use of Shapley values. Yet, despite their widespread adoption, Shapley-based methods often rest on ax…

AlphaGo Moment for Model Architecture Discovery

2025-07-24 · Yixiu Liu, Yang Nan, Weixian Xu, Xiangkun Hu 외 arxiv

While AI systems demonstrate exponentially improving capabilities, the pace of AI research itself remains linearly bounded by human cognitive capacity, creating an increasingly severe development bottleneck. We present A…

Neural Architecture Search

From Observation to Insight: Mechanistic World Models and the Quest for Autonomous Discovery

2026-07-14 · Ingmar Posner, Anson Lei, Bernhard Schölkopf arxiv

Recent advances in foundation models have transformed AI for Science, enabling remarkably accurate predictive performance across domains ranging from protein folding to weather forecasting. Yet prediction alone does not …

Representation LearningWeather Forecasting