paper-with-me

홈 › Papers

Towards a Benchmark for Scientific Understanding in Humans and Machines

2023-04-20 · Kristian Gonzalez Barman, Sascha Caron, Tom Claassen, Henk de Regt

Scientific understanding is a fundamental goal of science, allowing us to explain the world. There is currently no good way to measure the scientific understanding of agents, whether these be humans or Artificial Intelligence systems. Without a clear benchmark, it is challenging to evaluate and compare different levels of and approaches to scientific understanding. In this Roadmap, we propose a framework to create a benchmark for scientific understanding, utilizing tools from philosophy of science. We adopt a behavioral notion according to which genuine understanding should be recognized as an ability to perform certain tasks. We extend this notion by considering a set of questions that can gauge different levels of scientific understanding, covering information retrieval, the capability to arrange information to produce an explanation, and the ability to infer how things would be different under different circumstances. The Scientific Understanding Benchmark (SUB), which is formed by a set of these tests, allows for the evaluation and comparison of different approaches. Benchmarking plays a crucial role in establishing trust, ensuring quality control, and providing a basis for performance evaluation. By aligning machine and human scientific understanding we can improve their utility, ultimately advancing scientific understanding and helping to discover new insights within machines.

📄 PDF Abstract BibTeX arXiv:2304.10327

Code (0)

등록된 구현이 없습니다.

Tasks

BenchmarkingInformation RetrievalPhilosophyRetrieval

Similar Papers 제목 키워드 기반

Human vs. supervised machine learning: Who learns patterns faster?

2020-11-30 · Niklas Kühl, Marc Goutier, Lucas Baier, Clemens Wolff 외

The capabilities of supervised machine learning (SML), especially compared to human abilities, are being discussed in scientific research and in the usage of SML. This study provides an answer to how learning performance…

BIG-bench Machine Learning

Minds, Brains, AI

2024-04-21 · Jay Seitz

In the last year or so and going back many decades there has been extensive claims by major computational scientists, engineers, and others that AGI, artificial general intelligence, is five or ten years away, but withou…

Self-Driving Cars

Smarnet: Teaching Machines to Read and Comprehend Like Human

2017-10-08 · Zheqian Chen, Rongqin Yang, Bin Cao, Zhou Zhao 외

Machine Comprehension (MC) is a challenging task in Natural Language Processing field, which aims to guide the machine to comprehend a passage and answer the given question. Many existing approaches on MC task are suffer…

Question AnsweringReading ComprehensionTriviaQA

MEWL: Few-shot multimodal word learning with referential uncertainty

2023-06-01 · Guangyuan Jiang, Manjie Xu, Shiji Xin, Wei Liang 외

Without explicit feedback, humans can rapidly learn the meaning of words. Children can acquire a new word after just a few passive exposures, a process known as fast mapping. This word learning capability is believed to …

Do humans and machines have the same eyes? Human-machine perceptual differences on image classification

2023-04-18 · Minghao Liu, Jiaheng Wei, Yang Liu, James Davis

Trained computer vision models are assumed to solve vision tasks by imitating human behavior learned from training labels. Most efforts in recent vision research focus on measuring the model task performance using standa…

image-classificationImage Classification