paper-with-me

홈 › Papers

Common Metrics to Benchmark Human-Machine Teams (HMT): A Review

2020-08-11 · Praveen Damacharla, Ahmad Y. Javaid, Jennie J. Gallimore, Vijay K. Devabhaktuni

A significant amount of work is invested in human-machine teaming (HMT) across multiple fields. Accurately and effectively measuring system performance of an HMT is crucial for moving the design of these systems forward. Metrics are the enabling tools to devise a benchmark in any system and serve as an evaluation platform for assessing the performance, along with the verification and validation, of a system. Currently, there is no agreed-upon set of benchmark metrics for developing HMT systems. Therefore, identification and classification of common metrics are imperative to create a benchmark in the HMT field. The key focus of this review is to conduct a detailed survey aimed at identification of metrics employed in different segments of HMT and to determine the common metrics that can be used in the future to benchmark HMTs. We have organized this review as follows: identification of metrics used in HMTs until now, and classification based on functionality and measuring techniques. Additionally, we have also attempted to analyze all the identified metrics in detail while classifying them as theoretical, applied, real-time, non-real-time, measurable, and observable metrics. We conclude this review with a detailed analysis of the identified common metrics along with their usage to benchmark HMTs.

📄 PDF Abstract BibTeX arXiv:2008.04855

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Evaluation of Human-AI Teams for Learned and Rule-Based Agents in Hanabi

2021-07-15 · NeurIPS 2021 12 · Ho Chit Siu, Jaime D. Pena, Edenna Chen, Yutai Zhou 외

Deep reinforcement learning has generated superhuman AI in competitive games such as Go and StarCraft. Can similar learning techniques create a superior AI teammate for human-machine collaborative games? Will humans pref…

BenchmarkingDeep Reinforcement Learningreinforcement-learningReinforcement Learning+2

Who/What is My Teammate? Team Composition Considerations in Human-AI Teaming

2021-05-23 · Nathan J. McNeese, Beau G. Schelble, Lorenzo Barberis Canonico, Mustafa Demir

There are many unknowns regarding the characteristics and dynamics of human-AI teams, including a lack of understanding of how certain human-human teaming concepts may or may not apply to human-AI teams and how this comp…

Management

Lung-Originated Tumor Segmentation from Computed Tomography Scan (LOTUS) Benchmark

2022-01-03 · Parnian Afshar, Arash Mohammadi, Konstantinos N. Plataniotis, Keyvan Farahani 외

Lung cancer is one of the deadliest cancers, and in part its effective diagnosis and treatment depend on the accurate delineation of the tumor. Human-centered segmentation, which is currently the most common approach, is…

SegmentationTumor Segmentation

SemEval-2018 Task 11: Machine Comprehension Using Commonsense Knowledge

2018-06-01 · SEMEVAL 2018 6 · Simon Ostermann, Michael Roth, Ashutosh Modi, Stefan Thater 외

This report summarizes the results of the SemEval 2018 task on machine comprehension using commonsense knowledge. For this machine comprehension task, we created a new corpus, MCScript. It contains a high number of quest…

Reading ComprehensionStory Completion

Deep Generative Multi-Agent Imitation Model as a Computational Benchmark for Evaluating Human Performance in Complex Interactive Tasks: A Case Study in Football

2023-03-23 · Chaoyi Gu, Varuna De Silva

Evaluating the performance of human is a common need across many applications, such as in engineering and sports. When evaluating human performance in completing complex and interactive tasks, the most common way is to u…