Automated Testing of AI Models
The last decade has seen tremendous progress in AI technology and applications. With such widespread adoption, ensuring the reliability of the AI models is crucial. In past, we took the first step of creating a testing framework called AITEST for metamorphic properties such as fairness, robustness properties for tabular, time-series, and text classification models. In this paper, we extend the capability of the AITEST tool to include the testing techniques for Image and Speech-to-text models along with interpretability testing for tabular models. These novel extensions make AITEST a comprehensive framework for testing AI models.
Code (0)
등록된 구현이 없습니다.
Tasks
FairnessSpeech-to-Texttext-classificationText ClassificationTime SeriesTime Series AnalysisSimilar Papers 제목 키워드 기반
An Automated Testing Framework for Conversational Agents
Conversational agents are systems with a conversational interface that afford interaction in spoken language. These systems are becoming prevalent and are preferred in various contexts and for many users. Despite their i…
TERMINATOR: Better Automated UI Test Case Prioritization
Automated UI testing is an important component of the continuous integration process of software development. A modern web-based UI is an amalgam of reports from dozens of microservices written by multiple teams. Queries…
CPUPaving the Roadway for Safety of Automated Vehicles: An Empirical Study on Testing Challenges
The technology in the area of automated vehicles is gaining speed and promises many advantages. However, with the recent introduction of conditionally automated driving, we have also seen accidents. Test protocols for bo…
ACT: Automated CPS Testing for Open-Source Robotic Platforms
Open-source software for cyber-physical systems (CPS) often lacks robust testing involving robotic platforms, resulting in critical errors that remain undetected. This is especially challenging when multiple modules of C…
Automated Testing for Deep Learning Systems with Differential Behavior Criteria
In this work, we conducted a study on building an automated testing system for deep learning systems based on differential behavior criteria. The automated testing goals were achieved by jointly optimizing two objective …
Deep Learning