Systematic survey uncovers key properties, workflows, and evaluation methods across machine learning testing research, highlighting emerging challenges in autonomous systems.
This paper provides a comprehensive survey of techniques for testing machine learning systems; Machine Learning Testing (ML testing) research. It covers 144 papers on testing properties (e.g., correctness, robustness, and fairness), testing components (e.g., the data, learning program, and framework), testing workflow (e.g., test generation and test evaluation), and application scenarios (e.g., autonomous driving, machine translation). The paper also analyses trends concerning datasets, research trends, and research focus, concluding with research challenges and promising research directions in ML testing.
No takes yet. Share an insight, caveat, or question.
Zhang et al. (2020) studied this question.
Synapse has enriched 4 closely related papers on similar clinical questions. Consider them for comparative context: