arXiv · 2409.04504
Comment on Revisiting Neural Program Smoothing for Fuzzing
Abstract
MLFuzz, a work accepted at ACM FSE 2023, revisits the performance of a machine learning-based fuzzer, NEUZZ. We demonstrate that its main conclusion is entirely wrong due to several fatal bugs in the implementation and wrong evaluation setups, including an initialization bug in persistent mode, a program crash, an error in training dataset collection, and a mistake in fuzzing result collection. Additionally, MLFuzz uses noisy training datasets without sufficient data cleaning and preprocessing, which contributes to a drastic performance drop in NEUZZ. We address these issues and provide a corrected implementation and evaluation setup, showing that NEUZZ consistently performs well over AFL on the FuzzBench dataset. Finally, we reflect on the evaluation methods used in MLFuzz and offer practical advice on fair and scientific fuzzing evaluations.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Dongdong She, Kexin Pei, Junfeng Yang, Baishakhi Ray, Suman Jana. 2024-09-06. Comment on Revisiting Neural Program Smoothing for Fuzzing. https://arxiv.org/abs/2409.04504
Cite the original work for its findings. Save a collection to share your selection of sources.