arXiv:1911.02549 Abstract | arXiv Analytics

arXiv:1911.02549 [cs.LG]Abstract References Reviews Resources

MLPerf Inference Benchmark

Vijay Janapa Reddi, Christine Cheng, David Kanter, Peter Mattson, Guenther Schmuelling, Carole-Jean Wu, Brian Anderson, Maximilien Breughe, Mark Charlebois, William Chou, Ramesh Chukka, Cody Coleman, Sam Davis, Pan Deng, Greg Diamos, Jared Duke, Dave Fick, J. Scott Gardner, Itay Hubara, Sachin Idgunji, Thomas B. Jablin, Jeff Jiao, Tom St. John, Pankaj Kanwar, David Lee, Jeffery Liao, Anton Lokhmotov, Francisco Massa, Peng Meng, Paulius Micikevicius, Colin Osborne, Gennady Pekhimenko, Arun Tejusve Raghunath Rajan, Dilip Sequeira, Ashish Sirasao, Fei Sun, Hanlin Tang, Michael Thomson, Frank Wei, Ephrem Wu, Lingjie Xu, Koichi Yamada, Bing Yu, George Yuan, Aaron Zhong, Peizhao Zhang, Yuchen Zhou

Published 2019-11-06Version 1

Machine-learning (ML) hardware and software system demand is burgeoning. Driven by ML applications, the number of different ML inference systems has exploded. Over 100 organizations are building ML inference chips, and the systems that incorporate existing models span at least three orders of magnitude in power consumption and four orders of magnitude in performance; they range from embedded devices to data-center solutions. Fueling the hardware are a dozen or more software frameworks and libraries. The myriad combinations of ML hardware and ML software make assessing ML-system performance in an architecture-neutral, representative, and reproducible manner challenging. There is a clear need for industry-wide standard ML benchmarking and evaluation criteria. MLPerf Inference answers that call. Driven by more than 30 organizations as well as more than 200 ML engineers and practitioners, MLPerf implements a set of rules and practices to ensure comparability across systems with wildly differing architectures. In this paper, we present the method and design principles of the initial MLPerf Inference release. The first call for submissions garnered more than 600 inference-performance measurements from 14 organizations, representing over 30 systems that show a range of capabilities.

Categories: cs.LG, cs.PF, stat.ML

Keywords: mlperf inference benchmark, organizations, ml inference systems, software system demand, building ml inference chips

Related articles:

arXiv:2411.18122 [cs.LG] (Published 2024-11-27)

A Machine Learning-based Framework towards Assessment of Decision-Makers' Biases

Wanxue Dong, Maria De-arteaga, Maytal Saar-Tsechansky

arXiv Analytics

arXiv:1911.02549 [cs.LG]Abstract References Reviews Resources

MLPerf Inference Benchmark

Links

Toolbox

arXiv:1911.02549 [cs.LG]AbstractReferencesReviewsResources

MLPerf Inference Benchmark

Links

Toolbox

arXiv:1911.02549 [cs.LG]Abstract References Reviews Resources