학술논문

A demonstration of willump : a statistically-aware end-to-end optimizer for machine learning inference

Document Type

Academic Journal

Author

Kraft, Peter; Kang, Daniel; Narayanan, Deepak; Palkar, Shoumik; Bailis, Peter; Zaharia, Matei

Source

Proceedings of the VLDB Endowment. 13(12):2833-2836

Subject

Language

English

ISSN

2150-8097

Abstract

Systems for ML inference are widely deployed today, but they typically optimize ML inference workloads using techniques designed for conventional data serving workloads and miss critical opportunities to leverage the statistical nature of ML. In this demo, we present Willump, an optimizer for ML inference that introduces statistically-motivated optimizations targeting ML applications whose performance bottleneck is feature computation. Willump automatically cascades feature computation for classification queries: Willump classifies most data inputs using only high-value, low-cost features selected by a cost model, improving query performance by up to 5 x without statistically significant accuracy loss. In this demo, we use interactive and easily-downloadable Jupyter notebooks to show VLDB attendees which applications Willump can speed up, how to use Willump, and how Willump produces such large performance gains.

Online Access

Web of Science JCR 저널정보 Find it@PNU

이메일

부산대학교 도서관

Online Access

메일 발송