October 21, 2019Open Access

Parity models

Key Points

Key points are not available for this paper at this time.

Abstract

Machine learning models are becoming the primary work-horses for many applications. Services deploy models through prediction serving systems that take in queries and return predictions by performing inference on models. Prediction serving systems are commonly run on many machines in cluster settings, and thus are prone to slowdowns and failures that inflate tail latency. Erasure coding is a popular technique for achieving resource-efficient resilience to data unavailability in storage and communication systems. However, existing approaches for imparting erasure-coded resilience to distributed computation apply only to a severely limited class of functions, precluding their use for many serving workloads, such as neural network inference.

Bookmark

View Full Paper

Bookmark

View Full Paper

Parity models

Key Points

Abstract

Cite This Study