Llama: A heterogeneous & serverless framework for auto-tuning video analytics pipelines

Francisco Romero; Mark Zhao; Neeraja Yadwadkar; Christos Kozyrakis

doi:10.48550/arXiv.2102.01887

Llama: A heterogeneous & serverless framework for auto-tuning video analytics pipelines

Francisco Romero Stanford

Mark Zhao Stanford

Neeraja Yadwadkar Stanford

Christos Kozyrakis Stanford

ACM Symposium on Cloud Computing (SoCC), 2021

DOI: 10.48550/arXiv.2102.01887

Abstract

The proliferation of camera-enabled devices and large video repositories has led to a diverse set of video analytics applications. These applications rely on video pipelines, represented as DAGs of operations, to transform videos, process extracted metadata, and answer questions like, “Is this intersection congested?” The latency and resource efficiency of pipelines can be optimized using configurable knobs for each operation (e.g., sampling rate, batch size, or type of hardware used). However, determining efficient configurations is challenging because (a) the configuration search space is exponentially large, and (b) the optimal configuration depends on users’ desired latency and cost targets, (c) input video contents may exercise different paths in the DAG and produce a variable amount intermediate results. Existing video analytics and processing systems leave it to the users to manually configure operations and …

Hi, we are the Stanford MAST team

Llama: A heterogeneous & serverless framework for auto-tuning video analytics pipelines

Abstract

Materials

Bibtex