Pith. sign in

REVIEW 1 cited by

DeepCache: Principled Cache for Mobile Deep Vision

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1712.01670 v5 pith:DG3UW6ZJ submitted 2017-12-01 cs.CV

classification cs.CV
keywords deepcachemodelvideomobilecachedeepexploitingvision
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

We present DeepCache, a principled cache design for deep learning inference in continuous mobile vision. DeepCache benefits model execution efficiency by exploiting temporal locality in input video streams. It addresses a key challenge raised by mobile vision: the cache must operate under video scene variation, while trading off among cacheability, overhead, and loss in model accuracy. At the input of a model, DeepCache discovers video temporal locality by exploiting the video's internal structure, for which it borrows proven heuristics from video compression; into the model, DeepCache propagates regions of reusable results by exploiting the model's internal structure. Notably, DeepCache eschews applying video heuristics to model internals which are not pixels but high-dimensional, difficult-to-interpret data. Our implementation of DeepCache works with unmodified deep learning models, requires zero developer's manual effort, and is therefore immediately deployable on off-the-shelf mobile devices. Our experiments show that DeepCache saves inference execution time by 18% on average and up to 47%. DeepCache reduces system energy consumption by 20% on average.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Many Hands Make Light Work: Accelerating Edge Inference via Multi-Client Collaborative Caching

    cs.DC 2024-11 conditional novelty 5.0 of 10

    CoCa combines server-side global semantic caches with per-client dynamic cache allocation to cut edge inference latency by 23 to 45 percent with under 3 percent accuracy loss.

Pith tools