Story
arxiv_cs_lg ยท Sep 23, 2026 ยท paper
arxiv.orgSep 23, 2026
original source linked
In brief
Deploying deep learning models on edge CPUs is bottlenecked by computational and memory constraints. Mixed-precision quantization promises to reduce inference latency while preserving accuracy. However, quantization a...
Feed lens
eval