计算机科学
推论
GSM演进的增强数据速率
延迟(音频)
加速
边缘设备
修剪
班级(哲学)
人工智能
深层神经网络
架空(工程)
并行计算
作者
Maedeh Hemmat,Azadeh Davoodi,Yu Hen Hu
标识
DOI:10.1109/asp-dac52403.2022.9712496
摘要
We propose$\text{Edge}^{n}$AI, a framework to decompose a complex deep neural networks (DNN) over$n$available local edge devices with minimal communication overhead and overall latency. Our framework creates small DNNs (SNNs) from an original DNN by partitioning its classes across the edge devices, while taking into account their available resources. Class-aware pruning is applied to aggressively reduce the size of the SNN on each edge device. The SNNs perform inference in parallel, and are configured to generate a ‘Don't Know’ response when an unassigned class is identified. Our experiments show up to 17X inference speedup compared to a recent work, on devices of at most 150 MB memory when distributing a variant of VGG-16 over 20 parallel edge devices.
科研通智能强力驱动
Strongly Powered by AbleSci AI