FixyNN: Efficient Hardware for Mobile Computer Vision via Transfer Learning.

Paul N. Whatmough,Chuteng Zhou,Patrick Hansen,Shreyas K. Venkataramanaiah,Jae-sun Seo,Matthew Mattina

arXiv: Computer Vision and Pattern Recognition（2019）

引用 40|浏览154

暂无评分

摘要

The computational demands of computer vision tasks based on state-of-the-art Convolutional Neural Network (CNN) image classification far exceed the energy budgets of mobile devices. This paper proposes FixyNN, which consists of a fixed-weight feature extractor that generates ubiquitous CNN features, and a conventional programmable CNN accelerator which processes a dataset-specific CNN. Image classification models for FixyNN are trained end-to-end via transfer learning, with the common feature extractor representing the transfered part, and the programmable part being learnt on the target dataset. Experimental results demonstrate FixyNN hardware can achieve very high energy efficiencies up to 26.6 TOPS/W ($4.81 times$ better than iso-area programmable accelerator). Over a suite of six datasets we trained models via transfer learning with an accuracy loss of $u003c1%$ resulting in up to 11.2 TOPS/W - nearly $2 times$ more efficient than a conventional programmable CNN accelerator of the same area.

查看译文

AI 理解论文

溯源树

样例

生成溯源树，研究论文发展脉络

Chat Paper

正在生成论文摘要