EncoDeep

Abstract

This article proposes EncoDeep, an end-to-end framework that facilitates encoding, bitwidth customization, fine-tuning, and implementation of neural networks on FPGA platforms. EncoDeep incorporates nonlinear encoding to the computation flow of neural networks to save memory. The encoded features demand significantly lower storage compared to the raw full-precision activation values; therefore, the execution flow of EncoDeep hardware engine is completely performed within the FPGA using on-chip streaming buffers with no access to the off-chip DRAM. We further propose a fully automated optimization algorithm that determines the flexible encoding bitwidths across network layers. EncoDeep full-stack framework comprises a compiler that takes a high-level Python description of an arbitrary neural network. The compiler then instantiates the corresponding elements from EncoDeep Hardware library for FPGA implementation. Our evaluations on MNIST, SVHN, and CIFAR-10 datasets demonstrate an average of 4.65× throughput improvement compared to stand-alone weight encoding. We further compare EncoDeep with six FPGA accelerators on ImageNet, showing an average of 3.6× and 2.54× improvement in throughput and performance-per-watt, respectively.

Document Details

Document Type: Pub Defense Publication
Publication Date: Sep 29, 2020
Source ID: 10.1145/3391901

Entities

People

Farinaz Koushanfar
Mohammad Samragh
Mojan Javaheripi

Organizations

Army Research Office
Intelligence Advanced Research Projects Activity
Office of Naval Research
University of California, San Diego

EncoDeep

Abstract

Document Details

Entities

People

Organizations

Tags

Fields of Study

Readers

Technology Areas