2018 55th ACM/ESDA/IEEE Design Automation Conference (DAC) · 2018 · 44 citations · 16 references
EngineeringMachine LearningPhysicsPipeline DesignHardware AccelerationUnique Filter MappingHigh-performance ArchitectureApplied PhysicsComputer EngineeringComputer ArchitectureComputing SystemsDomain-specific AcceleratorQuantum DevicesComputer SciencePipeline BubblesDeep LearningNeural Architecture SearchAtomic Layer Computation
Although ReRAM-based convolutional neural network (CNN) accelerators have been widely studied, state-of-the-art solutions suffer from either incapability of training (e.g., ISSAC [1]) or inefficiency of inference (e.g., PipeLayer [2]) due to the pipeline design. In this work, we propose AtomLayer-a universal ReRAM-based accelerator to support both efficient CNN training and inference. AtomLayer uses the atomic layer computation which processes only one network layer each time to eliminate the pipeline related issues such as long latency, pipeline bubbles and large on-chip buffer overhead. For further optimization, we use a unique filter mapping and a data reuse system to minimize the cost of layer switching and DRAM access. Our experimental results show that AtomLayer can achieve higher power efficiency than ISSAC in inference (1.1×) and PipeLayer in training (1.6×), respectively, meanwhile reducing the footprint by 15 ×.
16
Deep Residual Learning for Image Recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren et al. · 2016 · 214.9K citations · Full text
Image Classification, Deep Neural Networks, Machine Vision +14
Convolutional Neural Networks for Sentence Classification
Yoon Kim · 2014 · 13.5K citations · Full text
Natural Language Processing, Llm Fine-tuning, Natural Language +14
In-Datacenter Performance Analysis of a Tensor Processing Unit
Norman P. Jouppi, Cliff Young, Nishant Patil et al. · 2017 · 4.3K citations