MARC | 대구한의대학교 도서관

MARC보기
LDR		00000nam u2200205 4500
001		000000433860
005		20200226105911
008		200131s2019 \|\|\|\|\|\|\|\|\|\|\|\|\|\|\|\|\| \|\|eng d
020		▼a 9781085691208
035		▼a (MiAaPQ)AAI22584467
040		▼a MiAaPQ ▼c MiAaPQ ▼d 247004
082	0	▼a 621.3
100	1	▼a Kadetotad, Deepak Vinayak.
245	10	▼a On-chip Learning and Inference Acceleration of Sparse Representations.
260		▼a [S.l.]: ▼b Arizona State University., ▼c 2019.
260	1	▼a Ann Arbor: ▼b ProQuest Dissertations & Theses, ▼c 2019.
300		▼a 160 p.
500		▼a Source: Dissertations Abstracts International, Volume: 81-03, Section: B.
500		▼a Advisor: Seo, Jae-sun.
502	1	▼a Thesis (Ph.D.)--Arizona State University, 2019.
506		▼a This item must not be sold to any third party vendors.
520		▼a The past decade has seen a tremendous surge in running machine learning (ML) functions on mobile devices, from mere novelty applications to now indispensable features for the next generation of devices. While the mobile platform capabilities range widely, long battery life and reliability are common design concerns that are crucial to remain competitive.Consequently, state-of-the-art mobile platforms have become highly heterogeneous by combining a powerful CPUs with GPUs to accelerate the computation of deep neural networks (DNNs), which are the most common structures to perform ML operations. But traditional von Neumann architectures are not optimized for the high memory bandwidth and massively parallel computation demands required by DNNs. Hence, propelling research into non-von Neumann architectures to support the demands of DNNs.The re-imagining of computer architectures to perform efficient DNN computations requires focusing on the prohibitive demands presented by DNNs and alleviating them. The two central challenges for efficient computation are (1) large memory storage and movement due to weights of the DNN and (2) massively parallel multiplications to compute the DNN output.Introducing sparsity into the DNNs, where certain percentage of either the weights or the outputs of the DNN are zero, greatly helps with both challenges. This along with algorithm-hardware co-design to compress the DNNs is demonstrated to provide efficient solutions to greatly reduce the power consumption of hardware that compute DNNs. Additionally, exploring emerging technologies such as non-volatile memories and 3-D stacking of silicon in conjunction with algorithm-hardware co-design architectures will pave the way for the next generation of mobile devices.Towards the objectives stated above, our specific contributions include (a) an architecture based on resistive crosspoint array that can update all values stored and compute matrix vector multiplication in parallel within a single cycle, (b) a framework of training DNNs with a block-wise sparsity to drastically reduce memory storage and total number of computations required to compute the output of DNNs, (c) the exploration of hardware implementations of sparse DNNs and architectural guidelines to reduce power consumption for the implementations in monolithic 3D integrated circuits, and (d) a prototype chip in 65nm CMOS accelerator for long-short term memory networks trained with the proposed block-wise sparsity scheme.
590		▼a School code: 0010.
650	4	▼a Electrical engineering.
690		▼a 0544
710	20	▼a Arizona State University. ▼b Electrical Engineering.
773	0	▼t Dissertations Abstracts International ▼g 81-03B.
773		▼t Dissertation Abstract International
790		▼a 0010
791		▼a Ph.D.
792		▼a 2019
793		▼a English
856	40	▼u http://www.riss.kr/pdu/ddodLink.do?id=T15492850 ▼n KERIS ▼z 이 자료의 원문은 한국교육학술정보원에서 제공합니다.
980		▼a 202002 ▼f 2020
990		▼a ***1008102
991		▼a E-BOOK