Repository logo
Log In(current)
  1. Home
  2. Colleges & Schools
  3. Graduate School
  4. Doctoral Dissertations
  5. Toward Generating Efficient Deep Neural Networks
Details

Toward Generating Efficient Deep Neural Networks

Date Issued
May 1, 2023
Author(s)
Li, Chengcheng
Advisor(s)
Hairong Qi
Additional Advisor(s)
Hairong Qi
Jinyuan Sun
Amir Sadovnik
Russell Zaretzki
Permanent URI
https://trace.tennessee.edu/handle/20.500.14382/29396
Abstract

Recent advances in deep neural networks have led to tremendous applications in various tasks, such as object classification and detection, image synthesis, natural language processing, game playing, and biological imaging. However, deploying these pre-trained networks on resource-limited devices poses a challenge, as most state-of- the-art networks contain millions of parameters, making them cumbersome and slow in real-world applications. To address this problem, numerous network compression and acceleration approaches, also known as efficient deep neural networks or efficient deep learning, have been investigated, in terms of hardware and software (algorithms), training, and inference. The aim of this dissertation is to study several algorithms, with a particular focus on network pruning and knowledge distillation, which have been identified as powerful techniques in enabling efficient processing without significant performance degradation. While these general methods are not limited to certain network structures, datasets, and tasks, we focus on compressing and accelerating popular convolutional neural networks (CNNs) for image classifications, for which these methods were originally developed and conventionally evaluated with extensive empirical study. In network pruning, we explore an important yet largely neglected aspect of network pruning, the efficiency of the pruning procedure itself. Our contributions to network pruning focus on post-training channel pruning, which is usually com- putationally intensive and heavily energy-consuming. A typical pruning procedure consists of iterative procedures of ranking, pruning, and fine-tuning. We challenge the common belief of the importance of ranking criteria with empirical studies and propose an efficient pipeline for pruning CNNs by integrating ranking and fine-tuning through computational re-usage. We evaluate the proposed method with extensive experimental studies. Our research in knowledge distillation, more precisely online knowledge distilla- tion, is motivated by the observation that training networks using existing online KD approaches is a highly dynamic procedure, in which each student network learns its parameters from scratch and acts as an instructor for other networks. Naturally, training with developing instructors tends to involve more uncertainty and fluctuation. To generate superior and robust knowledge, we focus on leveraging various information encoded in each peer’s learning trajectory to dynamically construct superior teachers to supervise other students, potentially improving the performance of students during inference.

Subjects

Neural Networks

Network Compression

Knowledge Distillatio...

Disciplines
Other Computer Engineering
Degree
Doctor of Philosophy
Major
Computer Engineering
File(s)
Thumbnail Image
Name

DissertationThesis_ChengchengLi_20230410.pdf

Size

6.91 MB

Format

Adobe PDF

Checksum (MD5)

0ef40e2d727d93e25d569e31031d3653


University Libraries

1015 Volunteer Boulevard
Knoxville, TN 37996
865-974-4351

Map & Directions
Donate to the Libraries
  • About
  • John C. Hodges Society
  • Speaking Volumes magazine
  • Outreach
  • Directory
  • Employment
  • Policies
  • Library Intranet
University of Tennessee power T logo

The University of Tennessee, Knoxville
Knoxville, Tennessee 37996
865-974-1000

Events
A-Z
Apply
Privacy
Map
Directory
Give to UT
Accessibility

Built with DSpace-CRIS software - Extension maintained and optimized by 4Science