Deep learning model compression and acceleration for resource constrained devices

dc.contributor.guideMishra, Vipul Kumar and Goswami, Anurag
dc.coverage.spatial
dc.creator.researcherChoudhary, Tejalal
dc.date.accessioned2023-09-05T10:38:12Z
dc.date.available2023-09-05T10:38:12Z
dc.date.awarded2022
dc.date.completed2022
dc.date.registered2016
dc.description.abstractIn recent years, machine learning and deep learning have shown remarkable improvement in newlinecomputer vision, natural language processing, stock prediction, weather forecasting, and audio newlineprocessing, to name a few. The performance of a deep neural network (DNN) is dependent newlineupon a significant number of weight parameters that need to be trained, which is a computational newlinebottleneck. DNNs are also known for their high resource requirements, weight redundancy, newlineand large-scale parameters. Hence, the utilization of DNNs is restricted to devices where newlineadequate resources required to execute them are not available, especially resource-constrained newlinedevices such as mobile phones, wearables, and other edge devices. For various practical applications, newlinethe trained models should be deployed on resource-constrained devices. Hence, it becomes newlineimperative to compress and accelerate these models before deploying them on resourceconstrained newlinedevices while making the least compromise on the model accuracy. To address newlinethese challenges, in the last couple of years, many researchers have suggested different techniques newlinesuch as pruning, quantization, low-rank factorization, and knowledge distillation for newlinemodel compression and acceleration. newlinePruning has emerged as an essential technique to reduce unimportant parameters and improve newlinethe model performance. However, finding the best pruning candidates and an optimal newlinenumber of parameters that can be pruned without significantly affecting the model performance newlineis time-consuming and requires a lot of manual tuning. Therefore, this thesis proposes novel newlinemethods for identifying and pruning the less important parameters of the trained deep learning newlinemodel. newline
dc.description.note
dc.format.accompanyingmaterialNone
dc.format.dimensions
dc.format.extent
dc.identifier.urihttp://hdl.handle.net/10603/510513
dc.languageEnglish
dc.publisher.institutionSchool of Computer Science Engineering and Technology
dc.publisher.placeGreater Noida
dc.publisher.universityBennett University
dc.relation
dc.rightsuniversity
dc.source.universityUniversity
dc.subject.keywordComputer Science
dc.subject.keywordComputer Science Software Engineering
dc.subject.keywordEngineering and Technology
dc.titleDeep learning model compression and acceleration for resource constrained devices
dc.title.alternative
dc.type.degreePh.D.

Files

Original bundle

Now showing 1 - 5 of 13
Loading...
Thumbnail Image
Name:
01_title.pdf
Size:
95.38 KB
Format:
Adobe Portable Document Format
Description:
Attached File
Loading...
Thumbnail Image
Name:
02_prelim pages.pdf
Size:
374.81 KB
Format:
Adobe Portable Document Format
Loading...
Thumbnail Image
Name:
03_contents.pdf
Size:
66.3 KB
Format:
Adobe Portable Document Format
Loading...
Thumbnail Image
Name:
04_abstract.pdf
Size:
51.4 KB
Format:
Adobe Portable Document Format
Loading...
Thumbnail Image
Name:
05_chapter 1.pdf
Size:
170.9 KB
Format:
Adobe Portable Document Format

License bundle

Now showing 1 - 1 of 1
Loading...
Thumbnail Image
Name:
license.txt
Size:
1.79 KB
Format:
Plain Text
Description: