Deep learning model compression and acceleration for resource constrained devices
| dc.contributor.guide | Mishra, Vipul Kumar and Goswami, Anurag | |
| dc.coverage.spatial | ||
| dc.creator.researcher | Choudhary, Tejalal | |
| dc.date.accessioned | 2023-09-05T10:38:12Z | |
| dc.date.available | 2023-09-05T10:38:12Z | |
| dc.date.awarded | 2022 | |
| dc.date.completed | 2022 | |
| dc.date.registered | 2016 | |
| dc.description.abstract | In recent years, machine learning and deep learning have shown remarkable improvement in newlinecomputer vision, natural language processing, stock prediction, weather forecasting, and audio newlineprocessing, to name a few. The performance of a deep neural network (DNN) is dependent newlineupon a significant number of weight parameters that need to be trained, which is a computational newlinebottleneck. DNNs are also known for their high resource requirements, weight redundancy, newlineand large-scale parameters. Hence, the utilization of DNNs is restricted to devices where newlineadequate resources required to execute them are not available, especially resource-constrained newlinedevices such as mobile phones, wearables, and other edge devices. For various practical applications, newlinethe trained models should be deployed on resource-constrained devices. Hence, it becomes newlineimperative to compress and accelerate these models before deploying them on resourceconstrained newlinedevices while making the least compromise on the model accuracy. To address newlinethese challenges, in the last couple of years, many researchers have suggested different techniques newlinesuch as pruning, quantization, low-rank factorization, and knowledge distillation for newlinemodel compression and acceleration. newlinePruning has emerged as an essential technique to reduce unimportant parameters and improve newlinethe model performance. However, finding the best pruning candidates and an optimal newlinenumber of parameters that can be pruned without significantly affecting the model performance newlineis time-consuming and requires a lot of manual tuning. Therefore, this thesis proposes novel newlinemethods for identifying and pruning the less important parameters of the trained deep learning newlinemodel. newline | |
| dc.description.note | ||
| dc.format.accompanyingmaterial | None | |
| dc.format.dimensions | ||
| dc.format.extent | ||
| dc.identifier.uri | http://hdl.handle.net/10603/510513 | |
| dc.language | English | |
| dc.publisher.institution | School of Computer Science Engineering and Technology | |
| dc.publisher.place | Greater Noida | |
| dc.publisher.university | Bennett University | |
| dc.relation | ||
| dc.rights | university | |
| dc.source.university | University | |
| dc.subject.keyword | Computer Science | |
| dc.subject.keyword | Computer Science Software Engineering | |
| dc.subject.keyword | Engineering and Technology | |
| dc.title | Deep learning model compression and acceleration for resource constrained devices | |
| dc.title.alternative | ||
| dc.type.degree | Ph.D. |
Files
Original bundle
1 - 5 of 13
Loading...
- Name:
- 01_title.pdf
- Size:
- 95.38 KB
- Format:
- Adobe Portable Document Format
- Description:
- Attached File
Loading...
- Name:
- 02_prelim pages.pdf
- Size:
- 374.81 KB
- Format:
- Adobe Portable Document Format
License bundle
1 - 1 of 1