An Optimized Deep Learning Framework for Continuous Sign Language Recognition

dc.contributor.guideGeetha M
dc.coverage.spatial
dc.creator.researcherNeena Aloysius
dc.date.accessioned2024-11-18T07:06:16Z
dc.date.available2024-11-18T07:06:16Z
dc.date.awarded2024
dc.date.completed2024
dc.date.registered2017
dc.description.abstractSign language is a form of movement language that conveys semantic information through hand and arm motions, facial expressions, and head/body postures, serving as a crucial communication medium for the deaf community. Researchers are motivated by the desire to integrate deaf newlineindividuals into mainstream society, leading to a growing interest in automatic sign language recognition systems. This recognition involves interpreting static or dynamic signing within the one-arm distance 3D space around the upper body of the signer.In this work, a comprehensive literature review is conducted within the domains of visionbased Continuous Sign Language Recognition (CSLR) and Sign Language Translation (SLT). The deep Learning (DL) strategy is adopted by all the recent works. Any DL-based CSLR framework has three main modules - feature extraction, sequence learning and alignment learning. Feature extraction is usually done by a CNN. Most of the works have used LSTMs for sequence learning. Notably, it has been observed that the latest Transformer model and its variants are under-explored for these tasks. Furthermore, there is a gap in the literature concerning the investigation of position encoding schemes specific to the Transformer architecture, which is particularly valuable as the architecture lacks inherent sequential information. Therefore, an extensive literature study is conducted on Transformers, their variants, and the available position encoding schemes.This research began with the exploration of new positioning schemes for the Transformer model within the context of CSLR and SLT. Consequently, a novel positioning scheme was introduced, utilizing Gated Recurrent Unit (GRU) as the relative position encoder, and the multi-head attention (MHA) mechanism was modified to integrate relative position embeddings. The resulting Transformer, incorporating both positioning schemes, is referred to as GRU-RST. Furthermore, it was demonstrated that relative positioning outperformed absolute position encoding for Transformer ...
dc.description.note
dc.format.accompanyingmaterialNone
dc.format.dimensions
dc.format.extentx, 109
dc.identifier.urihttp://hdl.handle.net/10603/601387
dc.languageEnglish
dc.publisher.institutionAmrita School of Computing
dc.publisher.placeCoimbatore
dc.publisher.universityAmrita Vishwa Vidyapeetham University
dc.relation
dc.rightsuniversity
dc.source.universityUniversity
dc.subject.keywordComputer Science
dc.subject.keywordComputer Science Artificial Intelligence; Deep Learning; Sign Language;Language Translation; E-Governance services; sign language; Continuous sign language; Vision based; Movement epenthesis
dc.subject.keywordEngineering and Technology
dc.titleAn Optimized Deep Learning Framework for Continuous Sign Language Recognition
dc.title.alternative
dc.type.degreePh.D.

Files

Original bundle

Now showing 1 - 5 of 17
Loading...
Thumbnail Image
Name:
01_title.pdf
Size:
337.07 KB
Format:
Adobe Portable Document Format
Description:
Attached File
Loading...
Thumbnail Image
Name:
02_prelim pages.pdf
Size:
1.05 MB
Format:
Adobe Portable Document Format
Loading...
Thumbnail Image
Name:
03_contents.pdf
Size:
66.95 KB
Format:
Adobe Portable Document Format
Loading...
Thumbnail Image
Name:
04_abstract.pdf
Size:
54.15 KB
Format:
Adobe Portable Document Format
Loading...
Thumbnail Image
Name:
05_chapter 1.pdf
Size:
2.7 MB
Format:
Adobe Portable Document Format

License bundle

Now showing 1 - 1 of 1
Loading...
Thumbnail Image
Name:
license.txt
Size:
1.79 KB
Format:
Plain Text
Description: