ADVANCES IN REAL-TIME AMERICAN SIGN LANGUAGE RECOGNITION SYSTEM USING DEEP LEARNING TECHNIQUES FOR ENHANCED ACCESSIBILITY

dc.contributor.advisorIlyas, Mohammad
dc.contributor.authorAlsharif, Bader
dc.date.accessioned2026-03-04T14:27:32Z
dc.date.issued2026
dc.description.abstractAdvancements in technology have significantly contributed to the development of innovative tools aimed at improving communication and accessibility for individuals with hearing impairments. This dissertation explores various machine learning and deep learning techniques for recognizing American Sign Language (ASL) gestures, focusing on enhancing accessibility and bridging the communication gap between hearing-impaired and hearing individuals. Traditional machine learning models, such as Random Forest, Support Vector Machines (SVM), and K-Nearest Neighbors (KNN), alongside deep learning architectures like AlexNet, ResNet-50, EfficientNet, ConvNeXt, and VisionTransformer, were investigated for their effectiveness. Experiments conducted on an extensive dataset of 87,000 ASL gesture images revealed exceptional recognition accuracy, with ResNet-50 achieving 99.98% and Random Forest reaching 99.55%, while other models performed within a range of 97% to 98%. Building on these findings, an innovative real-time recognition system was developed, integrating computer vision and deep learning techniques. The project initially utilized MediaPipe for precise hand movement tracking and YOLOv8, a state-of-the-art object detection model, to translate ASL gestures into text in real time. A comprehensive dataset of 29,820 annotated images was created to ensure strong generalization across diverse hand positions and lighting conditions. MediaPipe’s hand landmark annotations significantly enhanced input quality, improving the YOLOv8 models training accuracy. In addition, a more advanced framework was later designed that integrates YOLOv11 with MediaPipe for robust real-time ASL alphabet recognition. This system was trained on a large-scale dataset of 130,000 annotated images with custom keypoint-based annotations, enabling the model to capture subtle variations in hand and finger positions. Experimental evaluation demonstrated outstanding performance, achieving a mean Average Precision (mAP@0.5) of 98.2% with minimal latency, confirming its suitability for real-time applications in education, healthcare, and professional environments. Overall, the findings of this dissertation underscore the transformative potential of AI-driven solutions for ASL recognition. By bridging communication gaps through both traditional classification models and real-time deep learning frameworks, this work contributes to fostering inclusivity, accessibility, and independence for individuals with hearing impairments.
dc.format.extent151
dc.identifier.urihttps://hdl.handle.net/20.500.14154/78380
dc.language.isoen_US
dc.publisherSaudi Digital Library
dc.subjectAmerican Sign Language (ASL)
dc.subjectReal-time sign language recognition
dc.subjectDeep learning
dc.subjectComputer vision
dc.subjectHand landmark tracking
dc.subjectYOLO-based detection
dc.subjectTransfer learning
dc.subjectAssistive technology
dc.subjectAccessibility
dc.titleADVANCES IN REAL-TIME AMERICAN SIGN LANGUAGE RECOGNITION SYSTEM USING DEEP LEARNING TECHNIQUES FOR ENHANCED ACCESSIBILITY
dc.typeThesis
sdl.degree.departmentDepartment of Electrical Engineering and Computer Science
sdl.degree.disciplineComputer Engineering
sdl.degree.grantorFlorida Atlantic University
sdl.degree.nameDoctor of Philosophy

Files

Original bundle

Now showing 1 - 1 of 1
No Thumbnail Available
Name:
SACM-Dissertation.pdf
Size:
10.82 MB
Format:
Adobe Portable Document Format

License bundle

Now showing 1 - 1 of 1
No Thumbnail Available
Name:
license.txt
Size:
1.61 KB
Format:
Item-specific license agreed to upon submission
Description:

Copyright owned by the Saudi Digital Library (SDL) © 2026