American Sign Language Machine Learning based Translation Model
| dc.contributor.author | Zhanapiya, Abdolla | |
| dc.contributor.author | Zhumabayev, Alikhan | |
| dc.contributor.author | Sabitova, Sabina | |
| dc.contributor.author | Seilov, Sayat | |
| dc.contributor.author | Koldasbek, Yedil | |
| dc.date.accessioned | 2026-06-09T06:52:39Z | |
| dc.date.issued | 2026-04-29 | |
| dc.description.abstract | Reducing communication barriers between the hearing-impaired community and the broader society remains a critical challenge, as recognizing sign language gestures requires expertise that most people do not possess. This paper addresses this problem by evaluating and designing an ASL recognition system using deep learning approaches. The main objectives are to identify the most accurate models for both static alphabet-level and dynamic word-level gesture recognition, and to incorporate them into a real-time recognition system. For image-based recognition, CNN-based and hybrid models are trained and evaluated on the ASL Alphabet datasets. For video-based recognition, pretrained SlowFast, Slow-only, and ViViT models are evaluated on WLASL100 and WLASL300 subsets, and further tested on a custom-collected set of 300 sign language videos. The main results are as follows: (1) image-based experiments identify ResNet as best-performing model with accuracy >93 % for different configurations; (2) video-based experiments show ViViT achieving 51.9% and 27.5%, SlowFast achieving 65.38% and 61.38%, and Slow-only achieving 64.34% and 54.24% on WLASL100 and WLASL300 respectively; (3) inference on the custom test set yields the highest accuracy of (61%) with SlowFast Networks. These results confirm that CNN-based approaches remain state-of-the-art for static recognition, while spatiotemporal models are essential for dynamic word-level gesture recognition. | |
| dc.identifier.citation | Zhanapiya, Abdolla; Zhumabayev, Alikhan; Sabitova, Sabina; Seilov, Sayat; Koldasbek, Yedil (2026) American Sign Language Machine Learning based Translation Model. Nazarbayev University School of Engineering and Digital Sciences | |
| dc.identifier.uri | https://nur.nu.edu.kz/handle/123456789/18939 | |
| dc.language.iso | en | |
| dc.publisher | Nazarbayev University School of Engineering and Digital Sciences | |
| dc.rights | Attribution-NonCommercial-NoDerivs 3.0 United States | en |
| dc.rights.uri | http://creativecommons.org/licenses/by-nc-nd/3.0/us/ | |
| dc.subject | Sign Language Recognition | |
| dc.subject | Convolutional Neural Networks (CNNs) | |
| dc.subject | ViViT Model | |
| dc.subject | SlowFast Net works | |
| dc.subject | Slow-only Model | |
| dc.subject | Real-time Recognition | |
| dc.title | American Sign Language Machine Learning based Translation Model | |
| dc.type | Bachelor's Capstone project |
Files
Original bundle
1 - 2 of 2
Loading...
- Name:
- American Sign Language Machine Learning based Translation Model.pdf
- Size:
- 689.49 KB
- Format:
- Adobe Portable Document Format
- Description:
- Senior Project Report
Loading...
- Name:
- group47_final_presentation.pptx
- Size:
- 1.61 MB
- Format:
- Microsoft Powerpoint XML
- Description:
- Senior Project Presentation