Abstract:Efforts to make online media accessible to a regional audience have picked up pace in recent years with multilingual captioning and keyboards. However, techniques to extend this access to people with hearing loss are limited. Further, owing to a lack of structure in the education of hearing impaired and regional differences, the issue of standardization of Indian Sign Language (ISL) has been left unaddressed, forcing educators to rely on the local language to support the ISL structure, thereby creating an array of correlations for each object, hindering the language building skills of a student. This paper aims to present a useful technology that can be used to leverage online resources and make them accessible to the hearing-impaired community in their primary mode of communication. Our tool presents an avenue for the early development of language learning and communication skills essential for the education of children with a profound hearing loss. With the proposed technology, we aim to provide a standardized teaching and learning medium to a classroom setting that can utilize and promote ISL. The goals of our proposed system involve reducing the burden of teachers to act as a valuable teaching aid. The system allows for easy translation of any online video and correlation with ISL captioning using a 3D cartoonish avatar aimed to reinforce classroom concepts during the critical period. First, the video gets converted to text via subtitles and speech processing methods. The generated text is understood through NLP algorithms and then mapped to avatar captions which are then rendered to form a cohesive video alongside the original content. We validated our results through a 6-month period and a consequent 2-month study, where we recorded a 37% and 70% increase in performance of students taught using Sign captioned videos against student taught with English captioned videos. We also recorded a 73.08% increase in vocabulary acquisition through signed aided videos.

Modeling the Speed and Timing of American Sign Language to Generate Realistic Animations

Empirical Investigation of Users' Preferred Timing Parameters for American Sign Language Animations

Toward an example-based machine translation from written text to ASL using virtual agent animation

Text-driven Visual Prosody Generation for Embodied Conversational Agents

DiffSign: AI-Assisted Generation of Customizable Sign Language Videos With Enhanced Realism

Jointly Harnessing Prior Structures and Temporal Consistency for Sign Language Video Generation

Everybody Sign Now: Translating Spoken Language to Photo Realistic Sign Language Video

American Sign Language Video to Text Translation

Regression Analysis of Demographic and Technology-Experience Factors Influencing Acceptance of Sign Language Animation

Sign Language Production with Latent Motion Transformer

American Sign Language Translation Using Wearable Inertial and Electromyography Sensors for Tracking Hand Movements and Facial Expressions

Enhancing Portuguese Sign Language Animation with Dynamic Timing and Mouthing

Enhancing Bidirectional Sign Language Communication: Integrating YOLOv8 and NLP for Real-Time Gesture Recognition & Translation

Towards Fast and High-Quality Sign Language Production

A Classification Model Utilizing Facial Landmark Tracking to Determine Sentence Types for American Sign Language Recognition

Pose-Guided Fine-Grained Sign Language Video Generation

Fine-tuning of sign language recognition models: a technical report

Modeling Intensification for Sign Language Generation: A Computational Approach

SignSpeak: Open-Source Time Series Classification for ASL Translation

Automated 3D sign language caption generation for video

Neural Sign Actors: A diffusion model for 3D sign language production from text