Live capture / 01
For words, select Record a word and wait for the countdown. Signing alone does not start a recording.
Use a well-lit space and keep your palm facing the camera.
Explore sign-to-text and voice, one gesture at a time.
For words, select Record a word and wait for the countdown. Signing alone does not start a recording.
Use a well-lit space and keep your palm facing the camera.
Hold each sign steadyLetters are added after a stable detection.
Move between signsLower your hand briefly before repeating a letter.
No saved sessions yet.
Submit the study’s 5-point evaluation for accuracy, usability, and performance. Responses are stored anonymously.
Seven static letters are pretrained. Teach the 5 gesture in AI Studio to add it.
Seven letters come pretrained from ASL image examples. Start the camera to try them, or capture your own examples to personalize recognition. Teach the 5 gesture before using it. Uncertain poses are left blank instead of guessed. Camera frames stay on your device.
Start the camera above, hold the selected gesture, then capture samples.
Select Record a word beside the camera. It switches to Words in motion and starts a countdown. Sign one supported word during the capture, then review the result here. Personal training is optional.
Loading pretrained words…
Pretrained vocabulary will appear here.
Each recording has a 1.5-second countdown, then captures 2.5 seconds of movement.
Record at least three examples for a new word. Removing personal examples keeps the pretrained vocabulary.
Use only the listed words. Other signs can be mistaken for a supported word. This small PopSign ASL model has limited accuracy; review each suggestion before adding or speaking it. Personal examples stay in this browser. Evaluation results · Dataset credit
HandSpeak Web is a capstone prototype that uses computer vision to help bridge everyday communication between deaf, hard-of-hearing, and hearing individuals. It recognizes selected static letters and offers an experimental pretrained vocabulary for moving signs. Face landmarks are displayed for visual feedback; they are not used to interpret identity, emotions, or ASL grammar.