We wanted to explore how technology could help people who use sign language communicate with people who don’t understand it. As a beginner team, we set out to build a small prototype that recognizes a limited set of signs and turns them into speech.

Our web app uses a webcam to capture hand movements, recognizes signs we trained it on, and speaks the detected words aloud. It’s an early prototype focused on a small vocabulary.

We split the work across the interface, camera and hand tracking, sign recognition model, and speech integration. We used HTML, CSS, and JavaScript for the interface, MediaPipe for hand landmarks, and a trained model to recognize signs. We worked on connecting the recognized words to ElevenLabs for spoken output.

We learned how camera input, machine learning, and speech generation fit together. We also learned that recognizing a moving sign requires capturing a sequence of movements, rather than just one hand position. Working together taught us how to use GitHub branches and connect code developed separately.

Built With

Share this project:

Updates

Submission history