Overview

The goal of this project was intended to serve as a proof of concept for a machine learning model capable of receiving real time video input and displaying the digit or letter associated with that gesture from American Sign Language (ASL). This could be used in a variety of settings where waiting for an interpreter would not be feasible or degrade services rendered. One example would be the emergency room in a hospital.

Outcome

The outcome of this project was a model that was successfully able to take a live video stream and identify zero and one. While the model would stick on its last prediction when the current frame represents neither of the categories being evaluated, it is robust in its ability to output the correct value on the input. Due to time constraints the scope was scaled back as this working model already fulfilled the requirements of the class project.

Structure

More detailed breakdown to come…