A licensed dataset
Video recorded by volunteers who agreed, in writing, to how it would be used. No scraped faces and no ambiguous provenance.
We build the consented video data that visual speech recognition needs, and we train models on it from scratch. Every clip is licensed, every contributor is a volunteer, and every model is clean room.
Video recorded by volunteers who agreed, in writing, to how it would be used. No scraped faces and no ambiguous provenance.
Nothing pretrained and nothing borrowed from research code that forbids commercial use. What we ship, we can account for.
The models are made to sit behind other software, from captioning and conferencing to accessibility tools. Partners license the data, the model, or both.
People who have lost their voice to illness or surgery still form words. A camera can read them when a microphone cannot.
A factory floor defeats a microphone. So does an open office where speaking aloud is not an option.
Dropped packets, a muted mic, a bad connection. The picture keeps carrying the sentence after the sound stops.