SamaVaad

A working software preview of the two-way conversation loop — before the glove, this is the brain of SamaVaad, live in your browser.

ISL Sign Regional Speech Spoken Reply ISL Avatar

The SamaVaad Infinity Loop

Closing the conversational gap through continuous hardware and AI integration.

๐Ÿงค

Direction 1: Deaf → Hearing

The glove captures ISL signs, an on-device AI model converts them into a grammatically correct sentence in the hearing person's regional language, and it's spoken aloud.

Example Translation
"WATER + WANT"
→ hears "Mujhe paani chahiye" (Hindi)
๐Ÿ’ฌ

Direction 2: Hearing → Deaf

The hearing person speaks naturally, the AI transcribes and restructures it into ISL grammar, and an animated avatar signs the response back.

Example Translation
"Priya, have you eaten lunch?"
→ avatar signs "PRIYA EAT LUNCH FINISH YOU"

Step-by-Step Workflow

Deaf → Hearing

1 Sign in ISL
2 Flex sensors + IMU capture motion
3 On-device gesture recognition classifies the sign
4 Grammar bridge reconstructs a regional-language sentence
5 Text-to-speech plays the audio

Hearing → Deaf

1 Speak naturally
2 Speech recognized and transcribed
3 NLP simplifies it into ISL grammar
4 Sign sequence is mapped
5 Avatar animates the signs

The Six AI Layers

The proprietary architecture powering true bidirectional fluency.

Layer 1

ISL Gesture Recognition

A temporal CNN/LSTM model on-device classifies flex-sensor and motion data into ISL signs in real time, adapting to each signer's style.

Layer 2 Core Innovation

Cross-Linguistic Grammar Bridge

True bidirectional fluency requires a Cross-Linguistic Grammar Bridge, not mere word-swapping. Indian Sign Language has a fundamentally different grammar from spoken languages (lacks articles, tense markers, uses strict topic-comment order). Our Seq2Seq Transformer AI Engine strips tense markers and restructures syntax rules in real-time to ensure total fluency in both directions.

Layer 3

Regional Language Text-to-Speech

Converts the reconstructed sentence into natural speech in the target Indian language, with voice models cached on-device for offline use.

Layer 4

Speech Recognition + NLP Simplification

Transcribes the hearing person's speech and restructures it into clean ISL-compatible grammar, removing articles and tense markers.

Layer 5

ISL Avatar Animation Engine

Renders the sign sequence as a smooth animated avatar using motion interpolation between signs, rather than robotic pre-recorded clips.

Layer 6

Federated Personalization & Context

The model fine-tunes to each user's signing style on-device (federated learning, no personal data leaves) and uses recent conversation context to resolve ambiguous signs.

Curious what the physical glove looks like?

Explore the Hardware →