How Kiki Works

— and what she deliberately can't do. A plain-language system card for the AI your child talks to.

AI companies publish "system cards" — technical documents explaining what their systems can and cannot do. They are written for researchers. Ours is written for you, because our users are children, and we think a parent's right to understand the AI their child talks to should not depend on reading a research paper.

What happens when your child speaks

1

Listening

Speech recognition (Google Cloud) converts your child's voice to text. It is configured for how bilingual children actually speak — Hindi and English mixed in one sentence, high-pitched voices, mid-sentence restarts. Child speech is the hardest class of audio there is, and it is the class we measure and tune for.

2

Understanding and responding

Large language models (from Google and Anthropic) work inside a tightly scripted frame: evaluate the Hindi your child produced — crediting the attempt, not punishing the English — plan Kiki's reply for the current lesson topic and your child's level, and note any error for later practice. The models never converse freely; they fill a bounded role we define.

3

Speaking

Kiki's voice (ElevenLabs) was not picked by taste. We ran an acoustic study across 15 candidate voices, measured their pitch, and chose a warm, Hindi-native voice in the register of an elder sister — because that is the register children's most beloved characters actually live in.

Why we built it this way

The fashionable way to build a voice AI today is a speech-to-speech model: audio in, audio out, nothing in between. We deliberately chose the older, more inspectable architecture — because with no transcript in the middle, there is nothing to evaluate, nothing to show you in a weekly report, no error record to practise from, and no way to measure whether your child is actually improving. Every step of our pipeline is observable, correctable, and parent-visible. For a children's product, we believe that is not a limitation. It is the point.

What Kiki is designed to do

What Kiki deliberately can't do

Where it can go wrong — our known limitations

No AI system is perfect, and we would rather you hear the imperfections from us:

What we don't do — stated plainly

We do not train our own frontier AI models; we build carefully on top of the best available ones. We do not yet run a formal adversarial "red-teaming" programme — the conversation surface is narrow by construction, but as we scale, formal adversarial testing is on our roadmap, along with a fairness review of how our speech recognition handles different accents and dialects. We would rather name what is ahead of us than pretend it is behind us.

For what data is collected and who can see it, read our Child Safety & Data Promise. For the formal legal version, see the Privacy Policy & Terms.

Last updated: July 2026 · Hindi Tutor is a product of Prayaas Impact Technologies. Questions about anything on this page: shubham@hindispeakingtutor.in

← Back to Home