Storytime AI is a voice-first, multilingual storytelling platform for children aged 2–10. Built using Gemini, Whisper, LangChain, and ElevenLabs, it delivers personalized, screen-free narratives that adapt to each child’s voice input and language level. I led the product end-to-end — from concept to launch — optimizing for latency, education, and engagement.
Coming out soon on mobile and selected for Google’s AI Startup Program
Parents told us:
We built StoryTime to give children agency through voice and give parents peace of mind through privacy-first design.

Parent selects story length and theme preferences from our intuitive interface.
TTS greets the child: “Hi Kiaan! What story would you like to hear?”
Child responds aloud. STT captures the input.
Story branches dynamically across 2–5 interaction points based on age and time.
LLM uses token-based pacing to keep responses short, relevant, and engaging.
Each session ends with a positive moral and a soft voice cue. No data is stored. Fully COPPA/GDPR compliant.
Engineered for latency <5s to match child attention span
Loves trains and animals, craves screen-free playtime with educational content. Enjoys voice-based interactions and simple prompts.
Juggling work and parenting, seeks screen-free engagement for child, prioritizing learning and routine.
Before Storytime AI: Passive watching on YouTube. After Storytime AI: Personalized, interactive storytelling experience with metrics for improved engagement.
How parents and children experience StoryTime from first touch to final moral.
If our conversational AI voice-agent engages children ages 3–6 in adaptive, research-backed language education through personalized, interactive storytelling—while maintaining at least 80% alignment with vocabulary objectives and a 100% child-safe content score—then we can create measurable learning gains and parent satisfaction, opening monetization opportunities through B2C (freemium/subscription) and B2B (early learning centers).

Inputs: Child’s voice, name, language → transcribed by Whisper
Processing: Gemini 2.0 + LangChain (interaction logic)
Memory: MongoDB + LangSmith traces (session + preference)
Output: ElevenLabs voice (TTS) → personalized, multilingual
Loop: Real-time feedback & sentiment tracing refine future responses
*This is actively being updated and may not represent the current state (Moved from v1 - Universal and Open AI for STT and TTS)
Children consistently engage from start to finish
Successfully tested with families nationwide
Full COPPA and GDPR-K certification
Selected by Google's prestigious AI Startup Program for innovation in child-safe AI technology. Our token-safe Gemini and LangChain implementation sets new industry standards.
- Voice-to-Voice MVP (English)
- Gemini 2.0 Flash integration
- LangChain for interaction handling
- Whisper for voice transcription
- ElevenLabs for output voice
- Mobile app launch
- Support for multiple character and language voices
- Offline mode
- Parental control for voice and language selection
- Integrate RAG with LangChain
- Vector DB for reusable story components
- Support for multilingual user prompts
- Preference memory across sessions
- Checkpoint continuation logic
- LangSmith analytics traces
- Drop-off and interaction scoring
- Sentiment-driven fallback logic
- Adaptive prompt tuning and model refinement
- AI-driven voice style personalization
- Parent-facing dashboard and learning insights
- Story-level comprehension and engagement analytics
- Real-time multilingual narration powered by on-device inference
Storytime's roadmap, architecture, and core features are grounded in real-world voice-of-customer insights. Each user story and feature set reflects tested needs from parents and children across multiple environments — from bedtime routines to multilingual homes.
The following sections illustrate how user input shaped our design, from MVP to multilingual, adaptive storytelling.
As a child, I want to speak to the app and hear a response in a fun voice,
so that I can enjoy a magical story experience hands-free.
As a parent, I want the story to wait for my child's response before continuing,
so that it feels more interactive and less like a passive video.
As a parent, I want to select different story voices and language preferences,
so that my child can hear familiar or diverse voices and learn new languages.
As a parent, I want to quickly open the app and see recent stories and themes played,
so I can track what my child is engaging with.
As a child, I want the story to remember my favorite animal and my name,
so that it feels like it was made just for me.
As a parent, I want the story to help teach empathy or problem-solving,
so my child can learn while being entertained.
As a parent, I want to see how long my child engaged and where they dropped off,
so that I can understand attention span and improvement over time.
As a PM, I want to know which interaction points are working and which themes are skipped,
so that I can fine-tune our AI prompts for better outcomes.
As a multilingual parent, I want stories to be available in multiple languages,
so that my child can learn and connect in our native language.
As a parent, I want to ensure the story is safe, educational, and ad-free,
so I feel confident letting my child use it independently.
These stories directly shaped Storytime's product architecture, roadmap phases, and interaction point logic._
Interactive Story Generator, Voice-to-Voice Interface, Theme Selection Engine, Parent Dashboard
Multilingual Story Support, Sentiment-Adaptive Storytelling, Memory-Based Personalization
COPPA/GDPR compliance, audio moderation, token limits, age-appropriate interaction times
Principal Product Manager (Founder-led)
Led the full lifecycle — from concept → MVP → mobile platform with multilingual, voice-first experience. Defined architecture, roadmap, and metrics for production-quality AI integration.
Team: 4 PMs, 2 frontend devs, 1 ML engineer, 1 content designer
Stack: Gemini 2.0 Flash, LangChain, Whisper STT, ElevenLabs TTS, LangSmith, MongoDB, Vector DB
Tools: Notion, Figma, Retool, ElevenLabs, Latitude AI, Replit, Claude, DoveTail
"From bedtime stories to multilingual, intelligent narration — Storytime is voice-first AI storytelling that listens, adapts, and delights."
Used by families across 3 continents. Built for kids aged 2–10. Designed for high recall, repeat engagement, and voice-based learning.
🎖️ Designed to demonstrate Principal PM-level thinking, execution, and AI fluency across Gemini, LangChain, and RAG loops.
📩 Lets connect 🔗 LinkedIn
Storytime AI: Voice-First, Mobile, and Built for Imagination