# ChitraKatha_AI

- **Event:** [Google DeepMind Bangalore Hackathon](https://cerebralvalley.ai/e/google-deepmind-bangalore-hackathon)
- **When:** Sat, Jul 11 at 9:00 AM – 10:00 PM (GMT+5:30)
- **Where:** Marathahalli, Marathahalli Main Road
- **Team:** [Sahaja Pulluru](https://cerebralvalley.ai/u/Saaju)
- **GitHub:** https://github.com/SahajaPulluru/ChitraKatha_AI
- **Demo video:** https://youtu.be/827Hy094FFE
- **Gallery:** https://cerebralvalley.ai/e/google-deepmind-bangalore-hackathon/hackathon/gallery
- **Page:** https://cerebralvalley.ai/e/google-deepmind-bangalore-hackathon/hackathon/gallery/127

The Engagement Gap: Traditional educational media fails to capture children's short attention spans, whereas ChitraKatha provides personalized, interactive storybooks that turn screen time into an engaging learning experience.
The Character Consistency Bottleneck: Standard generative image models produce completely different-looking characters across pages; ChitraKatha solves this by utilizing a multimodal Vision Profiler to extract and enforce an immutable, repeatable visual signature of the child as the main character.
The High-Latency Production Cycle: Crafting a children's book usually takes weeks of sequential writing, illustrating, and voice-recording; ChitraKatha’s parallelized multi-agent orchestrator slashes this high-throughput pipeline down to under 20 seconds.
The Regional Language Divide: Over 80% of interactive children's media is locked behind English; ChitraKatha democratizes creative learning by generating and narrating premium stories in native Indian scripts (Hindi, Kannada, Tamil, Telugu).
Cultural Disconnection in AI: Global media lacks local folklore context; ChitraKatha bridges this gap with a "Mythological Twist" toggle that seamlessly weaves children as protagonists into rich Indian mythological landscapes (Panchatantra, Ramayana) to preserve cultural heritage.

gemini-3.5-flash (Multi-Agent Core): Powers Agent A: VisionProfiler to synthesize child character designs, Agent B: NarrativeArchitect to draft story script structures, and Agent D: LinguisticLocalizer to translate narration tracks into regional scripts.
gemini-omni-flash-preview (Agent C: Motion Orchestrator / Omni): Translates flat, static 2D story illustrations into beautiful, fluid, and temporally-stable cinematic motion video segments.
gemini-3.1-flash-lite-image (High-Throughput Imagen Engine): Executes high-velocity concurrent image generations to design the rich Ghibli-inspired, high-contrast background artwork.
gemini-3.1-flash-tts-preview (Narration Voice Engine): Synthesizes rich, expressive, human-like voice recordings of regional text narrations to provide a complete multisensory experience for children.

---

Markdown version of https://cerebralvalley.ai/e/google-deepmind-bangalore-hackathon/hackathon/gallery/127. Site index for agents: https://cerebralvalley.ai/llms.txt · full text: https://cerebralvalley.ai/llms-full.txt
