EdTech
Green & Gold Guru
Multimodal AI tutor that teaches step-by-step using voice, visuals, and a live whiteboard — without ever giving away the answer.
3
input modalities
01 — Problem
It's late at night, you're stuck on a complex problem, and office hours are over. Existing chatbots just give you the answer. Students need a tool that teaches through guided problem-solving across voice, text, and visual input.
02 — Approach
- 01
Streamlit-based multimodal interface with a live canvas whiteboard, voice input via speech recognition, and PDF analysis via pyPDF2.
- 02
OpenRouter routes to Gemini for step-by-step tutoring that guides without revealing the answer, adapting to whiteboard drawings and spoken questions.
- 03
Edge TTS provides natural voice responses; NumPy processes whiteboard images for real-time visual analysis.
03 — Results
3
modalities (voice, whiteboard, PDF)
Step-by-step
guided tutoring without giving answers
Stack
Next
Bowser→
Autonomous pediatric physical therapy robot that coaches exercises in real time using computer vision.