TubeTutor AI - Turn any video into visual notes, transcripts & mind maps

by
TubeTutor AI turns YouTube and TikTok videos into structured learning workspaces. It combines transcripts with visual analysis to capture on-screen steps, then creates summaries, chapters, multilingual notes, editable mind maps, and Notion-ready knowledge.

Add a comment

Replies

Best
Maker
📌
Hey Product Hunt 👋 I’m Chen, the maker of TubeTutor AI. I’ve worked in video editing and post-production for more than 15 years. One thing has always bothered me about learning from videos: some of the most useful information is visible on screen but never spoken aloud. Most video summarizers only read subtitles. That works for interviews, but it often falls short with editing tutorials, coding walkthroughs, software demos, silent videos, and lessons where the real instruction is hidden in cursor movements, menus, parameters, diagrams, or visual steps. So I built TubeTutor AI to combine transcript analysis with visual understanding. Paste a YouTube or TikTok link, and TubeTutor can turn it into: 🎬 Visual notes that capture important on-screen actions 📝 Concise summaries and timestamped chapters 🔍 Synchronized transcripts with clickable timestamps 🧠 Editable mind maps 🌍 Multilingual study notes and translated subtitles 📚 Reusable, Notion-ready knowledge TubeTutor is designed for students, creators, professionals, and language learners who want to understand videos without repeatedly scrubbing through the entire timeline. There’s a free plan, so you can test it with your own video. I’d especially love your feedback on: 1. Do the visual notes capture the moments you would normally pause to write down? 2. Which output is most useful to you: the summary, transcript, mind map, or notes? 3. What would make TubeTutor fit your learning workflow better? Try it with a tutorial where important steps happen on screen—that’s where TubeTutor is meant to stand out. Thanks for checking it out! I’ll be here throughout launch day and would genuinely love to hear what works and what doesn’t.

Half a coding tutorial is cursor movements and menu clicks nobody narrates, that part is true and underrated. How does the visual analysis tell an important on-screen action apart from someone just scrolling around?