This is absolutely fantastic, I love it!
I think I've found a (minor) bug, there seems to be an inaccuracy with the time stamps. I assume you're taking the chapter start and end times from the video themselves, but they don't align with the summary in this video:
https://www.videogist.co/videos/eevblog-102-diy-constant-cur...
For example, chapter 5, Heat Sink and Power Dissipation starts at 5:18 in your summary, but in the actual video it starts at 12:15.
EDIT:
I'm just now seeing that your summary only has 7 chapters, but the YouTube video is segmented into 13 chapters, so it appears you're not using them from the video after all. Are you doing the segmenting with the LLM as well? How do you get timestamps from that?
EDIT END
Anyway, thanks for this fantastic product, I was about to build something similar, but with a different focus:
Assume I've already seen a video and I want to look up a detail from that video, say how to do thermal calculations like in the example video, but I remember neither the video name nor the time stamp. I'm trying to build an app that generates embeddings for chunks / chapters of videos which I can then search semantically.
Do you have something similar planned? Because that's something I'd pay good money for.
Also, would you mind sharing a little bit about the tech stack? I'm assuming you're using yt-dlp to download videos with chapters and running whisper for a transcript, then something like gpt-3.5-turbo for summaries? Because that's how I'm doing it right now :D