Yeah, VideoDB is the next-gen infrastructure for videos and actually less costly than current video infrastructure.
But you can use any LLM for analysing the transcript.
But you can use any LLM for analysing the transcript.
Can you confirm if I could skip using VideoDB by using Whisper to transcribe the video, and then use that transcript with LLaMa to extract the important parts?