New here (hey 👋) and wanted to share something. We recently made public an ad intel app we'd been using internally in our agency. The app/system runs video ads through Gemini to pull out a transcript, the hook, and a few timed moments. It works fairly well - especially for ideation. The transcription is genuinely accurate.

A user opened a 2:22 video and saw a moment marked at 3:37.

Not a rounding error. Fifty percent past the end of a video the user was looking at.

The part that makes it interesting

The obvious guess is that the model didn't know how long the video was. That would be a reasonable bug and an easy fix.