AI Briefing
KO

Video LMM Benchmark 'TimeScope' Released

·2025.07.23 09:00

Key point

TimeScope, an open-source benchmark for measuring temporal understanding of long videos, has been released.

Details

Recently, video large multimodal models (LMMs) claim to handle long contexts, but in reality they often remain at the level of simple Visual Search.

TimeScope is an open-source benchmark that inserts short 5-10 second video clips (Needle) into long videos ranging from 1 minute to 8 hours to evaluate a model's actual temporal understanding capability.

The main evaluation items focus on the following three core capabilities:

  • Localized Retrieval: The ability to find a specific segment within a vast video and answer questions about it
  • Information Synthesis: The ability to gather and organize multiple pieces of detailed information across the timeline
  • Fine-Grained Temporal Perception: The ability to precisely analyze actions and events that require dense frame sampling

The purpose of this benchmark is to verify whether a model truly understands the flow of the video and the causal relationships between events, beyond simply skimming through frames.

This summary was generated automatically by AI. Check the original for the author's claims and context. Copyright belongs to the original author.

Our guide explains how the AI works. Report summary errors, attribution issues, or removal requests via Contact.