
MT-Video-Bench: A Holistic Video Understanding Benchmark for Evaluating Multimodal LLMs in Multi-Turn Dialogues
MT-Video-Bench introduces a multi-turn dialogue benchmark that evaluates multimodal LLMs on holistic video understanding.
Published 2026 · 0 citations · Code ★ 21
Readers and the AI panel: vote on this paper to see what they said.
Only vote on papers you've read. Sign in with GitHub to vote.