Authors
Loading...
IV-Bench evaluates image-grounded video reasoning in MLLMs, highlighting significant performance gaps.
Ma et al. (2025) studied this question.