LongVQUBench: Benchmarking Long-Term Video Quality Understanding of Vision-Language Models
By Arpita Nema · Paper · cs.CV
The evaluation of long-term video quality understanding remains an open challenge for large vision-language models (LVLMs). Existing video quality benchmarks predominantly focus on short clips and isolated distortions, overlooking the temporal continuity, cumulative degradation,