[논문리뷰] BVB: Benchmarking Agentic Video Understanding via Programmatic Reconstruction in Blender기존 Video Understanding 벤치마크는 주로 Question Answering (QA) 방식으로 모델을 평가하지만, 이러한 방식은 모델이 Answer Prior나 단일 프레임(Single Frame) 정보에 의존하여 정답을 맞출 수 있어 비디오의 Spatiotemporal한 이해를 온전히 입증하지 못한다.#Review#Agentic Video Understanding#Programmatic Reconstruction#Blender#Benchmark#Multimodal Agents#Dual VQA#Latent Similarity2026년 9월 14일댓글 수 로딩 중