[논문리뷰] CAPEval: A Decoupled Caption Evaluation across Understanding and Generation기존의 Caption Quality Evaluation 방식은 Caption 품질을 단일 Scalar Objective로 간주하여, Caption이 담는 Visual information의 범위(Coverage)와 주장하는 내용의 Factual correctness (Precision)라는 두 가지 distinct property를 혼동하는 문제점을 안고 있었다.#Review#Caption Evaluation#Multimodal Understanding#Text-to-Image Generation#Coverage#Precision#Downstream Utility#Vision-Language Models#Dataset Curation2026년 8월 4일댓글 수 로딩 중
[논문리뷰] CaptionQA: Is Your Caption as Useful as the Image Itself?본 논문은 기존 MLLM 평가 방식이 캡션의 실제 활용성, 즉 다운스트림 태스크에서 이미지를 대체할 수 있는 능력 을 간과한다고 지적합니다.#Review#Image Captioning#Caption Evaluation#Multimodal LLM#Utility-based Benchmark#Question Answering (QA)#Domain-specific Taxonomy#Hallucination#MLLM Evaluation2025년 11월 30일댓글 수 로딩 중