[논문리뷰] CLBench-V: Evaluating Multimodal Context Learning from Grounding to Knowledge Acquisition본 논문은 현대의 멀티모달 모델들이 실제 환경에서 제공되는 다중 모달 컨텍스트를 얼마나 효과적으로 학습하고 활용하는지 측정하기 위한 표준화된 벤치마크가 부재하다는 문제 의식에서 출발합니다.#Review#Multimodal Context Learning#Benchmark#Context Grounding#Information Application#Knowledge Acquisition#Multimodal LLMs#Long-context2026년 7월 29일댓글 수 로딩 중
[논문리뷰] Qwen-Image-Agent: Bridging the Context Gap in Real-World Image Generation본 논문은 기존 Text-to-Image (T2I) 모델이 실세계의 복잡하고 모호한 요청을 처리하는 데 겪는 구조적 한계를 해결하고자 합니다. T2I 모델은 일반적으로 완전히 명시된 프롬프트에 최적화되어 있으나, 실세계의 사용자 요청은 불완전하거나 맥락 정보를 필요로 하는 경우가 많습니다 .#Review#Agentic Image Generation#Context Gap#Context-Aware Planning#Context Grounding#IA-Bench#Multimodal Large Language Model (MLLM)2026년 6월 25일댓글 수 로딩 중