[논문리뷰] Generation as Auxiliary Supervision: Enhancing Visual Understanding at Zero Inference Overhead via Decoupled Embedding Prediction본 논문은 MLLM에서 시각적 이해와 생성이 서로 분리된 목표로 다뤄지며, 기존의 생성 학습이 모델의 고차원적 이해 능력을 저해하는 문제를 해결하고자 합니다.#Review#Multimodal Large Language Models#Visual Understanding#Auxiliary Supervision#Next Embedding Prediction#Mixture-of-Transformers#Representation Learning2026년 8월 16일댓글 수 로딩 중