[논문리뷰] Breaking the Vision-Action Shortcut: Latent Interface Training for Generalizable Robotics Foundation Models본 논문은 로봇 파운데이션 모델이 학습 데이터 내에서는 강력한 성능을 보이지만, 카메라 구도나 조명 등 시각적 분포가 변화할 때 성능이 급격히 저하되는 일반화 문제를 해결하고자 합니다.#Review#Robot Foundation Models#Vision-Action Shortcuts#Latent Interface Training#Generalization#Spatial-Goal-Conditioned#VLA#WAM2026년 9월 13일댓글 수 로딩 중