[논문리뷰] From Pretraining to Proficiency: Real-World Subtask RL for Long-Horizon Manipulation with Minimal Human Intervention본 논문은 긴 호흡의(Long-horizon) 로봇 조작 태스크에서 Pretrained policy가 전반적으로는 유능하지만 특정 핵심 Subtask에서 반복적으로 실패하는 문제를 해결하고자 합니다 .#Review#Robot Learning#Long-horizon Manipulation#Reinforcement Learning#Foundation Models#Residual Policy#Subtask Adaptation2026년 9월 20일댓글 수 로딩 중