[논문리뷰] Agent-G^2: Gaussian Guidance for Agentic Reinforcement Learning본 논문은 기존 Hint-based RL 방법론들이 Guidance Depth를 단일 스칼라 값으로 고정하여 작업 간의 난이도 차이를 반영하지 못하는 한계를 해결하고자 합니다.#Review#Agentic Reinforcement Learning#Hint-based RL#Guidance Depth#Adaptive Gaussian Schedule#Reward Sparsity#LLM Agents#GRPO2026년 8월 26일댓글 수 로딩 중