[논문리뷰] Scaling Large Reasoning Models beyond Human Supervision: A Path toward Superintelligence본 논문은 Large Reasoning Models (LRMs)가 인간의 Supervision이 모델 생성 경험의 규모와 복잡성을 따라잡기 어려워지는 상황에서도 추론 능력을 지속적으로 향상시킬 수 있는 방법을 모색합니다.#Review#Large Reasoning Models#Human Supervision#Superintelligence#Reinforcement Learning#Reward Models#Self-play#Autonomous Agents#Curriculum Learning2026년 8월 31일댓글 수 로딩 중
[논문리뷰] Guided Self-Evolving LLMs with Minimal Human Supervision본 논문은 기존의 자율 진화(self-evolving) 언어 모델(LLM)이 겪는 불안정성, 성능 정체, 개념 표류(concept drift) 및 다양성 붕괴(diversity collapse) 문제를 해결하고자 합니다.#Review#Self-Evolving LLMs#Self-Play#Reinforcement Learning#Curriculum Learning#Few-shot Learning#Human Supervision#Concept Drift#Diversity Collapse2025년 12월 2일댓글 수 로딩 중