[논문리뷰] Online Learning with LLM Experts from Limited FeedbackLarge Language Models (LLMs)는 다양한 태스크에서 널리 활용되고 있지만, 모델마다 비용과 능력이 상이하며, 특정 모델이 모든 태스크에서 다른 모델을 압도하지 못하는 경우가 많습니다.#Review#Online Learning#LLM Experts#Limited Feedback#Contextual Bandit#Regret Minimization#Adaptive Routing2026년 9월 13일댓글 수 로딩 중