[논문리뷰] Claim-Level Reliability Assessment for Efficient Test-Time Reasoning본 논문은 LLM의 Test-Time Scaling 과정에서 발생하는 신뢰성 평가의 불투명성과 계산 자원의 비효율성 문제를 해결하고자 합니다.#Review#Test-Time Reasoning#Reliability Assessment#Falsification#Chain-of-Thought#LLM#Consensus#Inference Scaling2026년 8월 16일댓글 수 로딩 중
[논문리뷰] GradCuit: Credit-Assigned Gradient Flow Enables Robust and Interpretable Test-Time Latent Reasoning본 논문은 기존 test-time latent reasoning 방식들이 겪는 구조적 한계인 불투명한 credit assignment 문제를 해결하고자 한다.#Review#Test-Time Reasoning#Latent Space#Gradient Flow#Transformer Circuits#Credit Assignment#Robustness2026년 8월 3일댓글 수 로딩 중
[논문리뷰] LongCat-Flash-Thinking-2601 Technical Report본 논문은 장기적인 상호작용과 추론이 요구되는 에이전트 태스크 에서 기존 모델들의 한계를 극복하고, 뛰어난 에이전트 추론 능력을 가진 오픈소스 MoE(Mixture-of-Experts) 대규모 언어 모델인 LongCat-Flash-Thinking-2601 을 개발하는 것을 목표로 합니다.#Review#Agentic AI#Large Language Models (LLMs)#Mixture-of-Experts (MoE)#Reinforcement Learning (RL)#Context Management#Scalable Training#Test-Time Reasoning#Open-Source Model2026년 1월 25일댓글 수 로딩 중