[논문리뷰] UniSwap: Streaming Audio-Visual Identity Swapping for Talking Videos본 연구는 말하는 영상의 캐릭터 교체 작업에서 시각적 모습과 목소리를 동시에 일관성 있게 바꾸는 통합 시스템의 부재라는 문제를 해결하고자 합니다. 기존 방식들은 비디오 캐릭터 교체와 음성 변환 모델을 개별적으로 최적화하여, 결과물에서 입 모양과 음성 간의 오디오-비주얼 불일치가 발생하기 쉽습니다.#Review#Audio-Visual Generation#Identity Swapping#Streaming Diffusion#Diffusion Transformer#Distillation#Talking Video2026년 8월 13일댓글 수 로딩 중
[논문리뷰] TailBooster: A Dual-Layer Generative Framework for Extreme Value Augmentation with Operational Validity Enforcement본 논문은 항공 운송 시스템에서 빈번하게 발생하는 극단적 지연이나 비정상적인 비행 시간과 같은 Extreme event를 예측하기 위한 데이터가 역사적 기록 내에서 부족하다는 점을 해결하고자 합니다.#Review#Generative AI#Anomaly Detection#Synthetic Data Augmentation#Extreme Value Generation#Extreme Event Prediction#Air Traffic Management2026년 8월 13일댓글 수 로딩 중
[논문리뷰] Specification-first convergence with an AI coding agent: a case study of dismantling a core architectural invariant across 189 files in a 717k-line codebase with no test oracle and no human code review본 논문은 AI 코딩 agent가 대규모 아키텍처 refactoring을 수행할 때 기존 인간 코드 리뷰의 확장성 한계를 해결하고자 합니다.#Review#AI coding agents#large-scale refactoring#architectural invariants#software maintenance#agentic software engineering#verification loops#specification-first protocol2026년 8월 13일댓글 수 로딩 중
[논문리뷰] Spatial Memory Agent: Experience-Grounded Procedure Memory for Spatial Intelligence본 논문은 frozen VLM의 spatial reasoning 능력을 파라미터 업데이트 없이, 외부 expert tool에 대한 의존성도 최소화하면서 향상시키는 방법을 탐구합니다.#Review#Spatial Intelligence#Embodied Agents#Vision-Language Models#Procedural Memory#Transfer Learning#Training-free Self-evolution2026년 8월 13일댓글 수 로딩 중
[논문리뷰] SKILLER: Language-Level Reinforcement Learning for Reusable Skill Extraction in Small Language Models본 논문은 강력한 closed-source 모델의 높은 추론 비용 문제와, 소형 모델(Small-scale LVLMs)에 기존 스킬을 직접 이식할 때 발생하는 Model-Mismatch 문제를 해결하고자 한다.#Review#Small Language Models#Reinforcement Learning#Agent Skills#Natural Language Policy#Model-Mismatch#Cost-Effectiveness#Agentic Deployment2026년 8월 13일댓글 수 로딩 중
[논문리뷰] PlayWorld: Benchmarking World Models with Agent Players over Long-Horizon Objectives본 논문은 기존의 월드 모델 평가 방식이 인간의 직관적인 평가 방식과 괴리가 있다는 점을 핵심 문제로 정의합니다 .#Review#World Models#Agent Player#Long-Horizon Objectives#Geometry Consistency#Interaction Fidelity#VQA Rubric#Benchmarking2026년 8월 13일댓글 수 로딩 중
[논문리뷰] PixSDS: Why Latent SDS Makes Noisy Pixels본 논문은 latent-diffusion 기반의 SDS 최적화 과정에서 발생하는 구조적 색상 Artifact와 고주파 텍스처 노이즈 문제를 해결하고자 합니다.#Review#Score Distillation Sampling#Latent Diffusion Models#Variational Autoencoders#Pixel-space Optimization#Gradient Repair#Text-to-3D Generation2026년 8월 13일댓글 수 로딩 중
[논문리뷰] OmniScientist: An Omni-Modal Omni-Discipline AI Scientist본 논문은 기존 AI 과학자 모델들이 과학 연구의 워크플로우를 자동화함에도 불구하고, 정작 발견의 핵심 근거가 되는 원시 데이터의 구조적 정보를 충분히 활용하지 못하는 Evidence-incomplete 문제를 해결하고자 합니다.#Review#AI Scientist#Multimodal Agents#Autonomous Research#Scientific Discovery#LLM Agents#Tool Use2026년 8월 13일댓글 수 로딩 중
[논문리뷰] Massive Activations in Hybrid Linear Attention Large Language Models: Pre-Attention Spikes and Inter-Spike Plateaus본 논문은 HLA LLM 내부의 계층별 하이브리드화가 Activation Dynamics를 어떻게 재구성하는지 규명하고자 합니다.#Review#Hybrid Linear Attention#Massive Activations#Pre-Attention Spikes#Inter-Spike Plateaus#Attention Sinks#Language Modeling2026년 8월 13일댓글 수 로딩 중
[논문리뷰] LycheeMemory V2: Efficient Long-Term Memory for LLM Agents via Semantic Segment-Level Consolidation본 논문은 LLM Agent의 장기 기억(Long-term memory) 시스템 구축 시 발생하는 높은 연산 비용과 메모리 요약 과정에서의 정보 손실 문제를 해결합니다. 기존의 Eager Consolidation 방식은 대화가 길어질수록 LLM 호출 빈도가 급격히 증가하여 비효율적인 Token 비용을 유발합니다 .#Review#LLM Agents#Long-Term Memory#Semantic Segmentation#Memory Consolidation#Memory Encoding#Structured Retrieval2026년 8월 13일댓글 수 로딩 중
[논문리뷰] LiveAnimate: Stable Long-Form Streaming Human Animation in Real-Time본 논문은 기존의 pose-driven human animation 시스템들이 지닌 실시간성 결여와 장기 생성 시의 품질 저하 문제를 해결하기 위해 LiveAnimate를 제안한다.#Review#Human Animation#Streaming#Real-Time#Diffusion Transformer#Long-Form Generation#KV-cache2026년 8월 13일댓글 수 로딩 중
[논문리뷰] LLMRouter: Unified Infrastructure for Developing, Evaluating, and Deploying LLM Routers본 논문은 현재 LLM 라우팅 연구가 파편화된 공식과 호환되지 않는 코드베이스로 인해 체계적인 비교와 발전이 저해되고 있다는 점을 핵심 문제로 정의합니다 .#Review#LLM Routing#Infrastructure#Sequential Decision Process#xRouteBench#Cost-effective Deployment#Personalized Routing2026년 8월 13일댓글 수 로딩 중
[논문리뷰] Knowing When to Quit: Diagnosing and Training LLMs to Abort Futile Reasoning본 논문은 LLM이 자신의 능력 범위를 인지하지 못하고, 해결 불가능한 문제에 대해 그럴듯한 거짓 추론(Futile Reasoning)을 지속하는 심각한 신뢰성 문제를 해결하고자 한다 .#Review#Futile Reasoning#Capability Alignment#Reinforcement Learning#Refusal#LLM Reliability#Reward Shaping2026년 8월 13일댓글 수 로딩 중
[논문리뷰] Intern-S2-Preview: Scientific Agentic Foundation Model본 논문은 현대의 과학적 탐구가 단순한 질의응답을 넘어, 이질적인 데이터 형식(multimodal)에 대한 이해와 도구(tools) 사용, 그리고 긴 작업 지평(long-horizon)에서의 지속적인 추론을 필요로 한다는 점에 주목합니다.#Review#Foundation Model#Scientific Agent#Multimodal Learning#Reinforcement Learning#Time Series Forecasting#Memory Decoder2026년 8월 13일댓글 수 로딩 중
[논문리뷰] How Can Rhetoric Reward-Hack AI Reviewers? Dissecting Rhetorical Sensitivity in AI-Based Peer Review본 연구는 AI가 과학 논문 리뷰에 활용됨에 따라 발생하는 수사적 표현에 의한 평가 왜곡(Reward Hacking) 문제를 해결하고자 합니다.#Review#AI Peer Review#Reward Hacking#Rhetorical Sensitivity#Large Language Models#Scientific Evaluation#Controlled Analysis2026년 8월 13일댓글 수 로딩 중
[논문리뷰] H2R-Bench: Benchmarking Human-to-Robot Manipulation Video Generation in World Models본 논문은 기존 비디오 생성 모델들이 인간의 시연 영상을 로봇 환경으로 전이(Transfer)하는 과정에서 발생하는 Embodiment 불일치 및 기능적 상호작용 실패 문제를 해결하고자 한다 .#Review#H2R-Bench#World Models#Video Generation#Cross-Embodiment Transfer#Robot Learning#Human-to-Robot2026년 8월 13일댓글 수 로딩 중
[논문리뷰] Full-bandwidth transformer본 논문은 기존 Autoregressive Transformer가 decoding 단계에서 겪는 수직적 정보 전달의 병목 현상(vertical feedback bottleneck) 문제를 해결하고자 합니다.#Review#Full-bandwidth transformer#Latent feedback decoding#Temporal parallelism#Multi-pass objective#Gated linear unit#Autoregressive decoding2026년 8월 13일댓글 수 로딩 중
[논문리뷰] DreamX-Phi 1.0: Action-Conditioned Video World Model for Robotic Manipulation기존 World Model들은 로봇 조작(Robotic Manipulation)을 위한 미래 관측 예측(Future Observation Prediction)에서 Photorealistic한 결과를 제공하지만, Action Fidelity 측면에서 중요한 한계점을 가집니다.#Review#Action-Conditioned Video World Model#Robotic Manipulation#SE(3) Transformation#Geometric Encoding#Depth Supervision#Object-Centric Consistency#Distribution Matching Distillation (DMD)#Action-Conditioned World Model#Robotic Manipulation#Video Prediction#Geometric Attention#Depth Supervision#Object Consistency#Distribution Matching Distillation2026년 8월 13일댓글 수 로딩 중
[논문리뷰] DarwinX: Evolving Agent Harnesses Through Natural Selection기존의 self-evolving agent 연구들은 단일 lineage(single-lineage)를 따라 개선을 시도하는 경향이 있어, 특정 작업에 대한 과적합이나 초기 편집에 따른 경로 의존성(path dependence) 문제에 직면한다.#Review2026년 8월 13일댓글 수 로딩 중
[논문리뷰] CW-BASS v2: Saturation-Aware Pseudo-Label Selection for Semi-Supervised Segmentation under Foundation-Model Teachers본 연구는 기존 SSSS 방법론들이 ResNet 기반의 비교적 약한 teacher를 기준으로 설계되어, Foundation-model의 등장으로 인한 새로운 regime에 적합하지 않다는 점을 문제로 지적합니다.#Review#Semi-Supervised Learning#Semantic Segmentation#Pseudo-labeling#Confidence Saturation#Foundation Models#Model Calibration2026년 8월 13일댓글 수 로딩 중