본문으로 건너뛰기

Review

[논문리뷰] Rethinking Saliency Maps: A Cognitive Human Aligned Taxonomy and Evaluation Framework for Explanations

댓글 수 로딩 중

[논문리뷰] Planning with Sketch-Guided Verification for Physics-Aware Video Generation

댓글 수 로딩 중

[논문리뷰] Parrot: Persuasion and Agreement Robustness Rating of Output Truth -- A Sycophancy Robustness Benchmark for LLMs

댓글 수 로딩 중

[논문리뷰] OpenMMReasoner: Pushing the Frontiers for Multimodal Reasoning with an Open and General Recipe

댓글 수 로딩 중

[논문리뷰] Multi-Faceted Attack: Exposing Cross-Model Vulnerabilities in Defense-Equipped Vision-Language Models

댓글 수 로딩 중

[논문리뷰] Mantis: A Versatile Vision-Language-Action Model with Disentangled Visual Foresight

댓글 수 로딩 중

[논문리뷰] Loomis Painter: Reconstructing the Painting Process

댓글 수 로딩 중

[논문리뷰] Downscaling Intelligence: Exploring Perception and Reasoning Bottlenecks in Small Multimodal Models

댓글 수 로딩 중

[논문리뷰] Diversity Has Always Been There in Your Visual Autoregressive Models

댓글 수 로딩 중

[논문리뷰] Video-as-Answer: Predict and Generate Next Video Event with Joint-GRPO

댓글 수 로딩 중

[논문리뷰] TurkColBERT: A Benchmark of Dense and Late-Interaction Models for Turkish Information Retrieval

댓글 수 로딩 중

[논문리뷰] Thinking-while-Generating: Interleaving Textual Reasoning throughout Visual Generation

댓글 수 로딩 중

[논문리뷰] Step-Audio-R1 Technical Report

댓글 수 로딩 중