본문으로 건너뛰기

Review

[논문리뷰] DIWALI - Diversity and Inclusivity aWare cuLture specific Items for India: Dataset and Assessment of LLMs for Cultural Text Adaptation in Indian Context

댓글 수 로딩 중

[논문리뷰] ContextFlow: Training-Free Video Object Editing via Adaptive Context Enrichment

댓글 수 로딩 중

[논문리뷰] ByteWrist: A Parallel Robotic Wrist Enabling Flexible and Anthropomorphic Motion for Confined Spaces

댓글 수 로딩 중

[논문리뷰] AuditoryBench++: Can Language Models Understand Auditory Knowledge without Hearing?

댓글 수 로딩 중

[논문리뷰] ARE: Scaling Up Agent Environments and Evaluations

댓글 수 로딩 중

[논문리뷰] WhisTLE: Deeply Supervised, Text-Only Domain Adaptation for Pretrained Speech Recognition Transformers

댓글 수 로딩 중

[논문리뷰] Video2Roleplay: A Multimodal Dataset and Framework for Video-Guided Role-playing Agents

댓글 수 로딩 중

[논문리뷰] SPATIALGEN: Layout-guided 3D Indoor Scene Generation

댓글 수 로딩 중

[논문리뷰] RPG: A Repository Planning Graph for Unified and Scalable Codebase Generation

댓글 수 로딩 중

[논문리뷰] RGB-Only Supervised Camera Parameter Optimization in Dynamic Scenes

댓글 수 로딩 중

[논문리뷰] MANZANO: A Simple and Scalable Unified Multimodal Model with a Hybrid Vision Tokenizer

댓글 수 로딩 중

[논문리뷰] Lynx: Towards High-Fidelity Personalized Video Generation

댓글 수 로딩 중

[논문리뷰] Latent Zoning Network: A Unified Principle for Generative Modeling, Representation Learning, and Classification

댓글 수 로딩 중

[논문리뷰] Do You Hear What I Mean? Quantifying the Instruction-Perception Gap in Instruction-Guided Expressive Text-To-Speech Systems

댓글 수 로딩 중