[논문리뷰] An Empirical Study of Training Pixel-Space Text-to-Image Diffusion Models본 논문은 대규모 환경에서 Pixel-Space Diffusion 모델이 Latent-Space 모델에 비해 학습 효율성이 떨어진다는 핵심 문제를 해결하고자 합니다. 기존 Latent-Space 모델은 VAE의 압축으로 인해 정보 손실이 발생하고 디코딩 지연(Latency)이 추가되는 한계가 있습니다.#Review#Pixel-Space Diffusion#Latent-to-Pixel Transition#Diffusion Transformer#Large-Scale Pre-training#Inference Efficiency#Step Distillation2026년 8월 17일댓글 수 로딩 중