[논문리뷰] Block Sparse Attention with Log-Linear Complexity대규모 언어 모델(Large Language Models, LLMs)을 더 긴 컨텍스트(Context)로 확장하는 것은 Full Self-Attention의 계산 복잡도가 시퀀스 길이에 대해 quadratic하게 증가하여 막대한 비용을 초래한다는 근본적인 문제에 직면해 있습니다.#Review#Sparse Attention#Log-Linear Complexity#Hierarchical Selection#LogSumExp Scoring#Triton Kernels#Long-Context Modeling2026년 9월 27일댓글 수 로딩 중