[논문리뷰] Massive Activations in Hybrid Linear Attention Large Language Models: Pre-Attention Spikes and Inter-Spike Plateaus본 논문은 HLA LLM 내부의 계층별 하이브리드화가 Activation Dynamics를 어떻게 재구성하는지 규명하고자 합니다.#Review#Hybrid Linear Attention#Massive Activations#Pre-Attention Spikes#Inter-Spike Plateaus#Attention Sinks#Language Modeling2026년 8월 13일댓글 수 로딩 중