[논문리뷰] When Context Bites: Detecting RAG Poisoning via Document-Level Attention Collapse본 논문은 RAG 시스템이 LLM의 지식 부족 문제를 효과적으로 해결하지만, Adversarial Documents 주입을 통한 Poisoning Attack에 취약하다는 핵심 문제를 다룹니다.#Review#Retrieval-augmented Generation#RAG Poisoning#Attention Collapse#Mechanistic Interpretability#LLM Security#Attack Detection#Document-Level Attention2026년 8월 17일댓글 수 로딩 중
[논문리뷰] Spider-Sense: Intrinsic Risk Sensing for Efficient Agent Defense with Hierarchical Adaptive Screening본 논문은 대규모 언어 모델(LLM) 기반 자율 에이전트의 보안 취약점을 해결하는 것을 목표로 합니다.#Review#LLM Agents#Agent Security#Intrinsic Risk Sensing#Adaptive Defense#Hierarchical Screening#Attack Detection#S2Bench Benchmark2026년 2월 5일댓글 수 로딩 중