[논문리뷰] Security Assessment of DeepSeek Harness with A.I.G: Evaluating Resistance to Indirect Prompt Injection본 논문은 DeepSeek Harness (DSH) 프레임워크를 대상으로 indirect prompt injection에 대한 보안 취약점을 체계적으로 평가하는 것을 목표로 한다.#Review#Indirect Prompt Injection#LLM Agent Security#DeepSeek Harness#A.I.G (AI-Infra-Guard)#Source-to-Sink Analysis#Red Teaming#Runtime Assessment2026년 8월 18일댓글 수 로딩 중
[논문리뷰] Benchmarks are Not Enough: RAMP for Runtime Assessing of Agentic Models in Production Systems본 논문은 기존의 LLM 에이전트 평가 방식이 정적이고 단기적인 작업에 치중되어 있어, 실제 프로덕션 환경에서 요구되는 복잡한 장기 워크플로우를 반영하지 못하는 문제를 해결하고자 합니다.#Review#Agentic Models#Runtime Assessment#Software Engineering#Long-horizon Workloads#Compiler Construction#Resurrection Protocol#Production Systems2026년 6월 3일댓글 수 로딩 중