论文 🔴 保留 · `Toward Secure LLM Agents: Threat Surfaces, Attacks, Defenses, and Benchmarking`
论文 WeClawArena: An Auditable Sandbox and Benchmark for Cross-User Agents Collaboration and Security in Human-Centered Agent Networks
论文 StealthBench: Measuring Operational Stealth in Autonomous Offensive-Security Agents
论文 Security and Privacy in Retrieval-Augmented Generation: Architectures, Threats, Defenses, and Future Directions for Building Trustworthy Systems
笔记 精读:AgentLAB — LLM Agent 长期攻击的系统性基准
2026-07-11
论文 From Controlled to the Wild: Evaluation of Pentesting Agents for the Real-World