Instant research discovery
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
| Episodic Memory in AI Agents Poses Risks that Should be Studied and Mitigated Chad DeChant Published: 2025-01-20Area: Agent SafetyCitations: 8 Tags: agent-safety, ai-safety, position | 2025-01-20 | Agent Safety | agent-safety, ai-safety, position | E5 / R3 (96%) | 8 |
| Infrastructure for AI Agents Alan Chan, Elija Perrier, Gillian K. Hadfield, Kevin Wei Published: 2025-01-17Area: Agent SafetyCitations: 24 Tags: agent-safety, ai-safety, position | 2025-01-17 | Agent Safety | agent-safety, ai-safety, position | E6 / R4 (96%) | 24 |
| The Task Shield: Enforcing Task Alignment to Defend Against Indirect Prompt Injection in LLM Agents Anna Squicciarini, Feiran Jia, Tong Wu, Xin Qin Published: 2024-12-21Area: Agent SafetyCitations: 26 Tags: agent-safety, ai-safety, alignment-training, empirical | 2024-12-21 | Agent Safety | agent-safety, ai-safety, alignment-training, empirical | E5 / R3 (96%) | 26 |
| Agent-SafetyBench: Evaluating the Safety of LLM Agents Hongning Wang, Jingzhuo Zhou, Junxiao Yang, Minlie Huang Published: 2024-12-19Area: Agent SafetyCitations: 107 Tags: agent-safety, ai-safety, benchmark | 2024-12-19 | Agent Safety | agent-safety, ai-safety, benchmark | E4 / R2 (96%) | 107 |
| MobileSafetyBench: Evaluating Safety of Autonomous Agents in Mobile Device Control Dongyoon Hahm, June Suk Choi, Juyong Lee, Kimin Lee Published: 2024-10-23Area: Agent SafetyCitations: 26 Tags: agent-safety, ai-safety, benchmark | 2024-10-23 | Agent Safety | agent-safety, ai-safety, benchmark | E4 / R3 (95%) | 26 |
| NetSafe: Exploring the Topological Safety of Multi-agent Networks Chenlong Yin, Guibin Zhang, Junyuan Mao, Kun Wang Published: 2024-10-21Area: Agent SafetyCitations: 26 Tags: adversarial-robustness, agent-safety, ai-safety, empirical | 2024-10-21 | Agent Safety | adversarial-robustness, agent-safety, ai-safety, empirical | E5 / R3 (94%) | 26 |
| Refusal-Trained LLMs Are Easily Jailbroken As Browser Agents Elaine Chang, Elaine Lau, Matt Fredrikson, Priyanshu Kumar Published: 2024-10-11Area: Agent SafetyCitations: 54 Tags: agent-safety, ai-safety, benchmark | 2024-10-11 | Agent Safety | agent-safety, ai-safety, benchmark | E5 / R3 (96%) | 54 |
| Permissive Information-Flow Analysis for Large Language Models Ahmed Salem, Andrew Paverd, Boris K枚pf, David Krueger Published: 2024-10-04Area: Agent SafetyCitations: 9 Tags: agent-safety, ai-safety, empirical | 2024-10-04 | Agent Safety | agent-safety, ai-safety, empirical | E5 / R3 (94%) | 9 |
| Agent Security Bench (ASB): Formalizing and Benchmarking Attacks and Defenses in LLM-based Agents Chenlu Zhan, Hanrong Zhang, Hongwei Wang, Jingyuan Huang Published: 2024-10-03Area: Agent SafetyCitations: 123 Tags: agent-safety, ai-safety, benchmark | 2024-10-03 | Agent Safety | agent-safety, ai-safety, benchmark | E7 / R4 (99%) | 123 |
| HAICOSYSTEM: An Ecosystem for Sandboxing Safety Risks in Human-AI Interactions Bill Yuchen Lin, Faeze Brahman, Frank Xu, Hao Zhu Published: 2024-09-24Area: Agent SafetyCitations: 34 Tags: agent-safety, ai-safety, benchmark | 2024-09-24 | Agent Safety | agent-safety, ai-safety, benchmark | E5 / R3 (95%) | 34 |
| RiskAwareBench: Towards Evaluating Physical Risk Awareness for High-level Planning of LLM-based Embodied Agents Baoyuan Wu, Bingzhe Wu, Zhengyou Zhang, Zihao Zhu Published: 2024-08-08Area: Agent SafetyCitations: 14 Tags: agent-safety, ai-safety, benchmark | 2024-08-08 | Agent Safety | agent-safety, ai-safety, benchmark | E5 / R3 (95%) | 14 |
| Caution for the Environment: Multimodal Agents are Susceptible to Environmental Distractions Aston Zhang, Hai Zhao, Tongxin Yuan, Xinbei Ma Published: 2024-08-05Area: Agent SafetyCitations: 47 Tags: adversarial-robustness, agent-safety, ai-safety, empirical | 2024-08-05 | Agent Safety | adversarial-robustness, agent-safety, ai-safety, empirical | E6 / R3 (95%) | 47 |
| Operationalizing Contextual Integrity in Privacy-Conscious Assistants Aneesh Pappu, Borja Balle, Chongyang Shi, Eugene Bagdasaryan Published: 2024-08-05Area: Agent SafetyCitations: 28 Tags: agent-safety, ai-safety, empirical | 2024-08-05 | Agent Safety | agent-safety, ai-safety, empirical | E4 / R3 (95%) | 28 |
| Breaking Agents: Compromising Autonomous LLM Agents Through Malfunction Amplification Ahmed Salem, Boyang Zhang, Michael Backes, Savvas Zannettou Published: 2024-07-30Area: Agent SafetyCitations: 65 Tags: agent-safety, ai-safety, empirical | 2024-07-30 | Agent Safety | agent-safety, ai-safety, empirical | E6 / R3 (93%) | 65 |
| Preemptive Detection and Correction of Misaligned Actions in LLM Agents Haishuo Fang, Iryna Gurevych, Xiaodan Zhu Published: 2024-07-16Area: Agent SafetyCitations: 6 Tags: agent-safety, ai-safety, empirical | 2024-07-16 | Agent Safety | agent-safety, ai-safety, empirical | E6 / R5 (95%) | 6 |
| Security Matrix for Multimodal Agents on Mobile Devices: A Systematic and Proof of Concept Study Chao Shen, Chenhao Lin, Shuaidong Li, Tianwei Zhang Published: 2024-07-12Area: Agent SafetyCitations: 2 Tags: agent-safety, ai-safety, empirical | 2024-07-12 | Agent Safety | agent-safety, ai-safety, empirical | E4 / R3 (97%) | 2 |
| Flooding Spread of Manipulated Knowledge in LLM-Based Multi-Agent Communities Gongshen Liu, Haodong Zhao, Jian Xie, Lifeng Liu Published: 2024-07-10Area: Agent SafetyCitations: 63 Tags: agent-safety, ai-safety, empirical | 2024-07-10 | Agent Safety | agent-safety, ai-safety, empirical | E6 / R3 (94%) | 63 |
| Towards shutdownable agents via stochastic choice Alexander Roman, Christos Ziakas, Elliott Thornley, Leyton Ho Published: 2024-06-30Area: Agent SafetyCitations: 1 Tags: agent-safety, ai-safety, empirical | 2024-06-30 | Agent Safety | agent-safety, ai-safety, empirical | E5 / R3 (95%) | 1 |
| AgentDojo: A Dynamic Environment to Evaluate Prompt Injection Attacks and Defenses for LLM Agents Edoardo Debenedetti, Florian Tramer, Jie Zhang, Luca Beurer-Kellner Published: 2024-06-19Area: Agent SafetyCitations: 94 Tags: agent-safety, ai-safety, benchmark | 2024-06-19 | Agent Safety | agent-safety, ai-safety, benchmark | E5 / R3 (98%) | 94 |
| GuardAgent: Safeguard LLM Agents by a Guard Agent via Knowledge-Enabled Reasoning Bo Li, Carl Yang, Chulin Xie, Dawn Song Published: 2024-06-13Area: Agent SafetyCitations: 69 Tags: agent-safety, ai-safety, empirical | 2024-06-13 | Agent Safety | agent-safety, ai-safety, empirical | E5 / R4 (97%) | 69 |
| Security of AI Agents Ethan Wang, Hao Chen, Yifeng He, Yuyang Rong Published: 2024-06-12Area: Agent SafetyCitations: 21 Tags: agent-safety, ai-safety, empirical | 2024-06-12 | Agent Safety | agent-safety, ai-safety, empirical | E5 / R3 (94%) | 21 |
| AI Agents Under Threat: A Survey of Key Security Challenges and Future Pathways Changzhou Han, Junwu Xiong, Sheng Wen, Wanlun Ma Published: 2024-06-04Area: Agent SafetyCitations: 152 Tags: agent-safety, ai-safety, survey | 2024-06-04 | Agent Safety | agent-safety, ai-safety, survey | E5 / R3 (94%) | 152 |
| AirGapAgent: Protecting Privacy-Conscious Conversational Agents Borja Balle, Daniel Ramage, Eugene Bagdasaryan, Marco Gruteser Published: 2024-05-08Area: Agent SafetyCitations: 53 Tags: agent-safety, ai-safety, empirical | 2024-05-08 | Agent Safety | agent-safety, ai-safety, empirical | E6 / R3 (93%) | 53 |
| LLM Agents can Autonomously Exploit One-day Vulnerabilities Akul Gupta, Daniel Kang, Richard Fang, Rohan Bindu Published: 2024-04-11Area: Agent SafetyCitations: 125 Tags: agent-safety, ai-safety, empirical | 2024-04-11 | Agent Safety | agent-safety, ai-safety, empirical | E5 / R3 (95%) | 125 |
| InjecAgent: Benchmarking Indirect Prompt Injections in Tool-Integrated LLM Agents Daniel Kang, Qiusi Zhan, Zhixiang Liang, Zifan Ying Published: 2024-03-05Area: Agent SafetyCitations: 251 Tags: agent-safety, ai-safety, benchmark | 2024-03-05 | Agent Safety | agent-safety, ai-safety, benchmark | E4 / R3 (96%) | 251 |
| A Trembling House of Cards? Mapping Adversarial Attacks against Language Agents Boyuan Zheng, Chaowei Xiao, Huan Sun, Lingbo Mo Published: 2024-02-15Area: Agent SafetyCitations: 23 Tags: adversarial-robustness, agent-safety, ai-safety, survey | 2024-02-15 | Agent Safety | adversarial-robustness, agent-safety, ai-safety, survey | E6 / R4 (97%) | 23 |
| Secret Collusion Among Generative AI Agents Christian Schroeder de Witt, Lewis Hammond, Martin Strohmeier, Mikhail Baranchuk Published: 2024-02-12Area: Agent SafetyCitations: 59 Tags: agent-safety, ai-safety, empirical, safety-evaluation | 2024-02-12 | Agent Safety | agent-safety, ai-safety, empirical, safety-evaluation | E5 / R3 (97%) | 59 |
| The Reasons that Agents Act: Intention and Instrumental Goals Francesca Toni, Francesco Belardinelli, Francis Rhys Ward, Matt MacDermott Published: 2024-02-11Area: Agent SafetyCitations: 22 Tags: agent-safety, ai-safety, theoretical | 2024-02-11 | Agent Safety | agent-safety, ai-safety, theoretical | E5 / R3 (96%) | 22 |
| Prioritizing Safeguarding Over Autonomy: Risks of LLM Agents for Science Arman Cohan, Jian Tang, Kunlun Zhu, Mark Gerstein Published: 2024-02-06Area: Agent SafetyCitations: 55 Tags: agent-safety, ai-safety, alignment-training, survey | 2024-02-06 | Agent Safety | agent-safety, ai-safety, alignment-training, survey | E5 / R3 (95%) | 55 |
| PsySafe: A Comprehensive Framework for Psychological-based Attack, Defense, and Evaluation of Multi-agent System Safety Feng Zhao, Hongzhi Gao, Huchuan Lu, Jing Shao Published: 2024-01-22Area: Agent SafetyCitations: 75 Tags: agent-safety, ai-safety, empirical, safety-evaluation | 2024-01-22 | Agent Safety | agent-safety, ai-safety, empirical, safety-evaluation | E6 / R4 (95%) | 75 |