Instant research discovery
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Showing 1-30 of 194 papers (page 1 of 7)路 43 ms
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
| AgentLeak: A Full-Stack Benchmark for Privacy Leakage in Multi-Agent LLM Systems Faouzi El Yagoubi, Godwin Badu-Marfo, Ranwa Al Mallah Published: 2026-02-12Area: Agent SafetyCitations: - Tags: agent-safety, ai-safety, benchmark | 2026-02-12 | Agent Safety | agent-safety, ai-safety, benchmark | E5 / R3 (95%) | - |
| AIR: Improving Agent Safety through Incident Response Junjie Chen, Jun Sun, Zibo Xiao Published: 2026-02-12Area: Agent SafetyCitations: - Tags: agent-safety, ai-safety, empirical | 2026-02-12 | Agent Safety | agent-safety, ai-safety, empirical | E5 / R3 (95%) | - |
| Authenticated Workflows: A Systems Approach to Protecting Agentic AI Mohan Rajagopalan, Vinay Rao Published: 2026-02-11Area: Agent SafetyCitations: - Tags: agent-safety, ai-safety, theoretical | 2026-02-11 | Agent Safety | agent-safety, ai-safety, theoretical | E5 / R4 (96%) | - |
| The Devil Behind Moltbook: Anthropic Safety is Always Vanishing in Self-Evolving AI Societies Chaozhuo Li, Chenxu Wang, Jinyu Hou, Ji Qi Published: 2026-02-10Area: Agent SafetyCitations: - Tags: agent-safety, ai-safety, alignment-training, empirical | 2026-02-10 | Agent Safety | agent-safety, ai-safety, alignment-training, empirical | E4 / R3 (95%) | - |
| Trustworthy Agentic AI Requires Deterministic Architectural Boundaries Manish Bhattarai, Minh Vu Published: 2026-02-10Area: Agent SafetyCitations: - Tags: agent-safety, ai-safety, alignment-training, theoretical | 2026-02-10 | Agent Safety | agent-safety, ai-safety, alignment-training, theoretical | E5 / R4 (97%) | - |
| A Behavioural and Representational Evaluation of Goal-Directedness in Language Model Agents Angelos Nalmpantis, Calum McNamara, Evgenii Kortukov, Fade Chen Published: 2026-02-09Area: Agent SafetyCitations: - Tags: agent-safety, ai-safety, empirical, safety-evaluation | 2026-02-09 | Agent Safety | agent-safety, ai-safety, empirical, safety-evaluation | E4 / R3 (94%) | - |
| When Actions Go Off-Task: Detecting and Correcting Misaligned Actions in Computer-Use Agents Chentao Ye, Huan Sun, Jaylen Jones, Junyi Li Published: 2026-02-09Area: Agent SafetyCitations: - Tags: agent-safety, ai-safety, benchmark | 2026-02-09 | Agent Safety | agent-safety, ai-safety, benchmark | E6 / R3 (94%) | - |
| AgentSys: Secure and Dynamic LLM Agents Through Explicit Hierarchical Memory Management Chaowei Xiao, Hao Li, Ning Zhang, Ruoyao Wen Published: 2026-02-07Area: Agent SafetyCitations: - Tags: agent-safety, ai-safety, empirical | 2026-02-07 | Agent Safety | agent-safety, ai-safety, empirical | E5 / R3 (96%) | - |
| Agentic Uncertainty Reveals Agentic Overconfidence Gb猫tondji Dovonon, Jean Kaddour, Leo Richter, Matt J. Kusner Published: 2026-02-06Area: Agent SafetyCitations: - Tags: adversarial-robustness, agent-safety, ai-safety, empirical | 2026-02-06 | Agent Safety | adversarial-robustness, agent-safety, ai-safety, empirical | E5 / R3 (96%) | - |
| TrajAD: Trajectory Anomaly Detection for Trustworthy LLM Agents Chong Zhang, Hansong Liu, Xiaoyan Wang, Yang Yu Published: 2026-02-06Area: Agent SafetyCitations: - Tags: agent-safety, ai-safety, empirical | 2026-02-06 | Agent Safety | agent-safety, ai-safety, empirical | E5 / R3 (96%) | - |
| MaMa: A Game-Theoretic Approach for Designing Safe Agentic Systems Adish Singla, Goran Radanovic, Jonathan N枚ther Published: 2026-02-04Area: Agent SafetyCitations: - Tags: agent-safety, ai-safety, empirical | 2026-02-04 | Agent Safety | agent-safety, ai-safety, empirical | E6 / R4 (93%) | - |
| LPS-Bench: Benchmarking Safety Awareness of Computer-Use Agents in Long-Horizon Planning under Benign and Adversarial Scenarios Chujia Hu, Dongrui Liu, Ge Gao, Tianyu Chen Published: 2026-02-03Area: Agent SafetyCitations: 1 Tags: adversarial-robustness, agent-safety, ai-safety, benchmark | 2026-02-03 | Agent Safety | adversarial-robustness, agent-safety, ai-safety, benchmark | E4 / R3 (95%) | 1 |
| Modular Safety Guardrails Are Necessary for Foundation-Model-Enabled Robots in the Real World Davood Soleymanzadeh, Fan Fei, Joonkyung Kim, Minghui Zheng Published: 2026-02-03Area: Agent SafetyCitations: - Tags: agent-safety, ai-safety, position | 2026-02-03 | Agent Safety | agent-safety, ai-safety, position | E5 / R4 (93%) | - |
| Risky-Bench: Probing Agentic Safety Risks under Real-World Deployment An Zhang, Bingnan Liu, Chaochao Lu, Chenhang Cui Published: 2026-02-03Area: Agent SafetyCitations: - Tags: agent-safety, ai-safety, benchmark | 2026-02-03 | Agent Safety | agent-safety, ai-safety, benchmark | E4 / R3 (95%) | - |
| To Defend Against Cyber Attacks, We Must Teach AI Agents to Hack Ruijie Meng, Terry Yue Zhuo, Wenbo Guo, Yangruibo Ding Published: 2026-02-01Area: Agent SafetyCitations: - Tags: agent-safety, ai-safety, position | 2026-02-01 | Agent Safety | agent-safety, ai-safety, position | E5 / R4 (94%) | - |
| Evolving Interpretable Constitutions for Multi-Agent Coordination Alice Saito, Hershraj Niranjani, Phan Xuan Tan, Rayan Yessou Published: 2026-01-31Area: Agent SafetyCitations: - Tags: agent-safety, ai-safety, empirical | 2026-01-31 | Agent Safety | agent-safety, ai-safety, empirical | E5 / R3 (93%) | - |
| Multi-Agent Systems Should be Treated as Principal-Agent Problems Mihaela van der Schaar, Paulius Rauba, Simonas Cepenas Published: 2026-01-30Area: Agent SafetyCitations: - Tags: agent-safety, ai-safety, position | 2026-01-30 | Agent Safety | agent-safety, ai-safety, position | E5 / R4 (94%) | - |
| StepShield: When, Not Whether to Intervene on Rogue Agents Gloria Felicia, Hemant Kumar, Jinfeng He, Michael Eniolade Published: 2026-01-29Area: Agent SafetyCitations: - Tags: agent-safety, ai-safety, benchmark | 2026-01-29 | Agent Safety | agent-safety, ai-safety, benchmark | E5 / R4 (97%) | - |
| Securing AI Agents in Cyber-Physical Systems: A Survey of Environmental Interactions, Deepfake Threats, and Defenses Hozefa Lakadawala, Mohsen Hatami, Van Tuan Pham, Yu Chen Published: 2026-01-28Area: Agent SafetyCitations: - Tags: agent-safety, ai-safety, survey | 2026-01-28 | Agent Safety | agent-safety, ai-safety, survey | E5 / R3 (95%) | - |
| Breaking the Protocol: Security Analysis of the Model Context Protocol Specification and Prompt Injection Vulnerabilities in Tool-Integrated LLM Agents Dmitry Namiot, Narek Maloyan Published: 2026-01-24Area: Agent SafetyCitations: - Tags: agent-safety, ai-safety, empirical | 2026-01-24 | Agent Safety | agent-safety, ai-safety, empirical | E5 / R3 (96%) | - |
| The Shadow Self: Intrinsic Value Misalignment in Large Language Model Agents Chen Chen, Kim Young Il, Kwok-Yan Lam, Qian Wang Published: 2026-01-24Area: Agent SafetyCitations: - Tags: agent-safety, ai-safety, alignment-training, benchmark | 2026-01-24 | Agent Safety | agent-safety, ai-safety, alignment-training, benchmark | E4 / R3 (96%) | - |
| MirrorGuard: Toward Secure Computer-Use Agents via Simulation-to-Real Reasoning Correction Changyue Jiang, Geng Hong, Jiarun Dai, Wenqi Zhang Published: 2026-01-19Area: Agent SafetyCitations: - Tags: agent-safety, ai-safety, empirical | 2026-01-19 | Agent Safety | agent-safety, ai-safety, empirical | E5 / R3 (95%) | - |
| Institutional AI: Governing LLM Collusion in Multi-Agent Cournot Markets via Public Governance Graphs Daniele Nardi, Federico Pierucci, Francesco Giarrusso, Marcantonio Bracale Syrnikov Published: 2026-01-16Area: Agent SafetyCitations: 1 Tags: agent-safety, ai-safety, empirical | 2026-01-16 | Agent Safety | agent-safety, ai-safety, empirical | E5 / R4 (92%) | 1 |
| AgentGuardian: Learning Access Control Policies to Govern AI Agent Behavior Asaf Shabtai, David Mimran, Denis Klimov, Gerard Levinov Published: 2026-01-15Area: Agent SafetyCitations: 2 Tags: agent-safety, ai-safety, empirical | 2026-01-15 | Agent Safety | agent-safety, ai-safety, empirical | E4 / R3 (96%) | 2 |
| Breaking Up with Normatively Monolithic Agency with GRACE: A Reason-Based Neuro-Symbolic Architecture for Safe and Ethical AI Alignment Felix Jahn, Kevin Baum, Lisa Dargasz, Patrick Schramowski Published: 2026-01-15Area: Agent SafetyCitations: - Tags: agent-safety, ai-safety, alignment-training, theoretical | 2026-01-15 | Agent Safety | agent-safety, ai-safety, alignment-training, theoretical | E6 / R4 (96%) | - |
| ToolSafe: Enhancing Tool Invocation Safety of LLM-based agents via Proactive Step-level Guardrail and Feedback Jing Shao, Lijun Li, Peiyang Liu, Shikun Zhang Published: 2026-01-15Area: Agent SafetyCitations: 2 Tags: agent-safety, ai-safety, empirical | 2026-01-15 | Agent Safety | agent-safety, ai-safety, empirical | E5 / R3 (95%) | 2 |
| Too Helpful to Be Safe: User-Mediated Attacks on Planning and Web-Use Agents Carsten Rudolph, Fengchao Chen, Tingmin Wu, Van Nguyen Published: 2026-01-14Area: Agent SafetyCitations: 2 Tags: agent-safety, ai-safety, empirical | 2026-01-14 | Agent Safety | agent-safety, ai-safety, empirical | E5 / R3 (93%) | 2 |
| Evaluating Implicit Regulatory Compliance in LLM Tool Invocation via Logic-Guided Synthesis Boqi Chen, Da Song, Foutse Khomh, Lei Ma Published: 2026-01-13Area: Agent SafetyCitations: - Tags: agent-safety, ai-safety, benchmark | 2026-01-13 | Agent Safety | agent-safety, ai-safety, benchmark | E4 / R3 (97%) | - |
| Safety Not Found (404): Hidden Risks of LLM-Based Robotics Decision Making Jaeyoon Seo, Jean Oh, Jihie Kim, Jua Han Published: 2026-01-09Area: Agent SafetyCitations: - Tags: agent-safety, ai-safety, empirical | 2026-01-09 | Agent Safety | agent-safety, ai-safety, empirical | E6 / R3 (98%) | - |
| Structural Representations for Cross-Attack Generalization in AI Agent Threat Detection Vignesh Iyer Published: 2026-01-05Area: Agent SafetyCitations: - Tags: agent-safety, ai-safety, empirical | 2026-01-05 | Agent Safety | agent-safety, ai-safety, empirical | E5 / R3 (96%) | - |