Instant research discovery
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
| Moving the Safety Barrier: Dynamic Routing Adaptive Alignment Against White-Box Attacks Chao Wu, Cheng Ji, Chuancheng Shi, Fei Shen Published: 2026-08-02Area: cs.CRCitations: - Tags: ai-safety, alignment-training, cscr, preprint | 2026-08-02 | cs.CR | ai-safety, alignment-training, cscr, preprint | E12 / R11 (91%) | - |
| Neuro-Symbolic Participation Governance for Verifiable AI Agents in Open Digital Twin Ecosystems Juan Li, Wei Cai, Yan Bai Published: 2026-08-02Area: cs.CRCitations: - Tags: ai-safety, cscr, preprint | 2026-08-02 | cs.CR | ai-safety, cscr, preprint | - | - |
| Permission Denied: Policy-Graded Evaluation of Coding Agents in Hardened Environments Dotan Davidovich, Hai Rozencwajg, Or Hiltch, Yair Amar Published: 2026-08-02Area: cs.CRCitations: - Tags: ai-safety, cscr, preprint, safety-evaluation | 2026-08-02 | cs.CR | ai-safety, cscr, preprint, safety-evaluation | E10 / R9 (91%) | - |
| Security-First Evaluation of Text-to-Terraform: Benchmarking LLMs and SLMs for Secure IaC Generation Diego Kreutz, Francis Luis Santos Vargas, Rodrigo Brand茫o Mansilha Published: 2026-08-02Area: cs.CRCitations: - Tags: ai-safety, cscr, preprint, safety-evaluation | 2026-08-02 | cs.CR | ai-safety, cscr, preprint, safety-evaluation | E12 / R6 (94%) | - |
| Vulnerabilities, Secrets and Misconfiguration in the Highest-Exposure Docker Hub Images Beatriz Machado, Cristhian Kapelinski, Diego Kreutz Published: 2026-08-02Area: cs.CRCitations: - Tags: ai-safety, cscr, preprint | 2026-08-02 | cs.CR | ai-safety, cscr, preprint | E12 / R14 (93%) | - |
| Why Formal Monitors Fail: Attack Distribution Entropy as a Coverage Bound for LTL-Based LLM Agent Safety Ruiyang Zhang Published: 2026-08-02Area: cs.CRCitations: - Tags: ai-safety, cscr, preprint | 2026-08-02 | cs.CR | ai-safety, cscr, preprint | - | - |
| $S^3$: Improving Agent Safety through Multi-Stage Defense Haoyu Wang, Jun Sun, Zibo Xiao Published: 2026-08-03Area: cs.CRCitations: 45 Tags: ai-safety, cscr, preprint | 2026-08-03 | cs.CR | ai-safety, cscr, preprint | E12 / R11 (90%) | 45 |
| Agentic Incident Response through Digital Twin-Enhanced Multiscale Planning Kim Hammar, Tao Li, Yiran Gao Published: 2026-08-03Area: cs.CRCitations: - Tags: ai-safety, cscr, preprint | 2026-08-03 | cs.CR | ai-safety, cscr, preprint | - | - |
| Antares: Foundation Models for Agentic Vulnerability Localization Aman Priyanshu, Amin Karbasi, Arthur Goldblatt, Baturay Saglam Published: 2026-08-03Area: cs.CRCitations: - Tags: ai-safety, cscr, preprint | 2026-08-03 | cs.CR | ai-safety, cscr, preprint | - | - |
| EntailLLM: Verifying LLM-Generated Vulnerability Discovery Paths with Domain Knowledge via Logic Programming Colton Payne, Dana Warmsley, Evelyn Kim, Jaikrishna Manojkumar Patil Published: 2026-08-03Area: cs.CRCitations: - Tags: ai-safety, cscr, preprint | 2026-08-03 | cs.CR | ai-safety, cscr, preprint | - | - |
| Evading Chain-of-Thought Monitoring Through Model Poisoning Amanda Minnich, Blake Bullwinkel, Giorgio Severi, Shujaat Mirza Published: 2026-08-03Area: cs.CRCitations: - Tags: ai-safety, cscr, preprint | 2026-08-03 | cs.CR | ai-safety, cscr, preprint | E11 / R8 (96%) | - |
| MNC: Scope-Bound Semantic Declassification for Private LLM-Agent Communication Hankai Liu, Jinghan Xu, Longze Fan, Xinjin Li Published: 2026-08-03Area: cs.CRCitations: - Tags: ai-safety, cscr, preprint | 2026-08-03 | cs.CR | ai-safety, cscr, preprint | - | - |
| MutMem: Cryptographically Authorized Mutation in Persistent Agent Memory Walid Saidi Published: 2026-08-03Area: cs.CRCitations: - Tags: ai-safety, cscr, preprint | 2026-08-03 | cs.CR | ai-safety, cscr, preprint | E9 / R8 (95%) | - |
| Privacy-Preserving AI Verification via Minimal Information Disclosure Gabriel Kulp, Sleem Abdelghafar Published: 2026-08-03Area: cs.CRCitations: 55 Tags: ai-safety, cscr, preprint | 2026-08-03 | cs.CR | ai-safety, cscr, preprint | E8 / R6 (90%) | 55 |
| Steganalysis of Adaptive Covert Collusion in Tool-Using Agent Populations: A Black-Box, Cross-Principal Approach Mohamed Chahine Ghanem Published: 2026-08-03Area: cs.CRCitations: - Tags: ai-safety, cscr, preprint | 2026-08-03 | cs.CR | ai-safety, cscr, preprint | E7 / R6 (93%) | - |
| AgentAntibody: An Adaptive Immune System for Defending LLM Agents against Prompt Injection Jiongchi Yu, Shihao Weng, Xiaofei Xie, Yang Feng Published: 2026-08-04Area: cs.CRCitations: 45 Tags: ai-safety, cscr, preprint | 2026-08-04 | cs.CR | ai-safety, cscr, preprint | E8 / R7 (89%) | 45 |
| AI Forensics Across White-, Grey-, and Black-Box Access: A Process Model and Research Agenda for Post-Incident Investigation of AI Systems Ali Dehghantanha, Sajad Homayoun Published: 2026-08-04Area: cs.CRCitations: - Tags: ai-safety, cscr, preprint | 2026-08-04 | cs.CR | ai-safety, cscr, preprint | - | - |
| AI Security Leaderboard: Methodology, Results and Minimal Standard Adam Gleave, Dan Zhao, Edward Yee, Grace Cheong Published: 2026-08-04Area: cs.CRCitations: - Tags: ai-safety, cscr, preprint | 2026-08-04 | cs.CR | ai-safety, cscr, preprint | E9 / R10 (95%) | - |
| AI Security Leaderboard: Methodology, Results and Minimal Standard Adam Gleave, Dan Zhao, Edward Yee, Grace Cheong Published: 2026-08-04Area: cs.CRCitations: 26 Tags: ai-safety, cscr, preprint | 2026-08-04 | cs.CR | ai-safety, cscr, preprint | E9 / R10 (95%) | 26 |
| An Inline Control Architecture for Language Models in Intelligent Transportation Systems Mounira Msahli, Narendra Kumar Dewangan Published: 2026-08-04Area: cs.CRCitations: - Tags: ai-safety, cscr, preprint | 2026-08-04 | cs.CR | ai-safety, cscr, preprint | E10 / R6 (88%) | - |
| A Security-Oriented Lifecycle Model for Large Language Model Systems Eleftherios Batzolis, George Drosatos, Konstantinos Rantos, Vassilis Katsouros Published: 2026-08-04Area: cs.CRCitations: - Tags: ai-safety, cscr, preprint | 2026-08-04 | cs.CR | ai-safety, cscr, preprint | - | - |
| Attribute-based Undetectable Watermarking for Generative AI Models Chung-Wei Lee, Er-Cheng Tang, Max Raffel, Miryam Mi-Ying Huang Published: 2026-08-04Area: cs.CRCitations: - Tags: ai-safety, cscr, preprint | 2026-08-04 | cs.CR | ai-safety, cscr, preprint | - | - |
| Behavioral Skill Reconstruction: Reconstructing Hidden Functionality from LLM Agent Skills Haoxuan Xu, Mengyuan Li, Peichun Hua Published: 2026-08-04Area: cs.CRCitations: 1 Tags: ai-safety, cscr, preprint | 2026-08-04 | cs.CR | ai-safety, cscr, preprint | E6 / R7 (90%) | 1 |
| Beyond the QBER Threshold: A Temporal QBER Based Machine Learning Framework for Multi Attack Detection in BB84 QKD Amit Shukla, Deepak Singh, Devesh Kumar, Isha Published: 2026-08-04Area: cs.CRCitations: 25 Tags: ai-safety, cscr, preprint | 2026-08-04 | cs.CR | ai-safety, cscr, preprint | E14 / R11 (93%) | 25 |
| DiagChain: A Diagnostic Benchmark for Evaluating LLM Agents on Evidence-Grounded Attack Chain Reconstruction Hai Wan, Jiabao Gao, Kai Chang, Tian Qiu Published: 2026-08-04Area: cs.CRCitations: - Tags: ai-safety, cscr, preprint | 2026-08-04 | cs.CR | ai-safety, cscr, preprint | - | - |
| FakeI2V-Bench: Benchmarking the Applicability of Image-level Deepfake Detectors for Deepfake Video Detection Delong Ran, Pei Li, Sihan Chen, Tianshuo Cong Published: 2026-08-04Area: cs.CRCitations: 150 Tags: ai-safety, cscr, preprint | 2026-08-04 | cs.CR | ai-safety, cscr, preprint | E13 / R15 (93%) | 150 |
| FBID: Adaptive Personalized Federated Learning for Robust Out-of-Distribution Attack Detection in IoT Networks An Khanh Bui, Cong Thanh Nguyen, Diep N. Nguyen, Hoang-Anh Pham Published: 2026-08-04Area: cs.CRCitations: 17 Tags: ai-safety, cscr, preprint | 2026-08-04 | cs.CR | ai-safety, cscr, preprint | E10 / R8 (93%) | 17 |
| Internalising the Identity Primitive: Cryptographic Individuality for an Autonomous Agent on a Public Blockchain Keisuke Suzuki Published: 2026-08-04Area: cs.CRCitations: 52 Tags: ai-safety, cscr, preprint | 2026-08-04 | cs.CR | ai-safety, cscr, preprint | E8 / R6 (94%) | 52 |
| Post-Hoc Trajectory-Risk Certification for Modular LLM-Based Security Agents Zhenpeng Li Published: 2026-08-04Area: cs.CRCitations: - Tags: ai-safety, cscr, preprint | 2026-08-04 | cs.CR | ai-safety, cscr, preprint | E7 / R6 (92%) | - |
| SparSEEty: Extracting Tokens from Sparsity-Exploiting LLM Serving Systems via Deterministic Side Channels Dokyung Song, Euihyun Lee, Jinyoung Park, Yongwan Jo Published: 2026-08-04Area: cs.CRCitations: - Tags: ai-safety, cscr, preprint | 2026-08-04 | cs.CR | ai-safety, cscr, preprint | E10 / R9 (91%) | - |