Instant research discovery
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
| Is Agentic Code Review Helpful? Mining Developers' Feedback to CodeRabbit Reviews in the Wild Hong Yi Lin, Kla Tantithamthavorn, Mingzhao Liang, Patanamon Thongtanunam Published: 2026-07-03Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-03 | cs.SE | ai-safety, csse, preprint | E9 / R9 (90%) | - |
| SkillOpt-Lite: Better and Faster Agent Self-evolution via One Line of Vibe Bo Li, Xinjie Zhang, Yifei Shen Published: 2026-07-03Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-03 | cs.SE | ai-safety, csse, preprint | E8 / R8 (93%) | - |
| Which Algorithm Specification Formats Help Language Models Implement Machine Learning Algorithms? Masahiro Kato, Taka Kato Published: 2026-07-03Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-03 | cs.SE | ai-safety, csse, preprint | E8 / R6 (91%) | - |
| AutoCedar: An Agentic Framework for Verifier-Guided Access Control Policy Synthesis Adarsh Vatsa, Sachi Shome, William Eiers, Yingming Zhou Published: 2026-07-04Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-04 | cs.SE | ai-safety, csse, preprint | E8 / R8 (89%) | - |
| Don't Blame the Large Language Model: How Scaffolding Evolution Shapes Coding Agent Quality Ahmed E. Hassan, Bram Adams, Hao Li, Oussama Ben Sghaier Published: 2026-07-04Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-04 | cs.SE | ai-safety, csse, preprint | E16 / R16 (94%) | - |
| Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents Abhishek Kumar, Carsten Maple Published: 2026-07-04Area: cs.SECitations: - Tags: adversarial-robustness, ai-safety, csse, preprint | 2026-07-04 | cs.SE | adversarial-robustness, ai-safety, csse, preprint | E12 / R14 (97%) | - |
| SkillFab: An Agent-Native Skill Production Platform Anjie Xu, Jingfan Chen, Leye Wang, Ruohan Xu Published: 2026-07-04Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-04 | cs.SE | ai-safety, csse, preprint | - | - |
| The "I Don't Know" Filter: Enhancing Agentic Reliability in Function Calling Boris Selitser, Mason del Rosario, Stefan Broecker, Thomas Strohmer Published: 2026-07-04Area: cs.SECitations: 28 Tags: ai-safety, csse, preprint | 2026-07-04 | cs.SE | ai-safety, csse, preprint | E12 / R11 (91%) | 28 |
| Why3-py: A Tool for Formal Verification of Hypothesis Testing and Meta-Analysis in Python Akira Tanaka, Yusuke Kawamoto Published: 2026-07-04Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-04 | cs.SE | ai-safety, csse, preprint | E14 / R9 (90%) | - |
| A Retrieval-Augmented Framework for Detecting and Resolving Pragmatic Ambiguities in Natural Language Requirements Pavithra PM Nair, Preethu Rose Anish Published: 2026-07-05Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-05 | cs.SE | ai-safety, csse, preprint | E9 / R8 (94%) | - |
| Benchmarking API Drift in LLM-Generated Quantum Code Across Successive SDK Versions Mohammad Arif Rasyidi, Syahirul Faiz Published: 2026-07-05Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-05 | cs.SE | ai-safety, csse, preprint | E8 / R10 (94%) | - |
| From Regulation to Requirements: An Automated Requirement Derivation and Explanation Pipeline Pavithra PM Nair, Preethu Rose Anish Published: 2026-07-05Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-05 | cs.SE | ai-safety, csse, preprint | E17 / R14 (95%) | - |
| Obey, Diverge, Collapse: Blind Obedience to Incorrect Instructions Drives Code LLMs to Irrecoverable Code Semantic Collapse Adi Bajpai, Anany Singh Divy, Rajiv Ratn Shah, Raj Jaiswal Published: 2026-07-05Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-05 | cs.SE | ai-safety, csse, preprint | E10 / R8 (95%) | - |
| An Exploration of Agentic Information Fusion for Test Maintenance Prediction Gregory Gay, Jingxiong Liu, Nasser Mohammadiha Published: 2026-07-06Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-06 | cs.SE | ai-safety, csse, preprint | E10 / R11 (92%) | - |
| Dashboard2Code: Evaluating Multimodal Models on Reconstructing Interactive Dashboards Baocai Shan, Hengjie Fang, Qiguang Chen, Qingfu Zhu Published: 2026-07-06Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-06 | cs.SE | ai-safety, csse, preprint | E12 / R8 (94%) | - |
| EvalLoop: A Methodology for Evaluation-Driven Iterative Improvement of Business AI Systems Danti Chen, Josh Fleischer, Kenneth Benavides Published: 2026-07-06Area: cs.SECitations: - Tags: ai-safety, csse, preprint, safety-evaluation | 2026-07-06 | cs.SE | ai-safety, csse, preprint, safety-evaluation | E20 / R19 (94%) | - |
| KAT-Coder-V2.5 Technical Report Bo Huang, Chao Wang, Cheng He, Cheng Luo Published: 2026-07-06Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-06 | cs.SE | ai-safety, csse, preprint | E10 / R11 (93%) | - |
| LLM-Based Test Oracles: Source-of-Authority Taxonomy -- A Systematic Literature Review Ali Hassaan Mughal, Muhammad Bilal Published: 2026-07-06Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-06 | cs.SE | ai-safety, csse, preprint | E8 / R6 (95%) | - |
| LLM-Driven CI-CD Workflow Intelligence for Cyber Systems Engineering Bonan Shen, Jiazhou Gao, Tao Ning, Wei-Jung Huang Published: 2026-07-06Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-06 | cs.SE | ai-safety, csse, preprint | E22 / R21 (93%) | - |
| Three-Phase Evaluation of AI-Assisted Software Development Life Cycle Carson Crockett, Jacob Viehe, Jason Ferraro, Joshua Strubel Published: 2026-07-06Area: cs.SECitations: - Tags: ai-safety, csse, preprint, safety-evaluation | 2026-07-06 | cs.SE | ai-safety, csse, preprint, safety-evaluation | E14 / R15 (91%) | - |
| What Do AI Agents Actually Change? An Empirical Taxonomy of Mutation Patterns in Performance-Improving Pull Requests Alexander Brownlee, Ashkan Sami, Illia Dovhoshliubnyi, Nima Soroush Published: 2026-07-06Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-06 | cs.SE | ai-safety, csse, preprint | E12 / R8 (92%) | - |
| Agents That Teach: Towards Designing Incidental Learning Back into AI-Assisted Software Development Adam P. Burden, Kapil Singi, Prithviraj K Tagadinamani, Rohit Mehra Published: 2026-07-07Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-07 | cs.SE | ai-safety, csse, preprint | E8 / R8 (94%) | - |
| Beyond Refusal: A Same-Lineage Study of Aligned and Abliterated LLMs for Vulnerability Analysis Heng Fan, Junhua Ding, Meikang Qiu, Mingchen Li Published: 2026-07-07Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-07 | cs.SE | ai-safety, csse, preprint | E14 / R10 (94%) | - |
| Evaluating Fine-Tuning and Metrics for Neural Decompilation of Dart AOT Binaries Amr G. Wassal, Ayman AboElhassan, Raafat Abualazm Published: 2026-07-07Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-07 | cs.SE | ai-safety, csse, preprint | E14 / R9 (95%) | - |
| Industry Classification of GitHub Repositories Using the North American Industry Classification System (NAICS) Alexander Quispe, Kevin Xu Published: 2026-07-07Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-07 | cs.SE | ai-safety, csse, preprint | E12 / R11 (92%) | - |
| Prompt Coach: An Empirical Evaluation of an Agentic Tutor for Learning Prompt Engineering in Software Development Adam P. Burden, Kapil Singi, Majd Sakr, Rohit Mehra Published: 2026-07-07Area: cs.SECitations: - Tags: ai-safety, csse, preprint, safety-evaluation | 2026-07-07 | cs.SE | ai-safety, csse, preprint, safety-evaluation | E12 / R11 (92%) | - |
| Property-Driven Synthetic Data Engineering for Data-Scarce Software Systems: Reflections from the Breast Cancer Domain Alberto Zambelli, Andrea Bombarda, Aurora Francesca Zanenga, Claudio Menghi Published: 2026-07-07Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-07 | cs.SE | ai-safety, csse, preprint | E12 / R8 (90%) | - |
| REFORGE: A Method for Benchmarking LLMs' Reverse Engineering Capabilities in Decompiled Binary Function Naming Andreas u. Schmidt, Nicolas Koller Published: 2026-07-07Area: cs.SECitations: 15 Tags: ai-safety, csse, preprint | 2026-07-07 | cs.SE | ai-safety, csse, preprint | E16 / R8 (91%) | 15 |
| Reliable and Developer-Aligned Evaluation of Agents for Software Engineering Razvan Mihai Popescu Published: 2026-07-07Area: cs.SECitations: - Tags: ai-safety, csse, preprint, safety-evaluation | 2026-07-07 | cs.SE | ai-safety, csse, preprint, safety-evaluation | E7 / R6 (89%) | - |
| RuBench: A Repository-Level Agentic Coding Benchmark with Natively Authored Russian Task Specifications Evgeny Shilov Published: 2026-07-07Area: cs.SECitations: 30 Tags: ai-safety, csse, preprint | 2026-07-07 | cs.SE | ai-safety, csse, preprint | E15 / R16 (94%) | 30 |