Instant research discovery
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
| When Does Restricting a Coding Agent to execute_code Help? A Regime $\times$ Agent-Design Ablation Hong Yang, Qi Yu, Travis Desell Published: 2026-07-12Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-12 | cs.SE | ai-safety, csse, preprint | E15 / R10 (93%) | - |
| A Large-Scale Dataset of MCP Implementations on GitHub Amine Barrak, Benny Toeppe, Emna Ksontini Published: 2026-07-11Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-11 | cs.SE | ai-safety, csse, preprint | E8 / R8 (91%) | - |
| Compaction as Epistemic Failure: How Agentic LLM Tools Fabricate Confirmed Results from Killed Processes Hiroki Tamba Published: 2026-07-11Area: cs.SECitations: 16 Tags: ai-safety, csse, preprint | 2026-07-11 | cs.SE | ai-safety, csse, preprint | E10 / R8 (94%) | 16 |
| Falsifiable Release Gates for Self-Improving Systems Deepak Soni Published: 2026-07-11Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-11 | cs.SE | ai-safety, csse, preprint | E10 / R15 (94%) | - |
| Mitigating LLM Sycophancy in Code Smell Detection Using Evidence-Guided Reasoning Prompts Istiaq Ahmed Fahad, Kamruzzaman Asif, Md. Nurul Ahad Tawhid Published: 2026-07-11Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-11 | cs.SE | ai-safety, csse, preprint | E11 / R10 (94%) | - |
| ML in a Box: Analyzing Containerization Practices in Open Source ML Projects Amine Barrak, Emna Ksontini, Faten Jebari, Wael Kessentini Published: 2026-07-11Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-11 | cs.SE | ai-safety, csse, preprint | - | - |
| Partial Contracts Suffice: Sound, LLM-Inferred Regression Verification Lucas C. Cordeiro, Rafael Menezes, Yiannis Charalambous, Youcheng Sun Published: 2026-07-11Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-11 | cs.SE | ai-safety, csse, preprint | E10 / R8 (94%) | - |
| Attention to Detail: Evaluating Energy, Performance, and Accuracy Trade-offs Across vLLM Configurations Cl茅ment Quinton, Nada Zine, Patricia Lago, Romain Rouvoy Published: 2026-07-10Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-10 | cs.SE | ai-safety, csse, preprint | E20 / R24 (95%) | - |
| Diversifying to Verify: When Task-Equivalent Programs Differ in Verifiability Ruben Martins, Shirley Yu Published: 2026-07-10Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-10 | cs.SE | ai-safety, csse, preprint | E10 / R6 (91%) | - |
| Do These Violent Delights Have Violent Ends? Measuring the Post-Merge Fate of Agentic Code Chunqiu Steven Xia, Courtney Miller Published: 2026-07-10Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-10 | cs.SE | ai-safety, csse, preprint | E10 / R8 (93%) | - |
| Failure as a Process: An Anatomy of CLI Coding Agent Trajectories Earl T. Barr, Federica Sarro, Han Li, He Ye Published: 2026-07-10Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-10 | cs.SE | ai-safety, csse, preprint | E8 / R6 (90%) | - |
| Git-Assistant: Planning-Based Support for Updating Git Repositories Alfredo Garrach贸n Ruiz, Daniel Borrajo, Tom谩s de la Rosa Published: 2026-07-10Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-10 | cs.SE | ai-safety, csse, preprint | E8 / R8 (94%) | - |
| Git-Assistant: Planning-Based Support for Updating Git Repositories Alfredo Garrach贸n Ruiz, Daniel Borrajo, Tom谩s de la Rosa Published: 2026-07-10Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-10 | cs.SE | ai-safety, csse, preprint | E10 / R10 (94%) | - |
| Inside the Skill Market: From Software Engineering Activities to Reusable Agent Skills Jialun Cao, Shing-Chi Cheung, Songqiang Chen, Xinru Yan Published: 2026-07-10Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-10 | cs.SE | ai-safety, csse, preprint | E9 / R8 (93%) | - |
| Practical Source Code Recovery from Binary Functions Using Anchor-Based Retrieval and LLM Reasoning Benjamin C. M. Fung, Charles Edward Gagnon, Philippe Charland, Steven H. H. Ding Published: 2026-07-10Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-10 | cs.SE | ai-safety, csse, preprint | E10 / R8 (94%) | - |
| Aleena: Alignment Agent for Research Software Engineering Collaborations Anant Mittal, Anshul Tambay, Carlos Garcia Jurado Suarez, Cordero Core Published: 2026-07-09Area: cs.SECitations: - Tags: ai-safety, alignment-training, csse, preprint | 2026-07-09 | cs.SE | ai-safety, alignment-training, csse, preprint | E18 / R15 (90%) | - |
| ProjAgent: Procedural Similarity Retrieval for Repository-Level Code Generation Aaron Imani, Iftekhar Ahmed, QiHong Chen Published: 2026-07-09Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-09 | cs.SE | ai-safety, csse, preprint | E10 / R8 (89%) | - |
| SCATE: Learning to Supervise Coding Agents for Cost-Effective Test Generation Ali Mesbah, Noor Nashid, Sijia Gu Published: 2026-07-09Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-09 | cs.SE | ai-safety, csse, preprint | E12 / R11 (90%) | - |
| The Patchwork Problem in LLM-Generated Code Reza M. Parizi, Viraaji Mothukuri Published: 2026-07-09Area: cs.SECitations: 28 Tags: ai-safety, csse, preprint | 2026-07-09 | cs.SE | ai-safety, csse, preprint | E24 / R17 (90%) | 28 |
| 3100 Opinions on Code Review in an AI World: Building Causal Theory from Practitioner Discourse Bogdan Vasilescu, Christian K盲stner, Courtney Miller, Shyam Agarwal Published: 2026-07-08Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-08 | cs.SE | ai-safety, csse, preprint | E6 / R7 (90%) | - |
| Progressive Crystallization: Turning Agent Exploration into Deterministic, Lower-Cost Workflows in Production Arun Malik Published: 2026-07-08Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-08 | cs.SE | ai-safety, csse, preprint | E10 / R10 (92%) | - |
| Agents That Teach: Towards Designing Incidental Learning Back into AI-Assisted Software Development Adam P. Burden, Kapil Singi, Prithviraj K Tagadinamani, Rohit Mehra Published: 2026-07-07Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-07 | cs.SE | ai-safety, csse, preprint | E8 / R8 (94%) | - |
| Beyond Refusal: A Same-Lineage Study of Aligned and Abliterated LLMs for Vulnerability Analysis Heng Fan, Junhua Ding, Meikang Qiu, Mingchen Li Published: 2026-07-07Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-07 | cs.SE | ai-safety, csse, preprint | E14 / R10 (94%) | - |
| Evaluating Fine-Tuning and Metrics for Neural Decompilation of Dart AOT Binaries Amr G. Wassal, Ayman AboElhassan, Raafat Abualazm Published: 2026-07-07Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-07 | cs.SE | ai-safety, csse, preprint | E14 / R9 (95%) | - |
| Industry Classification of GitHub Repositories Using the North American Industry Classification System (NAICS) Alexander Quispe, Kevin Xu Published: 2026-07-07Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-07 | cs.SE | ai-safety, csse, preprint | E12 / R11 (92%) | - |
| Prompt Coach: An Empirical Evaluation of an Agentic Tutor for Learning Prompt Engineering in Software Development Adam P. Burden, Kapil Singi, Majd Sakr, Rohit Mehra Published: 2026-07-07Area: cs.SECitations: - Tags: ai-safety, csse, preprint, safety-evaluation | 2026-07-07 | cs.SE | ai-safety, csse, preprint, safety-evaluation | E12 / R11 (92%) | - |
| Property-Driven Synthetic Data Engineering for Data-Scarce Software Systems: Reflections from the Breast Cancer Domain Alberto Zambelli, Andrea Bombarda, Aurora Francesca Zanenga, Claudio Menghi Published: 2026-07-07Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-07-07 | cs.SE | ai-safety, csse, preprint | E12 / R8 (90%) | - |
| REFORGE: A Method for Benchmarking LLMs' Reverse Engineering Capabilities in Decompiled Binary Function Naming Andreas u. Schmidt, Nicolas Koller Published: 2026-07-07Area: cs.SECitations: 15 Tags: ai-safety, csse, preprint | 2026-07-07 | cs.SE | ai-safety, csse, preprint | E16 / R8 (91%) | 15 |
| Reliable and Developer-Aligned Evaluation of Agents for Software Engineering Razvan Mihai Popescu Published: 2026-07-07Area: cs.SECitations: - Tags: ai-safety, csse, preprint, safety-evaluation | 2026-07-07 | cs.SE | ai-safety, csse, preprint, safety-evaluation | E7 / R6 (89%) | - |
| RuBench: A Repository-Level Agentic Coding Benchmark with Natively Authored Russian Task Specifications Evgeny Shilov Published: 2026-07-07Area: cs.SECitations: 30 Tags: ai-safety, csse, preprint | 2026-07-07 | cs.SE | ai-safety, csse, preprint | E15 / R16 (94%) | 30 |