Instant research discovery
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
| CAKE: Cloud Architecture Knowledge Evaluation of Large Language Models Florian Girardo Lukas, Krzysztof Sierszecki, Phongsakon Mark Konrad, Rahime Yilmaz Published: 2026-04-07Area: cs.SECitations: - Tags: ai-safety, csse, preprint, safety-evaluation | 2026-04-07 | cs.SE | ai-safety, csse, preprint, safety-evaluation | E6 / R4 (96%) | - |
| Does Pass Rate Tell the Whole Story? Evaluating Design Constraint Compliance in LLM-based Issue Resolution Chong Wang, Junhao Zeng, Junwei Liu, Kai Yu Published: 2026-04-07Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-04-07 | cs.SE | ai-safety, csse, preprint | E4 / R3 (92%) | - |
| "Don't Be Afraid, Just Learn": Insights from Industry Practitioners to Prepare Software Engineers in the Age of Generative AI Daniel Otten, Denys Poshyvanyk, Douglas Schmidt, Nathan Wintersgill Published: 2026-04-07Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-04-07 | cs.SE | ai-safety, csse, preprint | E4 / R3 (93%) | - |
| On the Role of Fault Localization Context for LLM-Based Program Repair Hadi Hemmati, Hung Viet Pham, Melika Sepidband Published: 2026-04-07Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-04-07 | cs.SE | ai-safety, csse, preprint | E4 / R2 (97%) | - |
| SemLink: A Semantic-Aware Automated Test Oracle for Hyperlink Verification using Siamese Sentence-BERT Farn Wang, Guan-Yan Yang, Kuo-Hui Yeh, Shu-Yuan Ku Published: 2026-04-07Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-04-07 | cs.SE | ai-safety, csse, preprint | E5 / R3 (97%) | - |
| Spec Kit Agents: Context-Grounded Agentic Workflows Pardis Taghavi, Santosh Bhavani Published: 2026-04-07Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-04-07 | cs.SE | ai-safety, csse, preprint | E4 / R2 (96%) | - |
| Compiled AI: Deterministic Code Generation for LLM-Based Workflow Automation Aaron Karlsberg, Anmol Sharma, Geert Trooskens, Gil Alterovitz Published: 2026-04-06Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-04-06 | cs.SE | ai-safety, csse, preprint | E5 / R3 (95%) | - |
| Edit, But Verify: An Empirical Audit of Instructed Code-Editing Benchmarks Amir M. Ebrahimi, Gopi Krishnan Rajbahadur Published: 2026-04-06Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-04-06 | cs.SE | ai-safety, csse, preprint | E5 / R3 (95%) | - |
| ENCRUST: Encapsulated Substitution and Agentic Refinement on a Live Scaffold for Safe C-to-Rust Translation Ali Shokri, Binoy Ravindran, Hohyun Sim, Hyeonjoong Cho Published: 2026-04-06Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-04-06 | cs.SE | ai-safety, csse, preprint | E6 / R3 (97%) | - |
| Nidus: Externalized Reasoning for AI-Assisted Engineering Danil Gorinevski Published: 2026-04-06Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-04-06 | cs.SE | ai-safety, csse, preprint | E6 / R3 (93%) | - |
| Planning to Explore: Curiosity-Driven Planning for LLM Test Generation Alfonso Amayuelas, Firas Laakom, J眉rgen Schmidhuber, Piotr Pi臋kos Published: 2026-04-06Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-04-06 | cs.SE | ai-safety, csse, preprint | E5 / R3 (97%) | - |
| Scaling Coding Agents via Atomic Skills Kelin Fu, Shing-Chi Cheung, Xinlong Yang, Yanhao Li Published: 2026-04-06Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-04-06 | cs.SE | ai-safety, csse, preprint | E5 / R3 (94%) | - |
| Architecture Without Architects: How AI Coding Agents Shape Software Architecture Phongsakon Mark Konrad, Riccardo Terrenzi, Serkan Ayvaz, Tim Lukas Adam Published: 2026-04-05Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-04-05 | cs.SE | ai-safety, csse, preprint | E5 / R3 (93%) | - |
| Closed-Loop Autonomous Software Development via Jira-Integrated Backlog Orchestration: A Case Study in Deterministic Control and Safety-Constrained Automation Elias Calboreanu Published: 2026-04-05Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-04-05 | cs.SE | ai-safety, csse, preprint | E6 / R4 (96%) | - |
| Toward a Sustainable Software Architecture Community: Evaluating ICSA's Environmental Impact Mahyar T. Moghaddam, Mikkel Baun Kj忙rgaard, Mina Alipour, Torben Worm Published: 2026-04-05Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-04-05 | cs.SE | ai-safety, csse, preprint | E4 / R3 (94%) | - |
| AgenticFlict: A Large-Scale Dataset of Merge Conflicts in AI Coding Agent Pull Requests on GitHub Daniel Ogenrwot, John Businge Published: 2026-04-04Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-04-04 | cs.SE | ai-safety, csse, preprint | E8 / R3 (98%) | - |
| AutoReSpec: A Framework for Generating Specification using Large Language Models Ragib Shahariar Ayon, Shibbir Ahmed Published: 2026-04-04Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-04-04 | cs.SE | ai-safety, csse, preprint | E6 / R3 (97%) | - |
| Measuring the Permission Gate: A Stress-Test Evaluation of Claude Code's Auto Mode Shuai Wang, Wenyuan Jiang, Yudong Gao, Zimo Ji Published: 2026-04-04Area: cs.SECitations: - Tags: ai-safety, csse, preprint, safety-evaluation | 2026-04-04 | cs.SE | ai-safety, csse, preprint, safety-evaluation | E5 / R3 (95%) | - |
| Persistent Cross-Attempt State Optimization for Repository-Level Code Generation Hongyu Zhang, Jiangshuai Wang, Linhao Wu, Lu Zhang Published: 2026-04-04Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-04-04 | cs.SE | ai-safety, csse, preprint | E6 / R4 (97%) | - |
| Squeez: Task-Conditioned Tool-Output Pruning for Coding Agents 脕d谩m Kov谩cs Published: 2026-04-04Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-04-04 | cs.SE | ai-safety, csse, preprint | E5 / R3 (99%) | - |
| Toward Executable Repository-Level Code Generation via Environment Alignment Hongyu Zhang, Junlei Shen, Linhao Wu, Lu Zhang Published: 2026-04-04Area: cs.SECitations: - Tags: ai-safety, alignment-training, csse, preprint | 2026-04-04 | cs.SE | ai-safety, alignment-training, csse, preprint | E5 / R3 (96%) | - |
| Agile Story-Point Estimation: Is RAG a Better Way to Go? Chanchal Roy, Lamyea Maha, Tajmilur Rahman Published: 2026-04-03Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-04-03 | cs.SE | ai-safety, csse, preprint | E6 / R3 (95%) | - |
| AI-Assisted Unit Test Writing and Test-Driven Code Refactoring: A Case Study Ema Smolic, Luka Hobor, Mario Brcic, Mihael Kovac Published: 2026-04-03Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-04-03 | cs.SE | ai-safety, csse, preprint | E5 / R3 (95%) | - |
| Beyond Isolated Tasks: A Framework for Evaluating Coding Agents on Sequential Software Evolution Chaitanya Devaguptapu, Ganesh Senrayan, KN Ajay Shastry, Pranoy Panda Published: 2026-04-03Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-04-03 | cs.SE | ai-safety, csse, preprint | E5 / R3 (96%) | - |
| Inside the Scaffold: A Source-Code Taxonomy of Coding Agent Architectures Benjamin Rombaut Published: 2026-04-03Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-04-03 | cs.SE | ai-safety, csse, preprint | E6 / R3 (97%) | - |
| Measuring LLM Trust Allocation Across Conflicting Software Artifacts Ahsanul Ameen Sabit, Noshin Ulfat, Soneya Binta Hossain Published: 2026-04-03Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-04-03 | cs.SE | ai-safety, csse, preprint | E4 / R2 (93%) | - |
| ProdCodeBench: A Production-Derived Benchmark for Evaluating AI Coding Agents Chandra Maddila, Matteo Paltenghi, Satish Chandra, Shubham Ugare Published: 2026-04-02Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-04-02 | cs.SE | ai-safety, csse, preprint | E5 / R3 (95%) | - |
| ToolMisuseBench: An Offline Deterministic Benchmark for Tool Misuse and Recovery in Agentic Systems Akshey Sigdel, Rista Baral Published: 2026-04-02Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-04-02 | cs.SE | ai-safety, csse, preprint | E5 / R3 (97%) | - |
| EvolveTool-Bench: Evaluating the Quality of LLM-Generated Tool Libraries as Software Artifacts Alibek T. Kaliyev, Artem Maryanskyy Published: 2026-04-01Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-04-01 | cs.SE | ai-safety, csse, preprint | E5 / R3 (97%) | - |
| Investigating Autonomous Agent Contributions in the Wild: Activity Patterns and Code Change over Time Andrei Botocan, David Gros, Maliheh Izadi, Prem Devanbu Published: 2026-04-01Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-04-01 | cs.SE | ai-safety, csse, preprint | E6 / R2 (96%) | - |