Instant research discovery
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Showing 1-30 of 745 papers (page 1 of 25)路 139 ms
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
| Beyond Execution: Auditing Experimental Fidelity in LLM-Driven Scientific Research Aimin Pan, Lezhi Yu, Shuibing He, Xiaogang Xu Published: 2026-08-27Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-08-27 | cs.SE | ai-safety, csse, preprint | E7 / R6 (91%) | - |
| FaultLens: Learning Compact Behavioral Test Suites for Generated Operational Programs Hang Lyu, Jingtao Zhang, Zeming Liu Published: 2026-08-27Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-08-27 | cs.SE | ai-safety, csse, preprint | E10 / R7 (90%) | - |
| From Static to Dynamic: Benchmarking Real-World Code Review with MCR-Bench Dewu Zheng, Hongyu Zhang, Kefeng Duan, Xilin Liu Published: 2026-08-27Area: cs.SECitations: 75 Tags: ai-safety, csse, preprint | 2026-08-27 | cs.SE | ai-safety, csse, preprint | E12 / R11 (95%) | 75 |
| Persona-Execution Separation: An Architecture Pattern for Evolving LLM Agents under Execution Audit Yisen Xi Published: 2026-08-27Area: cs.SECitations: 15 Tags: ai-safety, csse, preprint | 2026-08-27 | cs.SE | ai-safety, csse, preprint | E8 / R10 (89%) | 15 |
| SWE-Prime: Fewer Trajectories, Better Performance Dewu Zheng, Ensheng Shi, Hongyu Zhang, Jianxing Yu Published: 2026-08-27Area: cs.SECitations: 15 Tags: ai-safety, csse, preprint | 2026-08-27 | cs.SE | ai-safety, csse, preprint | E10 / R9 (91%) | 15 |
| A Few Pages of Markdown: Committed AI Configuration and Lower Quality Cost after Coding-Agent Adoption Bogdan Vasilescu, Brando Miranda, Hao He, Pavel Azaletskiy Published: 2026-08-26Area: cs.SECitations: 15 Tags: ai-safety, csse, preprint | 2026-08-26 | cs.SE | ai-safety, csse, preprint | E10 / R9 (91%) | 15 |
| A Hybrid Usability Approach for Rating Evaluation of M-Commerce Applications Ahmad Ibtisam, Arshad Ali, Bilal Khan Published: 2026-08-26Area: cs.SECitations: - Tags: ai-safety, csse, preprint, safety-evaluation | 2026-08-26 | cs.SE | ai-safety, csse, preprint, safety-evaluation | E21 / R20 (97%) | - |
| Fairness Invariants: A Relational Approach to Explaining and Mitigating Fairness Bugs Ashish Kumar, Gang Tan, Ranit Debnath Akash, Saeid Tizpaz-Niari Published: 2026-08-26Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-08-26 | cs.SE | ai-safety, csse, preprint | E10 / R7 (89%) | - |
| From General Agents to RCA Experts: A Self-Evolving Harness for Root Cause Analysis Haiyu Huang, Jiewei Lyu, Jinyang Liu, Michael R. Lyu Published: 2026-08-26Area: cs.SECitations: 46 Tags: ai-safety, csse, preprint | 2026-08-26 | cs.SE | ai-safety, csse, preprint | E14 / R10 (91%) | 46 |
| RotDroid: Cross-Orientation State Equivalence Testing for Detecting GUI Rotation Bugs in Android Apps Bo Jiang, Mengdi Qin Published: 2026-08-26Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-08-26 | cs.SE | ai-safety, csse, preprint | E8 / R6 (95%) | - |
| A Literate Programming Environment for Human and Machine Agents Adam T. Burke Published: 2026-08-25Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-08-25 | cs.SE | ai-safety, csse, preprint | - | - |
| ARISMA: Guidelines for AI- and LLM-Assisted Systematic Reviews, Scoping Reviews, and Mapping Studies Mahyar Tourchi Moghaddam, Mina Alipour Published: 2026-08-25Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-08-25 | cs.SE | ai-safety, csse, preprint | E17 / R8 (93%) | - |
| Automatic Model Card Generation Using an LLM Balreet Grewal, Cor-Paul Bezemer, Tajkia Rahman Toma Published: 2026-08-25Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-08-25 | cs.SE | ai-safety, csse, preprint | - | - |
| Evaluating and Preventing Security Smells in AI-Generated Ansible Code David Bainbridge, Junaid Haseeb, Pandu Ranga Reddy Konala, Vimal Kumar Published: 2026-08-25Area: cs.SECitations: 29 Tags: ai-safety, csse, preprint | 2026-08-25 | cs.SE | ai-safety, csse, preprint | E10 / R7 (95%) | 29 |
| Evaluating Language Models on Cross-Language Code Functional Equivalence Anderson Uch么a, Hui Sun, Rohit Gheyi, Wesley K. G. Assun莽茫o Published: 2026-08-25Area: cs.SECitations: 46 Tags: ai-safety, csse, preprint | 2026-08-25 | cs.SE | ai-safety, csse, preprint | E12 / R10 (93%) | 46 |
| LumiXAI: A Modular Full-Stack Framework for Feature Attribution Alfio Ferrara, Elisabetta Rocchetti, Lorenzo Gatta, Sergio Picascia Published: 2026-08-25Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-08-25 | cs.SE | ai-safety, csse, preprint | - | - |
| SPECMINE: A Large-Scale Corpus of Spec-Driven Development Artifacts Bogdan Vasilescu, Shyam Agarwal Published: 2026-08-25Area: cs.SECitations: 6 Tags: ai-safety, csse, preprint | 2026-08-25 | cs.SE | ai-safety, csse, preprint | E10 / R9 (93%) | 6 |
| The Empire, Long Divided, Must Unite: Architectural Convergence in Three LLM Agent Harnesses Dai Jiahong Published: 2026-08-25Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-08-25 | cs.SE | ai-safety, csse, preprint | E14 / R25 (92%) | - |
| Automated Synthesis of Cloud Emulators Ang Chen, Archit Bhatnagar, Sarah McClure, Sylvia Ratnasamy Published: 2026-08-24Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-08-24 | cs.SE | ai-safety, csse, preprint | E10 / R8 (94%) | - |
| Beyond Executable Models: The Pufibara Agent Harness and the Modelica Agent Workflow Benchmark for Physical System Modeling Zizhe Wang Published: 2026-08-24Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-08-24 | cs.SE | ai-safety, csse, preprint | E7 / R6 (94%) | - |
| Confidently Wrong, Silently So: Auditing Undetectable Failures of a Deployed On-Device Language Model Satwik Pandey, Shashwat Pandey, Suresh Raghu Published: 2026-08-24Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-08-24 | cs.SE | ai-safety, csse, preprint | E12 / R7 (92%) | - |
| Confidently Wrong, Silently So: Auditing Undetectable Failures of a Deployed On-Device Language Model Satwik Pandey, Shashwat Pandey, Suresh Raghu Published: 2026-08-24Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-08-24 | cs.SE | ai-safety, csse, preprint | E14 / R9 (93%) | - |
| Feedback That Backfires: Why Small Language Model Agents Repeat the Call They Just Watched Fail Esmail Gumaan Published: 2026-08-24Area: cs.SECitations: 15 Tags: ai-safety, csse, preprint | 2026-08-24 | cs.SE | ai-safety, csse, preprint | E12 / R12 (92%) | 15 |
| Do Not Copy/Paste: Soft Barriers for Copying in AI-Assisted Programming Alberick Euraste Djire, Iyiola E. Olatunji, Jacques Klein, Tegawend茅 F. Bissyand茅 Published: 2026-08-23Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-08-23 | cs.SE | ai-safety, csse, preprint | - | - |
| Evaluating Inference-Time Defenses Against Package Hallucination in LLM-Generated Code Alberick Euraste Djire, Earl T. Barr, Iyiola E. Olatunji, Jacques Klein Published: 2026-08-23Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-08-23 | cs.SE | ai-safety, csse, preprint | - | - |
| Learning from the Test: Self-Referential Differential Testing for Deep RL Agents David Lo, Jieke Shi, Junda He, Mingfei Cheng Published: 2026-08-23Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-08-23 | cs.SE | ai-safety, csse, preprint | - | - |
| ToolRobustBench: Stage-Wise Perturbation Evaluation and Failure Diagnosis for Tool-Calling Agents Yi Chang, YiShan Zheng, Yuan Wu Published: 2026-08-23Area: cs.SECitations: - Tags: ai-safety, csse, preprint, safety-evaluation | 2026-08-23 | cs.SE | ai-safety, csse, preprint, safety-evaluation | E15 / R12 (92%) | - |
| Architecture as Capability Equalizer for Coding Agents Arquimedes Canedo Published: 2026-08-22Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-08-22 | cs.SE | ai-safety, csse, preprint | E13 / R8 (95%) | - |
| Identifying Latent Declarative Representations of Code for Assisting Repository Migration Ashwin Srinivasan, Michael Bain, Shraddha Surana Published: 2026-08-22Area: cs.SECitations: - Tags: ai-safety, csse, preprint | 2026-08-22 | cs.SE | ai-safety, csse, preprint | E14 / R12 (93%) | - |
| Rebuild Dossier: Mechanically-Enforced Specs for Agentic App Rebuilds, and What Model-Tier Failures Reveal Parker Fawcett Published: 2026-08-22Area: cs.SECitations: 12 Tags: ai-safety, csse, preprint | 2026-08-22 | cs.SE | ai-safety, csse, preprint | E11 / R6 (89%) | 12 |