Instant research discovery
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
| AutoSupervision: Closing the Feedback Loop in Scientific Workflows with Grounded Revision Verification Ben Fei, Eunseo Jung, Fenghua Ling, Feng Liu Published: 2026-07-30Area: cs.CLCitations: 15 Tags: ai-safety, cscl, preprint | 2026-07-30 | cs.CL | ai-safety, cscl, preprint | E11 / R10 (90%) | 15 |
| Benchmarks Are Not Monolithic: Sample-Level Auditing and Orchestration for LLM Evaluation Jordan Sassoon, Philipp D. Siedler Published: 2026-07-30Area: cs.CLCitations: 1 Tags: ai-safety, cscl, preprint, safety-evaluation | 2026-07-30 | cs.CL | ai-safety, cscl, preprint, safety-evaluation | E16 / R15 (94%) | 1 |
| Beyond Borrowed Histories: Person-Aligned User Simulation for Interactive Role-Playing Evaluation Benfeng Xu, Hongtao Xie, Jie Gao, Lingyun Yu Published: 2026-07-30Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint, safety-evaluation | 2026-07-30 | cs.CL | ai-safety, cscl, preprint, safety-evaluation | E9 / R7 (90%) | - |
| Beyond Borrowed Histories: Person-Aligned User Simulation for Interactive Role-Playing Evaluation Benfeng Xu, Hongtao Xie, Jie Gao, Lingyun Yu Published: 2026-07-30Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint, safety-evaluation | 2026-07-30 | cs.CL | ai-safety, cscl, preprint, safety-evaluation | E8 / R7 (90%) | - |
| Beyond Similarity: Grounded Agentic Extraction and Expert-Adjudicated Evaluation of Intertextuality in Classical Chinese Histories Jun Wang, Wanyu Si, Zhaoji Wang Published: 2026-07-30Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint, safety-evaluation | 2026-07-30 | cs.CL | ai-safety, cscl, preprint, safety-evaluation | E8 / R5 (92%) | - |
| CACHE-UK: A Stability-Aware Memory Editor for Sequentially Updated Quantized LLMs in Finance Anubhav Lakra, Yue Feng Published: 2026-07-30Area: cs.CLCitations: 21 Tags: ai-safety, cscl, preprint | 2026-07-30 | cs.CL | ai-safety, cscl, preprint | E8 / R7 (92%) | 21 |
| Can Agents Deceive? Evaluating Reasoning and Deception in ParliamentBench using a Social Deduction Game Akiko Aizawa, Bela Gipp, Jan Philip Wahle, Lars Benedikt Kaesberg Published: 2026-07-30Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-30 | cs.CL | ai-safety, cscl, preprint | E12 / R12 (94%) | - |
| Can LVLMs Uncover the Truth Behind Visual Illusions? An Analysis of Perceptual and Reasoning Capabilities Da Li, Jianing Li, Jiaqing Lyu, Kexin Tang Published: 2026-07-30Area: cs.CLCitations: 45 Tags: ai-safety, cscl, preprint | 2026-07-30 | cs.CL | ai-safety, cscl, preprint | E10 / R7 (89%) | 45 |
| DualAnchor: Preserving Language Priors and Improving Lexical Fidelity in Gloss-Free Sign Language Translation Hongbin Zhang, Junhao Liu, Kehai Chen, Xuefeng Bai Published: 2026-07-30Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-30 | cs.CL | ai-safety, cscl, preprint | E10 / R8 (91%) | - |
| EMBL AI Librarian: Life-Sciences Knowledge Layer for AI Agents Daljit Nijjer, Fabio Petroni, Francesco Tabaro, Julio Saez-Rodriguez Published: 2026-07-30Area: cs.CLCitations: 15 Tags: ai-safety, cscl, preprint | 2026-07-30 | cs.CL | ai-safety, cscl, preprint | E10 / R7 (89%) | 15 |
| FinanceHarness: Autonomous Financial Deep Research Framework Burak Gokturk, Chen-Yu Lee, Ke Jiang, Rujun Han Published: 2026-07-30Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-30 | cs.CL | ai-safety, cscl, preprint | E6 / R4 (94%) | - |
| From Single- to Cross-Document: Benchmarking Multi-Granularity Event Analysis of Large Language Models Guannan Li, Jie Zou, Jinjie Qiu, Ke Qin Published: 2026-07-30Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-30 | cs.CL | ai-safety, cscl, preprint | E10 / R8 (92%) | - |
| ORCA-bench: How Ready Are Language Model Agents for Oncall? Abhineet Agarwal, Albert Gong, Anish Agarwal, Jason Schechner Published: 2026-07-30Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-30 | cs.CL | ai-safety, cscl, preprint | E14 / R8 (91%) | - |
| ORCA-bench: How Ready Are Language Model Agents for Oncall? Abhineet Agarwal, Albert Gong, Anish Agarwal, Jason Schechner Published: 2026-07-30Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-30 | cs.CL | ai-safety, cscl, preprint | E14 / R11 (90%) | - |
| Reasoning Consensus: Structural Ensembling of LLM Reasoning via Weighted DAG Aggregation Amruta Parulekar, Dilek Hakkani-T眉r, Hari Sundaram, Jinu Lee Published: 2026-07-30Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-30 | cs.CL | ai-safety, cscl, preprint | E9 / R7 (92%) | - |
| RepBench: Compiling Benchmarks into Capability Representations for Large Language Models Long Zhang, Shuman Liu, Xueru Bai, Yanshi Li Published: 2026-07-30Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-30 | cs.CL | ai-safety, cscl, preprint | E15 / R16 (92%) | - |
| Rethinking LLM-Judged Helpfulness as a Pedagogy Signal: A Pre-Registered Audit Across Tutor Models Boyuan Deng, Chongyang Gao, Hongyang Zhang, Jiale Liu Published: 2026-07-30Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-30 | cs.CL | ai-safety, cscl, preprint | E12 / R11 (93%) | - |
| Rethinking LLM-Judged Helpfulness as a Pedagogy Signal: A Pre-Registered Audit Across Tutor Models Boyuan Deng, Chongyang Gao, Hongyang Zhang, Jiale Liu Published: 2026-07-30Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-30 | cs.CL | ai-safety, cscl, preprint | E12 / R8 (92%) | - |
| Rolling With Resistance: Preference-Optimized LLM Counselors Can Trade Goal Persistence for Relational Attunement in Motivational Interviewing Junlong Shen, Weiying Chen, Zhexuan Tang Published: 2026-07-30Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-30 | cs.CL | ai-safety, cscl, preprint | E10 / R7 (93%) | - |
| Sample More, Reflect Less: Self-Refine and Reflexion Lose to Repeated Sampling at Equal Token Cost, from 1.5B to 7B Iliya Mirzaei Published: 2026-07-30Area: cs.CLCitations: 37 Tags: ai-safety, cscl, preprint | 2026-07-30 | cs.CL | ai-safety, cscl, preprint | E9 / R7 (95%) | 37 |
| Semantic-Aligned Structural Abstraction for Multimodal Sentiment Analysis Chuanxiang Ma, Feiyue Xue, Hui Liu, Junkai Li Published: 2026-07-30Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-30 | cs.CL | ai-safety, cscl, preprint | E11 / R10 (94%) | - |
| TextCloak: Thwarting Unauthorized LLM Exploitation via RL-Driven Unlearnable Text Bohan Jiang, Chengshuai Zhao, Dawei Li, Huan Liu Published: 2026-07-30Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-30 | cs.CL | ai-safety, cscl, preprint | E6 / R6 (91%) | - |
| The MADRS Pipeline: Supporting Depression Assessment in Clinical Trials Alex Boudreau, Francesco Periti, Katalin 脫csai, Mila Fodor Published: 2026-07-30Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-30 | cs.CL | ai-safety, cscl, preprint | E14 / R13 (95%) | - |
| Where and When to Commit: Candidate-Aware Decoding for Diffusion Language Models Chia-Ming Lee, Chih-Chung Hsu, Ming-Ching Chang, Xin Li Published: 2026-07-30Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-30 | cs.CL | ai-safety, cscl, preprint | E10 / R8 (91%) | - |
| AHA-Memes: A Fine-Grained Multimodal Benchmark for Understanding Hate in Arabic Memes Abul Hasnat, Ali Ezzat Shahroor, Firoj Alam, Md. Rafiul Biswas Published: 2026-07-29Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-29 | cs.CL | ai-safety, cscl, preprint | E26 / R26 (91%) | - |
| APEX-Accounting Austin Bennett, Bertie Vidgen, Brendan Foody, Charis Ching Published: 2026-07-29Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-29 | cs.CL | ai-safety, cscl, preprint | - | - |
| APEX-Accounting Austin Bennett, Bertie Vidgen, Brendan Foody, Charis Ching Published: 2026-07-29Area: cs.CLCitations: 15 Tags: ai-safety, cscl, preprint | 2026-07-29 | cs.CL | ai-safety, cscl, preprint | E11 / R9 (94%) | 15 |
| Automated Multilabel Mpox Research Classification with Explainable Transformer Models Tanjim Taharat Aurpa Published: 2026-07-29Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-29 | cs.CL | ai-safety, cscl, preprint | E10 / R5 (91%) | - |
| Belief Coevolution in a Social Network of Generalist and Specialist Large Language Models Courtney Maynard, Germans Savcisens, Samantha Dies, Tina Eliassi-Rad Published: 2026-07-29Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-29 | cs.CL | ai-safety, cscl, preprint | E13 / R9 (91%) | - |
| Benchmarking LLM Competence on Logical Inference over Probability Operators Alvin Grissom, Jack Greff, Nayera Hasan Published: 2026-07-29Area: cs.CLCitations: 45 Tags: ai-safety, cscl, preprint | 2026-07-29 | cs.CL | ai-safety, cscl, preprint | E9 / R5 (89%) | 45 |