Instant research discovery
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
| Towards More Standardized AI Evaluation: From Models to Agents Ali El Filali, In猫s Bedar Published: 2026-02-20Area: cs.CLCitations: 20 Tags: ai-safety, cscl, preprint, safety-evaluation | 2026-02-20 | cs.CL | ai-safety, cscl, preprint, safety-evaluation | E10 / R8 (92%) | 20 |
| Validating Political Position Predictions of Arguments Angus R. Williams, Anthony G. Cohn, Jordan Robinson, Katie Atkinson Published: 2026-02-20Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-02-20 | cs.CL | ai-safety, cscl, preprint | E13 / R7 (92%) | - |
| Vichara: Appellate Judgment Prediction and Explanation for the Indian Judicial System Pavithra PM Nair, Preethu Rose Anish Published: 2026-02-20Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-02-20 | cs.CL | ai-safety, cscl, preprint | E9 / R6 (93%) | - |
| ArabicNumBench: Evaluating Arabic Number Reading in Large Language Models Abdulaziz Alhammadi, Anas Alhumud, Muhammad Badruddin Khan Published: 2026-02-21Area: cs.CLCitations: 25 Tags: ai-safety, cscl, preprint | 2026-02-21 | cs.CL | ai-safety, cscl, preprint | E13 / R10 (91%) | 25 |
| DeepInnovator: Triggering the Innovative Capabilities of LLMs Bei Chen, Chao Huang, Chengen Huang, Fengji Zhang Published: 2026-02-21Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-02-21 | cs.CL | ai-safety, cscl, preprint | E11 / R7 (93%) | - |
| Rethinking Retrieval-Augmented Generation as a Cooperative Decision-Making Problem Lichang Song, Ting Long, Yi Chang Published: 2026-02-21Area: cs.CLCitations: 15 Tags: ai-safety, cscl, preprint | 2026-02-21 | cs.CL | ai-safety, cscl, preprint | E10 / R11 (91%) | 15 |
| Semantic Substrate Theory: An Operator-Theoretic Framework for Geometric Semantic Drift Stephen Russell Published: 2026-02-21Area: cs.CLCitations: 11 Tags: ai-safety, cscl, preprint | 2026-02-21 | cs.CL | ai-safety, cscl, preprint | E10 / R9 (88%) | 11 |
| Think$^{2}$: Grounded Metacognitive Reasoning in Large Language Models Abraham Paul Elenjical, Vasudeva Varma, Vivek Hruday Kavuri Published: 2026-02-21Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-02-21 | cs.CL | ai-safety, cscl, preprint | E10 / R9 (91%) | - |
| When Prompt Optimization Becomes Jailbreaking: Adaptive Red-Teaming of Large Language Models Nikhil Chekuru, Shivank Garg, Zachary Guzman, Zafir Shamsi Published: 2026-02-21Area: cs.CLCitations: 45 Tags: adversarial-robustness, ai-safety, cscl, preprint | 2026-02-21 | cs.CL | adversarial-robustness, ai-safety, cscl, preprint | E12 / R9 (94%) | 45 |
| Why Agent Caching Fails and How to Fix It: Structured Intent Canonicalization with Few-Shot Learning Abhinaba Basu Published: 2026-02-21Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-02-21 | cs.CL | ai-safety, cscl, preprint | E10 / R8 (94%) | - |
| Anatomy of Agentic Memory: Taxonomy and Empirical Analysis of Evaluation and System Limitations Alysa Zhao, Ayushi Kishore, Bingzhe Li, Dingyi Kang Published: 2026-02-22Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint, safety-evaluation | 2026-02-22 | cs.CL | ai-safety, cscl, preprint, safety-evaluation | E11 / R12 (94%) | - |
| How Do LLMs Encode Scientific Quality? An Empirical Study Using Monosemantic Features from Sparse Autoencoders Angelo Salatino, Enrico Motta, Francesco Osborne, Michael McCoubrey Published: 2026-02-22Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-02-22 | cs.CL | ai-safety, cscl, preprint | E11 / R8 (93%) | - |
| Learning to Reason for Multi-Step Retrieval of Personal Context in Personalized Question Answering Alireza Salemi, Hamed Zamani, Maryam Amirizaniani Published: 2026-02-22Area: cs.CLCitations: 15 Tags: ai-safety, cscl, preprint | 2026-02-22 | cs.CL | ai-safety, cscl, preprint | E10 / R9 (90%) | 15 |
| Next Reply Prediction X Dataset: Linguistic Discrepancies in Naively Generated Content Achim Rettinger, Kai Kugler, Michael Heseltine, Nils Schwager Published: 2026-02-22Area: cs.CLCitations: 12 Tags: ai-safety, cscl, preprint | 2026-02-22 | cs.CL | ai-safety, cscl, preprint | E9 / R8 (92%) | 12 |
| Prompt Injection as Role Confusion Charles Ye, Dylan Hadfield-Menell, Jasmine Cui Published: 2026-02-22Area: cs.CLCitations: 15 Tags: ai-safety, cscl, preprint | 2026-02-22 | cs.CL | ai-safety, cscl, preprint | E8 / R6 (93%) | 15 |
| Value Entanglement: Conflation Between Different Kinds of Good In (Some) Large Language Models Anna Leshinskaya, Junyi Li, Seong Hah Cho Published: 2026-02-22Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-02-22 | cs.CL | ai-safety, cscl, preprint | E9 / R7 (92%) | - |
| Agentic Framework for Political Biography Extraction Jiangnan Zhu, Junyan Jiang, Songpo Yang, Yifei Zhu Published: 2026-02-23Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-02-23 | cs.CL | ai-safety, cscl, preprint | E7 / R5 (92%) | - |
| AgenticSum: An Agentic Inference-Time Framework for Faithful Clinical Text Summarization Fahmida Liza Piya, Rahmatollah Beheshti Published: 2026-02-23Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-02-23 | cs.CL | ai-safety, cscl, preprint | E10 / R8 (93%) | - |
| Assessing Risks of Large Language Models in Mental Health Support: A Framework for Automated Clinical AI Red Teaming Ian Steenstra, Paola Pedrelli, Stacy Marsella, Timothy W. Bickmore Published: 2026-02-23Area: cs.CLCitations: 206 Tags: ai-safety, cscl, preprint, red-teaming | 2026-02-23 | cs.CL | ai-safety, cscl, preprint, red-teaming | E12 / R13 (91%) | 206 |
| Case-Aware LLM-as-a-Judge Evaluation for Enterprise-Scale RAG Systems Arush Verma, Luigi Medrano, Mukul Chhabra Published: 2026-02-23Area: cs.CLCitations: 15 Tags: ai-safety, cscl, preprint, safety-evaluation | 2026-02-23 | cs.CL | ai-safety, cscl, preprint, safety-evaluation | E9 / R7 (93%) | 15 |
| How Confident Is the First Token? An Uncertainty-Calibrated Prompt Optimization Framework for Large Language Model Classification and Understanding Guoyang Ju, Wei Chen, Yuanyuan Qi Published: 2026-02-23Area: cs.CLCitations: 39 Tags: ai-safety, cscl, preprint | 2026-02-23 | cs.CL | ai-safety, cscl, preprint | E7 / R6 (92%) | 39 |
| InterviewSim: A Scalable Framework for Interview-Grounded Personality Simulation Chien-Sheng Wu, Pranav Narayanan Venkit, Yada Pruksachatkun, Yu Li Published: 2026-02-23Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-02-23 | cs.CL | ai-safety, cscl, preprint | E7 / R5 (90%) | - |
| KNIGHT: Knowledge Graph-Driven Multiple-Choice Question Generation with Adaptive Hardness Calibration Behnam Bahrak, Erfan Shafiee Moghaddam, Farhan Farsi, Mahdi Noori Published: 2026-02-23Area: cs.CLCitations: 75 Tags: ai-safety, cscl, preprint | 2026-02-23 | cs.CL | ai-safety, cscl, preprint | E11 / R10 (91%) | 75 |
| Multilingual Large Language Models do not comprehend all natural languages to equal degrees Elena Pagliarini, Evelina Leivada, Ferdy Hubers, Fritz G眉nther Published: 2026-02-23Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-02-23 | cs.CL | ai-safety, cscl, preprint | E10 / R7 (90%) | - |
| NanoKnow: How to Know What Your Language Model Knows Jimmy Lin, Lingwei Gu, Nour Jedidi Published: 2026-02-23Area: cs.CLCitations: 28 Tags: ai-safety, cscl, preprint | 2026-02-23 | cs.CL | ai-safety, cscl, preprint | E10 / R11 (93%) | 28 |
| No One Size Fits All: QueryBandits for Hallucination Mitigation Alec Koppel, Manuela Veloso, Nicole Cho, Sumitra Ganesh Published: 2026-02-23Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-02-23 | cs.CL | ai-safety, cscl, preprint | E8 / R7 (90%) | - |
| Pyramid MoA: A Probabilistic Framework for Cost-Optimized Anytime Inference Arindam Khaled Published: 2026-02-23Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-02-23 | cs.CL | ai-safety, cscl, preprint | E15 / R10 (93%) | - |
| ReAttn: Improving Attention-based Re-ranking via Attention Re-weighting Fengran Mo, Jian-Yun Nie, Weixu Zhang, Yiyan Qi Published: 2026-02-23Area: cs.CLCitations: 15 Tags: ai-safety, cscl, preprint | 2026-02-23 | cs.CL | ai-safety, cscl, preprint | E13 / R11 (89%) | 15 |
| Temporal-Aware Heterogeneous Graph Reasoning with Multi-View Fusion for Temporal Question Answering Bowen Zhou, Jianting Liu, Jinwen Huang, Liang Li Published: 2026-02-23Area: cs.CLCitations: 27 Tags: ai-safety, cscl, preprint | 2026-02-23 | cs.CL | ai-safety, cscl, preprint | E8 / R6 (93%) | 27 |
| TherapyGym: Evaluating and Aligning Clinical Fidelity and Safety in Therapy Chatbots Arpandeep Khatua, Ehsan Adeli, Fangrui Huang, Kenan Ye Published: 2026-02-23Area: cs.CLCitations: 45 Tags: ai-safety, cscl, preprint | 2026-02-23 | cs.CL | ai-safety, cscl, preprint | E8 / R6 (91%) | 45 |