Instant research discovery
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
| Screen Before You Interpret: A Portable Validity Protocol for Benchmark-Based LLM Confidence Signals Jon-Paul Cacioli Published: 2026-04-20Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-04-20 | cs.CL | ai-safety, cscl, preprint | E16 / R10 (99%) | - |
| Semantic Needles in Document Haystacks: Sensitivity Testing of LLM-as-a-Judge Similarity Scoring Alexandra A. Sabrio, Erik VonKaenel, Lee Burke, Sinan G. Aksoy Published: 2026-04-20Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-04-20 | cs.CL | ai-safety, cscl, preprint | E10 / R4 (94%) | - |
| SPENCE: A Syntactic Probe for Detecting Contamination in NL2SQL Benchmarks Afshin Orojlooyjadid, Dan Roth, Graham Horwood, Hitesh Laxmichand Patel Published: 2026-04-20Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-04-20 | cs.CL | ai-safety, cscl, preprint | E17 / R8 (97%) | - |
| STaD: Scaffolded Task Design for Identifying Compositional Skill Gaps in LLMs Chad DeLuca, Hima Patel, Shailja Thakur, Sungeun An Published: 2026-04-20Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-04-20 | cs.CL | ai-safety, cscl, preprint | E10 / R6 (99%) | - |
| STaD: Scaffolded Task Design for Identifying Compositional Skill Gaps in LLMs Chad DeLuca, Hima Patel, Shailja Thakur, Sungeun An Published: 2026-04-20Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-04-20 | cs.CL | ai-safety, cscl, preprint | E10 / R6 (98%) | - |
| TLoRA: Task-aware Low Rank Adaptation of Large Language Models Jiawei Dang, Liang-Jie Zhang, Weicheng Lin, Yi Zhang Published: 2026-04-20Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-04-20 | cs.CL | ai-safety, cscl, preprint | E11 / R4 (97%) | - |
| Towards Intelligent Legal Document Analysis: CNN-Driven Classification of Case Law Texts Ahnaf Atef Choudhury, Jia Uddin, Moinul Hossain, Sadia Afrin Promi Published: 2026-04-20Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-04-20 | cs.CL | ai-safety, cscl, preprint | E10 / R8 (97%) | - |
| Transition-Matrix Regularization for Next Dialogue Act Prediction in Counselling Conversations Eric Rudolph, Jens Albrecht, Philipp Steigerwald Published: 2026-04-20Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-04-20 | cs.CL | ai-safety, cscl, preprint | E9 / R5 (96%) | - |
| Where Fake Citations Are Made: Tracing Field-Level Hallucination to Specific Neurons in LLMs Ruixiang Tang, Xiaodong Lin, Yihao Quan, Yuefei Chen Published: 2026-04-20Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-04-20 | cs.CL | ai-safety, cscl, preprint | E10 / R4 (97%) | - |
| $R^2$-dLLM: Accelerating Diffusion Large Language Models via Spatio-Temporal Redundancy Reduction Binfei Ji, Brucek Khailany, Kejing Xia, Nicolai Oswald Published: 2026-04-21Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-04-21 | cs.CL | ai-safety, cscl, preprint | E8 / R4 (99%) | - |
| Assessing Capabilities of Large Language Models in Social Media Analytics: A Multi-task Quest Hamid Karimi, Kartik Thakkar, Nazanin Donyapour, Ramtin Davoudi Published: 2026-04-21Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-04-21 | cs.CL | ai-safety, cscl, preprint | E11 / R5 (98%) | - |
| Chat2Workflow: A Benchmark for Generating Executable Visual Workflows with Natural Language Buqiang Xu, Guozhou Zheng, Ningyu Zhang, Shuofei Qiao Published: 2026-04-21Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-04-21 | cs.CL | ai-safety, cscl, preprint | E7 / R5 (96%) | - |
| Cross-Model Consistency of AI-Generated Exercise Prescriptions: A Repeated Generation Study Across Three Large Language Models Kihyuk Lee Published: 2026-04-21Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-04-21 | cs.CL | ai-safety, cscl, preprint | E7 / R4 (96%) | - |
| Cross-Model Consistency of AI-Generated Exercise Prescriptions: A Repeated Generation Study Across Three Large Language Models Kihyuk Lee Published: 2026-04-21Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-04-21 | cs.CL | ai-safety, cscl, preprint | E7 / R4 (96%) | - |
| CulturALL: Benchmarking Multilingual and Multicultural Competence of LLMs on Grounded Tasks Adrian Neo Sau Xun, Alham Fikri Aji, Baotian Hu, Bo Zeng Published: 2026-04-21Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-04-21 | cs.CL | ai-safety, cscl, preprint | E8 / R4 (98%) | - |
| Depression Risk Assessment in Social Media via Large Language Models Giorgia Gulino, Manuel Petrucci Published: 2026-04-21Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-04-21 | cs.CL | ai-safety, cscl, preprint | E8 / R4 (96%) | - |
| Detecting Hallucinations in SpeechLLMs at Inference Time Using Attention Maps Bashar Awwad Shiekh Hasan, Evgenii Tsymbalov, Jonas Waldendorf Published: 2026-04-21Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-04-21 | cs.CL | ai-safety, cscl, preprint | E11 / R4 (98%) | - |
| From Signal Degradation to Computation Collapse: Uncovering the Two Failure Modes of LLM Quantization Bohan Yu, Chenxi Zhou, Jiang Li, Jinyu Ye Published: 2026-04-21Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-04-21 | cs.CL | ai-safety, cscl, preprint | E8 / R5 (97%) | - |
| How Do Answer Tokens Read Reasoning Traces? Self-Reading Patterns in Thinking LLMs for Quantitative Reasoning Chengfu Huo, Haoyang Chen, Jianzhi Shao, Tao Zhang Published: 2026-04-21Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-04-21 | cs.CL | ai-safety, cscl, preprint | E7 / R6 (95%) | - |
| Impact of large language models on peer review opinions from a fine-grained perspective: Evidence from top conference proceedings in AI Chengzhi Zhang, Tong Bao, Wenqing Wu, Yi Zhao Published: 2026-04-21Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-04-21 | cs.CL | ai-safety, cscl, preprint | E9 / R3 (97%) | - |
| IndiaFinBench: An Evaluation Benchmark for Large Language Model Performance on Indian Financial Regulatory Text Rajveer Singh Pall Published: 2026-04-21Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint, safety-evaluation | 2026-04-21 | cs.CL | ai-safety, cscl, preprint, safety-evaluation | E9 / R4 (98%) | - |
| LePREC: Reasoning as Classification over Structured Factors for Assessing Relevance of Legal Issues Aashish Srivastava, Adnan Trakic, Chetan Arora, Fanyu Wang Published: 2026-04-21Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-04-21 | cs.CL | ai-safety, cscl, preprint | E7 / R4 (97%) | - |
| Location Not Found: Exposing Implicit Local and Global Biases in Multilingual LLMs Adi Mayrav Gilady, Avinatan Hassidim, Guy Mor-Lan, Idan Szpektor Published: 2026-04-21Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-04-21 | cs.CL | ai-safety, cscl, preprint | E7 / R4 (97%) | - |
| Product-of-Experts Training Reduces Dataset Artifacts in Natural Language Inference Aby Mammen Mathew Published: 2026-04-21Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-04-21 | cs.CL | ai-safety, cscl, preprint | E7 / R4 (97%) | - |
| RARE: Redundancy-Aware Retrieval Evaluation Framework for High-Similarity Corpora Hanjun Cho, Jay-Yoon Lee Published: 2026-04-21Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint, safety-evaluation | 2026-04-21 | cs.CL | ai-safety, cscl, preprint, safety-evaluation | E9 / R5 (98%) | - |
| Rethinking Scale: Deployment Trade-offs of Small Language Models under Agent Paradigms Mats Brorsson, Xinlin Wang Published: 2026-04-21Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-04-21 | cs.CL | ai-safety, cscl, preprint | E10 / R4 (97%) | - |
| RoLegalGEC: Legal Domain Grammatical Error Detection and Correction Dataset for Romanian Dumitru-Clementin Cercel, Mircea Timpuriu Published: 2026-04-21Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-04-21 | cs.CL | ai-safety, cscl, preprint | E9 / R5 (96%) | - |
| RoLegalGEC: Legal Domain Grammatical Error Detection and Correction Dataset for Romanian Dumitru-Clementin Cercel, Mihaela-Claudia Cercel, Mircea Timpuriu Published: 2026-04-21Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-04-21 | cs.CL | ai-safety, cscl, preprint | E8 / R5 (96%) | - |
| SAHM: A Benchmark for Arabic Financial and Shari'ah-Compliant Reasoning Ahmed Heakl, Dani Bouch, Jimin Huang, Marwa Elsaid khalil Published: 2026-04-21Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-04-21 | cs.CL | ai-safety, cscl, preprint | E7 / R4 (97%) | - |
| SAMoRA: Semantic-Aware Mixture of LoRA Experts for Task-Adaptive Learning Boyan Shi, Huaiyu Wan, Junfeng Shen, Shaojiang Wang Published: 2026-04-21Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-04-21 | cs.CL | ai-safety, cscl, preprint | E9 / R5 (97%) | - |