Instant research discovery
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
| Closed-Loop Validation-Repair for Healthcare Interoperability: A Multi-Model Study of Schema Compliance in Clinical LLMs Jianru Shen Published: 2026-07-27Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-27 | cs.CL | ai-safety, cscl, preprint | - | - |
| CogArena: A Multimethod Evaluation of Cognitive Ability Structure in Large Language Models Dengzhe Hou, Fangzhou Lin, Kazunori D Yamada, Lingyu Jiang Published: 2026-07-27Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint, safety-evaluation | 2026-07-27 | cs.CL | ai-safety, cscl, preprint, safety-evaluation | E15 / R16 (91%) | - |
| DataOrchestra: Learning to Orchestrate Per-Example Curation of Pretraining Data Pengfei Liu, Shijie Xia, Yikun Wang, Zhen Huang Published: 2026-07-27Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-27 | cs.CL | ai-safety, cscl, preprint | - | - |
| D-Score: A Spectral Hidden-State Signal for Hallucination Detection in Large Language Models Bianca Raimondi, Davide Evangelista, Elena Loli Piccolomini, Maurizio Gabbrielli Published: 2026-07-27Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-27 | cs.CL | ai-safety, cscl, preprint | - | - |
| DS@GT ARC at CheckThat! 2026: LLM-Based Trace Ranking and Grouped Reward Modeling for Multilingual Numerical Claim Verification Sagnik Sinha, Shreyas Shrestha Published: 2026-07-27Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-27 | cs.CL | ai-safety, cscl, preprint | E17 / R9 (90%) | - |
| Evaluating Communicative Belief Updates in Large Language Models via Implicature Recognition and Cancellation Cesare Spinoso-Di Piano, Jackie Chi Kit Cheung, Marius Mosbach, Verna Dankers Published: 2026-07-27Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-27 | cs.CL | ai-safety, cscl, preprint | E10 / R7 (93%) | - |
| Evaluating Communicative Belief Updates in Large Language Models via Implicature Recognition and Cancellation Cesare Spinoso-Di Piano, Jackie Chi Kit Cheung, Marius Mosbach, Verna Dankers Published: 2026-07-27Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-27 | cs.CL | ai-safety, cscl, preprint | E10 / R9 (91%) | - |
| LEX-EC: A Lexical Evidence-Channel Audit Framework for Zero-Shot LLM Personality Classification in Black-Box Settings Ashok K. Goel, Brittany Harbison Published: 2026-07-27Area: cs.CLCitations: 15 Tags: ai-safety, cscl, preprint | 2026-07-27 | cs.CL | ai-safety, cscl, preprint | E19 / R13 (91%) | 15 |
| LEX-EC: A Lexical Evidence-Channel Audit Framework for Zero-Shot LLM Personality Classification in Black-Box Settings Ashok K. Goel, Brittany Harbison Published: 2026-07-27Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-27 | cs.CL | ai-safety, cscl, preprint | - | - |
| Looping Is Not Reliability: State-Bound Evidence and Typed Revision Contracts for Agentic Code Repair Jianwei Yang, Qiang Yang, Xueping Gao Published: 2026-07-27Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-27 | cs.CL | ai-safety, cscl, preprint | - | - |
| Self-Authored Verification Is Unreliable in Heuristic Self-Improving Agents Cong Cao, Dakui Wang, Diandian Guo, Fangfang Yuan Published: 2026-07-27Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-27 | cs.CL | ai-safety, cscl, preprint | E7 / R5 (94%) | - |
| StanceFlip: A Comprehensive Multi-Dimensional Benchmark for Multimodal Conversational Stance Flipping Forecasting Chaoyang Li, Hao Chen, Heyan Chai, Jianyang Qin Published: 2026-07-27Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-27 | cs.CL | ai-safety, cscl, preprint | - | - |
| SyRuP: Enhancing System-Prompt Following via Reward-Guided Prediction in LLM Decoding Jaehyung Kim, Minjae Kang, Seoyeon Kim Published: 2026-07-27Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-27 | cs.CL | ai-safety, cscl, preprint | - | - |
| Tag Questions and the Generational Reversal of Sycophancy Across 45 Language Models Tapan Parikh Published: 2026-07-27Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-27 | cs.CL | ai-safety, cscl, preprint | - | - |
| The Physics of Multi-Turn Long-Horizon Planning: From Pre-training to Post-training via Single- and Multi-Teacher On-Policy Agentic Distillation Jun Zhao, Kang Liu, Tianyi Men, Zhuoran Jin Published: 2026-07-27Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-27 | cs.CL | ai-safety, cscl, preprint | - | - |
| The Tokenizer Tax: Quantifying and Explaining the Cross-Lingual Cost of Subword Tokenization for Indian Languages Priyansh Srivastava Published: 2026-07-27Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-27 | cs.CL | ai-safety, cscl, preprint | - | - |
| Understanding Tone-Dependent Inference Cost in Large Language Models Akhil Kumar, Om Dobariya Published: 2026-07-27Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-27 | cs.CL | ai-safety, cscl, preprint | - | - |
| A Frozen 12B Beats Frontier Models on Verified Work: 100% Accuracy, 0 Tokens, Bit-Exact, Forever Sietse Schelpe Published: 2026-07-26Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-26 | cs.CL | ai-safety, cscl, preprint | - | - |
| An empirical investigation into the properties of standard word embeddings Salomon Kabongo Published: 2026-07-26Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-26 | cs.CL | ai-safety, cscl, preprint | - | - |
| Do Diagrams Help Large Language Models Reason? Evidence from Syllogistic Reasoning Koji Mineshima, Risako Ando Published: 2026-07-26Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-26 | cs.CL | ai-safety, cscl, preprint | - | - |
| Do Small Models Use the Law You Give Them? Context-Injected Fine-Tuning for Legal QA in Bangladesh Abrar Mohammed Tanzim Alam, Md Adnan Arefeen, Mir Mohammad Asif Abdullah, Moniruzzaman Mahadi Published: 2026-07-26Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-26 | cs.CL | ai-safety, cscl, preprint | - | - |
| Earnings25: A Comprehensive 500-Hour Speech Benchmark for Finance Anshul Wadhawan, Brendan Fahy, David Weisberg, Denglin Jiang Published: 2026-07-26Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-26 | cs.CL | ai-safety, cscl, preprint | - | - |
| How Context Attribution Handles What the Model Already Knows Lin Zhu, Quoc-Huy Trinh, Sebastian Szyller Published: 2026-07-26Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-26 | cs.CL | ai-safety, cscl, preprint | - | - |
| Indic DiarBench: A Multilingual Joint Diarization and ASR Benchmark for Indian Languages Aditya Mehndiratta, Deovrat Mehendale, Dhruv Rathi, Kaushal Bhogale Published: 2026-07-26Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-26 | cs.CL | ai-safety, cscl, preprint | - | - |
| Novel Claim or D茅j脿 Vu? Rethinking "Contamination-Free'' Dynamic Evaluation for Multimodal Automated Fact-Checking Dacheng Wen, Francis C. M. Lau, Haorui He, Reynold Cheng Published: 2026-07-26Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint, safety-evaluation | 2026-07-26 | cs.CL | ai-safety, cscl, preprint, safety-evaluation | - | - |
| Reasoning or Memorization: Can LLMs Understand and Generate Chinese Xiehouyu Riddles? Chongtian Shao, Hai Hu, Kejia Zhang, Siyuan Song Published: 2026-07-26Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-26 | cs.CL | ai-safety, cscl, preprint | - | - |
| ADAGE: A Language-Agnostic Pipeline for Analogical Reasoning Evaluation Ahmed Haj Ahmed, Alvin Grissom Published: 2026-07-25Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint, safety-evaluation | 2026-07-25 | cs.CL | ai-safety, cscl, preprint, safety-evaluation | - | - |
| Attention-Guided Layer Selection for Contrastive Decoding in Large Language Models Kiyoaki Shirai, Natthawut Kertkeidkachorn, Yusuke Sakai Published: 2026-07-25Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-25 | cs.CL | ai-safety, cscl, preprint | - | - |
| Beyond a Global Norm: Personalizing Toxicity Sensitivity in Language Models Without Retraining Alina Florea, Andrei B. Trache, Anne Arzberger, Enrico Liscio Published: 2026-07-25Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-25 | cs.CL | ai-safety, cscl, preprint | - | - |
| Beyond Direct Answering: Aligning Educational LLMs as Socratic Guides via Heuristic Reinforcement Learning Siyu Song, Wentao Liu, Xiaodong Zou, Xiaokun Wang Published: 2026-07-25Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-07-25 | cs.CL | ai-safety, cscl, preprint | - | - |