Instant research discovery
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
| ArgLLM-App: An Interactive System for Argumentative Reasoning with Large Language Models Adam Dejl, Deniz Gorur, Francesca Toni Published: 2026-02-27Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-02-27 | cs.CL | ai-safety, cscl, preprint | E10 / R8 (94%) | - |
| ARGUS: Seeing the Influence of Narrative Features on Persuasion in Argumentative Texts Federico Pianzola, Khalid Al-Khatib, Malvina Nissim, Sara Nabhani Published: 2026-02-27Area: cs.CLCitations: 15 Tags: ai-safety, cscl, preprint | 2026-02-27 | cs.CL | ai-safety, cscl, preprint | E17 / R21 (89%) | 15 |
| A unified foundational framework for knowledge injection and evaluation of Large Language Models in Combustion Science Han Li, QingGuo Zhou, Runze Mao, Tianhao Wu Published: 2026-02-27Area: cs.CLCitations: 16 Tags: ai-safety, cscl, preprint, safety-evaluation | 2026-02-27 | cs.CL | ai-safety, cscl, preprint, safety-evaluation | E10 / R8 (93%) | 16 |
| BRIDGE the Gap: Mitigating Bias Amplification in Automated Scoring of English Language Learners via Inter-group Data Augmentation Jingyuan Huang, Lei Liu, Ninghao Liu, Xiaoming Zhai Published: 2026-02-27Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-02-27 | cs.CL | ai-safety, cscl, preprint | E6 / R5 (93%) | - |
| CoME: Empowering Channel-of-Mobile-Experts with Informative Hybrid-Capabilities Reasoning Bo Du, Changyu Chen, Jiankun Zhao, Jian Luan Published: 2026-02-27Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-02-27 | cs.CL | ai-safety, cscl, preprint | E12 / R9 (92%) | - |
| Do LLMs Benefit From Their Own Words? Jacob Andreas, Jenny Y. Huang, Leshem Choshen, Ramon Astudillo Published: 2026-02-27Area: cs.CLCitations: 45 Tags: ai-safety, cscl, preprint | 2026-02-27 | cs.CL | ai-safety, cscl, preprint | E10 / R6 (90%) | 45 |
| Evaluating LLM-Based Translation of a Low-Resource Technical Language: The Medical and Philosophical Greek of Galen Cameron Pattison, Jacob Wu, James L. Zainaldin, Manuela Marai Published: 2026-02-27Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-02-27 | cs.CL | ai-safety, cscl, preprint | E11 / R9 (91%) | - |
| From Leaky Thoughts to Private Reasoning: Controlling What LRMs Say to Themselves Haonan Li, Haritz Puerto, Iryna Gurevych, Timothy Baldwin Published: 2026-02-27Area: cs.CLCitations: 15 Tags: ai-safety, cscl, preprint | 2026-02-27 | cs.CL | ai-safety, cscl, preprint | E12 / R7 (94%) | 15 |
| From Static Benchmarks to Dynamic Protocol: Agent-Centric Text Anomaly Detection for Evaluating LLM Reasoning Dongmin Kim, Junhyun Lee, Sanghyu Yoon, Seungdong Yoa Published: 2026-02-27Area: cs.CLCitations: 15 Tags: ai-safety, cscl, preprint | 2026-02-27 | cs.CL | ai-safety, cscl, preprint | E14 / R15 (90%) | 15 |
| Full-Stack Domain Enhancement for Combustion LLMs: Construction and Optimization Qingguo Zhou, Quanjia Xiao, Runze Mao, Tianhao Wu Published: 2026-02-27Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-02-27 | cs.CL | ai-safety, cscl, preprint | E10 / R8 (91%) | - |
| Induced Numerical Instability: Hidden Costs in Multimodal Large Language Models Arunesh Sinha, Jun Sun, Wai Tuck Wong Published: 2026-02-27Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-02-27 | cs.CL | ai-safety, cscl, preprint | E13 / R7 (94%) | - |
| LFQA-HP-1M: A Large-Scale Human Preference Dataset for Long-Form Question Answering Fahmid Shahriar Iqbal, Rafid Ishrak Jahan, Sagnik Ray Choudhury Published: 2026-02-27Area: cs.CLCitations: 45 Tags: ai-safety, cscl, preprint | 2026-02-27 | cs.CL | ai-safety, cscl, preprint | E13 / R9 (90%) | 45 |
| LLM-Driven Multi-Turn Task-Oriented Dialogue Synthesis for Realistic Reasoning Kai Yang, Yu Zhu Published: 2026-02-27Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-02-27 | cs.CL | ai-safety, cscl, preprint | E7 / R7 (92%) | - |
| LSR: Linguistic Safety Robustness Benchmark for Low-Resource West African Languages Godwin Abuh Faruna Published: 2026-02-27Area: cs.CLCitations: 9 Tags: ai-safety, cscl, preprint | 2026-02-27 | cs.CL | ai-safety, cscl, preprint | E12 / R10 (93%) | 9 |
| Preference Packing: Efficient Preference Optimization for Large Language Models Jaekyung Cho Published: 2026-02-27Area: cs.CLCitations: 14 Tags: ai-safety, cscl, preprint | 2026-02-27 | cs.CL | ai-safety, cscl, preprint | E12 / R10 (93%) | 14 |
| Query Disambiguation via Answer-Free Context: Doubling Performance on Humanity's Last Exam Cynthia Matuszek, Michael Majurski Published: 2026-02-27Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-02-27 | cs.CL | ai-safety, cscl, preprint | E6 / R5 (95%) | - |
| Stepwise Penalization for Length-Efficient Chain-of-Thought Reasoning Besnik Fetahu, Bing Yin, Chia-Yuan Chang, Hejie Cui Published: 2026-02-27Area: cs.CLCitations: 15 Tags: ai-safety, cscl, preprint | 2026-02-27 | cs.CL | ai-safety, cscl, preprint | E10 / R9 (92%) | 15 |
| Task-Centric Acceleration of Small-Language Models Dor Tsur, Ran Levy, Sharon Adar Published: 2026-02-27Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-02-27 | cs.CL | ai-safety, cscl, preprint | E8 / R7 (91%) | - |
| Task Complexity Matters: An Empirical Study of Reasoning in LLMs for Sentiment Analysis Donghao Huang, Zhaoxia Wang Published: 2026-02-27Area: cs.CLCitations: 16 Tags: ai-safety, cscl, preprint | 2026-02-27 | cs.CL | ai-safety, cscl, preprint | E10 / R7 (93%) | 16 |
| Transformers are Stateless Differentiable Neural Computers Bo Tang, Weiwei Xie Published: 2026-02-27Area: cs.CLCitations: 1 Tags: ai-safety, cscl, preprint | 2026-02-27 | cs.CL | ai-safety, cscl, preprint | E7 / R5 (92%) | 1 |
| TRIZ-RAGNER: A Retrieval-Augmented Large Language Model for TRIZ-Aware Named Entity Recognition in Patent-Based Contradiction Mining Yue Zhao, Yuqing Wu, Zitong Xu Published: 2026-02-27Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-02-27 | cs.CL | ai-safety, cscl, preprint | E16 / R12 (91%) | - |
| Universal Conceptual Structure in Neural Translation: Probing NLLB-200's Multilingual Geometry Kyle Elliott Mathewson Published: 2026-02-27Area: cs.CLCitations: 45 Tags: ai-safety, cscl, preprint | 2026-02-27 | cs.CL | ai-safety, cscl, preprint | E8 / R7 (93%) | 45 |
| When Metrics Disagree: Automatic Similarity vs. LLM-as-a-Judge for Clinical Dialogue Evaluation Bian Sun, Orvill de la Torre, Zhenjian Wang, Zirui Wang Published: 2026-02-27Area: cs.CLCitations: 15 Tags: ai-safety, cscl, preprint, safety-evaluation | 2026-02-27 | cs.CL | ai-safety, cscl, preprint, safety-evaluation | E9 / R7 (94%) | 15 |
| A Comprehensive Evaluation of LLM Unlearning Robustness under Multi-Turn Interaction Ruihao Pan, Suhang Wang Published: 2026-02-28Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint, safety-evaluation | 2026-02-28 | cs.CL | ai-safety, cscl, preprint, safety-evaluation | E10 / R8 (93%) | - |
| CIRCUS: Circuit Consensus under Uncertainty via Stability Ensembles Swapnil Parekh Published: 2026-02-28Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-02-28 | cs.CL | ai-safety, cscl, preprint | E8 / R6 (89%) | - |
| Constitutional Black-Box Monitoring for Scheming in LLM Agents James Peters-Gill, Marius Hobbhahn, Rich Barton-Cooper, Simon Storf Published: 2026-02-28Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-02-28 | cs.CL | ai-safety, cscl, preprint | E7 / R7 (95%) | - |
| CURE: A Multimodal Benchmark for Clinical Understanding and Retrieval Evaluation Linjie Mu, Shaoting Zhang, Xiaofan Zhang, Xizhuo Zhang Published: 2026-02-28Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint, safety-evaluation | 2026-02-28 | cs.CL | ai-safety, cscl, preprint, safety-evaluation | E8 / R9 (91%) | - |
| Evaluating Progress in Graph Foundation Models: A Comprehensive Benchmark and New Insights Chang Zhou, Hong Cheng, Ruijuan Liang, Shenghua Ye Published: 2026-02-28Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-02-28 | cs.CL | ai-safety, cscl, preprint | E44 / R44 (91%) | - |
| From Flat to Structural: Enhancing Automated Short Answer Grading with GraphRAG Hang Li, Haoyu Han, Hui Liu, Joseph Krajcik Published: 2026-02-28Area: cs.CLCitations: 1 Tags: ai-safety, cscl, preprint | 2026-02-28 | cs.CL | ai-safety, cscl, preprint | E10 / R7 (90%) | 1 |
| Improving Automatic Summarization of Radiology Reports through Mid-Training of Large Language Models Cheng Peng, Jieting Li Lu, Mengxian Lyu, Mengyuan Zhang Published: 2026-02-28Area: cs.CLCitations: 25 Tags: ai-safety, cscl, preprint | 2026-02-28 | cs.CL | ai-safety, cscl, preprint | E10 / R8 (93%) | 25 |