Instant research discovery
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Showing 1-30 of 1000+ papers (page 1 of 34)路 1098 ms
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
| A Content-Based Framework for Cybersecurity Refusal Decisions in Large Language Models Edan Maor, Gil Gekker, Meirav Segal, Noa Linder Published: 2026-02-17Area: cs.CLCitations: 15 Tags: ai-safety, cscl, preprint | 2026-02-17 | cs.CL | ai-safety, cscl, preprint | E9 / R9 (95%) | 15 |
| Avey-B Devang Acharya, Mohammad Hammoud Published: 2026-02-17Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-02-17 | cs.CL | ai-safety, cscl, preprint | E15 / R14 (90%) | - |
| Beyond Binary Classification: Detecting Fine-Grained Sexism in Social Media Videos Danae S谩nchez Villegas, Desmond Elliott, Laura De Grazia, Mariona Taul茅 Published: 2026-02-17Area: cs.CLCitations: 15 Tags: ai-safety, cscl, preprint | 2026-02-17 | cs.CL | ai-safety, cscl, preprint | E11 / R10 (93%) | 15 |
| Beyond Static Pipelines: Learning Dynamic Workflows for Text-to-SQL Peiyu Liu, Runyu Chen, Wei Xu, Yihan Wang Published: 2026-02-17Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-02-17 | cs.CL | ai-safety, cscl, preprint | E15 / R10 (91%) | - |
| ChartEditBench: Evaluating Grounded Multi-Turn Chart Editing in Multimodal Language Models Atharva Naik, Carolyn Ros茅, Lawanya Baghel, Manav Nitin Kapadnis Published: 2026-02-17Area: cs.CLCitations: 15 Tags: ai-safety, cscl, preprint | 2026-02-17 | cs.CL | ai-safety, cscl, preprint | E12 / R10 (89%) | 15 |
| DocSplit: A Comprehensive Benchmark Dataset and Evaluation Approach for Document Packet Recognition and Splitting Bob Strahan, Boyi Xie, Diego A. Socolinsky, Md Mofijul Islam Published: 2026-02-17Area: cs.CLCitations: 15 Tags: ai-safety, cscl, preprint, safety-evaluation | 2026-02-17 | cs.CL | ai-safety, cscl, preprint, safety-evaluation | E11 / R9 (90%) | 15 |
| Far Out: Evaluating Language Models on Slang in Australian and Indian English Aditya Joshi, Deniz Kaya Dilsiz, Dipankar Srirag Published: 2026-02-17Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-02-17 | cs.CL | ai-safety, cscl, preprint | E17 / R13 (91%) | - |
| FERRET: Framework for Expansion Reliant Red Teaming Joanna Bitton, Maya Pavlova, Ninareh Mehrabi, Vitor Albiero Published: 2026-02-17Area: cs.CLCitations: 15 Tags: ai-safety, cscl, preprint, red-teaming | 2026-02-17 | cs.CL | ai-safety, cscl, preprint, red-teaming | E12 / R11 (93%) | 15 |
| GATech at AbjadMed: Bidirectional Encoders vs. Causal Decoders: Insights from 82-Class Arabic Medical Classification Ahmed Khaled Khamis Published: 2026-02-17Area: cs.CLCitations: 15 Tags: ai-safety, cscl, preprint | 2026-02-17 | cs.CL | ai-safety, cscl, preprint | E8 / R7 (92%) | 15 |
| Language Statistics and False Belief Reasoning: Evidence from 41 Open-Weight LMs Cameron Jones, James A. Michaelov, Pamela D. Rivi猫re, Samuel Taylor Published: 2026-02-17Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-02-17 | cs.CL | ai-safety, cscl, preprint | E10 / R7 (91%) | - |
| NeuroSymActive: Differentiable Neural-Symbolic Reasoning with Active Exploration for Knowledge Graph Question Answering Jiekai Wu, Rong Fu, Shuaishuai Cao, Simon Fong Published: 2026-02-17Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-02-17 | cs.CL | ai-safety, cscl, preprint | E8 / R6 (92%) | - |
| Orchestration-Free Customer Service Automation: A Privacy-Preserving and Flowchart-Guided Framework Chen Jason Zhang, Di Jiang, Hanlin Gu, Li Qing Published: 2026-02-17Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-02-17 | cs.CL | ai-safety, cscl, preprint | E7 / R6 (91%) | - |
| Revisiting Northrop Frye's Four Myths Theory with Large Language Models Antonio L. Furtado, Edirlei Soares de Lima, Marco A. Casanova Published: 2026-02-17Area: cs.CLCitations: 1 Tags: ai-safety, cscl, preprint | 2026-02-17 | cs.CL | ai-safety, cscl, preprint | E22 / R11 (92%) | 1 |
| SENS-ASR: Semantic Embedding injection in Neural-transducer for Streaming Automatic Speech Recognition Anthony Larcher, Elys Allesiardo, Valentin Vielzeuf, Youness Dkhissi Published: 2026-02-17Area: cs.CLCitations: 45 Tags: ai-safety, cscl, preprint | 2026-02-17 | cs.CL | ai-safety, cscl, preprint | E12 / R9 (92%) | 45 |
| STAPO: Stabilizing Reinforcement Learning for LLMs by Silencing Rare Spurious Tokens Bo Zhang, Guojian Zhan, Jiang Wu, Jingliang Duan Published: 2026-02-17Area: cs.CLCitations: 30 Tags: ai-safety, cscl, preprint | 2026-02-17 | cs.CL | ai-safety, cscl, preprint | E9 / R6 (93%) | 30 |
| Updating Parametric Knowledge with Context Distillation Retains Post-Training Capabilities Mustafa Omer Gul, Shankar Padmanabhan, Tanya Goyal Published: 2026-02-17Area: cs.CLCitations: 15 Tags: ai-safety, cscl, preprint | 2026-02-17 | cs.CL | ai-safety, cscl, preprint | E10 / R10 (90%) | 15 |
| What Is Missing: Interpretable Ratings for Large Language Model Outputs Nicholas Stranges, Yimin Yang Published: 2026-02-17Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-02-17 | cs.CL | ai-safety, cscl, preprint | E7 / R5 (92%) | - |
| Aligning Language Models from User Interactions Andreas Krause, Barna P谩sztor, Giorgia Ramponi, Idan Shenfeld Published: 2026-02-18Area: cs.CLCitations: 15 Tags: ai-safety, cscl, preprint | 2026-02-18 | cs.CL | ai-safety, cscl, preprint | E9 / R10 (91%) | 15 |
| Align Once, Benefit Multilingually: Enforcing Multilingual Consistency for LLM Safety Alignment Juntao Dai, Xiaohao Liu, Yaodong Yang, Yuyan Bu Published: 2026-02-18Area: cs.CLCitations: 1 Tags: ai-safety, alignment-training, cscl, preprint | 2026-02-18 | cs.CL | ai-safety, alignment-training, cscl, preprint | E10 / R7 (92%) | 1 |
| Are LLMs Ready to Replace Bangla Annotators? Md. Najib Hasan, Souvika Sarkar, Touseef Hasan Published: 2026-02-18Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-02-18 | cs.CL | ai-safety, cscl, preprint | - | - |
| Balancing Faithfulness and Performance in Reasoning via Multi-Listener Soft Execution Ali Payani, Elias Stengel-Eskin, Hyunji Lee, Mohit Bansal Published: 2026-02-18Area: cs.CLCitations: 45 Tags: ai-safety, cscl, preprint | 2026-02-18 | cs.CL | ai-safety, cscl, preprint | E15 / R12 (93%) | 45 |
| Beyond Learning: A Training-Free Alternative to Model Adaptation Hwangnam Kim, Kyeonghyun Yoo, Namkyung Yoon, Sanghong Kim Published: 2026-02-18Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-02-18 | cs.CL | ai-safety, cscl, preprint | E9 / R10 (93%) | - |
| Calibrate-Then-Act: Cost-Aware Exploration in LLM Agents Greg Durrett, Nicholas Tomlin, Wenxuan Ding Published: 2026-02-18Area: cs.CLCitations: 68 Tags: ai-safety, cscl, preprint | 2026-02-18 | cs.CL | ai-safety, cscl, preprint | E10 / R8 (89%) | 68 |
| Diagnosing Retrieval Bias Under Multiple In-Context Knowledge Updates in Large Language Models Boyu Qiao, Kun Li, Sean Guo, Songlin Hu Published: 2026-02-18Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-02-18 | cs.CL | ai-safety, cscl, preprint | E15 / R12 (90%) | - |
| Eigenmood Space: Uncertainty-Aware Spectral Graph Analysis of Psychological Patterns in Classical Persian Poetry Kourosh Shahnazari, Mohammadali Keshtparvar, Seyed Moein Ayyoubzadeh Published: 2026-02-18Area: cs.CLCitations: 35 Tags: ai-safety, cscl, preprint | 2026-02-18 | cs.CL | ai-safety, cscl, preprint | E12 / R7 (89%) | 35 |
| Evaluating Monolingual and Multilingual Large Language Models for Greek Question Answering: The DemosQA Benchmark Charalampos Mastrokostas, Nikolaos Giarelis, Nikos Karacapilidis Published: 2026-02-18Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-02-18 | cs.CL | ai-safety, cscl, preprint | E11 / R10 (92%) | - |
| Explainable AI: Context-Aware Layer-Wise Integrated Gradients for Explaining Transformer Models Jugal Kalita, Melkamu Abay Mersha Published: 2026-02-18Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-02-18 | cs.CL | ai-safety, cscl, preprint | E10 / R10 (93%) | - |
| Flow Map Language Models: One-step Language Modeling via Continuous Denoising Aditi Raghunathan, Chanhyuk Lee, Jaehoon Yoo, Jerry Huang Published: 2026-02-18Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-02-18 | cs.CL | ai-safety, cscl, preprint | E7 / R6 (93%) | - |
| From Growing to Looping: A Unified View of Iterative Computation in LLMs Emmanouil Angelis, Ferdinand Kapl, Johannes von Oswald, Kaitlin Maile Published: 2026-02-18Area: cs.CLCitations: 46 Tags: ai-safety, cscl, preprint | 2026-02-18 | cs.CL | ai-safety, cscl, preprint | E8 / R7 (93%) | 46 |
| IndicEval: A Bilingual Indian Educational Evaluation Framework for Large Language Models Abhinaw Jagtap, Gaurav Azad, Nachiket Tapas, Saurabh Bharti Published: 2026-02-18Area: cs.CLCitations: 9 Tags: ai-safety, cscl, preprint, safety-evaluation | 2026-02-18 | cs.CL | ai-safety, cscl, preprint, safety-evaluation | E19 / R20 (94%) | 9 |