Instant research discovery
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
| Prompting is not enough: supervised baselines and leakage control for measuring shared decision-making with LLMs in pediatric encounters Angela Fagerlin, Angela Zhu, Bernardo Modenesi, Jody Lin Published: 2026-08-14Area: cs.CLCitations: 16 Tags: ai-safety, cscl, preprint | 2026-08-14 | cs.CL | ai-safety, cscl, preprint | E6 / R6 (94%) | 16 |
| SimpleOPD: Simple Tokenizer-Agnostic On-Policy Distillation for Long-Context Reasoning Fangchen Yu, Ganqu Cui, Haodi Lei, Haonan He Published: 2026-08-14Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-08-14 | cs.CL | ai-safety, cscl, preprint | E10 / R8 (93%) | - |
| Writing Style Similarity Reflects Academic Genealogy Cameron Manzo Published: 2026-08-14Area: cs.CLCitations: 12 Tags: ai-safety, cscl, preprint | 2026-08-14 | cs.CL | ai-safety, cscl, preprint | E8 / R5 (93%) | 12 |
| Amplified Does Not Mean Predictive: Reasoning Behaviors in Thinking Models Graham Neubig, Jean de Dieu Nyandwi, Leena Mathur, Robert Hawkins Published: 2026-08-13Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-08-13 | cs.CL | ai-safety, cscl, preprint | E22 / R10 (93%) | - |
| AQuA: Recursively Self-Improving Quantitative Trading Research Agents Jason Ge, Jiacheng Guo, Mengdi Wang, Suozhi Huang Published: 2026-08-13Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-08-13 | cs.CL | ai-safety, cscl, preprint | E11 / R8 (91%) | - |
| AQuA: Recursively Self-Improving Quantitative Trading Research Agents Jiacheng Guo, Jian Ge, Mengdi Wang, Suozhi Huang Published: 2026-08-13Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-08-13 | cs.CL | ai-safety, cscl, preprint | E12 / R10 (91%) | - |
| Are You Sure You're Sure? On the Impact of Instruction Tuning on Confidence and Lexical Diversity Irina Proskurina, Mayank Kumar, Oyindolapo O. Komolafe Published: 2026-08-13Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-08-13 | cs.CL | ai-safety, cscl, preprint | E12 / R16 (97%) | - |
| Better Decomposition, Free Aggregation: A Synthesizer-Folding Framework for Multilingual Multi-Hop Question Answering Jingbo Zhu, Shi Feng, Tong Xiao, Weidong Bao Published: 2026-08-13Area: cs.CLCitations: 30 Tags: ai-safety, cscl, preprint | 2026-08-13 | cs.CL | ai-safety, cscl, preprint | E14 / R11 (91%) | 30 |
| CLAIR-Fin: An Adversarial Multi-Agent Framework for Claim-Level Verification and Adaptive Debate in Cross-Modal Financial QA Fatema Tuj Johora Faria, Jubayer Al Mahmud, Md. Alam Hossain, M. F. Mridha Published: 2026-08-13Area: cs.CLCitations: - Tags: adversarial-robustness, ai-safety, cscl, preprint | 2026-08-13 | cs.CL | adversarial-robustness, ai-safety, cscl, preprint | E12 / R9 (93%) | - |
| CLAIR-Fin: An Adversarial Multi-Agent Framework for Claim-Level Verification and Adaptive Debate in Cross-Modal Financial QA Fatema Tuj Johora Faria, Jubayer Al Mahmud, Md. Alam Hossain, M. F. Mridha Published: 2026-08-13Area: cs.CLCitations: - Tags: adversarial-robustness, ai-safety, cscl, preprint | 2026-08-13 | cs.CL | adversarial-robustness, ai-safety, cscl, preprint | E10 / R9 (91%) | - |
| Class Imbalance and Batch Effects in LLM-Based Screening for Systematic Reviews Clayton Suguio Hida, Danilo Monteiro Ribeiro, Gilberto Sussumu Hida Published: 2026-08-13Area: cs.CLCitations: 26 Tags: ai-safety, cscl, preprint | 2026-08-13 | cs.CL | ai-safety, cscl, preprint | E10 / R7 (95%) | 26 |
| CRAFT: LLM-Based Iterative Refinement for Temporal Reasoning over Clinical Narratives Chengyang He, Lijing Wang, Marko Zivkovic, Ping Wang Published: 2026-08-13Area: cs.CLCitations: 37 Tags: ai-safety, cscl, preprint | 2026-08-13 | cs.CL | ai-safety, cscl, preprint | E9 / R6 (92%) | 37 |
| DFM Mimir v1: An Open HRM Delivering Frontier Performance at 1B Parameters Using Only Permissible Post-Training Data Gianluca Barmina, Jacob Nielsen, Kenneth Enevoldsen, Lukas Galke Poech Published: 2026-08-13Area: cs.CLCitations: 12 Tags: ai-safety, cscl, preprint | 2026-08-13 | cs.CL | ai-safety, cscl, preprint | E11 / R8 (92%) | 12 |
| ERSkill: Evolving for Skill-Guided Adaptive Memory Retrieval Guanrxu Zhu, Haolong Chen, Lei Xue, Liang Zhang Published: 2026-08-13Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-08-13 | cs.CL | ai-safety, cscl, preprint | E14 / R12 (91%) | - |
| Falsehood and Impossibility Are Different Directions in an AI's Representation of Language Yoon Pyo Lee Published: 2026-08-13Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-08-13 | cs.CL | ai-safety, cscl, preprint | E8 / R8 (91%) | - |
| Falsehood and Impossibility Are Different Directions in an AI's Representation of Language Yoon Pyo Lee Published: 2026-08-13Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-08-13 | cs.CL | ai-safety, cscl, preprint | E8 / R6 (90%) | - |
| From Atomic Evidence to Logical Composition: Structured Compositional Reasoning over Compound Answer Options Maria Leonor Pacheco, Obed Junias Published: 2026-08-13Area: cs.CLCitations: 45 Tags: ai-safety, cscl, preprint | 2026-08-13 | cs.CL | ai-safety, cscl, preprint | E11 / R9 (95%) | 45 |
| GEM: A Generative Embedding Model Bridging Reasoning and Retrieval Craig Macdonald, Zhili Shen Published: 2026-08-13Area: cs.CLCitations: 45 Tags: ai-safety, cscl, preprint | 2026-08-13 | cs.CL | ai-safety, cscl, preprint | E15 / R13 (91%) | 45 |
| GEM: A Generative Embedding Model Bridging Reasoning and Retrieval Craig Macdonald, Zhili Shen Published: 2026-08-13Area: cs.CLCitations: 45 Tags: ai-safety, cscl, preprint | 2026-08-13 | cs.CL | ai-safety, cscl, preprint | E8 / R7 (92%) | 45 |
| How Do VLMs Behave When Blind or Misled? Behavioral Evaluation of VLMs on Scientific Figures Ananya Mukherjee, Christian Greisinger, Owusu-Banahene Osei, Paul Osemudiame Oamen Published: 2026-08-13Area: cs.CLCitations: 15 Tags: ai-safety, cscl, preprint, safety-evaluation | 2026-08-13 | cs.CL | ai-safety, cscl, preprint, safety-evaluation | E10 / R8 (95%) | 15 |
| It's How You Ask: Gender-Associated Linguistic Bias in LLMs Anjalie Field, Katherine Van Koevering Published: 2026-08-13Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-08-13 | cs.CL | ai-safety, cscl, preprint | E15 / R10 (94%) | - |
| LigBench: A Unified and Human-Aligned Benchmark for LLM-based Research Idea Generation Chenrun Wang, Kai Yu, Lu Chen, Mingxuan Zhu Published: 2026-08-13Area: cs.CLCitations: 15 Tags: ai-safety, cscl, preprint | 2026-08-13 | cs.CL | ai-safety, cscl, preprint | E9 / R8 (90%) | 15 |
| LittleLearner: Language Models Under Pedagogically Controlled Knowledge Exposure Fanfei Li, Jana Zeller, Manuel Prada-Corral, Prasanna Mayilvahanan Published: 2026-08-13Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-08-13 | cs.CL | ai-safety, cscl, preprint | E9 / R8 (94%) | - |
| Measuring Fairness in Large Audio Language Models via Semantic-Aware Bias Estimation Zhe Liu Published: 2026-08-13Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-08-13 | cs.CL | ai-safety, cscl, preprint | E10 / R8 (93%) | - |
| Mixture of Training: Recombining Small-Scale Scaffolded Pretraining Runs into a Larger Language Model Keith Rush, Lucio Dery, Mohammed Sabry, Sean Augenstein Published: 2026-08-13Area: cs.CLCitations: 15 Tags: ai-safety, cscl, preprint | 2026-08-13 | cs.CL | ai-safety, cscl, preprint | E8 / R6 (90%) | 15 |
| PatientAct: Theory-Grounded Mental Health Client Simulation Guanqun Bi, Jialu Zhao, Minlie Huang, Sahand Sabour Published: 2026-08-13Area: cs.CLCitations: 15 Tags: ai-safety, cscl, preprint | 2026-08-13 | cs.CL | ai-safety, cscl, preprint | E8 / R6 (90%) | 15 |
| Self-Referential Induction Increases Response Instability Relative to Unresolvable and Verifiable Questions in Large Language Models Paras Balani, Subhrakanta Panda Published: 2026-08-13Area: cs.CLCitations: 12 Tags: ai-safety, cscl, preprint | 2026-08-13 | cs.CL | ai-safety, cscl, preprint | E8 / R6 (90%) | 12 |
| TeachMateGPT: A Multi-Agent Knowledge-Grounded Framework for Pedagogical Assessment Generation from Science Curriculum Materials Fatema Tuj Johora Faria, Jubayer Al Mahmud, M. F. Mridha, Mukaffi Bin Moin Published: 2026-08-13Area: cs.CLCitations: - Tags: ai-safety, cscl, preprint | 2026-08-13 | cs.CL | ai-safety, cscl, preprint | E8 / R7 (92%) | - |
| Toward a Gricean Retreat: Probing LLMs for Knowledge Boundaries and Referent Specificity Dananjay Srinivas, Maria Pacheco, Saksham Khatwani Published: 2026-08-13Area: cs.CLCitations: 12 Tags: ai-safety, cscl, preprint | 2026-08-13 | cs.CL | ai-safety, cscl, preprint | E10 / R7 (95%) | 12 |
| When Do LLMs Actually Help? Evaluating LLMs as Data Quality Annotators Praphulla Lal Shrestha Published: 2026-08-13Area: cs.CLCitations: 13 Tags: ai-safety, cscl, preprint | 2026-08-13 | cs.CL | ai-safety, cscl, preprint | E8 / R6 (93%) | 13 |