Instant research discovery
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
| Coarse Indexing, Fine Evidence: Decoupling Temporal Granularity in Long-Video RAG Bin Zheng, Huihua Yang, Junhua Fang, Zhe Jin Published: 2026-08-24Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-08-24 | cs.CV | ai-safety, cscv, preprint | - | - |
| Cultural Moment Benchmark: Evaluating Video Cultural Reasoning and Grounding in Southeast Asia Burak Satar, Cheng Yu-Tong, Chong-Wah Ngo, Huy Hoang Tran Published: 2026-08-24Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-08-24 | cs.CV | ai-safety, cscv, preprint | - | - |
| DF-MoE: Generalizable Deepfake Detection via Multimodal Sparse Mixture-of-Experts A. Sophia Koepke, Florinel Alin Croitoru, Iuliana Georgescu, Radu Tudor Ionescu Published: 2026-08-24Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-08-24 | cs.CV | ai-safety, cscv, preprint | - | - |
| E2S-Pruner: Progressive Two-Stage Evidence Fusion for Visual Token Pruning in Vision-Language Models Daqian Shi, Hualong Yu, Qi Wang, Shang Gao Published: 2026-08-24Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-08-24 | cs.CV | ai-safety, cscv, preprint | - | - |
| EG-ARSA: An Expert-Grounded Open Model for Visual Road Safety Auditing in Low-Resource Settings Md Thamed Bin Zaman Chowdhury, Moazzem Hossain Published: 2026-08-24Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-08-24 | cs.CV | ai-safety, cscv, preprint | - | - |
| Hyperbolic Hierarchical Clustering for Visual Representation Learning Chunchao Guo, Guikun Chen, Jianan Wei, Wenguan Wang Published: 2026-08-24Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-08-24 | cs.CV | ai-safety, cscv, preprint | - | - |
| Infant Care Video Dataset for Classification of Interventions Using Transformers Igor Bogdanov, James Green Published: 2026-08-24Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-08-24 | cs.CV | ai-safety, cscv, preprint | E6 / R5 (92%) | - |
| LUCAID: Agentic Multimodal AI for Lung Cancer Precision Pathology Adam Narai, Alexander M枚llers, Andreas Kunft, Binru Yang Published: 2026-08-24Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-08-24 | cs.CV | ai-safety, cscv, preprint | E10 / R10 (93%) | - |
| Object-Uni: A Unified Model for Object-Centric Spatial Understanding and Controllable Generation DanDan Zheng, Jingdong Chen, Libin Wang, Mining Tan Published: 2026-08-24Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-08-24 | cs.CV | ai-safety, cscv, preprint | - | - |
| Predicting Radiologist Expertise from 3D Gaze Patterns During CT Interpretation Anna Anikina, Bulat Ibragimov, Claudia Mello-Thoms, Leila Khaertdinova Published: 2026-08-24Area: cs.CVCitations: 23 Tags: ai-safety, cscv, preprint | 2026-08-24 | cs.CV | ai-safety, cscv, preprint | E11 / R9 (93%) | 23 |
| Progressively Learning Heterogeneous Skills in a Unified Latent Space Linpu He, Ming Gong, Wei-Shi Zheng, Yue-Yi Zhang Published: 2026-08-24Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-08-24 | cs.CV | ai-safety, cscv, preprint | E7 / R5 (93%) | - |
| Restoring Without Forgetting: Continual Learning Across Image Degradations Alif Ashrafee, Bartosz Krawczyk Published: 2026-08-24Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-08-24 | cs.CV | ai-safety, cscv, preprint | E11 / R10 (92%) | - |
| Thinking Beyond Videos: Unifying Video Reasoning and Deep Research for Open-World Video Agents Bin Wen, Bohan Hou, Changyi Liu, Han Li Published: 2026-08-24Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-08-24 | cs.CV | ai-safety, cscv, preprint | E11 / R10 (91%) | - |
| Thinking Beyond Videos: Unifying Video Reasoning and Deep Research for Open-World Video Agents Bin Wen, Bohan Hou, Changyi Liu, Han Li Published: 2026-08-24Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-08-24 | cs.CV | ai-safety, cscv, preprint | - | - |
| Too much of a good thing -- when knowledge distillation promotes overfitting, and how to avoid it Christian Wagner, Daniel Molina, Irene Trigueros-Lorca, Isaac Triguero Published: 2026-08-24Area: cs.CVCitations: 15 Tags: ai-safety, cscv, preprint | 2026-08-24 | cs.CV | ai-safety, cscv, preprint | E9 / R6 (91%) | 15 |
| Towards Comprehensive Basketball Understanding Jiayuan Rao, Shangzhe Di, Weidi Xie, Yirong Hu Published: 2026-08-24Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-08-24 | cs.CV | ai-safety, cscv, preprint | - | - |
| WildHandBench: A Benchmark for Handwritten Text Understanding that Challenges MLLMs and Humans Changda Zhou, Cheng Cui, Jianwen Yang, Jianying Qu Published: 2026-08-24Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-08-24 | cs.CV | ai-safety, cscv, preprint | - | - |
| AI-based worker guidance in assembly and disassembly operations using multimodal ego/exo-centric data capture and structured task knowledge J枚rg Kr眉ger, Vivek Chavan Published: 2026-08-23Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-08-23 | cs.CV | ai-safety, cscv, preprint | - | - |
| A Lightweight Multimodal Vision-Language Framework for Early-Stage Anatomical Green Fruit Classification in Commercial Orchards Chen Chen, Manoj Karkee, Ranjan Sapkota, William Bu Published: 2026-08-23Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-08-23 | cs.CV | ai-safety, cscv, preprint | E10 / R10 (94%) | - |
| Fusing Perceptual Vision Experts with Multimodal Large Language Models for Explainable Plant Disease Diagnosis: From Benchmark Imagery to Real-World Robotic Field Validation Konstantinos I. Roumeliotis, Lirong Xiang, Manoj Karkee, Nikolaos D. Tselikas Published: 2026-08-23Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-08-23 | cs.CV | ai-safety, cscv, preprint | E12 / R14 (93%) | - |
| GAN-Diff : Coupling Pretrained WGAN-GP Features with Conditional Diffusion U-Nets Ahmed Faizul Haque Dhrubo, Ashadulla Hil Galib, Mohammad Abdul Qayum, Mohammad Ashrafuzzaman Khan Published: 2026-08-23Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-08-23 | cs.CV | ai-safety, cscv, preprint | - | - |
| Multimodal examination answer data with expert-designed Outcome-Based Education rubrics for criterion-level assessment A. K. M. Masudur Rahman, A K Z Rasel Rahman, Jahangir Alam SM, Md Khalid Syfullah Published: 2026-08-23Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-08-23 | cs.CV | ai-safety, cscv, preprint | - | - |
| OVIBench: Benchmarking Online Video Question Answering under Interruption Bowen Liu, Naiming Liu, Shuning Wang, Tie Zhang Published: 2026-08-23Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-08-23 | cs.CV | ai-safety, cscv, preprint | - | - |
| Pre-Decoding Acoustic Triage for Budgeted Vision-Language Captioning of Untrimmed Egocentric Video Changyi Li, Masoud Jalayer, Yu Xiao Published: 2026-08-23Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-08-23 | cs.CV | ai-safety, cscv, preprint | - | - |
| Training-Free VLM Personalization via Calibrated Residual Decoding Ang Li, Jiaao Yu, Pengran Wang, Xianming Hu Published: 2026-08-23Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-08-23 | cs.CV | ai-safety, cscv, preprint | - | - |
| TransHands: Repurposing Human Pose Encoders as Hand Pose Encoders Claudia Ferraris, Gabriella Olmo, Gianluca Amprimo, Milo Piccioli Published: 2026-08-23Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-08-23 | cs.CV | ai-safety, cscv, preprint | - | - |
| Vision-Language Models for Occupational Physical Exposure Assessment: Estimating External Hand Forces in Manual Material Handling Tasks from RGB Video Aanuoluwapo Ojelade, Maury A. Nussbaum, Mohammad Sadra Rajabi, Sunwook Kim Published: 2026-08-23Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-08-23 | cs.CV | ai-safety, cscv, preprint | - | - |
| ADMIL: Attention-Distilled Multiple Instance Learning for Selective Foundation Model Inference in Pathology Duncan Stothers, Ren-Chin Wu, William Lotter Published: 2026-08-22Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-08-22 | cs.CV | ai-safety, cscv, preprint | E8 / R8 (96%) | - |
| A Scalable Vector Graphics Latent Space Elia Frigieri, Leonardo Zini, Lorenzo Baraldi Published: 2026-08-22Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-08-22 | cs.CV | ai-safety, cscv, preprint | E10 / R10 (90%) | - |
| GuardPaint:SpeculativeSafetyDecodingforText-to-ImageGeneration Aman Chadha, Amitava Das, Aranbi Dutta, Arsh Abbas Naqvi Published: 2026-08-22Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-08-22 | cs.CV | ai-safety, cscv, preprint | E14 / R13 (93%) | - |