Instant research discovery
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
| Fast-WAM: Do World Action Models Need Test-time Future Imagination? Hang Zhao, Tianyuan Yuan, Yicheng Liu, Zibin Dong Published: 2026-03-17Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-03-17 | cs.CV | ai-safety, cscv, preprint | E5 / R4 (97%) | - |
| FSMC-Pose: Frequency and Spatial Fusion with Multiscale Self-calibration for Cattle Mounting Pose Estimation Fangjing Li, Haiyang Liu, Ming Jin, Ronghua Gao Published: 2026-03-17Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-03-17 | cs.CV | ai-safety, cscv, preprint | E7 / R4 (98%) | - |
| GATS: Gaussian Aware Temporal Scaling Transformer for Invariant 4D Spatio-Temporal Point Cloud Representation Jiayi Tian, Jiaze Wang Published: 2026-03-17Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-03-17 | cs.CV | ai-safety, cscv, preprint | E6 / R5 (98%) | - |
| Generalist Multimodal LLMs Gain Biometric Expertise via Human Salience Adam Czajka, Byron Dowling, Christopher Sweet, Jacob Piland Published: 2026-03-17Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-03-17 | cs.CV | ai-safety, cscv, preprint | E5 / R3 (97%) | - |
| Grounding the Score: Explicit Visual Premise Verification for Reliable Vision-Language Process Reward Models Dai Guan, Erchao Zhao, Guanjun Jiang, Junxin Wang Published: 2026-03-17Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-03-17 | cs.CV | ai-safety, cscv, preprint | E5 / R3 (95%) | - |
| Hidden Clones: Exposing and Fixing Family Bias in Vision-Language Model Ensembles Zacharie Bugaud Published: 2026-03-17Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-03-17 | cs.CV | ai-safety, cscv, preprint | E5 / R3 (97%) | - |
| Homogeneous and Heterogeneous Consistency progressive Re-ranking for Visible-Infrared Person Re-identification Yiming Wang Published: 2026-03-17Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-03-17 | cs.CV | ai-safety, cscv, preprint | E4 / R3 (95%) | - |
| HopChain: Multi-Hop Data Synthesis for Generalizable Vision-Language Reasoning An Yang, Binghai Wang, Bowen Yu, Chang Gao Published: 2026-03-17Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-03-17 | cs.CV | ai-safety, cscv, preprint | E4 / R3 (95%) | - |
| Interact3D: Compositional 3D Generation of Interactive Objects Hui Shan, Keyang Luo, Ming Li, Sizhe Zheng Published: 2026-03-17Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-03-17 | cs.CV | ai-safety, cscv, preprint | E5 / R3 (94%) | - |
| IOSVLM: A 3D Vision-Language Model for Unified Dental Diagnosis from Intraoral Scans Chenyi Zhou, Huimin Xiong, Tianxiang Hu, Yang Feng Published: 2026-03-17Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-03-17 | cs.CV | ai-safety, cscv, preprint | E5 / R3 (96%) | - |
| Kestrel: Grounding Self-Refinement for LVLM Hallucination Mitigation Cihang Xie, Haoqin Tu, Hardy Chen, Huaxiu Yao Published: 2026-03-17Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-03-17 | cs.CV | ai-safety, cscv, preprint | E6 / R4 (95%) | - |
| LICA: Layered Image Composition Annotations for Graphic Design Research Elad Hirsch, Mohit Garg, Purvanshi Mehta, Shubham Yadav Published: 2026-03-17Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-03-17 | cs.CV | ai-safety, cscv, preprint | E5 / R3 (95%) | - |
| MessyKitchens: Contact-rich object-level 3D scene reconstruction Fabio Pizzati, Ivan Laptev, Junaid Ahmed Ansari, Ran Ding Published: 2026-03-17Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-03-17 | cs.CV | ai-safety, cscv, preprint | E5 / R3 (95%) | - |
| MLLM-based Textual Explanations for Face Comparison Anil K Jain, Redwan Sony, Ross Arun Published: 2026-03-17Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-03-17 | cs.CV | ai-safety, cscv, preprint | E5 / R3 (96%) | - |
| MSRAMIE: Multimodal Structured Reasoning Agent for Multi-instruction Image Editing Ken Chen, Sachith Seneviratne, Saman Halgamuge, Xiangwei Wang Published: 2026-03-17Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-03-17 | cs.CV | ai-safety, cscv, preprint | E6 / R4 (96%) | - |
| Parallel In-context Learning for Large Vision Language Models Daiki Chijiwa, Shin'ya Yamaguchi, Taku Hasegawa, Tamao Sakao Published: 2026-03-17Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-03-17 | cs.CV | ai-safety, cscv, preprint | E5 / R3 (96%) | - |
| PathGLS: Evaluating Pathology Vision-Language Models without Ground Truth through Multi-Dimensional Consistency Fei Su, Minbing Chen, Zhu Meng Published: 2026-03-17Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-03-17 | cs.CV | ai-safety, cscv, preprint | E5 / R3 (95%) | - |
| PhysQuantAgent: An Inference Pipeline of Mass Estimation for Vision-Language Models Hisayuki Yokomizo, Nakamasa Inoue, Shuhei Kurita, Taiki Miyanishi Published: 2026-03-17Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-03-17 | cs.CV | ai-safety, cscv, preprint | E5 / R4 (95%) | - |
| RASLF: Representation-Aware State Space Model for Light Field Super-Resolution Kai Jin, Kuan Song, Min Xu, Wenlong Chen Published: 2026-03-17Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-03-17 | cs.CV | ai-safety, cscv, preprint | E5 / R3 (95%) | - |
| REFORGE: Multi-modal Attacks Reveal Vulnerable Concept Unlearning in Image Generation Models Fanxiao Li, Haoran Li, Li Tang, Renyang Liu Published: 2026-03-17Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-03-17 | cs.CV | ai-safety, cscv, preprint | E6 / R3 (93%) | - |
| Retrieving Counterfactuals Improves Visual In-Context Learning Aidong Zhang, Guangzhi Xiong, Sanchit Sinha, Zhenghao He Published: 2026-03-17Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-03-17 | cs.CV | ai-safety, cscv, preprint | E5 / R3 (94%) | - |
| SA-CycleGAN-2.5D: Self-Attention CycleGAN with Tri-Planar Context for Multi-Site MRI Harmonization Chunwei Liu, Ishrith Gowda Published: 2026-03-17Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-03-17 | cs.CV | ai-safety, cscv, preprint | E6 / R3 (94%) | - |
| SF-Mamba: Rethinking State Space Model for Vision Masakazu Yoshimura, Takeshi Ohashi, Teruaki Hayashi, Wei-Yao Wang Published: 2026-03-17Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-03-17 | cs.CV | ai-safety, cscv, preprint | E5 / R3 (95%) | - |
| SOMA: Unifying Parametric Human Body Models Edy Lim, Ehsan Hassani, Hyejin Moon, Jan Kautz Published: 2026-03-17Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-03-17 | cs.CV | ai-safety, cscv, preprint | E6 / R4 (98%) | - |
| SparkVSR: Interactive Video Super-Resolution via Sparse Keyframe Propagation Akshay Gadde, Balu Adsumilli, Jiongze Yu, Pooja Verlani Published: 2026-03-17Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-03-17 | cs.CV | ai-safety, cscv, preprint | E5 / R3 (95%) | - |
| Understanding Pruning Regimes in Vision-Language Models Through Domain-Aware Layer Selection Kamal Ginotra, Nima Safaei, Saeed Khaki Published: 2026-03-17Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-03-17 | cs.CV | ai-safety, cscv, preprint | E35 / R30 (96%) | - |
| Unlearning for One-Step Generative Models via Unbalanced Optimal Transport Hyundo Choi, Jaewoong Choi, Jinseong Park, Junhyeong An Published: 2026-03-17Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-03-17 | cs.CV | ai-safety, cscv, preprint | E7 / R3 (95%) | - |
| V-Co: A Closer Look at Visual Representation Alignment via Co-Denoising Chu Wang, Han Lin, Jaemin Cho, Mohit Bansal Published: 2026-03-17Area: cs.CVCitations: - Tags: ai-safety, alignment-training, cscv, preprint | 2026-03-17 | cs.CV | ai-safety, alignment-training, cscv, preprint | E5 / R3 (97%) | - |
| VisBrowse-Bench: Benchmarking Visual-Native Search for Multimodal Browsing Agents Bo Zhang, Changtao Miao, Feier Wu, Jinbo Su Published: 2026-03-17Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-03-17 | cs.CV | ai-safety, cscv, preprint | E4 / R3 (94%) | - |
| Visual Prompt Discovery via Semantic Exploration Jaechang Kim, Jungseul Ok, Kuang-Da Wang, Shingo Takamatsu Published: 2026-03-17Area: cs.CVCitations: - Tags: ai-safety, cscv, preprint | 2026-03-17 | cs.CV | ai-safety, cscv, preprint | E5 / R3 (94%) | - |