Instant research discovery
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
| A Neuromorphic Trigger for Efficient Audio Event Detection Benjamin Hatton, Luca Peres, Oliver Rhodes Published: 2026-06-16Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-06-16 | cs.SD | ai-safety, cssd, preprint | E9 / R5 (98%) | - |
| Descriptor: Certus Caliber Classification Gunshot Dataset (C3GD) Ryan Quinn, Sinclair Gurny Published: 2026-06-16Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-06-16 | cs.SD | ai-safety, cssd, preprint | E7 / R4 (97%) | - |
| L-Proto: Language-Aware Episodic Prototypical Training for Multilingual Speaker Verification Deok-Hyeon Cho, Hyung-Seok Oh, Seong-Whan Lee, Seung-Bin Kim Published: 2026-06-16Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-06-16 | cs.SD | ai-safety, cssd, preprint | E6 / R4 (97%) | - |
| MagpieTTS-LF: Inference-Time Long-Form Speech Generation Without Training on Long-Form data Jason Li, Paarth Neekhara, Roy Fejgin, Ryan Langman Published: 2026-06-16Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-06-16 | cs.SD | ai-safety, cssd, preprint | E9 / R5 (97%) | - |
| ArtBoost: Synthetic Articulatory Data Augmentation for Acoustic-to-Articulatory Inversion Byungchan Hwang, Hak Gu Kim, Hyung Kyu Kim Published: 2026-06-15Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-06-15 | cs.SD | ai-safety, cssd, preprint | E9 / R5 (98%) | - |
| ArtNet: A JEPA-Like Articulatory Predictive Framework for Robust Zero-Shot Phoneme Recognition Fuliang Weng, Shu Shang, Yaqian Zhou, Zeqian Hu Published: 2026-06-15Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-06-15 | cs.SD | ai-safety, cssd, preprint | E9 / R5 (99%) | - |
| Dual-Granularity Orthogonal Disentanglement for Generalizable Audio Deepfake Detection Chunhong Yuan, Hugen Lv, Xiangyu Li, Zhuodong Liu Published: 2026-06-15Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-06-15 | cs.SD | ai-safety, cssd, preprint | E15 / R7 (97%) | - |
| MuVAP: Multimodal Multiparty Voice Activity Projection for Turn-taking Prediction in the Wild Gabriel Skantze, Haotian Qi Published: 2026-06-15Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-06-15 | cs.SD | ai-safety, cssd, preprint | E9 / R4 (98%) | - |
| Probing Low Frame Rate Degradation in Neural Audio Codecs Alex Gichamba, Moise Busogi Published: 2026-06-15Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-06-15 | cs.SD | ai-safety, cssd, preprint | E12 / R4 (97%) | - |
| TuneJury: An Open Metric for Improving Music Generation Preference Alignment Chris Donahue, Haiwen Xia, Junghyun Koo, Junwon Lee Published: 2026-06-15Area: cs.SDCitations: - Tags: ai-safety, alignment-training, cssd, preprint | 2026-06-15 | cs.SD | ai-safety, alignment-training, cssd, preprint | E9 / R5 (99%) | - |
| Vibrato Expression Control for Singing Voice Conversion with Improving Independent Control Dong-Min Byun, Joon-Seung Choi, Seong-Whan Lee Published: 2026-06-15Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-06-15 | cs.SD | ai-safety, cssd, preprint | E8 / R5 (98%) | - |
| AP-GRPO: Anchor-Gated Phonetic Alignment with Policy Optimization for Pathological Speech Reconstruction Amir M. Rahmani, Henry Peng Zou, Hoang H Nguyen, Honghui Xu Published: 2026-06-14Area: cs.SDCitations: - Tags: ai-safety, alignment-training, cssd, preprint | 2026-06-14 | cs.SD | ai-safety, alignment-training, cssd, preprint | E10 / R4 (98%) | - |
| Improving Code-Switching ASR with Code-Mixing Guided Synthetic Speech Eng Siong Chng, Haoyang Li, Hardik B. Sailor, Hexin Liu Published: 2026-06-14Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-06-14 | cs.SD | ai-safety, cssd, preprint | E8 / R4 (97%) | - |
| NVMOS: Non-Verbal Vocalization Quality Assessment in Speech Jialong Mai, Jinxin Ji, Wencui Liu, Xiangmin Xu Published: 2026-06-14Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-06-14 | cs.SD | ai-safety, cssd, preprint | E7 / R4 (96%) | - |
| All That Glitters Is Not Audio: Rethinking Text Priors and Audio Reliance in Audio-Language Evaluation Chen-An Li, Chih-Kai Yang, Hung-yi Lee, Ke-Han Lu Published: 2026-04-27Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint, safety-evaluation | 2026-04-27 | cs.SD | ai-safety, cssd, preprint, safety-evaluation | E10 / R4 (100%) | - |
| RAS: a Reliability Oriented Metric for Automatic Speech Recognition Bohan Li, Hankun Wang, Jing Peng, Kai Yu Published: 2026-04-27Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-04-27 | cs.SD | ai-safety, cssd, preprint | E9 / R5 (97%) | - |
| Speech Enhancement Based on Drifting Models Bastiaan Kleijn, Diego Caviedes-Nozal, Liang Xu, Longfei Felix Yan Published: 2026-04-27Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-04-27 | cs.SD | ai-safety, cssd, preprint | E10 / R5 (99%) | - |
| ATIR: Towards Audio-Text Interleaved Contextual Retrieval Chenghao Zhang, Tong Zhao, Yutao Zhu, Zhicheng Dou Published: 2026-04-22Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-04-22 | cs.SD | ai-safety, cssd, preprint | E7 / R5 (97%) | - |
| Enhancing Speaker Verification with Whispered Speech via Post-Processing Magdalena Go艂臋biowska, Piotr Syga Published: 2026-04-22Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-04-22 | cs.SD | ai-safety, cssd, preprint | E9 / R4 (97%) | - |
| ONOTE: Benchmarking Omnimodal Notation Processing for Expert-level Music Intelligence Fanhong Meng, Haoran Luo, Luu Anh Tuan, Menghe Ma Published: 2026-04-22Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-04-22 | cs.SD | ai-safety, cssd, preprint | E11 / R4 (97%) | - |
| BEAT: Tokenizing and Generating Symbolic Music by Uniform Temporal Steps Haoyu Gu, Jingwei Zhao, Lekai Qian, Ziyu Wang Published: 2026-04-21Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-04-21 | cs.SD | ai-safety, cssd, preprint | E7 / R4 (96%) | - |
| Environmental Sound Deepfake Detection Using Deep-Learning Framework Alexander Schindler, Dat Tran, David Fischinger, Khoi Vu Published: 2026-04-21Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-04-21 | cs.SD | ai-safety, cssd, preprint | E8 / R4 (95%) | - |
| HalluAudio: A Comprehensive Benchmark for Hallucination Detection in Large Audio-Language Models Daipeng Zhang, Feiyu Zhao, Jianguo Wei, Wenhuan Lu Published: 2026-04-21Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-04-21 | cs.SD | ai-safety, cssd, preprint | E9 / R6 (97%) | - |
| Tadabur: A Large-Scale Quran Audio Dataset Faisal Alherran Published: 2026-04-21Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-04-21 | cs.SD | ai-safety, cssd, preprint | E8 / R4 (98%) | - |
| Towards Streaming Target Speaker Extraction via Chunk-wise Interleaved Splicing of Autoregressive Language Model Guiping Zhong, Haiyun Li, Hui Lu, Huimeng Wang Published: 2026-04-21Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-04-21 | cs.SD | ai-safety, cssd, preprint | E10 / R5 (97%) | - |
| A novel LSTM music generator based on the fractional time-frequency feature extraction Chen Chaofan, Chen Wei, Deng Xinyi, Li Xiulai Published: 2026-04-20Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-04-20 | cs.SD | ai-safety, cssd, preprint | E9 / R6 (96%) | - |
| Latent Fourier Transform Cheng-Zhi Anna Huang, Mason Wang Published: 2026-04-20Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-04-20 | cs.SD | ai-safety, cssd, preprint | E7 / R4 (96%) | - |
| Video-Robin: Autoregressive Diffusion Planning for Intent-Grounded Video-to-Music Generation Aryan Vijay Bhosale, Dinesh Manocha, Gouthaman KV, Lie Lu Published: 2026-04-19Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-04-19 | cs.SD | ai-safety, cssd, preprint | E8 / R5 (98%) | - |
| Video-Robin: Autoregressive Diffusion Planning for Intent-Grounded Video-to-Music Generation Aryan Vijay Bhosale, Dinesh Manocha, Gouthaman KV, Lie Lu Published: 2026-04-19Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-04-19 | cs.SD | ai-safety, cssd, preprint | E10 / R7 (99%) | - |
| AST: Adaptive, Seamless, and Training-Free Precise Speech Editing Jianwei Yin, Jinshan Zhang, Jintao Chen, Meng Xi Published: 2026-04-17Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-04-17 | cs.SD | ai-safety, cssd, preprint | E8 / R8 (95%) | - |