Instant research discovery
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
Search and browse arXiv CS/AI/ML papers, enriched with AI-generated insights.
Generate novel research ideas grounded in real arXiv papers with Brainstorm.
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
| Paper | Published | Area | Tags | Intel | Citations |
|---|---|---|---|---|---|
| Controllable Singing Style Conversion with Boundary-Aware Information Bottleneck Jihua Zhu, Wenyu Wang, Xin Gao, Yiquan Zhou Published: 2026-04-07Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-04-07 | cs.SD | ai-safety, cssd, preprint | E5 / R4 (95%) | - |
| Generating Synthetic Doctor-Patient Conversations for Long-form Audio Summarization Adam Rothschild, Ahmed Hassoon, Andrew Perrault, David Gr眉nert Published: 2026-04-07Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-04-07 | cs.SD | ai-safety, cssd, preprint | E5 / R3 (95%) | - |
| Towards Real-Time Human-AI Musical Co-Performance: Accompaniment Generation with Latent Diffusion Models and MAX/MSP Shlomo Dubnov, Tornike Karchkhadze Published: 2026-04-08Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-04-08 | cs.SD | ai-safety, cssd, preprint | E5 / R3 (98%) | - |
| AT-ADD: All-Type Audio Deepfake Detection Challenge Evaluation Plan Guangtao Zhai, Haonan Cheng, Hengyan Huang, Jian Liu Published: 2026-04-09Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint, safety-evaluation | 2026-04-09 | cs.SD | ai-safety, cssd, preprint, safety-evaluation | E5 / R3 (96%) | - |
| Selective Attention System (SAS): Device-Addressed Speech Detection for Real-Time On-Device Voice AI Bonny Banerjee, Daniyal Anjum, David Joohun Kim, Omar Abbasi Published: 2026-04-09Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-04-09 | cs.SD | ai-safety, cssd, preprint | E4 / R3 (96%) | - |
| AudioGuard: Toward Comprehensive Audio Safety Protection Across Diverse Threat Models Bo Li, Chen Fang, Mintong Kang Published: 2026-04-10Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-04-10 | cs.SD | ai-safety, cssd, preprint | E5 / R3 (97%) | - |
| DDSP-QbE++: Improving Speech Quality for Speech Anonymisation for Atypical Speech Sebastian Stober, Suhita Ghosh, Yamini Sinha Published: 2026-04-10Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-04-10 | cs.SD | ai-safety, cssd, preprint | E4 / R3 (96%) | - |
| GRM: Utility-Aware Jailbreak Attacks on Audio LLMs via Gradient-Ratio Masking Di Wu, Guocong Quan, Hengyuan Na, Miao Hu Published: 2026-04-10Area: cs.SDCitations: - Tags: adversarial-robustness, ai-safety, cssd, preprint | 2026-04-10 | cs.SD | adversarial-robustness, ai-safety, cssd, preprint | E5 / R3 (94%) | - |
| Noise-Aware In-Context Learning for Hallucination Mitigation in ALLMs Khalid Zaman, Masashi Unoki, Qixuan Huang Published: 2026-04-10Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-04-10 | cs.SD | ai-safety, cssd, preprint | E5 / R3 (95%) | - |
| Audio-Omni: Extending Multi-modal Understanding to Versatile Audio Generation and Editing Binxin Yang, Chen Li, Hubery Yin, Jiexuan Zhang Published: 2026-04-12Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-04-12 | cs.SD | ai-safety, cssd, preprint | E5 / R4 (96%) | - |
| Cross-Cultural Bias in Mel-Scale Representations: Evidence and Alternatives from Speech and Music Ajay Pundhir, Shivam Chauhan Published: 2026-04-12Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-04-12 | cs.SD | ai-safety, cssd, preprint | E6 / R4 (96%) | - |
| MeloTune: On-Device Arousal Learning and Peer-to-Peer Mood Coupling for Proactive Music Curation Hongwei Xu Published: 2026-04-12Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-04-12 | cs.SD | ai-safety, cssd, preprint | E6 / R4 (97%) | - |
| MeloTune: On-Device Arousal Learning and Peer-to-Peer Mood Coupling for Proactive Music Curation Hongwei Xu Published: 2026-04-12Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-04-12 | cs.SD | ai-safety, cssd, preprint | E6 / R4 (96%) | - |
| VidAudio-Bench: Benchmarking V2A and VT2A Generation across Four Audio Categories Qian Zhang, Xiongkuo Min, Yixuan Gao, Yuqin Cao Published: 2026-04-12Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-04-12 | cs.SD | ai-safety, cssd, preprint | E5 / R3 (96%) | - |
| ActorMind: Emulating Human Actor Reasoning for Speech Role-Playing Wei Xue, Xi Chen, Yike Guo Published: 2026-04-13Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-04-13 | cs.SD | ai-safety, cssd, preprint | E7 / R3 (96%) | - |
| Audio Flamingo Next: Next-Generation Open Audio-Language Models for Speech, Sound, and Music Arushi Goel, Aya Aljafari, Bryan Catanzaro, Chao-Han Huck Yang Published: 2026-04-13Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-04-13 | cs.SD | ai-safety, cssd, preprint | E5 / R3 (96%) | - |
| Audio Source Separation in Reverberant Environments using $尾$-divergence based Nonnegative Factorization Mahmoud Fakhry, Maurizio Omologo, Piergiorgio Svaizer Published: 2026-04-14Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-04-14 | cs.SD | ai-safety, cssd, preprint | E5 / R3 (95%) | - |
| Elastic Net Regularization and Gabor Dictionary for Classification of Heart Sound Signals using Deep Learning Ascensi贸n Gallardo-Antol铆n, Mahmoud Fakhry Published: 2026-04-14Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-04-14 | cs.SD | ai-safety, cssd, preprint | E5 / R3 (95%) | - |
| Comparison of window shapes and lengths in short-time feature extraction for classification of heart sound signals Abeer FathAllah Brery, Mahmoud Fakhry Published: 2026-04-15Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-04-15 | cs.SD | ai-safety, cssd, preprint | E5 / R3 (95%) | - |
| Towards Fine-grained Temporal Perception: Post-Training Large Audio-Language Models with Audio-Side Time Prompt Ian McLoughlin, Jun Liu, Lirong Dai, Nan Jiang Published: 2026-04-15Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-04-15 | cs.SD | ai-safety, cssd, preprint | E5 / R3 (95%) | - |
| Temporal Contrastive Decoding: A Training-Free Method for Large Audio-Language Models Martin Tak谩膷, Salem Lahlou, Yanda Li, Yuhan Liu Published: 2026-04-16Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-04-16 | cs.SD | ai-safety, cssd, preprint | E9 / R4 (98%) | - |
| AST: Adaptive, Seamless, and Training-Free Precise Speech Editing Jianwei Yin, Jinshan Zhang, Jintao Chen, Meng Xi Published: 2026-04-17Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-04-17 | cs.SD | ai-safety, cssd, preprint | E8 / R8 (95%) | - |
| Video-Robin: Autoregressive Diffusion Planning for Intent-Grounded Video-to-Music Generation Aryan Vijay Bhosale, Dinesh Manocha, Gouthaman KV, Lie Lu Published: 2026-04-19Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-04-19 | cs.SD | ai-safety, cssd, preprint | E8 / R5 (98%) | - |
| Video-Robin: Autoregressive Diffusion Planning for Intent-Grounded Video-to-Music Generation Aryan Vijay Bhosale, Dinesh Manocha, Gouthaman KV, Lie Lu Published: 2026-04-19Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-04-19 | cs.SD | ai-safety, cssd, preprint | E10 / R7 (99%) | - |
| A novel LSTM music generator based on the fractional time-frequency feature extraction Chen Chaofan, Chen Wei, Deng Xinyi, Li Xiulai Published: 2026-04-20Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-04-20 | cs.SD | ai-safety, cssd, preprint | E9 / R6 (96%) | - |
| Latent Fourier Transform Cheng-Zhi Anna Huang, Mason Wang Published: 2026-04-20Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-04-20 | cs.SD | ai-safety, cssd, preprint | E7 / R4 (96%) | - |
| BEAT: Tokenizing and Generating Symbolic Music by Uniform Temporal Steps Haoyu Gu, Jingwei Zhao, Lekai Qian, Ziyu Wang Published: 2026-04-21Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-04-21 | cs.SD | ai-safety, cssd, preprint | E7 / R4 (96%) | - |
| Environmental Sound Deepfake Detection Using Deep-Learning Framework Alexander Schindler, Dat Tran, David Fischinger, Khoi Vu Published: 2026-04-21Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-04-21 | cs.SD | ai-safety, cssd, preprint | E8 / R4 (95%) | - |
| HalluAudio: A Comprehensive Benchmark for Hallucination Detection in Large Audio-Language Models Daipeng Zhang, Feiyu Zhao, Jianguo Wei, Wenhuan Lu Published: 2026-04-21Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-04-21 | cs.SD | ai-safety, cssd, preprint | E9 / R6 (97%) | - |
| Tadabur: A Large-Scale Quran Audio Dataset Faisal Alherran Published: 2026-04-21Area: cs.SDCitations: - Tags: ai-safety, cssd, preprint | 2026-04-21 | cs.SD | ai-safety, cssd, preprint | E8 / R4 (98%) | - |