论文解读

Dissecting Performance Degradation in Audio Source Separation under Sampling Frequency Mismatch

音乐源分离 | 7.5/10

 · 更新于 2026-09-25 · 约 9 分钟 · 4275 字 阅读 →
论文解读

DOMA: Leveraging Diffusion Language Models with Adaptive Prior for Intent Classification and Slot Filling

语音对话系统 | 8.5/10

 · 更新于 2026-09-25 · 约 10 分钟 · 4529 字 阅读 →
论文解读

DSRMS-TransUnet: A Decentralized Non-Shifted Transunet for Shallow Water Acoustic Source Range Estimation

声源定位 | 8.0/10

 · 更新于 2026-09-25 · 约 10 分钟 · 4613 字 阅读 →
论文解读

DSSR: Decoupling Salient and Subtle Representations Under Missing Modalities for Multimodal Emotion Recognition

情感识别 | 7.5/10

 · 更新于 2026-09-25 · 约 11 分钟 · 5035 字 阅读 →
论文解读

Dynamic Noise-Aware Multi Lora Framework Towards Real-World Audio Deepfake Detection

音频深度伪造检测 | 8.0/10

 · 更新于 2026-09-25 · 约 9 分钟 · 4267 字 阅读 →
论文解读

Enhancing Noise Robustness for Neural Speech Codecs Through Resource-Efficient Progressive Quantization Perturbation Simulation

语音增强 | 7.5/10

 · 更新于 2026-09-25 · 约 9 分钟 · 4043 字 阅读 →
论文解读

Fake Speech Wild: Detecting Deepfake Speech on Social Media Platform

语音伪造检测 | 7.0/10

 · 更新于 2026-09-25 · 约 9 分钟 · 4224 字 阅读 →
论文解读

Fine-Tuning Bigvgan-V2 for Robust Musical Tuning Preservation

音乐生成 | 7.5/10

 · 更新于 2026-09-25 · 约 8 分钟 · 3887 字 阅读 →
论文解读

Frontend Token Enhancement for Token-Based Speech Recognition

语音识别 | 8.0/10

 · 更新于 2026-09-25 · 约 11 分钟 · 5397 字 阅读 →
论文解读

Gdiffuse: Diffusion-Based Speech Enhancement with Noise Model Guidance

语音增强 | 7.0/10

 · 更新于 2026-09-25 · 约 10 分钟 · 4722 字 阅读 →
论文解读

Generalizability of Predictive and Generative Speech Enhancement Models to Pathological Speakers

语音增强 | 7.0/10

 · 更新于 2026-09-25 · 约 10 分钟 · 4542 字 阅读 →
论文解读

Graph-based Modality Alignment for Robustness in Conversational Emotion Recognition

语音情感识别 | 8.0/10

 · 更新于 2026-09-25 · 约 11 分钟 · 5275 字 阅读 →
论文解读

GRNet: Graph Reconstruction Network for Robust Multimodal Sentiment Analysis

多模态情感分析 | 7.5/10

 · 更新于 2026-09-25 · 约 9 分钟 · 4441 字 阅读 →
论文解读

Hanui: Harnessing Distributional Discrepancies for Singing Voice Deepfake Detection

音频深度伪造检测 | 8.0/10

 · 更新于 2026-09-25 · 约 9 分钟 · 4470 字 阅读 →
论文解读

HVAC-EAR: Eavesdropping Human Speech Using HVAC Systems

音频安全 | 8.5/10

 · 更新于 2026-09-25 · 约 11 分钟 · 5225 字 阅读 →
论文解读

I-DCCRN-VAE: An Improved Deep Representation Learning Framework for Complex VAE-Based Single-Channel Speech Enhancement

语音增强 | 7.5/10

 · 更新于 2026-09-25 · 约 10 分钟 · 4966 字 阅读 →
论文解读

Improving Automatic Speech Recognition by Mitigating Distortions Introduced by Speech Enhancement Under Drone Noise

语音识别 | 6.5/10

 · 更新于 2026-09-25 · 约 10 分钟 · 4511 字 阅读 →
论文解读

Improving Binaural Distance Estimation in Reverberant Rooms Through Contrastive And Multi-Task Learning

声源定位 | 7.0/10

 · 更新于 2026-09-25 · 约 10 分钟 · 4718 字 阅读 →
论文解读

Input-Adaptive Differentiable Filterbanks via Hypernetworks for Robust Speech Processing

语音识别 | 7.5/10

 · 更新于 2026-09-25 · 约 9 分钟 · 4225 字 阅读 →
论文解读

Joint Estimation of Primary and Secondary Paths for Personalized Hearable Applications

主动降噪 | 7.5/10

 · 更新于 2026-09-25 · 约 8 分钟 · 3603 字 阅读 →