论文解读

Acoustic Teleportation Via Disentangled Neural Audio Codec Representations

语音增强 | 7.0/10

 · 更新于 2026-09-24 · 约 10 分钟 · 4710 字 阅读 →
论文解读

Adapting Diarization-Conditioned Whisper for End-to-End Multi-Talker Speech Recognition

语音识别 | 7.5/10

 · 更新于 2026-09-24 · 约 11 分钟 · 5056 字 阅读 →
论文解读

Adaptive Deterministic Flow Matching for Target Speaker Extraction

目标说话人提取 | 8.0/10

 · 更新于 2026-09-24 · 约 11 分钟 · 5050 字 阅读 →
论文解读

Adaptive Embedding Fusion with Contrastive Learning for Robust Fully Few-Shot Class-Incremental Audio Classification

音频分类 | 7.5/10

 · 更新于 2026-09-24 · 约 11 分钟 · 5041 字 阅读 →
论文解读

Adaptive Per-Channel Energy Normalization Front-End for Robust Audio Signal Processing

音频分类 | 7.5/10

 · 更新于 2026-09-24 · 约 10 分钟 · 4700 字 阅读 →
论文解读

Adaptive Rotary Steering with Joint Autoregression for Robust Extraction of Closely Moving Speakers in Dynamic Scenarios

语音分离 | 8.5/10

 · 更新于 2026-09-24 · 约 12 分钟 · 5752 字 阅读 →
论文解读

Adaptive Spectral Weighting in Sagittal-Plane Sound Localization: A Reliability-Driven Approach

声源定位 | 6.5/10

 · 更新于 2026-09-24 · 约 9 分钟 · 4116 字 阅读 →
论文解读

Adaptive Task-Incremental Learning For Underwater Acoustic Recognition Based on Mixture-of-Experts Adapter

水下声学目标识别 | 7.0/10

 · 更新于 2026-09-24 · 约 9 分钟 · 4274 字 阅读 →
论文解读

Addressing Gradient Misalignment in Data-Augmented Training for Robust Speech Deepfake Detection

语音伪造检测 | 7.0/10

 · 更新于 2026-09-24 · 约 8 分钟 · 3953 字 阅读 →
论文解读

ADH-VA: Adaptive Directed-Hypergraph Convolution with VA Contrastive Learning for Multimodal Conversational Emotion Recognition

语音情感识别 | 7.5/10

 · 更新于 2026-09-24 · 约 10 分钟 · 4727 字 阅读 →
论文解读

Advanced modeling of interlanguage speech intelligibility benefit with L1-L2 multi-task learning using differentiable K-means for accent-robust discrete token-based ASR

语音识别 | 7.0/10

 · 更新于 2026-09-24 · 约 9 分钟 · 4144 字 阅读 →
论文解读

Advancing LLM-Based Multi-Channel Multi-Speaker Speech Recognition with Global Cross-Channel Attention and Sentence-Ordered First-In First-Out Serialized Output Training

语音识别 | 7.5/10

 · 更新于 2026-09-24 · 约 11 分钟 · 5058 字 阅读 →
论文解读

Advancing Semi-Supervised Child Speech Recognition with Omni-Temporal Classification under Label Noise

语音识别 | 7.5/10

 · 更新于 2026-09-24 · 约 10 分钟 · 4912 字 阅读 →
论文解读

Advancing Speech Summarization in Multi-Modal LLMs with Reinforcement Learning

音频问答 | 7.0/10

 · 更新于 2026-09-24 · 约 9 分钟 · 4272 字 阅读 →
论文解读

Advancing Speech Understanding in Speech-Aware Language Models with GRPO

语音问答 | 7.0/10

 · 更新于 2026-09-24 · 约 7 分钟 · 3396 字 阅读 →
论文解读

Adversarial Defense via Generative Speech Enhancement Module

语音增强 对抗防御 | 7.5/10

 · 更新于 2026-09-24 · 约 8 分钟 · 3619 字 阅读 →
论文解读

Adversarial Fine-Tuning on Speech Foundation Model with Vulnerable Attention Consistency Regularization for Robust Speech Recognition

语音识别 | 7.5/10

 · 更新于 2026-09-24 · 约 9 分钟 · 4509 字 阅读 →
论文解读

Adversarial Rivalry Learning for Music Classification

音乐分类 | 6.5/10

 · 更新于 2026-09-24 · 约 10 分钟 · 4912 字 阅读 →
论文解读

Affect-Jigsaw: Integrating Core and Peripheral Emotions for Harmonious Fine-Grained Multimodal Emotion Recognition

语音情感识别 | 8.0/10

 · 更新于 2026-09-24 · 约 10 分钟 · 4583 字 阅读 →
论文解读

AFT: An Exemplar-Free Class Incremental Learning Method for Environmental Sound Classification

音频分类 | 7.0/10

 · 更新于 2026-09-24 · 约 9 分钟 · 4408 字 阅读 →