论文解读

Identifying Birdsong Syllables without Labelled Data

生物声学 | 7.0/10

 · 更新于 2026-09-06 · 约 8 分钟 · 3972 字 阅读 →
论文解读

Identifying the Minimal and Maximal Phonetic Subspace of Speech Representations

语音识别 | 8.0/10

 · 更新于 2026-09-06 · 约 10 分钟 · 4910 字 阅读 →
论文解读

Identity Leakage Through Accent Cues in Voice Anonymisation

语音匿名化 | 7.0/10

 · 更新于 2026-09-06 · 约 9 分钟 · 4208 字 阅读 →
论文解读

Impact of Phonetics on Speaker Identity in Adversarial Voice Attack

说话人验证 | 7.0/10

 · 更新于 2026-09-06 · 约 8 分钟 · 3657 字 阅读 →
论文解读

Improving Active Learning for Melody Estimation by Disentangling Uncertainties

音乐信息检索 | 7.5/10

 · 更新于 2026-09-06 · 约 9 分钟 · 4455 字 阅读 →
论文解读

Improving Anomalous Sound Detection with Attribute-Aware Representation from Domain-Adaptive Pre-Training

音频事件检测 | 8.0/10

 · 更新于 2026-09-06 · 约 11 分钟 · 5175 字 阅读 →
论文解读

Improving Audio Event Recognition with Consistency Regularization

音频事件检测 | 7.0/10

 · 更新于 2026-09-06 · 约 9 分钟 · 4143 字 阅读 →
论文解读

Improving Audio Question Answering with Variational Inference

音频问答 | 7.5/10

 · 更新于 2026-09-06 · 约 9 分钟 · 4080 字 阅读 →
论文解读

Improving Automatic Speech Recognition by Mitigating Distortions Introduced by Speech Enhancement Under Drone Noise

语音识别 | 6.5/10

 · 更新于 2026-09-06 · 约 10 分钟 · 4511 字 阅读 →
论文解读

Improving Binaural Distance Estimation in Reverberant Rooms Through Contrastive And Multi-Task Learning

声源定位 | 7.0/10

 · 更新于 2026-09-06 · 约 10 分钟 · 4718 字 阅读 →
论文解读

Improving Contextual Asr Via Multi-Grained Fusion With Large Language Models

语音识别 | 8.5/10

 · 更新于 2026-09-06 · 约 8 分钟 · 3976 字 阅读 →
论文解读

Improving Interpretability in Generative Multitimbral DDSP Frameworks via Semantically-Disentangled Musical Attributes

音频生成 | 7.5/10

 · 更新于 2026-09-06 · 约 12 分钟 · 5654 字 阅读 →
论文解读

Improving Multimodal Brain Encoding Model with Dynamic Subject-Awareness Routing

脑信号编码 | 8.0/10

 · 更新于 2026-09-06 · 约 12 分钟 · 5527 字 阅读 →
论文解读

Improving the Speaker Anonymization Evaluation’s Robustness to Target Speakers with Adversarial Learning

语音匿名化 | 7.5/10

 · 更新于 2026-09-06 · 约 12 分钟 · 5692 字 阅读 →
论文解读

In-Sync: Adaptation of Speech Aware Large Language Models for ASR with Word level timestamp predictions

语音识别 | 7.0/10

 · 更新于 2026-09-06 · 约 10 分钟 · 4573 字 阅读 →
论文解读

InconVAD: A Two-Stage Dual-Tower Framework for Multimodal Emotion Inconsistency Detection

语音情感识别 | 7.5/10

 · 更新于 2026-09-06 · 约 10 分钟 · 5000 字 阅读 →
论文解读

Incremental Learning for Audio Classification with Hebbian Deep Neural Networks

音频分类 | 7.5/10

 · 更新于 2026-09-06 · 约 8 分钟 · 4000 字 阅读 →
论文解读

Individualize the HRTF Neural Field Using Anthropometric Parameters Weighted by Direction-Attention

空间音频 | 7.0/10

 · 更新于 2026-09-06 · 约 10 分钟 · 4851 字 阅读 →
论文解读

Influence of Clean Speech Characteristics on Speech Enhancement Performance

语音增强 | 8.0/10

 · 更新于 2026-09-06 · 约 9 分钟 · 4057 字 阅读 →
论文解读

Influence-Aware Curation and Active Selection for Industrial and Surveillance Sound Events

音频事件检测 | 7.0/10

 · 更新于 2026-09-06 · 约 10 分钟 · 5004 字 阅读 →