论文解读

Easy Turn: Integrating Acoustic and Linguistic Modalities for Robust Turn-Taking in Full-Duplex Spoken Dialogue Systems

语音对话系统 | 7.0/10

 · 更新于 2026-09-25 · 约 10 分钟 · 4864 字 阅读 →
论文解读

EchoFake: A Replay-Aware Dataset For Practical Speech Deepfake Detection

音频深度伪造检测 | 8.5/10

 · 更新于 2026-09-25 · 约 8 分钟 · 3666 字 阅读 →
论文解读

EEG and Eye-Tracking Driven Dynamic Target Speaker Extraction with Spontaneous Attention Switching

语音分离 | 7.0/10

 · 更新于 2026-09-25 · 约 10 分钟 · 4652 字 阅读 →
论文解读

Emilia-NV: A Non-Verbal Speech Dataset with Word-Level Annotation for Human-Like Speech Modeling

语音识别 | 7.5/10

 · 更新于 2026-09-25 · 约 10 分钟 · 4512 字 阅读 →
论文解读

Enabling Multi-Species Bird Classification on Low-Power Bioacoustic Loggers

生物声学 | 8.0/10

 · 更新于 2026-09-25 · 约 10 分钟 · 4664 字 阅读 →
论文解读

Evaluating Bias in Spoken Dialogue LLMs for Real-World Decisions and Recommendations

模型评估 | 7.0/10

 · 更新于 2026-09-25 · 约 8 分钟 · 3876 字 阅读 →
论文解读

Evaluating Compositional Structure in Audio Representations

模型评估 | 7.0/10

 · 更新于 2026-09-25 · 约 9 分钟 · 4315 字 阅读 →
论文解读

Evaluating Disentangled Representations for Controllable Music Generation

音乐生成 | 7.5/10

 · 更新于 2026-09-25 · 约 8 分钟 · 4005 字 阅读 →
论文解读

Evaluating Emotion Recognition in Spoken Language Models on Emotionally Incongruent Speech

语音情感识别 | 7.5/10

 · 更新于 2026-09-25 · 约 7 分钟 · 3322 字 阅读 →
论文解读

Evaluating High-Resolution Piano Sustain Pedal Depth Estimation with Musically Informed Metrics

音乐信息检索 | 8.0/10

 · 更新于 2026-09-25 · 约 9 分钟 · 4473 字 阅读 →
论文解读

Evaluating Pretrained Speech Embedding Systems for Dysarthria Detection Across Heterogenous Datasets

语音生物标志物 | 7.5/10

 · 更新于 2026-09-25 · 约 8 分钟 · 3927 字 阅读 →
论文解读

Generalizability of Predictive and Generative Speech Enhancement Models to Pathological Speakers

语音增强 | 7.0/10

 · 更新于 2026-09-25 · 约 10 分钟 · 4542 字 阅读 →
论文解读

Generative Audio Extension and Morphing

音频生成 | 7.5/10

 · 更新于 2026-09-25 · 约 12 分钟 · 5736 字 阅读 →
论文解读

Hair Noise Analysis and Mitigation for Smart Glasses Audio Captures

语音增强 | 7.5/10

 · 更新于 2026-09-25 · 约 8 分钟 · 3856 字 阅读 →
论文解读

HiFi-HARP: A High-Fidelity 7th-Order Ambisonic Room Impulse Response Dataset

数据集 | 7.5/10

 · 更新于 2026-09-25 · 约 11 分钟 · 5023 字 阅读 →
论文解读

High-Fidelity Speech Enhancement Via Discrete Audio Tokens

语音增强 | 7.5/10

 · 更新于 2026-09-25 · 约 9 分钟 · 4455 字 阅读 →
论文解读

How to Label Resynthesized Audio: The Dual Role of Neural Audio Codecs in Audio Deepfake Detection

音频深度伪造检测 | 7.5/10

 · 更新于 2026-09-25 · 约 9 分钟 · 4091 字 阅读 →
论文解读

Human-1 by Josh Talks: A Full-Duplex Conversational Modeling Framework in Hindi using Real-World Conversations

语音对话系统 | 7.5/10

 · 更新于 2026-09-25 · 约 12 分钟 · 5611 字 阅读 →
会议任务专题

ICASSP 2026 - 数据集

共 3 篇 ICASSP 2026 数据集 方向论文

 · 更新于 2026-09-25 · 约 10 分钟 · 4906 字 阅读 →
论文解读

Interpretable Music Harmonic Analysis Through Multilinear Mixture of Experts

音乐理解 | 7.5/10

 · 更新于 2026-09-25 · 约 10 分钟 · 4568 字 阅读 →