论文解读

TARNet: A Temporal-Aware Multi-Scale Architecture for Closed-Set Speaker Identification

TARNet: A Temporal-Aware Multi-Scale Architecture for Closed-Set Speaker Identification

 · 更新于 2026-09-24 · 约 18 分钟 · 8941 字 阅读 →
论文解读

Advancing automatic speech recognition using feature fusion with self-supervised learning features: A case study on Fearless Steps Apollo corpus

语音识别 | 7.0/10

 · 更新于 2026-09-24 · 约 11 分钟 · 5309 字 阅读 →
论文解读

Mind the Shift: Using Delta SSL Embeddings to Enhance Child ASR

语音识别 | 7.0/10

 · 更新于 2026-09-24 · 约 8 分钟 · 3909 字 阅读 →
论文解读

Recovering Performance in Speech Emotion Recognition from Discrete Tokens Via Multi-Layer Fusion and Paralinguistic Feature Integration

语音情感识别 | 6.5/10

 · 更新于 2026-09-24 · 约 10 分钟 · 4949 字 阅读 →
论文解读

Robust Deepfake Audio Detection via Multi-Level Intermediate Feature Fusion

音频深度伪造检测 | 7.5/10

 · 更新于 2026-09-24 · 约 8 分钟 · 3781 字 阅读 →
论文解读

Advancing automatic speech recognition using feature fusion with self-supervised learning features: A case study on Fearless Steps Apollo corpus

语音识别 | 7.0/10

 · 更新于 2026-09-24 · 约 11 分钟 · 5316 字 阅读 →