论文解读

DSSR: Decoupling Salient and Subtle Representations Under Missing Modalities for Multimodal Emotion Recognition

情感识别 | 7.5/10

 · 更新于 2026-09-06 · 约 11 分钟 · 5035 字 阅读 →
论文解读

Dual Contrastive Learning for Semi-Supervised Domain Adaptation in Bi-Modal Depression Recognition

语音生物标志物 | 7.0/10

 · 更新于 2026-09-06 · 约 9 分钟 · 4365 字 阅读 →
论文解读

Dual Data Scaling for Robust Two-Stage User-Defined Keyword Spotting

语音活动检测 | 7.5/10

 · 更新于 2026-09-06 · 约 11 分钟 · 5107 字 阅读 →
论文解读

Dual-Perspective Multimodal Sentiment Analysis with MoE Fusion: Representation Learning via Semantic Resonance and Divergence

多模态情感分析 | 7.0/10

 · 更新于 2026-09-06 · 约 9 分钟 · 4367 字 阅读 →
论文解读

Dual-Strategy-Enhanced Conbimamba for Neural Speaker Diarization

说话人分离 | 8.0/10

 · 更新于 2026-09-06 · 约 10 分钟 · 4642 字 阅读 →
论文解读

Dynamic Balanced Cross-Modal Attention with Gated Sequence Restoration: Towards Robust Multimodal Sentiment Analysis

跨模态 | 7.5/10

 · 更新于 2026-09-06 · 约 9 分钟 · 4284 字 阅读 →
论文解读

Dynamic Noise-Aware Multi Lora Framework Towards Real-World Audio Deepfake Detection

音频深度伪造检测 | 8.0/10

 · 更新于 2026-09-06 · 约 9 分钟 · 4267 字 阅读 →
论文解读

Dynamic Spectrogram Analysis with Local-Aware Graph Networks for Audio Anti-Spoofing

音频深度伪造检测 | 8.5/10

 · 更新于 2026-09-06 · 约 10 分钟 · 4971 字 阅读 →
论文解读

Dynamically Slimmable Speech Enhancement Network with Metric-Guided Training

语音增强 | 7.5/10

 · 更新于 2026-09-06 · 约 9 分钟 · 4160 字 阅读 →
论文解读

E2E-AEC: Implementing An End-To-End Neural Network Learning Approach for Acoustic Echo Cancellation

语音增强 | 7.5/10

 · 更新于 2026-09-06 · 约 9 分钟 · 4163 字 阅读 →
论文解读

Easy Turn: Integrating Acoustic and Linguistic Modalities for Robust Turn-Taking in Full-Duplex Spoken Dialogue Systems

语音对话系统 | 7.0/10

 · 更新于 2026-09-06 · 约 10 分钟 · 4864 字 阅读 →
论文解读

ECHO: Frequency-Aware Hierarchical Encoding for Variable-Length Signals

音频分类 | 9.5/10

 · 更新于 2026-09-06 · 约 9 分钟 · 4374 字 阅读 →
论文解读

EchoFake: A Replay-Aware Dataset For Practical Speech Deepfake Detection

音频深度伪造检测 | 8.5/10

 · 更新于 2026-09-06 · 约 8 分钟 · 3666 字 阅读 →
论文解读

EchoRAG: A Two-Stage Framework for Audio-Text Retrieval and Temporal Grounding

音频检索 | 7.5/10

 · 更新于 2026-09-06 · 约 11 分钟 · 5031 字 阅读 →
论文解读

ECSA: Dual-Branch Emotion Compensation for Emotion-Consistent Speaker Anonymization

语音匿名化 | 8.5/10

 · 更新于 2026-09-06 · 约 9 分钟 · 4486 字 阅读 →
论文解读

EdgeSpot: Efficient and High-Performance Few-Shot Model for Keyword Spotting

语音活动检测 | 7.5/10

 · 更新于 2026-09-06 · 约 9 分钟 · 4076 字 阅读 →
论文解读

EEG and Eye-Tracking Driven Dynamic Target Speaker Extraction with Spontaneous Attention Switching

语音分离 | 7.0/10

 · 更新于 2026-09-06 · 约 10 分钟 · 4652 字 阅读 →
论文解读

EEND-SAA: Enrollment-Less Main Speaker Voice Activity Detection Using Self-Attention Attractors

语音活动检测 | 7.5/10

 · 更新于 2026-09-06 · 约 10 分钟 · 4559 字 阅读 →
论文解读

Efficient Audio-Visual Inference Via Token Clustering And Modality Fusion

音频问答 | 7.5/10

 · 更新于 2026-09-06 · 约 10 分钟 · 4866 字 阅读 →
论文解读

Efficient Depression Detection from Speech via Language-Independent Prompt-Driven Reprogramming

语音生物标志物 | 7.5/10

 · 更新于 2026-09-06 · 约 9 分钟 · 4393 字 阅读 →