论文解读

DSpAST: Disentangled Representations for Spatial Audio Reasoning with Large Language Models

音频问答 | 8.0/10

 · 更新于 2026-09-14 · 约 10 分钟 · 4732 字 阅读 →
论文解读

DSRMS-TransUnet: A Decentralized Non-Shifted Transunet for Shallow Water Acoustic Source Range Estimation

声源定位 | 8.0/10

 · 更新于 2026-09-14 · 约 10 分钟 · 4613 字 阅读 →
论文解读

DSSR: Decoupling Salient and Subtle Representations Under Missing Modalities for Multimodal Emotion Recognition

情感识别 | 7.5/10

 · 更新于 2026-09-14 · 约 11 分钟 · 5035 字 阅读 →
论文解读

Dual Contrastive Learning for Semi-Supervised Domain Adaptation in Bi-Modal Depression Recognition

语音生物标志物 | 7.0/10

 · 更新于 2026-09-14 · 约 9 分钟 · 4365 字 阅读 →
论文解读

Dual Data Scaling for Robust Two-Stage User-Defined Keyword Spotting

语音活动检测 | 7.5/10

 · 更新于 2026-09-14 · 约 11 分钟 · 5107 字 阅读 →
论文解读

Dual-Perspective Multimodal Sentiment Analysis with MoE Fusion: Representation Learning via Semantic Resonance and Divergence

多模态情感分析 | 7.0/10

 · 更新于 2026-09-14 · 约 9 分钟 · 4367 字 阅读 →
论文解读

Dual-Strategy-Enhanced Conbimamba for Neural Speaker Diarization

说话人分离 | 8.0/10

 · 更新于 2026-09-14 · 约 10 分钟 · 4642 字 阅读 →
论文解读

缺失与偏置并存时如何做鲁棒的多模态情感分析:门控序列修复与平衡跨模态注意的协同

📄 缺失与偏置并存时如何做鲁棒的多模态情感分析:门控序列修复与平衡跨模态注意的协同 会议论文 ID:conference:icassp:2026:icassp-arnumber:11460389

 · 更新于 2026-09-14 · 约 16 分钟 · 7690 字 阅读 →
论文解读

Dynamic Noise-Aware Multi Lora Framework Towards Real-World Audio Deepfake Detection

音频深度伪造检测 | 8.0/10

 · 更新于 2026-09-14 · 约 9 分钟 · 4267 字 阅读 →
论文解读

Dynamic Spectrogram Analysis with Local-Aware Graph Networks for Audio Anti-Spoofing

音频深度伪造检测 | 8.5/10

 · 更新于 2026-09-14 · 约 10 分钟 · 4971 字 阅读 →
论文解读

Dynamically Slimmable Speech Enhancement Network with Metric-Guided Training

语音增强 | 7.5/10

 · 更新于 2026-09-14 · 约 9 分钟 · 4160 字 阅读 →
论文解读

E2E-AEC: Implementing An End-To-End Neural Network Learning Approach for Acoustic Echo Cancellation

语音增强 | 7.5/10

 · 更新于 2026-09-14 · 约 9 分钟 · 4163 字 阅读 →
论文解读

Easy Turn: Integrating Acoustic and Linguistic Modalities for Robust Turn-Taking in Full-Duplex Spoken Dialogue Systems

语音对话系统 | 7.0/10

 · 更新于 2026-09-14 · 约 10 分钟 · 4864 字 阅读 →
论文解读

ECHO: Frequency-Aware Hierarchical Encoding for Variable-Length Signals

音频分类 | 9.5/10

 · 更新于 2026-09-14 · 约 9 分钟 · 4374 字 阅读 →
论文解读

EchoFake: A Replay-Aware Dataset For Practical Speech Deepfake Detection

音频深度伪造检测 | 8.5/10

 · 更新于 2026-09-14 · 约 8 分钟 · 3666 字 阅读 →
论文解读

EchoRAG: A Two-Stage Framework for Audio-Text Retrieval and Temporal Grounding

音频检索 | 7.5/10

 · 更新于 2026-09-14 · 约 11 分钟 · 5031 字 阅读 →
论文解读

ECSA: Dual-Branch Emotion Compensation for Emotion-Consistent Speaker Anonymization

语音匿名化 | 8.5/10

 · 更新于 2026-09-14 · 约 9 分钟 · 4486 字 阅读 →
论文解读

EdgeSpot: Efficient and High-Performance Few-Shot Model for Keyword Spotting

语音活动检测 | 7.5/10

 · 更新于 2026-09-14 · 约 9 分钟 · 4076 字 阅读 →
论文解读

EEG and Eye-Tracking Driven Dynamic Target Speaker Extraction with Spontaneous Attention Switching

语音分离 | 7.0/10

 · 更新于 2026-09-14 · 约 10 分钟 · 4652 字 阅读 →
论文解读

EEND-SAA: Enrollment-Less Main Speaker Voice Activity Detection Using Self-Attention Attractors

语音活动检测 | 7.5/10

 · 更新于 2026-09-14 · 约 10 分钟 · 4559 字 阅读 →