论文解读

VowelPrompt: Hearing Speech Emotions from Text via Vowel-level Prosodic Augmentation

语音情感识别 | 7.5/10

 · 更新于 2026-09-25 · 约 11 分钟 · 5344 字 阅读 →
论文解读

EmoTransCap: Dataset and Pipeline for Emotion Transition-Aware Speech Captioning in Discourses

语音情感识别 | 7.5/10

 · 更新于 2026-09-25 · 约 11 分钟 · 5250 字 阅读 →
论文解读

The False Resonance: A Critical Examination of Emotion Embedding Similarity for Speech Generation Evaluation

语音情感识别 模型评估 | 7.0/10

 · 更新于 2026-09-25 · 约 11 分钟 · 5334 字 阅读 →
论文解读

Acoustic and Facial Markers of Perceived Conversational Success in Spontaneous Speech

语音情感识别 | 6.0/10

 · 更新于 2026-09-25 · 约 8 分钟 · 3562 字 阅读 →
论文解读

ADH-VA: Adaptive Directed-Hypergraph Convolution with VA Contrastive Learning for Multimodal Conversational Emotion Recognition

语音情感识别 | 7.5/10

 · 更新于 2026-09-25 · 约 10 分钟 · 4727 字 阅读 →
论文解读

Affect-Jigsaw: Integrating Core and Peripheral Emotions for Harmonious Fine-Grained Multimodal Emotion Recognition

语音情感识别 | 8.0/10

 · 更新于 2026-09-25 · 约 10 分钟 · 4583 字 阅读 →
论文解读

AMBER2: Dual Ambiguity-Aware Emotion Recognition Applied to Speech and Text

语音情感识别 | 8.0/10

 · 更新于 2026-09-25 · 约 9 分钟 · 4014 字 阅读 →
论文解读

APKD: Aligned And Paced Knowledge Distillation Towards Lightweight Heterogeneous Multimodal Emotion Recognition

情感识别 | 7.5/10

 · 更新于 2026-09-25 · 约 8 分钟 · 3585 字 阅读 →
论文解读

Attention-Weighted Centered Kernel Alignment for Knowledge Distillation in Large Audio-Language Models Applied To Speech Emotion Recognition

语音情感识别 | 8.0/10

 · 更新于 2026-09-25 · 约 12 分钟 · 5893 字 阅读 →
论文解读

B-GRPO: Unsupervised Speech Emotion Recognition Based on Batched-Group Relative Policy Optimization

语音情感识别 | 6.5/10

 · 更新于 2026-09-25 · 约 9 分钟 · 4316 字 阅读 →
论文解读

Behind the Scenes: Mechanistic Interpretability of Lora-Adapted Whisper for Speech Emotion Recognition

语音情感识别 | 7.5/10

 · 更新于 2026-09-25 · 约 8 分钟 · 3951 字 阅读 →
论文解读

Bimodal Fusion Framework for Dynamic Facial Expression Recognition In-The-Wild

语音情感识别 | 7.0/10

 · 更新于 2026-09-25 · 约 8 分钟 · 3875 字 阅读 →
论文解读

Clue2Emo: A Brain-Inspired Framework for Open-Vocabulary Multimodal Emotion Recognition

语音情感识别 | 8.5/10

 · 更新于 2026-09-25 · 约 12 分钟 · 6003 字 阅读 →
论文解读

Context-Aware Dynamic Graph Learning for Multimodal Emotion Recognition with Missing Modalities

语音情感识别 | 8.8/10

 · 更新于 2026-09-25 · 约 9 分钟 · 4275 字 阅读 →
论文解读

DDSR-Net: Robust Multimodal Sentiment Analysis via Dynamic Modality Reliability Assessment

语音情感识别 | 6.5/10

 · 更新于 2026-09-25 · 约 17 分钟 · 8029 字 阅读 →
论文解读

DGSDNet: Dual-Graph Spectral Diffusion Network for Incomplete Multimodal Emotion Recognition in Conversations

语音情感识别 | 8.0/10

 · 更新于 2026-09-25 · 约 11 分钟 · 5031 字 阅读 →
论文解读

Diffemotalk: Audio-Driven Facial Animation with Fine-Grained Emotion Control via Diffusion Models

语音情感识别 | 7.5/10

 · 更新于 2026-09-25 · 约 11 分钟 · 5341 字 阅读 →
论文解读

Do You Hear What I Mean? Quantifying the Instruction-Perception GAP in Instruction-Guided Expressive Text-to-Speech Systems

语音合成 | 8.0/10

 · 更新于 2026-09-25 · 约 9 分钟 · 4453 字 阅读 →
论文解读

缺失与偏置并存时如何做鲁棒的多模态情感分析:门控序列修复与平衡跨模态注意的协同

📄 缺失与偏置并存时如何做鲁棒的多模态情感分析:门控序列修复与平衡跨模态注意的协同 会议论文 ID:conference:icassp:2026:icassp-arnumber:11460389

 · 更新于 2026-09-25 · 约 16 分钟 · 7690 字 阅读 →
论文解读

ECSA: Dual-Branch Emotion Compensation for Emotion-Consistent Speaker Anonymization

语音匿名化 | 8.5/10

 · 更新于 2026-09-25 · 约 9 分钟 · 4486 字 阅读 →