论文解读

AuEmoChat: Authentic Emotion Understanding and Rendering for Conversational Speech Synthesis

语音合成 | 6.9/10

 · 更新于 2026-09-25 · 约 18 分钟 · 8526 字 阅读 →
论文解读

Rethinking Speech Foundation Model Fine-tuning: Better SFT or Better Match?

语音情感识别 | 6.7/10

 · 更新于 2026-09-25 · 约 18 分钟 · 8730 字 阅读 →
论文解读

Do We Really Need Multimodal Emotion Language Models Larger Than 1B Parameters?

语音情感识别 | 7.4/10

 · 更新于 2026-09-25 · 约 14 分钟 · 6817 字 阅读 →
论文解读

SHAP-Weighted Cross-Modal Expert Fusion for Emotion and Sentiment Recognition: Evidence and Limits

语音情感识别 | 7.7/10

 · 更新于 2026-09-25 · 约 12 分钟 · 5806 字 阅读 →
论文解读

Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts

语音情感识别 | 6.9/10

 · 更新于 2026-09-25 · 约 17 分钟 · 8349 字 阅读 →
论文解读

Layer-wise Cross-Lingual Depression Detection from Speech: Analysis with Contrastive Alignment

语音情感识别 | 5.5/10

 · 更新于 2026-09-25 · 约 15 分钟 · 7228 字 阅读 →
论文解读

ADEPT: RL-Aligned Agentic Decoding of Emotion via Evidence Probing Tools — From Consensus Learning to Ambiguity-Driven Emotion Reasoning

语音情感识别 | 6.5/10

 · 更新于 2026-09-25 · 约 20 分钟 · 9724 字 阅读 →
论文解读

AgentSteerTTS: A Multi-Agent Closed-Loop Framework for Composite-Instruction Text-to-Speech

语音合成 | 7.9/10

 · 更新于 2026-09-25 · 约 17 分钟 · 8042 字 阅读 →
论文解读

Sparse Autoencoders for Interpretable Emotion Control in Text-to-Speech

语音合成 | 7/10

 · 更新于 2026-09-25 · 约 14 分钟 · 6988 字 阅读 →
论文解读

Automatic Detection of Stress from Speech in the Trier Social Stress Test

语音情感识别 | 7.4/10

 · 更新于 2026-09-25 · 约 15 分钟 · 7439 字 阅读 →
论文解读

A Fair and Transparent Framework for Speech-Based Depression Detection: Balancing Interpretability and Performance

语音情感识别 | 7.4/10

 · 更新于 2026-09-25 · 约 18 分钟 · 8609 字 阅读 →
论文解读

Gated Multi-Graph Fusion via Graph Attention Networks for Alzheimer's Disease Detection

语音情感识别 | 5.2/10

 · 更新于 2026-09-25 · 约 11 分钟 · 5054 字 阅读 →
论文解读

SIGMA: Saliency-Guided Sparse Mask Attacks for Speech Emotion Recognition

语音情感识别 | 7.1/10

 · 更新于 2026-09-25 · 约 11 分钟 · 5240 字 阅读 →
论文解读

Learning from Annotation Uncertainty: Entropy-Aware Curriculum for Speech Emotion Recognition

语音情感识别 | 7.4/10

 · 更新于 2026-09-25 · 约 11 分钟 · 5469 字 阅读 →
论文解读

EmotionAI: A Privacy-Preserving Computational Intelligence Pipeline for Speech-Emotion-Grounded Conversational Analysis

语音情感识别 | 6.9/10

 · 更新于 2026-09-25 · 约 11 分钟 · 5369 字 阅读 →
论文解读

Comparative Reasoning: Making an Audio Language Model Better at Comparing Emotions

语音情感识别 | 7.5/10

 · 更新于 2026-09-25 · 约 14 分钟 · 6591 字 阅读 →
论文解读

Backdoor Attacks on Speech Emotion Recognition via TTS-Generated Poisoning

语音情感识别 | 7/10

 · 更新于 2026-09-25 · 约 11 分钟 · 5489 字 阅读 →
论文解读

CAAD: Contrastive Audio-Aware Distillation for Efficient Speech Language Models

语音识别 | 8.9/10

 · 更新于 2026-09-25 · 约 11 分钟 · 5368 字 阅读 →
论文解读

Scaling Audio Models Efficiently: A Joint Study of Compute Constraints and Optimization Behavior

语音识别 | 7.2/10

 · 更新于 2026-09-25 · 约 11 分钟 · 5153 字 阅读 →
论文解读

Reading between the Lines: Leveraging Large Language Models for Global Dementia and Depression Assessment from Clinical Interviews

语音情感识别 | 6.8/10

 · 更新于 2026-09-25 · 约 28 分钟 · 13976 字 阅读 →