论文解读

Efficient Solutions for Mitigating Initialization Bias in Unsupervised Self-Adaptive Auditory Attention Decoding

听觉注意解码 | 8.5/10

 · 更新于 2026-09-06 · 约 8 分钟 · 3989 字 阅读 →
论文解读

EMG-to-Speech with Fewer Channels

语音合成 | 7.5/10

 · 更新于 2026-09-06 · 约 9 分钟 · 4171 字 阅读 →
论文解读

Emilia-NV: A Non-Verbal Speech Dataset with Word-Level Annotation for Human-Like Speech Modeling

语音识别 | 7.5/10

 · 更新于 2026-09-06 · 约 10 分钟 · 4512 字 阅读 →
论文解读

Emo-TTA: Improving Test-Time Adaptation of Audio-Language Models for Speech Emotion Recognition

语音情感识别 | 7.0/10

 · 更新于 2026-09-06 · 约 10 分钟 · 4628 字 阅读 →
论文解读

EMORL-TTS: Reinforcement Learning for Fine-Grained Emotion Control in LLM-based TTS

语音合成 | 8.5/10

 · 更新于 2026-09-06 · 约 11 分钟 · 5127 字 阅读 →
论文解读

EmoShift: Lightweight Activation Steering for Enhanced Emotion-Aware Speech Synthesis

语音合成 | 7.0/10

 · 更新于 2026-09-06 · 约 9 分钟 · 4419 字 阅读 →
论文解读

Emotion-Aligned Generation in Diffusion Text to Speech Models Via Preference-Guided Optimization

语音合成 | 8.0/10

 · 更新于 2026-09-06 · 约 10 分钟 · 4704 字 阅读 →
论文解读

Emotional Damage: Investigating Safety Vulnerabilities of Large Audio-Language Models Under Speaker Emotional Variations

音频安全 | 7.5/10

 · 更新于 2026-09-06 · 约 8 分钟 · 3817 字 阅读 →
论文解读

Emotional Dimension Control in Language Model-Based Text-To-Speech: Spanning a Broad Spectrum of Human Emotions

语音合成 | 7.5/10

 · 更新于 2026-09-06 · 约 9 分钟 · 4379 字 阅读 →
论文解读

EmoTri-RL: Emotion- and Cause-Aware Reinforcement Learning for Multi-Modal Empathetic Dialogue

语音情感识别 | 7.0/10

 · 更新于 2026-09-06 · 约 10 分钟 · 4826 字 阅读 →
论文解读

Empowering Multimodal Respiratory Sound Classification with Counterfactual Adversarial Debiasing for Out-of-Distribution Robustness

音频分类 | 7.0/10

 · 更新于 2026-09-06 · 约 9 分钟 · 4255 字 阅读 →
论文解读

Enabling Multi-Species Bird Classification on Low-Power Bioacoustic Loggers

生物声学 | 8.0/10

 · 更新于 2026-09-06 · 约 10 分钟 · 4664 字 阅读 →
论文解读

Encoding Emotion Through Self-Supervised Eye Movement Reconstruction

语音情感识别 | 7.5/10

 · 更新于 2026-09-06 · 约 11 分钟 · 5306 字 阅读 →
论文解读

Enhanced Generative Machine Listener

音频分类 | 7.0/10

 · 更新于 2026-09-06 · 约 9 分钟 · 4321 字 阅读 →
论文解读

Enhancing Audio Question-Answering Performance Through Log-Likelihood Guided Reward Functions

音频问答 | 8.5/10

 · 更新于 2026-09-06 · 约 9 分钟 · 4398 字 阅读 →
论文解读

Enhancing Automatic Drum Transcription with Online Dynamic Few-Shot Learning

音乐信息检索 | 7.0/10

 · 更新于 2026-09-06 · 约 8 分钟 · 3991 字 阅读 →
论文解读

Enhancing Dialogue-Related Speech Tasks with Generated Spoken Dialogues

语音对话系统 | 6.5/10

 · 更新于 2026-09-06 · 约 10 分钟 · 4549 字 阅读 →
论文解读

Enhancing Noise Robustness for Neural Speech Codecs Through Resource-Efficient Progressive Quantization Perturbation Simulation

语音增强 | 7.5/10

 · 更新于 2026-09-06 · 约 9 分钟 · 4043 字 阅读 →
论文解读

Enhancing Speaker Verification with w2v-BERT 2.0 and Knowledge Distillation Guided Structured Pruning

说话人验证 | 7.5/10

 · 更新于 2026-09-06 · 约 10 分钟 · 4903 字 阅读 →
论文解读

Enhancing Speech Intelligibility Prediction for Hearing Aids with Complementary Speech Foundation Model Representations

语音增强 | 7.5/10

 · 更新于 2026-09-06 · 约 8 分钟 · 3521 字 阅读 →