论文解读

A Bimodal Approach for Detecting Fatigue Using Speech and Personal Assessments in College Students

A Bimodal Approach for Detecting Fatigue Using Speech and Personal Assessments in College Students

 · 更新于 2026-09-25 · 约 8 分钟 · 3846 字 阅读 →
论文解读

A Dynamic Gated Cross-Attention Framework for Audio-Text Apparent Personality Analysis

音频分类 | 7.0/10

 · 更新于 2026-09-25 · 约 9 分钟 · 4050 字 阅读 →
论文解读

ACIR-MACL: Effective Multimodal Sentiment Analysis via Attention-Based Causal Intervention Regularization and Multi-Aspect Contrastive Learning

情感分析 | 7.0/10

 · 更新于 2026-09-25 · 约 11 分钟 · 5347 字 阅读 →
论文解读

Acoustic and Facial Markers of Perceived Conversational Success in Spontaneous Speech

语音情感识别 | 6.0/10

 · 更新于 2026-09-25 · 约 8 分钟 · 3562 字 阅读 →
论文解读

Acoustic Feedback Cancellation in Hearing Aids Exploiting an Inertial Sensor

音频分类 | 7.0/10

 · 更新于 2026-09-25 · 约 10 分钟 · 4915 字 阅读 →
论文解读

ADH-VA: Adaptive Directed-Hypergraph Convolution with VA Contrastive Learning for Multimodal Conversational Emotion Recognition

语音情感识别 | 7.5/10

 · 更新于 2026-09-25 · 约 10 分钟 · 4727 字 阅读 →
论文解读

Advancing Speech Summarization in Multi-Modal LLMs with Reinforcement Learning

音频问答 | 7.0/10

 · 更新于 2026-09-25 · 约 9 分钟 · 4272 字 阅读 →
论文解读

Affect-Jigsaw: Integrating Core and Peripheral Emotions for Harmonious Fine-Grained Multimodal Emotion Recognition

语音情感识别 | 8.0/10

 · 更新于 2026-09-25 · 约 10 分钟 · 4583 字 阅读 →
论文解读

ALMA-Chor: Leveraging Audio-Lyric Alignment with Mamba for Chorus Detection

音乐信息检索 | 7.0/10

 · 更新于 2026-09-25 · 约 9 分钟 · 4379 字 阅读 →
论文解读

AMBER2: Dual Ambiguity-Aware Emotion Recognition Applied to Speech and Text

语音情感识别 | 8.0/10

 · 更新于 2026-09-25 · 约 9 分钟 · 4014 字 阅读 →
论文解读

An Anomaly-Aware and Audio-Enhanced Dual-Pathway Framework for Alzheimer’s Disease Progression Classification

语音生物标志物 | 7.0/10

 · 更新于 2026-09-25 · 约 12 分钟 · 5640 字 阅读 →
论文解读

An End-to-End Multimodal System for Subtitle Recognition and Chinese-Japanese Translation in Short Dramas

多模态模型 | 7.0/10

 · 更新于 2026-09-25 · 约 11 分钟 · 5044 字 阅读 →
论文解读

An Unsupervised Alignment Feature Fusion System for Spoken Language-Based Dementia Detection

语音生物标志物 | 7.0/10

 · 更新于 2026-09-25 · 约 9 分钟 · 4201 字 阅读 →
论文解读

AnimalCLAP: Taxonomy-Aware Language-Audio Pretraining for Species Recognition and Trait Inference

音频分类 | 8.0/10

 · 更新于 2026-09-25 · 约 10 分钟 · 4960 字 阅读 →
论文解读

APKD: Aligned And Paced Knowledge Distillation Towards Lightweight Heterogeneous Multimodal Emotion Recognition

情感识别 | 7.5/10

 · 更新于 2026-09-25 · 约 8 分钟 · 3585 字 阅读 →
论文解读

AQUA-Bench: Beyond finding answers to knowing when there are None in Audio Question Answering

音频问答 | 7.0/10

 · 更新于 2026-09-25 · 约 8 分钟 · 3971 字 阅读 →
论文解读

Attention-Weighted Centered Kernel Alignment for Knowledge Distillation in Large Audio-Language Models Applied To Speech Emotion Recognition

语音情感识别 | 8.0/10

 · 更新于 2026-09-25 · 约 12 分钟 · 5893 字 阅读 →
论文解读

Attentive AV-Fusionnet: Audio-Visual Quality Prediction with Hybrid Attention

音视频 | 7.0/10

 · 更新于 2026-09-25 · 约 10 分钟 · 4767 字 阅读 →
论文解读

Audience-Aware Co-speech Gesture Generation in Public Speaking via Anticipation Tokens

音频生成 | 8.0/10

 · 更新于 2026-09-25 · 约 11 分钟 · 5224 字 阅读 →
论文解读

Audio-Guided Multimodal Approach for Fine-Grained Alignment and Boundary Modeling in Active Speaker Detection

说话人检测 | 7.5/10

 · 更新于 2026-09-25 · 约 9 分钟 · 4161 字 阅读 →