论文解读

Teacher-Student Structure for Domain Adaptation in Ensemble Audio-Visual Video Deepfake Detection

多模态模型 | 7.4/10

 · 更新于 2026-09-24 · 约 12 分钟 · 5649 字 阅读 →
论文解读

Frozen Multimodal Embeddings for Personality and Cognitive Ability Assessment in Asynchronous Video Interviews

语音情感识别 | 6.7/10

 · 更新于 2026-09-24 · 约 13 分钟 · 6098 字 阅读 →
论文解读

VISA: A Visual Information Strengthened Audio-Reasoning System for the Interspeech 2026 ARC Agent Track

音频问答 | 3.9/10

 · 更新于 2026-09-24 · 约 12 分钟 · 5843 字 阅读 →
论文解读

Neck-Learn: Attention-Based Multiple Instance Learning and Ensemble Framework for Ecological Momentary Assessment

语音生物标志物 | 7.0/10

 · 更新于 2026-09-24 · 约 11 分钟 · 5210 字 阅读 →
论文解读

Voting-Based Pitch Estimation with Temporal and Frequential Alignment and Correlation Aware Selection

语音识别 | 8.0/10

 · 更新于 2026-09-24 · 约 11 分钟 · 5178 字 阅读 →
论文解读

Meta-Ensemble Learning with Diverse Data Splits for Improved Respiratory Sound Classification

音频分类 | 8.0/10

 · 更新于 2026-09-24 · 约 9 分钟 · 4218 字 阅读 →