论文解读Teacher-Student Structure for Domain Adaptation in Ensemble Audio-Visual Video Deepfake Detection多模态模型 | 7.4/10
论文解读Frozen Multimodal Embeddings for Personality and Cognitive Ability Assessment in Asynchronous Video Interviews语音情感识别 | 6.7/10
论文解读VISA: A Visual Information Strengthened Audio-Reasoning System for the Interspeech 2026 ARC Agent Track音频问答 | 3.9/10
论文解读Neck-Learn: Attention-Based Multiple Instance Learning and Ensemble Framework for Ecological Momentary Assessment语音生物标志物 | 7.0/10
论文解读Voting-Based Pitch Estimation with Temporal and Frequential Alignment and Correlation Aware Selection语音识别 | 8.0/10
论文解读Meta-Ensemble Learning with Diverse Data Splits for Improved Respiratory Sound Classification音频分类 | 8.0/10