论文解读
Light-weight Pronunciation Assessment via Discrete Speech Token Surprisal
自监督学习 | 6.4/10
自监督学习 | 6.4/10
语音识别 | 8.7/10
语音合成 | 5.7/10
音频检索 | 5.7/10
NEST: Narrative Event Structures in Time for Long Video Understanding
多模态模型 | 8.1/10
说话人验证 | 8.6/10
语音合成 | 7.4/10
Pitch Spelling Jazz Lead Sheets, Solo Transcriptions, Classical Piano and Monophonic Scores
PolSeT: Polish Semantics of Timbre Dataset
语音质量评估 | 7.3/10
Prismriver: Formalization of Music Theory and Algorithmic Composition in Lean 4
语音合成 | 6.2/10
语音合成 | 7.9/10
语音转换 | 8/10
语音识别 | 8.7/10
对比学习 | 6.5/10
Stuttering Classification and Segmentation with Attention-Based Multiple Instance Learning
语音识别 | 8.3/10
语音增强 | 7.0/10