论文解读
Pixel-TTS: Image based Text Rendering for Robust Text-to-Speech
语音合成 | 7.5/10
语音合成 | 7.5/10
语音生成 | 8.6/10
自监督学习 | 8.5/10
音频深度伪造检测 | 6.7/10
语音识别 | 6.2/10
语音识别 | 7.6/10
大语言模型 | 7.3/10
Semi-Supervised Speech Confidence Detection using Pseudo-Labelling and Whisper Embeddings
自监督学习 | 7.4/10
语音翻译 | 7.6/10
说话人验证 | 7.9/10
多模态模型 | 7.4/10
TMASC: Transmasculine Attitude and Speech Corpus
语音增强 | 5.9/10
多模态模型 | 9.7/10
音频生成 | 8.7/10
多模态模型 | 6.2/10
自适应滤波 | 8/10
鲁棒性 | 9.4/10
When the Same Musical Knowledge Forgets Differently: A Clean Probe of Pathway-Dependent Forgetting