论文解读
Next-Turn: Duration-Aware Streaming Endpoint Detection via Time-to-Next-Speech-Onset Prediction
语音合成 | 7.9/10
语音合成 | 7.9/10
Non-Autoregressive Minimum Bayes' Risk Decoding for Fast Speech Recognition
多模态模型 | 5.6/10
语音合成 | 9.3/10
语音识别 | 7.7/10
语音增强 | 7.6/10
语音情感识别 | 6.8/10
语音增强 | 7/10
语音识别 | 7.6/10
Synergizing Zero-Shot Cross-Lingual Alzheimer Detection with Language-Invariant Multimodal Bi-Geometric Adversarial Learning
音频分类 | 6.4/10
音频分类 | 7.4/10
Vibrato Expression Control for Singing Voice Conversion with Improving Independent Control
语音识别 | 9.1/10
共分析 35 篇语音/AI 论文
音频分类 | 8.3/10
声源定位 | 9.7/10
音频生成 | 8.7/10
音乐信息检索 | 8.0/10
An auscultation location specific study on the relationship between expiratory-to-inspiratory acoustic patterns and spirometric airflow limitation across age and gender in …