论文解读
SARA: A Dual-Stream VAE for High-Fidelity Speech Generation via Integrating Semantic and Acoustic Representations
语音合成 | 7.9/10
语音合成 | 7.9/10
音频生成 | 7.2/10
音乐信息检索 | 7.1/10
模型压缩 | 7.7/10
Steering Where to Listen: Instruction-Based Activation Steering Redirects Temporal Attention in Large Audio-Language Models
语音合成 | 8.1/10
语音合成 | 7.5/10
低资源 | 9.6/10
语音识别 | 6.4/10
语音合成 | 8.1/10
语音识别 | 7.5/10
共分析 36 篇语音/AI 论文
说话人验证 | 6/10
语音质量评估 | 8/10
多语言 | 8/10
AudioProcessBench: Benchmark for Identifying Process Errors in Audio-Grounded Reasoning
语音问答 | 7.5/10
说话人日志 | 6/10
语音编码 | 7.9/10
语音识别 | 7.5/10