论文解读PairAlign: A Framework for Sequence Tokenization via Self-Alignment with Applications to Audio Tokenization音频编码 | 7.0/10
论文解读Arbitrarily Settable Frame Rate Neural Speech Codec with Content Adaptive Variable Length Segmentation音频生成 | 7.0/10
论文解读Cross-Architecture Knowledge Distillation of WavLM for Lightweight Speaker Verification说话人验证 | 8.0/10
论文解读Phonological Tokenizer: Prosody-Aware Phonetic Token Via Multi-Objective Fine-Tuning with Differentiable K-Means语音表示学习 | 8.0/10
论文解读The Role of Prosodic and Lexical Cues in Turn-Taking with Self-Supervised Speech Representations语音对话系统 | 7.5/10