论文解读

Accurate Decoding of Natural Sentences from Non-Invasive Brain Recordings

语音识别 | 7.5/10

 · 更新于 2026-09-25 · 约 10 分钟 · 5004 字 阅读 →
论文解读

Understanding Multilingual Medical ASR Adaptation Through Layer-Wise Analysis

语音识别 | 7.2/10

 · 更新于 2026-09-25 · 约 11 分钟 · 5145 字 阅读 →
论文解读

Cached LLM Probability Retrieval for Speech Recognition

语音识别 | 6.1/10

 · 更新于 2026-09-25 · 约 15 分钟 · 7301 字 阅读 →
论文解读

DuplexGen: Decoupling Content, Timing, and Acoustics for Synthetic Dialogue Speech

语音合成 | 6.1/10

 · 更新于 2026-09-25 · 约 13 分钟 · 6343 字 阅读 →
论文解读

Navigating Speech Enhancement for Real-Time MRI: A Systematic Assessment of Signal Quality, Source Preservation, and Downstream Tasks

语音增强 | 5.7/10

 · 更新于 2026-09-25 · 约 18 分钟 · 8798 字 阅读 →
论文解读

The Null Token Knows: Reducing Message-Free Hallucination in ASR and NMT

语音识别 | 6.5/10

 · 更新于 2026-09-25 · 约 15 分钟 · 7035 字 阅读 →
论文解读

Unadapted Multilingual ASR on a Garrusi Kurdish Evaluation Set: A Common-Reference Staged Normalization Analysis

语音识别 | 5.6/10

 · 更新于 2026-09-25 · 约 14 分钟 · 6794 字 阅读 →
论文解读

Leading-Silence Augmentation and Multi-Stage Synthetic Supervision for the Second MLC-SLM Challenge

说话人日志 | 6.5/10

 · 更新于 2026-09-25 · 约 14 分钟 · 6807 字 阅读 →
论文解读

Measuring Fairness in Large Audio Language Models via Semantic-Aware Bias Estimation

语音识别 | 5.6/10

 · 更新于 2026-09-25 · 约 15 分钟 · 7337 字 阅读 →
论文解读

StreamHear: Domain-Adapted Pseudo-Labeling for Semi-Supervised Streaming Speech Recognition

语音识别 | 6.4/10

 · 更新于 2026-09-25 · 约 12 分钟 · 5783 字 阅读 →
论文解读

Alignment Drift in Single-Model Speculative Decoding for ASR: Mechanism, Correction, and Cost

语音识别 | 7.3/10

 · 更新于 2026-09-25 · 约 14 分钟 · 6965 字 阅读 →
论文解读

Comparative Analysis of Multilingual Pre-trained Models for Nepali Automatic Speech Recognition

语音识别 | 7.6/10

 · 更新于 2026-09-25 · 约 15 分钟 · 7343 字 阅读 →
论文解读

DonorRank: Donor Language Selection for Low-Resource Cross-Lingual Speech Recognition

语音识别 | 6.6/10

 · 更新于 2026-09-25 · 约 10 分钟 · 4656 字 阅读 →
论文解读

Easper: An Accessible ASR Pipeline for Language Documentation

语音识别 | 7.0/10

 · 更新于 2026-09-25 · 约 12 分钟 · 5955 字 阅读 →
论文解读

On-Policy Self-Distillation for Multi-Dialect ASR: Mastering Dialects, Retaining Mandarin

语音识别 | 7.0/10

 · 更新于 2026-09-25 · 约 5 分钟 · 2303 字 阅读 →
论文解读

The SLT 2026 SmartGlasses Challenge: Benchmarking Egocentric Multi-Talker Speech Recognition and Understanding with Audio-Language Models

语音识别 | 7.5/10

 · 更新于 2026-09-25 · 约 14 分钟 · 6748 字 阅读 →
论文解读

ASR-Roundtrip Evaluation Can Mask Context- and Convention-Dependent Reading Errors in Chinese News TTS

语音合成 | 6.5/10

 · 更新于 2026-09-25 · 约 13 分钟 · 6433 字 阅读 →
论文解读

Edge Phoneme Recognition for Children's Speech through Age-Aware Training

语音识别 | 6.3/10

 · 更新于 2026-09-25 · 约 14 分钟 · 6776 字 阅读 →
论文解读

myMediWhisper: Construction of Burmese Medical Speech Corpus and Whisper Fine-Tuning for Clinical Dialogue ASR

语音识别 | 6.7/10

 · 更新于 2026-09-25 · 约 17 分钟 · 8115 字 阅读 →
论文解读

Seeds Before Objectives: Rethinking Evaluation for Low-Resource Garhwali ASR

语音识别 | 7.9/10

 · 更新于 2026-09-25 · 约 19 分钟 · 9335 字 阅读 →