论文解读

Leveraging Multiple Speech Enhancers for Non-Intrusive Intelligibility Prediction for Hearing-Impaired Listeners

模型评估 | 7.5/10

 · 更新于 2026-09-06 · 约 12 分钟 · 5540 字 阅读 →
论文解读

Leveraging prediction entropy for Automatic prompt weighting in Zero-Shot Audio-Language Classification

音频分类 | 7.5/10

 · 更新于 2026-09-06 · 约 9 分钟 · 4393 字 阅读 →
论文解读

Leveraging Segment-Level Speech Representations for LLM-Based Speech Recognition

语音识别 | 7.0/10

 · 更新于 2026-09-06 · 约 11 分钟 · 5225 字 阅读 →
论文解读

Leveraging Text-to-Speech and Voice Conversion as Data Augmentation for Alzheimer's Disease Detection from Spontaneous Speech

语音生物标志物 | 7.0/10

 · 更新于 2026-09-06 · 约 11 分钟 · 5134 字 阅读 →
论文解读

Leveraging Whisper Embeddings For Audio-Based Lyrics Matching

音乐信息检索 | 7.0/10

 · 更新于 2026-09-06 · 约 10 分钟 · 4853 字 阅读 →
论文解读

Lightweight and Generalizable Acoustic Scene Representations Via Contrastive Fine-Tuning and Distillation

音频场景理解 | 8.0/10

 · 更新于 2026-09-06 · 约 11 分钟 · 5226 字 阅读 →
论文解读

Lightweight and Perceptually-Guided Voice Conversion for Electro-Laryngeal Speech

语音转换 | 7.5/10

 · 更新于 2026-09-06 · 约 13 分钟 · 6301 字 阅读 →
论文解读

Lightweight Implicit Neural Network for Binaural Audio Synthesis

空间音频 | 7.0/10

 · 更新于 2026-09-06 · 约 13 分钟 · 6286 字 阅读 →
论文解读

Lightweight Phoneme-Conditioned Bandwidth Extension for Body-Conducted Speech

语音增强 | 7.5/10

 · 更新于 2026-09-06 · 约 10 分钟 · 4701 字 阅读 →
论文解读

Lingometer: On-Device Personal Speech Word Counting System

语音活动检测 | 8.0/10

 · 更新于 2026-09-06 · 约 12 分钟 · 5825 字 阅读 →
论文解读

Linguard: Authenticating Speech Recordings Using Speech Recognition and Watermark

音频安全 | 6.5/10

 · 更新于 2026-09-06 · 约 9 分钟 · 4307 字 阅读 →
论文解读

LipsAM: Lipschitz-Continuous Amplitude Modifier for Audio Signal Processing and its Application to Plug-And-Play Dereverberation

语音增强 | 7.5/10

 · 更新于 2026-09-06 · 约 10 分钟 · 4661 字 阅读 →
论文解读

Lisa: Lightweight Yet Superb Neural Speech Coding

语音编码 | 8.5/10

 · 更新于 2026-09-06 · 约 10 分钟 · 4932 字 阅读 →
论文解读

Listen, But Don't Leak: Sensitive Data Protection for Privacy Aware Automatic Speech Recognition with Acoustic Triggers

语音识别 | 7.5/10

 · 更新于 2026-09-06 · 约 9 分钟 · 4187 字 阅读 →
论文解读

LLAC: Learned Lossless Audio Codec

音频无损编码 | 7.5/10

 · 更新于 2026-09-06 · 约 8 分钟 · 3743 字 阅读 →
论文解读

LLM-Based Post-ASR Error Correction for Disordered Speech

语音识别 | 7.5/10

 · 更新于 2026-09-06 · 约 11 分钟 · 5031 字 阅读 →
论文解读

Localizing Speech Deepfakes Beyond Transitions via Segment-Aware Learning

音频深度伪造检测 | 8.0/10

 · 更新于 2026-09-06 · 约 9 分钟 · 4261 字 阅读 →
论文解读

LongSpeech: A Scalable Benchmark for Transcription, Translation and Understanding in Long Speech

基准测试 | 7.8/10

 · 更新于 2026-09-06 · 约 8 分钟 · 3975 字 阅读 →
论文解读

Look, Listen and Segment: Towards Weakly Supervised Audio-Visual Semantic Segmentation

音视频 | 7.0/10

 · 更新于 2026-09-06 · 约 9 分钟 · 4187 字 阅读 →
论文解读

Loose Coupling of Spectral and Spatial Models for Multi-Channel Diarization and Enhancement of Meetings in Dynamic Environments

说话人日志 语音分离 | 7.2/10

 · 更新于 2026-09-06 · 约 8 分钟 · 3877 字 阅读 →