论文解读

Tatemae: Detecting Alignment Faking via Tool Selection in LLMs

大语言模型 | 7.5/10

 · 更新于 2026-09-25 · 约 10 分钟 · 4820 字 阅读 →
论文解读

The False Resonance: A Critical Examination of Emotion Embedding Similarity for Speech Generation Evaluation

语音情感识别 模型评估 | 7.0/10

 · 更新于 2026-09-25 · 约 11 分钟 · 5334 字 阅读 →
论文解读

A Feature-Optimized Audio Watermarking Algorithm with Adaptive Embedding Strength

音频安全 | 7.5/10

 · 更新于 2026-09-25 · 约 9 分钟 · 4296 字 阅读 →
论文解读

A Framework for Controlled Multi-Speaker Audio Synthesis for Robustness Evaluation of Speaker Diarisation Systems

说话人日志 | 7.5/10

 · 更新于 2026-09-25 · 约 9 分钟 · 4437 字 阅读 →
论文解读

A Robust KNN Approach for Multi-Class Laryngeal Disease Detection using MFCC Features

音频分类 | 7.5/10

 · 更新于 2026-09-25 · 约 8 分钟 · 3778 字 阅读 →
论文解读

A Superb-Style Benchmark of Self-Supervised Speech Models for Audio Deepfake Detection

音频深度伪造检测 | 7.0/10

 · 更新于 2026-09-25 · 约 9 分钟 · 4149 字 阅读 →
论文解读

A Unified SVD-Modal Solution for Sparse Sound Field Reconstruction with Hybrid Spherical-Linear Microphone Arrays

声源定位 | 6.5/10

 · 更新于 2026-09-25 · 约 10 分钟 · 4607 字 阅读 →
论文解读

AccLID: Accent-aware Language Identification for Robust Multilingual Speech Recognition

语音识别 | 7.0/10

 · 更新于 2026-09-25 · 约 9 分钟 · 4432 字 阅读 →
论文解读

Adaptive Per-Channel Energy Normalization Front-End for Robust Audio Signal Processing

音频分类 | 7.5/10

 · 更新于 2026-09-25 · 约 10 分钟 · 4700 字 阅读 →
论文解读

Addressing Gradient Misalignment in Data-Augmented Training for Robust Speech Deepfake Detection

语音伪造检测 | 7.0/10

 · 更新于 2026-09-25 · 约 8 分钟 · 3953 字 阅读 →
论文解读

Advanced modeling of interlanguage speech intelligibility benefit with L1-L2 multi-task learning using differentiable K-means for accent-robust discrete token-based ASR

语音识别 | 7.0/10

 · 更新于 2026-09-25 · 约 9 分钟 · 4144 字 阅读 →
论文解读

Adversarial Defense via Generative Speech Enhancement Module

语音增强 对抗防御 | 7.5/10

 · 更新于 2026-09-25 · 约 8 分钟 · 3619 字 阅读 →
论文解读

Adversarial Fine-Tuning on Speech Foundation Model with Vulnerable Attention Consistency Regularization for Robust Speech Recognition

语音识别 | 7.5/10

 · 更新于 2026-09-25 · 约 9 分钟 · 4509 字 阅读 →
论文解读

AFT: An Exemplar-Free Class Incremental Learning Method for Environmental Sound Classification

音频分类 | 7.0/10

 · 更新于 2026-09-25 · 约 9 分钟 · 4408 字 阅读 →
论文解读

AI-Generated Music Detection in Broadcast Monitoring

音频深度伪造检测 | 7.0/10

 · 更新于 2026-09-25 · 约 9 分钟 · 4218 字 阅读 →
论文解读

AMBER2: Dual Ambiguity-Aware Emotion Recognition Applied to Speech and Text

语音情感识别 | 8.0/10

 · 更新于 2026-09-25 · 约 9 分钟 · 4014 字 阅读 →
论文解读

AmbiDrop: Array-Agnostic Speech Enhancement Using Ambisonics Encoding and Dropout-Based Learning

语音增强 | 7.0/10

 · 更新于 2026-09-25 · 约 62 分钟 · 30885 字 阅读 →
论文解读

AnyAccomp: Generalizable Accompaniment Generation Via Quantized Melodic Bottleneck

音乐生成 | 8.0/10

 · 更新于 2026-09-25 · 约 9 分钟 · 4201 字 阅读 →
论文解读

AnyRIR: Robust Non-Intrusive Room Impulse Response Estimation in the Wild

空间音频 | 7.0/10

 · 更新于 2026-09-25 · 约 9 分钟 · 4482 字 阅读 →
论文解读

AQUA-Bench: Beyond finding answers to knowing when there are None in Audio Question Answering

音频问答 | 7.0/10

 · 更新于 2026-09-25 · 约 8 分钟 · 3971 字 阅读 →