论文解读

A Text-To-Text Alignment Algorithm for Better Evaluation of Modern Speech Recognition Systems

模型评估 | 7.5/10

 · 更新于 2026-09-11 · 约 8 分钟 · 3720 字 阅读 →
论文解读

A Unified SVD-Modal Solution for Sparse Sound Field Reconstruction with Hybrid Spherical-Linear Microphone Arrays

声源定位 | 6.5/10

 · 更新于 2026-09-11 · 约 10 分钟 · 4607 字 阅读 →
论文解读

A Unsupervised Domain Adaptation Framework For Semi-Supervised Melody Extraction Using Confidence Matrix Replace and Nearest Neighbour Supervision

音乐信息检索 | 8.0/10

 · 更新于 2026-09-11 · 约 9 分钟 · 4391 字 阅读 →
论文解读

ACAVCaps: Enabling Large-Scale Training for Fine-Grained and Diverse Audio Understanding

音频分类 | 8.5/10

 · 更新于 2026-09-11 · 约 8 分钟 · 3768 字 阅读 →
论文解读

Accelerating Regularized Attention Kernel Regression for Spectrum Cartography

频谱测绘 | 8.5/10

 · 更新于 2026-09-11 · 约 9 分钟 · 4353 字 阅读 →
论文解读

AccLID: Accent-aware Language Identification for Robust Multilingual Speech Recognition

语音识别 | 7.0/10

 · 更新于 2026-09-11 · 约 9 分钟 · 4432 字 阅读 →
论文解读

ACIR-MACL: Effective Multimodal Sentiment Analysis via Attention-Based Causal Intervention Regularization and Multi-Aspect Contrastive Learning

情感分析 | 7.0/10

 · 更新于 2026-09-11 · 约 11 分钟 · 5347 字 阅读 →
论文解读

Acoustic and Facial Markers of Perceived Conversational Success in Spontaneous Speech

语音情感识别 | 6.0/10

 · 更新于 2026-09-11 · 约 8 分钟 · 3562 字 阅读 →
论文解读

Acoustic Feedback Cancellation in Hearing Aids Exploiting an Inertial Sensor

音频分类 | 7.0/10

 · 更新于 2026-09-11 · 约 10 分钟 · 4915 字 阅读 →
论文解读

Acoustic Non-Stationarity Objective Assessment with Hard Label Criteria for Supervised Learning Models

音频分类 | 7.0/10

 · 更新于 2026-09-11 · 约 8 分钟 · 3845 字 阅读 →
论文解读

Acoustic Teleportation Via Disentangled Neural Audio Codec Representations

语音增强 | 7.0/10

 · 更新于 2026-09-11 · 约 10 分钟 · 4710 字 阅读 →
论文解读

Adapting Diarization-Conditioned Whisper for End-to-End Multi-Talker Speech Recognition

语音识别 | 7.5/10

 · 更新于 2026-09-11 · 约 11 分钟 · 5056 字 阅读 →
论文解读

Adaptive Deterministic Flow Matching for Target Speaker Extraction

目标说话人提取 | 8.0/10

 · 更新于 2026-09-11 · 约 11 分钟 · 5050 字 阅读 →
论文解读

Adaptive Embedding Fusion with Contrastive Learning for Robust Fully Few-Shot Class-Incremental Audio Classification

音频分类 | 7.5/10

 · 更新于 2026-09-11 · 约 11 分钟 · 5041 字 阅读 →
论文解读

Adaptive Per-Channel Energy Normalization Front-End for Robust Audio Signal Processing

音频分类 | 7.5/10

 · 更新于 2026-09-11 · 约 10 分钟 · 4700 字 阅读 →
论文解读

Adaptive Rotary Steering with Joint Autoregression for Robust Extraction of Closely Moving Speakers in Dynamic Scenarios

语音分离 | 8.5/10

 · 更新于 2026-09-11 · 约 12 分钟 · 5752 字 阅读 →
论文解读

Adaptive Spectral Weighting in Sagittal-Plane Sound Localization: A Reliability-Driven Approach

声源定位 | 6.5/10

 · 更新于 2026-09-11 · 约 9 分钟 · 4116 字 阅读 →
论文解读

Adaptive Task-Incremental Learning For Underwater Acoustic Recognition Based on Mixture-of-Experts Adapter

水下声学目标识别 | 7.0/10

 · 更新于 2026-09-11 · 约 9 分钟 · 4274 字 阅读 →
论文解读

Addressing Gradient Misalignment in Data-Augmented Training for Robust Speech Deepfake Detection

语音伪造检测 | 7.0/10

 · 更新于 2026-09-11 · 约 8 分钟 · 3953 字 阅读 →
论文解读

ADH-VA: Adaptive Directed-Hypergraph Convolution with VA Contrastive Learning for Multimodal Conversational Emotion Recognition

语音情感识别 | 7.5/10

 · 更新于 2026-09-11 · 约 10 分钟 · 4727 字 阅读 →