论文解读

AccLID: Accent-aware Language Identification for Robust Multilingual Speech Recognition

语音识别 | 7.0/10

 · 更新于 2026-09-13 · 约 9 分钟 · 4432 字 阅读 →
论文解读

ACIR-MACL: Effective Multimodal Sentiment Analysis via Attention-Based Causal Intervention Regularization and Multi-Aspect Contrastive Learning

情感分析 | 7.0/10

 · 更新于 2026-09-13 · 约 11 分钟 · 5347 字 阅读 →
论文解读

Acoustic and Facial Markers of Perceived Conversational Success in Spontaneous Speech

语音情感识别 | 6.0/10

 · 更新于 2026-09-13 · 约 8 分钟 · 3562 字 阅读 →
论文解读

Acoustic Feedback Cancellation in Hearing Aids Exploiting an Inertial Sensor

音频分类 | 7.0/10

 · 更新于 2026-09-13 · 约 10 分钟 · 4915 字 阅读 →
论文解读

Acoustic Non-Stationarity Objective Assessment with Hard Label Criteria for Supervised Learning Models

音频分类 | 7.0/10

 · 更新于 2026-09-13 · 约 8 分钟 · 3845 字 阅读 →
论文解读

Acoustic Teleportation Via Disentangled Neural Audio Codec Representations

语音增强 | 7.0/10

 · 更新于 2026-09-13 · 约 10 分钟 · 4710 字 阅读 →
论文解读

Adapting Diarization-Conditioned Whisper for End-to-End Multi-Talker Speech Recognition

语音识别 | 7.5/10

 · 更新于 2026-09-13 · 约 11 分钟 · 5056 字 阅读 →
论文解读

Adaptive Deterministic Flow Matching for Target Speaker Extraction

目标说话人提取 | 8.0/10

 · 更新于 2026-09-13 · 约 11 分钟 · 5050 字 阅读 →
论文解读

Adaptive Embedding Fusion with Contrastive Learning for Robust Fully Few-Shot Class-Incremental Audio Classification

音频分类 | 7.5/10

 · 更新于 2026-09-13 · 约 11 分钟 · 5041 字 阅读 →
论文解读

Adaptive Per-Channel Energy Normalization Front-End for Robust Audio Signal Processing

音频分类 | 7.5/10

 · 更新于 2026-09-13 · 约 10 分钟 · 4700 字 阅读 →
论文解读

Adaptive Rotary Steering with Joint Autoregression for Robust Extraction of Closely Moving Speakers in Dynamic Scenarios

语音分离 | 8.5/10

 · 更新于 2026-09-13 · 约 12 分钟 · 5752 字 阅读 →
论文解读

Adaptive Spectral Weighting in Sagittal-Plane Sound Localization: A Reliability-Driven Approach

声源定位 | 6.5/10

 · 更新于 2026-09-13 · 约 9 分钟 · 4116 字 阅读 →
论文解读

Adaptive Task-Incremental Learning For Underwater Acoustic Recognition Based on Mixture-of-Experts Adapter

水下声学目标识别 | 7.0/10

 · 更新于 2026-09-13 · 约 9 分钟 · 4274 字 阅读 →
论文解读

Addressing Gradient Misalignment in Data-Augmented Training for Robust Speech Deepfake Detection

语音伪造检测 | 7.0/10

 · 更新于 2026-09-13 · 约 8 分钟 · 3953 字 阅读 →
论文解读

ADH-VA: Adaptive Directed-Hypergraph Convolution with VA Contrastive Learning for Multimodal Conversational Emotion Recognition

语音情感识别 | 7.5/10

 · 更新于 2026-09-13 · 约 10 分钟 · 4727 字 阅读 →
论文解读

Advanced modeling of interlanguage speech intelligibility benefit with L1-L2 multi-task learning using differentiable K-means for accent-robust discrete token-based ASR

语音识别 | 7.0/10

 · 更新于 2026-09-13 · 约 9 分钟 · 4144 字 阅读 →
论文解读

Advancing LLM-Based Multi-Channel Multi-Speaker Speech Recognition with Global Cross-Channel Attention and Sentence-Ordered First-In First-Out Serialized Output Training

语音识别 | 7.5/10

 · 更新于 2026-09-13 · 约 11 分钟 · 5058 字 阅读 →
论文解读

Advancing Semi-Supervised Child Speech Recognition with Omni-Temporal Classification under Label Noise

语音识别 | 7.5/10

 · 更新于 2026-09-13 · 约 10 分钟 · 4912 字 阅读 →
论文解读

Advancing Speech Summarization in Multi-Modal LLMs with Reinforcement Learning

音频问答 | 7.0/10

 · 更新于 2026-09-13 · 约 9 分钟 · 4272 字 阅读 →
论文解读

Advancing Speech Understanding in Speech-Aware Language Models with GRPO

语音问答 | 7.0/10

 · 更新于 2026-09-13 · 约 7 分钟 · 3396 字 阅读 →