论文解读

Spatial Power Estimation via Riemannian Covariance Matching

声源定位 | 6.5/10

 · 更新于 2026-09-24 · 约 15 分钟 · 7380 字 阅读 →
论文解读

Online Segmented Beamforming via Dynamic Programming

声源定位 | 6.0/10

 · 更新于 2026-09-24 · 约 14 分钟 · 6970 字 阅读 →
论文解读

NDF+: Joint Neural Directional Filtering and Diffuse Sound Extraction

空间音频 | 6.5/10

 · 更新于 2026-09-24 · 约 11 分钟 · 5284 字 阅读 →
论文解读

Adaptive Diagonal Loading for Norm Constrained Beamforming

波束成形 | 7.0/10

 · 更新于 2026-09-24 · 约 8 分钟 · 3730 字 阅读 →
论文解读

Spatial-Magnifier: Spatial upsampling for multichannel speech enhancement

语音增强 | 7.0/10

 · 更新于 2026-09-24 · 约 11 分钟 · 5050 字 阅读 →
论文解读

A Learning-Based Automotive Sound Field Reproduction Method Using Plane-Wave Decomposition and Multi-Position Constraint

空间音频 | 7.5/10

 · 更新于 2026-09-24 · 约 10 分钟 · 4592 字 阅读 →
论文解读

Beamforming Using Virtual Microphones for Hearing Aid Applications

语音增强 | 7.5/10

 · 更新于 2026-09-24 · 约 8 分钟 · 3819 字 阅读 →
论文解读

Equipping Large Language Model with Directional Speech Understanding Capabilities

语音识别 语音翻译 | 7.0/10

 · 更新于 2026-09-24 · 约 9 分钟 · 4467 字 阅读 →
论文解读

Loose Coupling of Spectral and Spatial Models for Multi-Channel Diarization and Enhancement of Meetings in Dynamic Environments

说话人日志 语音分离 | 7.2/10

 · 更新于 2026-09-24 · 约 8 分钟 · 3877 字 阅读 →
论文解读

Low-Latency Audio Front-End Region-of-Interest Beamforming for Smart Glasses

语音增强 | 7.0/10

 · 更新于 2026-09-24 · 约 9 分钟 · 4308 字 阅读 →
论文解读

Mixture To Beamformed Mixture: Leveraging Beamformed Mixture As Weak-Supervision for Speech Enhancement and Noise-Robust ASR

语音增强 | 8.0/10

 · 更新于 2026-09-24 · 约 9 分钟 · 4195 字 阅读 →
论文解读

Mixture-of-Experts Framework for Field-of-View Enhanced Signal-Dependent Binauralization of Moving Talkers

空间音频 | 6.5/10

 · 更新于 2026-09-24 · 约 8 分钟 · 3871 字 阅读 →
论文解读

Multi-Channel Speech Enhancement for Cocktail Party Speech Emotion Recognition

语音情感识别 | 7.5/10

 · 更新于 2026-09-24 · 约 12 分钟 · 5585 字 阅读 →
论文解读

Neural Network-Based Time-Frequency-Bin-Wise Linear Combination of Beamformers for Underdetermined Target Source Extraction

语音分离 | 7.0/10

 · 更新于 2026-09-24 · 约 11 分钟 · 5398 字 阅读 →
论文解读

On The Design of Efficient Neural Methods for Geometry-Agnostic Multichannel Speech Enhancement

语音增强 | 6.5/10

 · 更新于 2026-09-24 · 约 9 分钟 · 4364 字 阅读 →
论文解读

On the Design of Higher-Order Time-Intensity Microphone Arrays for Panoramic Audio Recording and Reproduction

空间音频 | 7.0/10

 · 更新于 2026-09-24 · 约 9 分钟 · 4217 字 阅读 →
论文解读

Reference Microphone Selection for Guided Source Separation Based on The Normalized L-P Norm

语音增强 | 7.0/10

 · 更新于 2026-09-24 · 约 8 分钟 · 3975 字 阅读 →
论文解读

Sequential and Simultaneous Optimization of Microphone Array Geometry and Region-of-Interest Beamforming

声源定位 | 7.5/10

 · 更新于 2026-09-24 · 约 10 分钟 · 4818 字 阅读 →
论文解读

SIRUP: A Diffusion-Based Virtual Upmixer of Steering Vectors for Highly-Directive Spatialization with First-Order Ambisonics

声源定位 | 7.0/10

 · 更新于 2026-09-24 · 约 9 分钟 · 4466 字 阅读 →
论文解读

Spatial Covariance Matrix Reconstruction for Speech Enhancement in Reverberant Multi-Source Environments

语音增强 | 7.5/10

 · 更新于 2026-09-24 · 约 10 分钟 · 4858 字 阅读 →