论文解读

Input-Adaptive Differentiable Filterbanks via Hypernetworks for Robust Speech Processing

语音识别 | 7.5/10

 · 更新于 2026-09-24 · 约 9 分钟 · 4225 字 阅读 →
论文解读

Is Phase Really Needed for Weakly-Supervised Dereverberation?

语音增强 | 6.0/10

 · 更新于 2026-09-24 · 约 8 分钟 · 3528 字 阅读 →
论文解读

Joint Estimation of Piano Dynamics and Metrical Structure with a Multi-Task Multi-Scale Network

音乐理解 | 7.5/10

 · 更新于 2026-09-24 · 约 10 分钟 · 4795 字 阅读 →
论文解读

Korean aegyo speech shows systematic F1 increase to signal childlike qualities

语音情感识别 | 6.0/10

 · 更新于 2026-09-24 · 约 5 分钟 · 2463 字 阅读 →
论文解读

Learnable Mel-Frontend for Robust Underwater Acoustic Target Detection under Non-Target Interference

音频分类 | 6.5/10

 · 更新于 2026-09-24 · 约 10 分钟 · 4531 字 阅读 →
论文解读

Mambaformer: State-Space Augmented Self-Attention with Downup Sampling for Monaural Speech Enhancement

语音增强 | 7.0/10

 · 更新于 2026-09-24 · 约 10 分钟 · 4866 字 阅读 →
论文解读

Non-Line-of-Sight Vehicle Detection via Audio-Visual Fusion

音频分类 | 8.0/10

 · 更新于 2026-09-24 · 约 9 分钟 · 4086 字 阅读 →
论文解读

Poly-SVC: Polyphony-Aware Singing Voice Conversion with Harmonic Modeling

歌唱语音转换 | 6.5/10

 · 更新于 2026-09-24 · 约 10 分钟 · 4949 字 阅读 →
论文解读

Random Matrix-Driven Graph Representation Learning For Bioacoustic Recognition

生物声学 | 7.5/10

 · 更新于 2026-09-24 · 约 9 分钟 · 4382 字 阅读 →
论文解读

RMODGDF: A Robust STFT-Derived Feature for Musical Instrument Recognition

音乐信息检索 | 7.0/10

 · 更新于 2026-09-24 · 约 9 分钟 · 4290 字 阅读 →
论文解读

Snore Sound Classification Based on Physiological Features and Adaptive Loss Function

音频分类 | 6.5/10

 · 更新于 2026-09-24 · 约 11 分钟 · 5116 字 阅读 →
论文解读

Spectrogram Event Based Feature Representation for Generalizable Automatic Music Transcription

音乐信息检索 | 7.5/10

 · 更新于 2026-09-24 · 约 10 分钟 · 5000 字 阅读 →
论文解读

Subgraph Localization in the Subbands for Partially Spoofed Speech Detection

音频深度伪造检测 | 8.0/10

 · 更新于 2026-09-24 · 约 8 分钟 · 3790 字 阅读 →
论文解读

Subspace Hybrid Adaptive Filtering for Phonocardiogram Signal Denoising

音频增强 | 7.0/10

 · 更新于 2026-09-24 · 约 8 分钟 · 3758 字 阅读 →
论文解读

UMV: A Mixture-Of-Experts Vision Transformer with Multi-Spectrogram Fusion for Underwater Ship Noise Classification

音频分类 | 7.5/10

 · 更新于 2026-09-24 · 约 9 分钟 · 4212 字 阅读 →
论文解读

UNMIXX: Untangling Highly Correlated Singing Voices Mixtures

语音分离 | 8.5/10

 · 更新于 2026-09-24 · 约 9 分钟 · 4383 字 阅读 →
论文解读

Unsupervised Discovery and Analysis of the Vocal Repertoires and Patterns of Select Corvid Species

生物声学 | 7.5/10

 · 更新于 2026-09-24 · 约 10 分钟 · 4918 字 阅读 →
论文解读

USVexplorer: Robust Detection of Ultrasonic Vocalizations with Cross Species Generalization

音频事件检测 | 8.0/10

 · 更新于 2026-09-24 · 约 12 分钟 · 5686 字 阅读 →
论文解读

Voting-Based Pitch Estimation with Temporal and Frequential Alignment and Correlation Aware Selection

语音识别 | 8.0/10

 · 更新于 2026-09-24 · 约 11 分钟 · 5178 字 阅读 →
论文解读

WaveSP-Net: Learnable Wavelet-Domain Sparse Prompt Tuning for Speech Deepfake Detection

语音伪造检测 | 8.0/10

 · 更新于 2026-09-24 · 约 13 分钟 · 6181 字 阅读 →