论文解读

SCRAPL: Scattering Transform with Random Paths for Machine Learning

音频生成 | 8.5/10

 · 更新于 2026-09-24 · 约 11 分钟 · 5085 字 阅读 →
论文解读

Learnable Fractional Superlets with a Spectro-Temporal Emotion Encoder for Speech Emotion Recognition

语音情感识别 | 7.5/10

 · 更新于 2026-09-24 · 约 9 分钟 · 4089 字 阅读 →
论文解读

Query-Guided Spatial–Temporal–Frequency Interaction for Music Audio–Visual Question Answering

音频问答 | 7.0/10

 · 更新于 2026-09-24 · 约 10 分钟 · 4714 字 阅读 →
论文解读

SCRAPL: Scattering Transform with Random Paths for Machine Learning

音频生成 | 8.5/10

 · 更新于 2026-09-24 · 约 10 分钟 · 4976 字 阅读 →
论文解读

Earable Platform with Integrated Simultaneous EEG Sensing and Auditory Stimulation

音频事件检测 | 5.5/10

 · 更新于 2026-09-24 · 约 6 分钟 · 2641 字 阅读 →
论文解读

Spectrographic Portamento Gradient Analysis: A Quantitative Method for Historical Cello Recordings with Application to Beethoven's Piano and Cello Sonatas, 1930--2012

音乐信息检索 | 7.5/10

 · 更新于 2026-09-24 · 约 7 分钟 · 3274 字 阅读 →
论文解读

Recurrence-Based Nonlinear Vocal Dynamics as Digital Biomarkers for Depression Detection from Conversational Speech

语音生物标志物 | 6.5/10

 · 更新于 2026-09-24 · 约 7 分钟 · 3316 字 阅读 →
论文解读

A Noniterative Phase Retrieval Considering the Zeros of STFT Magnitude

信号处理 | 7.5/10

 · 更新于 2026-09-24 · 约 7 分钟 · 3218 字 阅读 →
论文解读

Acoustic Non-Stationarity Objective Assessment with Hard Label Criteria for Supervised Learning Models

音频分类 | 7.0/10

 · 更新于 2026-09-24 · 约 8 分钟 · 3845 字 阅读 →
论文解读

An Audio-Visual Speech Separation Network with Joint Cross-Attention and Iterative Modeling

语音分离 | 7.5/10

 · 更新于 2026-09-24 · 约 10 分钟 · 4638 字 阅读 →
论文解读

An Event-Based Sequence Modeling Approach to Recognizing Non-Triad Chords with Oversegmentation Minimization

音乐信息检索 | 7.5/10

 · 更新于 2026-09-24 · 约 10 分钟 · 4973 字 阅读 →
论文解读

AR-BSNet: Towards Ultra-Low Complexity Autoregressive Target Speaker Extraction With Band-Split Modeling

语音分离 | 7.0/10

 · 更新于 2026-09-24 · 约 10 分钟 · 4535 字 阅读 →
论文解读

Audio Deepfake Detection at the First Greeting: "Hi!"

音频深度伪造检测 | 7.5/10

 · 更新于 2026-09-24 · 约 9 分钟 · 4507 字 阅读 →
论文解读

BioSEN: A Bio-Acoustic Signal Enhancement Network for Animal Vocalizations

生物声学 | 7.5/10

 · 更新于 2026-09-24 · 约 11 分钟 · 5221 字 阅读 →
论文解读

BSMP-SENet:Band-Split Magnitude-Phase Network for Speech Enhancement

语音增强 | 7.0/10

 · 更新于 2026-09-24 · 约 9 分钟 · 4436 字 阅读 →
论文解读

Coupling Acoustic Geometry and Visual Semantics for Robust Depth Estimation

空间音频 | 7.5/10

 · 更新于 2026-09-24 · 约 20 分钟 · 9555 字 阅读 →
论文解读

Cross-Cultural Bias in Mel-Scale Representations: Evidence and Alternatives from Speech and Music

语音识别 | 7.0/10

 · 更新于 2026-09-24 · 约 9 分钟 · 4247 字 阅读 →
论文解读

Enabling Multi-Species Bird Classification on Low-Power Bioacoustic Loggers

生物声学 | 8.0/10

 · 更新于 2026-09-24 · 约 10 分钟 · 4664 字 阅读 →
论文解读

H-nnPBFDAF: Hierarchical Neural Network Partitioned Block Frequency Domain Adaptive Filter with Novel Block Activation Probability

语音增强 | 7.5/10

 · 更新于 2026-09-24 · 约 10 分钟 · 4613 字 阅读 →
论文解读

HVAC-EAR: Eavesdropping Human Speech Using HVAC Systems

音频安全 | 8.5/10

 · 更新于 2026-09-24 · 约 11 分钟 · 5225 字 阅读 →