论文解读

Solving the Helmholtz Equation Via Physics-Informed Neural Networks with an Adaptive Weighting Strategy

声学建模 | 6.5/10

 · 更新于 2026-09-16 · 约 8 分钟 · 3874 字 阅读 →
论文解读

SONAR: Self-Distilled Continual Pre-Training for Domain Adaptive Audio Representation

音频事件检测 | 7.0/10

 · 更新于 2026-09-16 · 约 8 分钟 · 3901 字 阅读 →
论文解读

SoundCompass: Navigating Target Sound Extraction with Effective Directional Clue Integration in Complex Acoustic Scenes

语音分离 | 7.5/10

 · 更新于 2026-09-16 · 约 10 分钟 · 4540 字 阅读 →
论文解读

Sounding Highlights: Dual-Pathway Audio Encoders for Audio-Visual Video Highlight Detection

视频高光检测 | 8.5/10

 · 更新于 2026-09-16 · 约 13 分钟 · 6335 字 阅读 →
论文解读

Sounds that Shape: Audio-Driven 3D Mesh Generation with Attribute-Decoupled Score Distillation Sampling

音频生成 | 7.0/10

 · 更新于 2026-09-16 · 约 8 分钟 · 3847 字 阅读 →
论文解读

Source Separation For A Cappella Music

语音分离 | 6.5/10

 · 更新于 2026-09-16 · 约 9 分钟 · 4161 字 阅读 →
论文解读

SP-MCQA: Evaluating Intelligibility of TTS Beyond the Word Level

语音合成 | 7.0/10

 · 更新于 2026-09-16 · 约 9 分钟 · 4202 字 阅读 →
论文解读

SPADE: Structured Pruning and Adaptive Distillation for Efficient LLM-TTS

语音合成 | 7.5/10

 · 更新于 2026-09-16 · 约 10 分钟 · 4626 字 阅读 →
论文解读

SPAM: Style Prompt Adherence Metric for Prompt-Based TTS

语音合成 | 7.0/10

 · 更新于 2026-09-16 · 约 10 分钟 · 4628 字 阅读 →
论文解读

Sparse Autoencoders Make Audio Foundation Models More Explainable

模型评估 | 6.5/10

 · 更新于 2026-09-16 · 约 9 分钟 · 4339 字 阅读 →
论文解读

Sparse-View Visual-Acoustic Latent Learning for Novel-View Audio Synthesis

空间音频 | 7.5/10

 · 更新于 2026-09-16 · 约 11 分钟 · 5274 字 阅读 →
论文解读

Spatial Covariance Matrix Reconstruction for Speech Enhancement in Reverberant Multi-Source Environments

语音增强 | 7.5/10

 · 更新于 2026-09-16 · 约 10 分钟 · 4858 字 阅读 →
论文解读

Spatial-CLAP: Learning Spatially-Aware Audio–Text Embeddings for Multi-Source Conditions

空间音频 | 8.5/10

 · 更新于 2026-09-16 · 约 10 分钟 · 4905 字 阅读 →
论文解读

Spatially Aware Self-Supervised Models for Multi-Channel Neural Speaker Diarization

说话人分离 | 8.0/10

 · 更新于 2026-09-16 · 约 11 分钟 · 5315 字 阅读 →
论文解读

SpatialNet-Echo: Real-Time Acoustic Echo Cancellation via Integrated Narrow-Band and Cross-Band Processing

语音增强 | 7.5/10

 · 更新于 2026-09-16 · 约 10 分钟 · 4681 字 阅读 →
论文解读

Speaker Anonymisation for Speech-Based Suicide Risk Detection

语音匿名化 | 7.5/10

 · 更新于 2026-09-16 · 约 8 分钟 · 3992 字 阅读 →
论文解读

Speaking Clearly: A Simplified Whisper-Based Codec for Low-Bitrate Speech Coding

语音编码 | 7.5/10

 · 更新于 2026-09-16 · 约 10 分钟 · 4729 字 阅读 →
论文解读

Spectral or Spatial? Leveraging Both for Speaker Extraction in Challenging Data Conditions

语音分离 | 7.0/10

 · 更新于 2026-09-16 · 约 11 分钟 · 5058 字 阅读 →
论文解读

Spectrogram Event Based Feature Representation for Generalizable Automatic Music Transcription

音乐信息检索 | 7.5/10

 · 更新于 2026-09-16 · 约 10 分钟 · 5000 字 阅读 →
论文解读

Speech Emotion Recognition based on Hierarchical Transformer with Shifted Windows

语音情感识别 | 8.0/10

 · 更新于 2026-09-16 · 约 9 分钟 · 4324 字 阅读 →