论文解读

Phase-Space Signal Processing of Acoustic Data for Advanced Manufacturing In-Situ Monitoring

音频事件检测 | 7.0/10

 · 更新于 2026-09-25 · 约 7 分钟 · 3474 字 阅读 →
论文解读

RASD-SR: A Robust Anomalous Sound Detection Framework with Score Recalibration

异常声音检测 | 8.5/10

 · 更新于 2026-09-25 · 约 11 分钟 · 5155 字 阅读 →
论文解读

Refgen: Reference-Guided Synthetic Data Generation for Anomalous Sound Detection

音频事件检测 | 7.5/10

 · 更新于 2026-09-25 · 约 8 分钟 · 3682 字 阅读 →
论文解读

Representation-Based Data Quality Audits for Audio

数据集 | 7.5/10

 · 更新于 2026-09-25 · 约 9 分钟 · 4054 字 阅读 →
论文解读

SELD-MOHA: A Fine-Tuning Method with the Mixture of Heterogeneous Adapters for Sound Event Localization and Detection

音频事件检测 | 7.0/10

 · 更新于 2026-09-25 · 约 11 分钟 · 5325 字 阅读 →
论文解读

Shared Representation Learning for Reference-Guided Targeted Sound Detection

音频事件检测 | 8.5/10

 · 更新于 2026-09-25 · 约 10 分钟 · 4764 字 阅读 →
论文解读

SONAR: Self-Distilled Continual Pre-Training for Domain Adaptive Audio Representation

音频事件检测 | 7.0/10

 · 更新于 2026-09-25 · 约 8 分钟 · 3901 字 阅读 →
论文解读

Task-Oriented Sound Privacy Preservation for Sound Event Detection Via End-to-End Adversarial Multi-Task Learning

音频事件检测 | 7.5/10

 · 更新于 2026-09-25 · 约 12 分钟 · 5739 字 阅读 →
论文解读

Temporally Heterogeneous Graph Contrastive Learning for Multimodal Acoustic Event Classification

音频事件检测 | 8.5/10

 · 更新于 2026-09-25 · 约 8 分钟 · 3902 字 阅读 →
论文解读

Tldiffgan: A Latent Diffusion-Gan Framework with Temporal Information Fusion for Anomalous Sound Detection

音频事件检测 | 7.5/10

 · 更新于 2026-09-25 · 约 10 分钟 · 4622 字 阅读 →
论文解读

Toward Faithful Explanations in Acoustic Anomaly Detection

音频事件检测 | 7.5/10

 · 更新于 2026-09-25 · 约 8 分钟 · 3867 字 阅读 →
论文解读

Triad: Tri-Head with Auxiliary Duplicating Permutation Invariant Training for Multi-Task Sound Event Localization and Detection

音频事件检测 | 7.5/10

 · 更新于 2026-09-25 · 约 9 分钟 · 4020 字 阅读 →
论文解读

USVexplorer: Robust Detection of Ultrasonic Vocalizations with Cross Species Generalization

音频事件检测 | 8.0/10

 · 更新于 2026-09-25 · 约 12 分钟 · 5686 字 阅读 →
论文解读

Earable Platform with Integrated Simultaneous EEG Sensing and Auditory Stimulation

音频事件检测 | 5.5/10

 · 更新于 2026-09-25 · 约 6 分钟 · 2648 字 阅读 →
论文解读

Disentangling Damage from Operational Variability: A Label-Free Self-Supervised Representation Learning Framework for Output-Only Structural Damage Identification

本文针对结构健康监测中损伤信号易被环境与操作变异掩盖的核心挑战,提出了一种**无标签、自监督的解缠表示学习框架**。该框架采用双流自编码器架构,通过**时间序列重构损失**确保信息完整性,并利用**VICReg自监督损失**(基于假设损伤状态不变的基线期数据)强制损伤敏感表征(`z_dmg`)对操作

 · 更新于 2026-09-25 · 约 11 分钟 · 5106 字 阅读 →
论文解读

Sky-Ear: An Unmanned Aerial Vehicle-Enabled Victim Sound Detection and Localization System

本文针对无人机搜救任务中视觉系统受遮蔽、能耗高的问题,提出了一个名为“Sky-Ear”的音频驱动受害者检测与定位系统。核心方法是设计了一个基于环形麦克风阵列的两阶段处理框架:在“哨兵阶段”,系统利用单

 · 更新于 2026-09-25 · 约 13 分钟 · 6117 字 阅读 →
论文解读

SpotSound: Enhancing Large Audio-Language Models with Fine-Grained Temporal Grounding

本文旨在解决大型音频语言模型在**细粒度音频事件时间定位**上的不足。现有模型因训练数据缺乏精确时间戳、基准测试过于简单,导致在长音频中定位短暂事件(“大海捞针”)时表现不可靠。为此,作者提出了**S

 · 更新于 2026-09-25 · 约 11 分钟 · 5348 字 阅读 →
论文解读

Towards Fine-grained Temporal Perception: Post-Training Large Audio-Language Models with Audio-Side Time Prompt

这篇论文旨在解决大型音频语言模型(LALM)在细粒度时间感知(如精确定位声音事件的起止时间)上的不足。作者提出了**TimePro-RL**框架,其核心是两步走策略:首先,提出**音频侧时间提示(AS

 · 更新于 2026-09-25 · 约 10 分钟 · 4592 字 阅读 →
论文解读

Transformer Based Machine Fault Detection From Audio Input

本文旨在探讨基于Transformer的架构在机器故障音频检测任务上相对于传统卷积神经网络(CNN)的潜在优势。**要解决的问题**是传统CNN在处理频谱图时固有的局部性和平移不变性等归纳偏置,可能并

 · 更新于 2026-09-25 · 约 6 分钟 · 2593 字 阅读 →