论文解读

A Text-Steerable Instrument for Sketching Procedural Soundscapes via Language Models

音乐生成 | 6.8/10

 · 更新于 2026-09-25 · 约 11 分钟 · 5326 字 阅读 →
论文解读

LeVo 2: Stable and Melodious Song Generation via Hierarchical Representation Modeling and Progressive Post-Training

音乐生成 | 9.4/10

 · 更新于 2026-09-25 · 约 4 分钟 · 1658 字 阅读 →
论文解读

Generative AI and Copyright Infringement: A Legal-Technical Analysis of AI Music Generation Systems Under 17 U.S.C. Title 17

音乐生成 | 6.0/10

 · 更新于 2026-09-25 · 约 10 分钟 · 4812 字 阅读 →
论文解读

Attractive and Repulsive Pattern Control in Sequence Generation

音乐生成 | 8.1/10

 · 更新于 2026-09-25 · 约 13 分钟 · 6487 字 阅读 →
论文解读

Digital Revival: Acoustic Documentation and Digital Reactivation of Historical Woodwind Instruments

音乐生成 | 5.3/10

 · 更新于 2026-09-25 · 约 11 分钟 · 5086 字 阅读 →
论文解读

AudioCALM: Continuous Autoregressive Language Modeling for Universal Audio Generation

语音合成 | 7.9/10

 · 更新于 2026-09-25 · 约 15 分钟 · 7085 字 阅读 →
论文解读

Improving Text-to-Music Generation with Human Preference Rewards

音乐生成 | 8.5/10

 · 更新于 2026-09-25 · 约 13 分钟 · 6462 字 阅读 →
论文解读

Libretto: Giving LLM Agents a Sense of Musical Structure

音乐生成 | 9.2/10

 · 更新于 2026-09-25 · 约 14 分钟 · 6981 字 阅读 →
论文解读

LK Jam: System Architecture and Implementation of a Real-Time Human-AI Interactive Music Generation System using Role-Aware GRU

音乐生成 | 7.0/10

 · 更新于 2026-09-25 · 约 15 分钟 · 7390 字 阅读 →
论文解读

Co-policy: Responsive Human-Robot Co-Creation for Musical Performances

音乐生成 | 8.5/10

 · 更新于 2026-09-25 · 约 14 分钟 · 6648 字 阅读 →
每日研究速递

语音/音乐/音频论文速递 2026-06-22

共分析 1 篇语音/AI 论文

 · 更新于 2026-09-25 · 约 4 分钟 · 1684 字 阅读 →
论文解读

Closing the Loop: PID Feedback Control for Interpretable Activation Steering in Symbolic Music Generation

音乐生成 | 8.7/10

 · 更新于 2026-09-25 · 约 11 分钟 · 5239 字 阅读 →
论文解读

AudioX-Turbo: A Unified Framework for Efficient Anything-to-Audio Generation

音频生成 | 9/10

 · 更新于 2026-09-25 · 约 14 分钟 · 6874 字 阅读 →
论文解读

Generative Modeling of Bach-Style Symbolic Music: A Comparative Study of Autoregressive, Latent-Variable, and Adversarial Approaches

音乐生成 | 5.7/10

 · 更新于 2026-09-25 · 约 10 分钟 · 4808 字 阅读 →
论文解读

PianoKontext: Expressive Performance Rendering from Deadpan Context

音乐生成 | 9.1/10

 · 更新于 2026-09-25 · 约 9 分钟 · 4271 字 阅读 →
论文解读

Can LLMs understand LilyPond? A benchmark for symbolic music generation and understanding

音乐生成 | 7/10

 · 更新于 2026-09-25 · 约 11 分钟 · 5101 字 阅读 →
论文解读

Entropy as a Structural Prior: How a Log-Barrier on DiT Belief Space Drives Musical Diversity and Development

音乐生成 | 4.2/10

 · 更新于 2026-09-25 · 约 16 分钟 · 7818 字 阅读 →
论文解读

Exploring LLMs for South Asian Music Understanding and Generation

音乐生成 | 7.7/10

 · 更新于 2026-09-25 · 约 11 分钟 · 5025 字 阅读 →
论文解读

Foley-Omni: A Unified Multimodal Generation Model from Task-Level Audio Synthesis to Complete Video Soundtrack Generation

音频生成 | 7/10

 · 更新于 2026-09-25 · 约 14 分钟 · 6522 字 阅读 →
论文解读

SegTune: Structured and Fine-Grained Control for Song Generation

音乐生成 | 8.5/10

 · 更新于 2026-09-25 · 约 14 分钟 · 6788 字 阅读 →