查看文章

researchgate.net 中的 [PDF]

A context encoder for audio inpainting

作者

Andrés Marafioti, Nathanaël Perraudin, Nicki Holighaus, Piotr Majdak

发表日期

2019/10/14

期刊

IEEE/ACM Transactions on Audio, Speech, and Language Processing

卷号

期号

页码范围

2362-2372

出版商

IEEE

简介

In this article, we study the ability of deep neural networks (DNNs) to restore missing audio content based on its context, i.e., inpaint audio gaps. We focus on a condition which has not received much attention yet: gaps in the range of tens of milliseconds. We propose a DNN structure that is provided with the signal surrounding the gap in the form of time-frequency (TF) coefficients. Two DNNs with either complex-valued TF coefficient output or magnitude TF coefficient output were studied by separately training them on inpainting two types of audio signals (music and musical instruments) having 64-ms long gaps. The magnitude DNN outperformed the complex-valued DNN in terms of signal-to-noise ratios and objective difference grades. Although, for instruments, a reference inpainting obtained through linear predictive coding performed better in both metrics, it performed worse than the magnitude DNN for music …

引用总数

被引用次数：85

20182019202020212022202320241 5 16 17 13 20 11

学术搜索中的文章

A context encoder for audio inpainting

A Marafioti, N Perraudin, N Holighaus, P Majdak - IEEE/ACM Transactions on Audio, Speech, and …, 2019

被引用次数：85 相关文章所有 7 个版本