PESQNet (Loss) 是否需要干净的参考输入? 原始的 PESQ does, 但是 ACR 听觉测试 (Does a PESQNet (Loss) Require a Clean Reference Input? The Original PESQ Does, But ACR Listening Tests Don't) - 专知论文

会员服务 ·

0

DNS · 损失 · 原点 · INTERSPEECH · 循环神经网络 ·

2022 年 5 月 4 日

Does a PESQNet (Loss) Require a Clean Reference Input? The Original PESQ Does, But ACR Listening Tests Don't

翻译：PESQNet (Loss) 是否需要干净的参考输入? 原始的 PESQ does, 但是 ACR 听觉测试

Ziyi Xu,Maximilian Strake,Tim Fingscheidt

Perceptual evaluation of speech quality (PESQ) requires a clean speech reference as input, but predicts the results from (reference-free) absolute category rating (ACR) tests. In this work, we train a fully convolutional recurrent neural network (FCRN) as deep noise suppression (DNS) model, with either a non-intrusive or an intrusive PESQNet, where only the latter has access to a clean speech reference. The PESQNet is used as a mediator providing a perceptual loss during the DNS training to maximize the PESQ score of the enhanced speech signal. For the intrusive PESQNet, we investigate two topologies, called early-fusion (EF) and middle-fusion (MF) PESQNet, and compare to the non-intrusive PESQNet to evaluate and to quantify the benefits of employing a clean speech reference input during DNS training. Detailed analyses show that the DNS trained with the MF-intrusive PESQNet outperforms the Interspeech 2021 DNS Challenge baseline and the same DNS trained with an MSE loss by 0.23 and 0.12 PESQ points, respectively. Furthermore, we can show that only marginal benefits are obtained compared to the DNS trained with the non-intrusive PESQNet. Therefore, as ACR listening tests, the PESQNet does not necessarily require a clean speech reference input, opening the possibility of using real data for DNS training.

翻译：对语言质量(PESQ)的感知性评价(PESQ)要求将语言质量(PESQ)作为投入,但预测了(无参考)绝对等级(ACR)测试的结果。在这项工作中,我们将完全进化的经常性神经网络(FCRN)作为深噪抑制模式,使用非侵入性或侵扰性的PESQNet,只有后者才能获得清洁语音参考,只有后者才能获得清洁语音参考。PESQNet被用作在DNS培训期间提供感知性损失的调解人,以最大限度地提高PESQ强化语音信号的评分。在侵入性PESQNet中,我们调查了两种称为早期聚集(EF)和中聚(MF)PESQNet)的全演常态性神经网络(FCRN),我们用经过培训的MESPES 2021 DNS 挑战性基准和相同的DNSDNS(DNS),我们用经过培训的MSEAR 12 来评估清洁性测试的MESA 。

0

相关内容

DNS

域名系统（英文： Domain Name System， DNS）是因特网的一项核心服务，它作为可以将域名和IP地址相互映射的一个分布式数据库，能够使人更方便的访问互联网，而不用去记住能够被机器直接读取的IP数串。

【跨语言BERT模型大集合】Transfer learning is increasingly going multilingual with language-specific BERT models

专知会员服务

54+阅读 · 2020年1月30日

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

专知会员服务

19+阅读 · 2019年10月22日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

50+阅读 · 2019年10月17日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

专知会员服务

65+阅读 · 2019年10月9日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

ACM TOMM Call for Papers

ACM TOMM Call for Papers

CCF多媒体专委会

2+阅读 · 2022年3月23日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

中国图象图形学学会CSIG

0+阅读 · 2021年11月16日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium3

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium3

中国图象图形学学会CSIG

0+阅读 · 2021年11月9日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium2

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium2

中国图象图形学学会CSIG

0+阅读 · 2021年11月8日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium1

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium1

中国图象图形学学会CSIG

0+阅读 · 2021年11月3日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

【推荐】RNN/LSTM时序预测

【推荐】RNN/LSTM时序预测

机器学习研究会

25+阅读 · 2017年9月8日

ARB抑制miR-193a表达促进早期糖尿病肾病壁层上皮细胞-足细胞转分化研究

国家自然科学基金

0+阅读 · 2015年12月31日

猪肌内脂肪沉积过程中miRNA调控机理研究

国家自然科学基金

0+阅读 · 2014年12月31日

一种脉冲电子束物理气相沉积梯度MCrAlY包覆涂层组织结构演变及性能研究

国家自然科学基金

0+阅读 · 2013年12月31日

proBDNF通过P75NTR/sortilin受体促进心肌缺血再灌注损伤的作用及机制研究

国家自然科学基金

0+阅读 · 2012年12月31日

函数域中的Vinogradov中值定理

国家自然科学基金

0+阅读 · 2012年12月31日

Au1-xM'x/3DOM MOy (M' = Pd, Pt; M = Cr, Mn, Co, Fe)的可控制备及催化CO和VOC氧化的性能研究

国家自然科学基金

0+阅读 · 2012年12月31日

脂联素介导绵羊肌内、内脏和皮下脂肪细胞生脂的基因表达模式与分子调控机理研究

国家自然科学基金

0+阅读 · 2012年12月31日

大白菜抗根肿病基因CRb的分离克隆及其功能鉴定

国家自然科学基金

0+阅读 · 2011年12月31日

SLC22A3-Histamin-LDL途径介导冠心病的分子机制研究

国家自然科学基金

0+阅读 · 2011年12月31日

基于list-mode数据的快速SART真3D PET断层重建算法的研究

国家自然科学基金

0+阅读 · 2011年12月31日

Neural Networks as Paths through the Space of Representations

Arxiv

0+阅读 · 2022年6月22日

A Study on the Evaluation of Generative Models

Arxiv

0+阅读 · 2022年6月22日

Talk-to-Resolve: Combining scene understanding and spatial dialogue to resolve granular task ambiguity for a collocated robot

Arxiv

0+阅读 · 2022年6月22日

Shifted Compression Framework: Generalizations and Improvements

Arxiv

0+阅读 · 2022年6月21日

$Solving linear systems of the form $(A + γUU^T)\, {\bf x} = {\bf b}$ by preconditioned iterative methods$

Solving linear systems of the form $(A + γUU^T)\, {\bf x} = {\bf b}$ by preconditioned iterative methods

Arxiv

0+阅读 · 2022年6月21日

Discretization and index-robust error analysis for constrained high-index saddle dynamics on high-dimensional sphere

Arxiv

0+阅读 · 2022年6月21日

When Does Re-initialization Work?

Arxiv

0+阅读 · 2022年6月20日

PR-SZZ: How pull requests can support the tracing of defects in software repositories

Arxiv

0+阅读 · 2022年6月20日

Coverage Analysis of LEO Satellite Downlink Networks: Orbit Geometry Dependent Approach

Arxiv

0+阅读 · 2022年6月19日

CoSe-Co: Text Conditioned Generative CommonSense Contextualizer

Arxiv

0+阅读 · 2022年6月17日

VIP会员

文章信息

相关主题

循环神经网络

相关VIP内容

【跨语言BERT模型大集合】Transfer learning is increasingly going multilingual with language-specific BERT models

专知会员服务

54+阅读 · 2020年1月30日

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

专知会员服务

19+阅读 · 2019年10月22日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

50+阅读 · 2019年10月17日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

专知会员服务

65+阅读 · 2019年10月9日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

大语言模型智能体强化学习：全景综述

《城市滨海地区：理解复杂多变环境下的指挥控制框架》50页报告

【伯克利博士论文】从推理服务到训练：面向大规模 LLM 智能体的高效系统

美空军“顶点2025”实验：推进AI在C2、动态目标锁定与联盟集成中的应用

相关资讯

ACM TOMM Call for Papers

ACM TOMM Call for Papers

CCF多媒体专委会

2+阅读 · 2022年3月23日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

中国图象图形学学会CSIG

0+阅读 · 2021年11月16日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium3

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium3

中国图象图形学学会CSIG

0+阅读 · 2021年11月9日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium2

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium2

中国图象图形学学会CSIG

0+阅读 · 2021年11月8日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium1

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium1

中国图象图形学学会CSIG

0+阅读 · 2021年11月3日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

【推荐】RNN/LSTM时序预测

【推荐】RNN/LSTM时序预测

机器学习研究会

25+阅读 · 2017年9月8日

相关论文

Neural Networks as Paths through the Space of Representations

Arxiv

0+阅读 · 2022年6月22日

A Study on the Evaluation of Generative Models

Arxiv

0+阅读 · 2022年6月22日

Talk-to-Resolve: Combining scene understanding and spatial dialogue to resolve granular task ambiguity for a collocated robot

Arxiv

0+阅读 · 2022年6月22日

Shifted Compression Framework: Generalizations and Improvements

Arxiv

0+阅读 · 2022年6月21日

$Solving linear systems of the form $(A + γUU^T)\, {\bf x} = {\bf b}$ by preconditioned iterative methods$

Solving linear systems of the form $(A + γUU^T)\, {\bf x} = {\bf b}$ by preconditioned iterative methods

Arxiv

0+阅读 · 2022年6月21日

Discretization and index-robust error analysis for constrained high-index saddle dynamics on high-dimensional sphere

Arxiv

0+阅读 · 2022年6月21日

When Does Re-initialization Work?

Arxiv

0+阅读 · 2022年6月20日

PR-SZZ: How pull requests can support the tracing of defects in software repositories

Arxiv

0+阅读 · 2022年6月20日

Coverage Analysis of LEO Satellite Downlink Networks: Orbit Geometry Dependent Approach

Arxiv

0+阅读 · 2022年6月19日

CoSe-Co: Text Conditioned Generative CommonSense Contextualizer

Arxiv

0+阅读 · 2022年6月17日

相关基金

ARB抑制miR-193a表达促进早期糖尿病肾病壁层上皮细胞-足细胞转分化研究

国家自然科学基金

0+阅读 · 2015年12月31日

猪肌内脂肪沉积过程中miRNA调控机理研究

国家自然科学基金

0+阅读 · 2014年12月31日

一种脉冲电子束物理气相沉积梯度MCrAlY包覆涂层组织结构演变及性能研究

国家自然科学基金

0+阅读 · 2013年12月31日

proBDNF通过P75NTR/sortilin受体促进心肌缺血再灌注损伤的作用及机制研究

国家自然科学基金

0+阅读 · 2012年12月31日

函数域中的Vinogradov中值定理

国家自然科学基金

0+阅读 · 2012年12月31日

Au1-xM'x/3DOM MOy (M' = Pd, Pt; M = Cr, Mn, Co, Fe)的可控制备及催化CO和VOC氧化的性能研究

国家自然科学基金

0+阅读 · 2012年12月31日

脂联素介导绵羊肌内、内脏和皮下脂肪细胞生脂的基因表达模式与分子调控机理研究

国家自然科学基金

0+阅读 · 2012年12月31日

大白菜抗根肿病基因CRb的分离克隆及其功能鉴定

国家自然科学基金

0+阅读 · 2011年12月31日

SLC22A3-Histamin-LDL途径介导冠心病的分子机制研究

国家自然科学基金

0+阅读 · 2011年12月31日

基于list-mode数据的快速SART真3D PET断层重建算法的研究

国家自然科学基金

0+阅读 · 2011年12月31日

微信扫码咨询专知VIP会员