EERRO: ESPnet 不受监督的ASR开放源码工具包 (EURO: ESPnet Unsupervised ASR Open-source Toolkit) - 专知论文

会员服务 ·

0

无监督 · 语音识别 · state-of-the-art · Extensibility · MoDELS ·

2022 年 11 月 30 日

EURO: ESPnet Unsupervised ASR Open-source Toolkit

翻译：EERRO: ESPnet 不受监督的ASR开放源码工具包

Dongji Gao,Jiatong Shi,Shun-Po Chuang,Leibny Paola Garcia,Hung-yi Lee,Shinji Watanabe,Sanjeev Khudanpur

This paper describes the ESPnet Unsupervised ASR Open-source Toolkit (EURO), an end-to-end open-source toolkit for unsupervised automatic speech recognition (UASR). EURO adopts the state-of-the-art UASR learning method introduced by the Wav2vec-U, originally implemented at FAIRSEQ, which leverages self-supervised speech representations and adversarial training. In addition to wav2vec2, EURO extends the functionality and promotes reproducibility for UASR tasks by integrating S3PRL and k2, resulting in flexible frontends from 27 self-supervised models and various graph-based decoding strategies. EURO is implemented in ESPnet and follows its unified pipeline to provide UASR recipes with a complete setup. This improves the pipeline's efficiency and allows EURO to be easily applied to existing datasets in ESPnet. Extensive experiments on three mainstream self-supervised models demonstrate the toolkit's effectiveness and achieve state-of-the-art UASR performance on TIMIT and LibriSpeech datasets. EURO will be publicly available at https://github.com/espnet/espnet, aiming to promote this exciting and emerging research area based on UASR through open-source activity.

翻译：本文介绍ESPnet 不受监督的ASR开放源码工具包(EURO),这是一个用于不受监督的自动语音识别的端到端开放源码工具包(UASR)。欧洲区域办事处采用Wav2vec-U采用的由Wav2vec-U采用的最新UASR学习方法,该方法最初由Wav2vec-U在FAIRSEQ实施,利用自我监督的语音陈述和对抗性培训。除了wav2vec2外,欧洲区域办事处还扩展了该功能,通过整合S3PRL和K2,促进对UASR任务的再传播。这导致27个自我监督模式和各种基于图表的解码战略的灵活前端。欧洲区域办事处在ESPnet网实施并遵循其统一的管道,向UASR食谱提供全套的UASR食谱。这提高了管道的效率,使EURO易于应用于ESPnet网中的现有数据集。关于三种主流自我监督模式的广泛实验展示了该工具包的有效性,并实现了在TIMIMEX和Listripest/Ex数据库的这一新出现的公共活动领域实现欧盟艺术USR的绩效。

0

相关内容

无监督

NeurlPS 2022 | 自然语言处理相关论文分类整理

NeurlPS 2022 | 自然语言处理相关论文分类整理

专知会员服务

51+阅读 · 2022年10月2日

史上最全！358篇机器学习&自然语言处理综述论文！都这儿了

专知会员服务

129+阅读 · 2020年7月18日

50+篇《神经架构搜索NAS》2020论文合集

专知会员服务

61+阅读 · 2020年3月19日

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

专知会员服务

165+阅读 · 2020年3月18日

CVPR 2020 论文开源项目合集

专知会员服务

110+阅读 · 2020年3月12日

【中科院自动化所】序列到序列语音识别的无监督预训练（Unsupervised pre-training for sequence to sequence speech recognition）

【中科院自动化所】序列到序列语音识别的无监督预训练（Unsupervised pre-training for sequence to sequence speech recognition）

专知会员服务

33+阅读 · 2020年1月5日

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

专知会员服务

160+阅读 · 2019年10月12日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium9

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium9

中国图象图形学学会CSIG

0+阅读 · 2021年12月17日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium3

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium3

中国图象图形学学会CSIG

0+阅读 · 2021年11月9日

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

专知

133+阅读 · 2020年3月18日

【Github】All4NLP：自然语言处理相关资源整理

【Github】All4NLP：自然语言处理相关资源整理

AINLP

23+阅读 · 2019年8月9日

BERT/Transformer/迁移学习NLP资源大列表

BERT/Transformer/迁移学习NLP资源大列表

专知

19+阅读 · 2019年6月9日

BERT/注意力机制/Transformer/迁移学习NLP资源大列表：awesome-bert-nlp

BERT/注意力机制/Transformer/迁移学习NLP资源大列表：awesome-bert-nlp

AINLP

40+阅读 · 2019年6月9日

强化学习的Unsupervised Meta-Learning

强化学习的Unsupervised Meta-Learning

CreateAMind

18+阅读 · 2019年1月7日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

【推荐】MXNet深度情感分析实战

【推荐】MXNet深度情感分析实战

机器学习研究会

16+阅读 · 2017年10月4日

CoFe2O4/BaSrTiO3复合势垒多铁隧道结的制备及隧穿特性研究

国家自然科学基金

0+阅读 · 2013年12月31日

基于主动微波遥感的土壤盐分定量反演研究

国家自然科学基金

0+阅读 · 2013年12月31日

HIPIMS制备环境自适应C/MoSx复合润滑薄膜的界面结构调控及摩擦机理

国家自然科学基金

0+阅读 · 2012年12月31日

基于Ontology的藏文语料库检索关键技术研究

国家自然科学基金

0+阅读 · 2012年12月31日

多组元共掺杂TiO2陶瓷晶界偏析及晶界势垒结构研究

国家自然科学基金

0+阅读 · 2012年12月31日

汉藏双语跨语言语音转换中的关键技术研究

国家自然科学基金

0+阅读 · 2012年12月31日

铬渣中Cr(VI)在土壤-地下水中的微界面过程与迁移机制研究

国家自然科学基金

0+阅读 · 2012年12月31日

基于树的句法翻译模型关键技术研究

国家自然科学基金

0+阅读 · 2012年12月31日

稀土基富集磷酸化肽的磁性核壳结构材料的合成与质谱检测

国家自然科学基金

0+阅读 · 2011年12月31日

JIP1通过蛋白结合互作调控tau蛋白磷酸化及神经纤维缠结的分子生物学机制

国家自然科学基金

0+阅读 · 2011年12月31日

TopoBERT: Plug and Play Toponym Recognition Module Harnessing Fine-tuned BERT

Arxiv

0+阅读 · 2023年1月31日

The Efficacy of Self-Supervised Speech Models for Audio Representations

Arxiv

0+阅读 · 2023年1月31日

Content-aware Warping for View Synthesis

Arxiv

0+阅读 · 2023年1月30日

Learning to Speak from Text: Zero-Shot Multilingual Text-to-Speech with Unsupervised Text Pretraining

Arxiv

0+阅读 · 2023年1月30日

Fast Correlation Function Calculator -- A high-performance pair counting toolkit

Arxiv

0+阅读 · 2023年1月29日

Pre-training for Speech Translation: CTC Meets Optimal Transport

Arxiv

0+阅读 · 2023年1月27日

Learning to Unlearn: Instance-wise Unlearning for Pre-trained Classifiers

Arxiv

0+阅读 · 2023年1月27日

Multi-Prompt Alignment for Multi-Source Unsupervised Domain Adaptation

Arxiv

0+阅读 · 2023年1月27日

Cross-Domain Adaptive Clustering for Semi-Supervised Domain Adaptation

Cross-Domain Adaptive Clustering for Semi-Supervised Domain Adaptation

Arxiv

19+阅读 · 2021年4月19日

Bridging the Gap Between Spectral and Spatial Domains in Graph Neural Networks

Bridging the Gap Between Spectral and Spatial Domains in Graph Neural Networks

Arxiv

15+阅读 · 2020年3月26日

VIP会员

文章信息

相关主题

state-of-the-art

相关VIP内容

NeurlPS 2022 | 自然语言处理相关论文分类整理

NeurlPS 2022 | 自然语言处理相关论文分类整理

专知会员服务

51+阅读 · 2022年10月2日

史上最全！358篇机器学习&自然语言处理综述论文！都这儿了

专知会员服务

129+阅读 · 2020年7月18日

50+篇《神经架构搜索NAS》2020论文合集

专知会员服务

61+阅读 · 2020年3月19日

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

专知会员服务

165+阅读 · 2020年3月18日

CVPR 2020 论文开源项目合集

专知会员服务

110+阅读 · 2020年3月12日

【中科院自动化所】序列到序列语音识别的无监督预训练（Unsupervised pre-training for sequence to sequence speech recognition）

【中科院自动化所】序列到序列语音识别的无监督预训练（Unsupervised pre-training for sequence to sequence speech recognition）

专知会员服务

33+阅读 · 2020年1月5日

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

专知会员服务

160+阅读 · 2019年10月12日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

《多智能体不确定环境追逃博弈研究》216页

美智库最新发布《解放军"人机编组协同作战"发展路径：理论与实践》53页

现代战争"杀伤区"理论：空间尺度与结构特征、控制手段与毁伤机制、生存策略与战线转移

《俄军无人机创新技术或已在乌克兰达成"战场空中封锁"作战效果》最新18页报告

相关资讯

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium9

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium9

中国图象图形学学会CSIG

0+阅读 · 2021年12月17日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium3

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium3

中国图象图形学学会CSIG

0+阅读 · 2021年11月9日

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

专知

133+阅读 · 2020年3月18日

【Github】All4NLP：自然语言处理相关资源整理

【Github】All4NLP：自然语言处理相关资源整理

AINLP

23+阅读 · 2019年8月9日

BERT/Transformer/迁移学习NLP资源大列表

BERT/Transformer/迁移学习NLP资源大列表

专知

19+阅读 · 2019年6月9日

BERT/注意力机制/Transformer/迁移学习NLP资源大列表：awesome-bert-nlp

BERT/注意力机制/Transformer/迁移学习NLP资源大列表：awesome-bert-nlp

AINLP

40+阅读 · 2019年6月9日

强化学习的Unsupervised Meta-Learning

强化学习的Unsupervised Meta-Learning

CreateAMind

18+阅读 · 2019年1月7日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

【推荐】MXNet深度情感分析实战

【推荐】MXNet深度情感分析实战

机器学习研究会

16+阅读 · 2017年10月4日

相关论文

TopoBERT: Plug and Play Toponym Recognition Module Harnessing Fine-tuned BERT

Arxiv

0+阅读 · 2023年1月31日

The Efficacy of Self-Supervised Speech Models for Audio Representations

Arxiv

0+阅读 · 2023年1月31日

Content-aware Warping for View Synthesis

Arxiv

0+阅读 · 2023年1月30日

Learning to Speak from Text: Zero-Shot Multilingual Text-to-Speech with Unsupervised Text Pretraining

Arxiv

0+阅读 · 2023年1月30日

Fast Correlation Function Calculator -- A high-performance pair counting toolkit

Arxiv

0+阅读 · 2023年1月29日

Pre-training for Speech Translation: CTC Meets Optimal Transport

Arxiv

0+阅读 · 2023年1月27日

Learning to Unlearn: Instance-wise Unlearning for Pre-trained Classifiers

Arxiv

0+阅读 · 2023年1月27日

Multi-Prompt Alignment for Multi-Source Unsupervised Domain Adaptation

Arxiv

0+阅读 · 2023年1月27日

Cross-Domain Adaptive Clustering for Semi-Supervised Domain Adaptation

Cross-Domain Adaptive Clustering for Semi-Supervised Domain Adaptation

Arxiv

19+阅读 · 2021年4月19日

Bridging the Gap Between Spectral and Spatial Domains in Graph Neural Networks

Bridging the Gap Between Spectral and Spatial Domains in Graph Neural Networks

Arxiv

15+阅读 · 2020年3月26日

相关基金

CoFe2O4/BaSrTiO3复合势垒多铁隧道结的制备及隧穿特性研究

国家自然科学基金

0+阅读 · 2013年12月31日

基于主动微波遥感的土壤盐分定量反演研究

国家自然科学基金

0+阅读 · 2013年12月31日

HIPIMS制备环境自适应C/MoSx复合润滑薄膜的界面结构调控及摩擦机理

国家自然科学基金

0+阅读 · 2012年12月31日

基于Ontology的藏文语料库检索关键技术研究

国家自然科学基金

0+阅读 · 2012年12月31日

多组元共掺杂TiO2陶瓷晶界偏析及晶界势垒结构研究

国家自然科学基金

0+阅读 · 2012年12月31日

汉藏双语跨语言语音转换中的关键技术研究

国家自然科学基金

0+阅读 · 2012年12月31日

铬渣中Cr(VI)在土壤-地下水中的微界面过程与迁移机制研究

国家自然科学基金

0+阅读 · 2012年12月31日

基于树的句法翻译模型关键技术研究

国家自然科学基金

0+阅读 · 2012年12月31日

稀土基富集磷酸化肽的磁性核壳结构材料的合成与质谱检测

国家自然科学基金

0+阅读 · 2011年12月31日

JIP1通过蛋白结合互作调控tau蛋白磷酸化及神经纤维缠结的分子生物学机制

国家自然科学基金

0+阅读 · 2011年12月31日

微信扫码咨询专知VIP会员