学习ASR途径:一种稀少的多语言ASR模式 (Learning ASR pathways: A sparse multilingual ASR model) - 专知论文

会员服务 ·

0

语音识别 · Learning · Performer · 剪枝 · MoDELS ·

2022 年 9 月 13 日

Learning ASR pathways: A sparse multilingual ASR model

翻译：学习ASR途径:一种稀少的多语言ASR模式

Mu Yang,Andros Tjandra,Chunxi Liu,David Zhang,Duc Le,John H. L. Hansen,Ozlem Kalinli

from arxiv, 5 pages, 3 figures

Neural network pruning can be effectively applied to compress automatic speech recognition (ASR) models. However, in multilingual ASR, performing language-agnostic pruning may lead to severe performance degradation on some languages because language-agnostic pruning masks may not fit all languages and discard important language-specific parameters. In this work, we present ASR pathways, a sparse multilingual ASR model that activates language-specific sub-networks ("pathways"), such that the parameters for each language are learned explicitly. With the overlapping sub-networks, the shared parameters can also enable knowledge transfer for lower resource languages via joint multilingual training. We propose a novel algorithm to learn ASR pathways, and evaluate the proposed method on 4 languages with a streaming RNN-T model. Our proposed ASR pathways outperform both dense models (-5.0% average WER) and a language-agnostically pruned model (-21.4% average WER), and provide better performance on low-resource languages compared to the monolingual sparse models.

翻译：神经网络运行可以有效地应用于压缩自动语音识别(ASR)模型。但是,在多语言的ASR中,进行语言-不可知性剪裁可能会导致某些语言的性能严重退化,因为语言-不可知性剪裁面罩可能不适应所有语言,并抛弃重要的语言特有参数。在这项工作中,我们介绍了ASR路径,即一种稀疏的多语种ASR模式,可以激活语言专用子网络(“路径”),这样可以明确了解每种语言的参数。在重叠的子网络中,共享参数还可以通过联合多语种培训为较低资源语言提供知识转让。我们提出了一种新式算法,学习ASR路径,并用流流式RNN-T模型评价4种语言的拟议方法。我们提议的ASR路径比密集模式(-5.0%平均WER)和语言-敏感型小网络模式(21.4%平均WER)都优于单一语言稀有模式,并且能够更好地表现低资源语言。

0

相关内容

语音识别

语音识别是计算机科学和计算语言学的一个跨学科子领域，它发展了一些方法和技术，使计算机可以将口语识别和翻译成文本。它也被称为自动语音识别（ASR），计算机语音识别或语音转文本（STT）。它整合了计算机科学，语言学和计算机工程领域的知识和研究。

【图机器学习进展与趋势@ICML2022】Graph Machine Learning @ ICML 2022

【图机器学习进展与趋势@ICML2022】Graph Machine Learning @ ICML 2022

专知会员服务

40+阅读 · 2022年7月25日

不可错过！《机器学习100讲》课程，UBC Mark Schmidt讲授

不可错过！《机器学习100讲》课程，UBC Mark Schmidt讲授

专知会员服务

75+阅读 · 2022年6月28日

纽约大学最新《语音识别Speech Recognition》2020课程，不可错过！

纽约大学最新《语音识别Speech Recognition》2020课程，不可错过！

专知会员服务

44+阅读 · 2020年11月2日

50+篇《神经架构搜索NAS》2020论文合集

专知会员服务

61+阅读 · 2020年3月19日

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

专知会员服务

165+阅读 · 2020年3月18日

【跨语言BERT模型大集合】Transfer learning is increasingly going multilingual with language-specific BERT models

专知会员服务

54+阅读 · 2020年1月30日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

开源书：PyTorch深度学习起步

开源书：PyTorch深度学习起步

专知会员服务

51+阅读 · 2019年10月11日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

RoBERTa for Chinese：大规模中文预训练RoBERTa模型

RoBERTa for Chinese：大规模中文预训练RoBERTa模型

AINLP

30+阅读 · 2019年9月8日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

ICLR2019最佳论文出炉

ICLR2019最佳论文出炉

专知

12+阅读 · 2019年5月6日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

【论文推荐】最新七篇图像分割相关论文—Attention U-Net、对抗结构匹配损失、卷积CRFs、对抗样本、弱监督分割

【论文推荐】最新七篇图像分割相关论文—Attention U-Net、对抗结构匹配损失、卷积CRFs、对抗样本、弱监督分割

专知

19+阅读 · 2018年5月31日

【论文推荐】最新5篇度量学习（Metric Learning）相关论文—人脸验证、BIER、自适应图卷积、注意力机制、单次学习

【论文推荐】最新5篇度量学习（Metric Learning）相关论文—人脸验证、BIER、自适应图卷积、注意力机制、单次学习

专知

17+阅读 · 2018年2月11日

【论文推荐】最新5篇语音识别（ASR）相关论文—音频对抗样本、对抗性语音识别系统、声学模型、序列到序列、口语可理解性矫正

【论文推荐】最新5篇语音识别（ASR）相关论文—音频对抗样本、对抗性语音识别系统、声学模型、序列到序列、口语可理解性矫正

专知

14+阅读 · 2018年2月4日

【论文】变分推断（Variational inference)的总结

【论文】变分推断（Variational inference)的总结

机器学习研究会

39+阅读 · 2017年11月16日

基于1,2,4-三唑-5-硫酮有机小分子的合成与光伏性能研究

国家自然科学基金

0+阅读 · 2015年12月31日

激光加载对NEA半导体光阴极激活层的破坏机理研究

国家自然科学基金

0+阅读 · 2014年12月31日

基于四嗪为受体单元共轭聚合物的设计、合成及光伏性能研究

国家自然科学基金

0+阅读 · 2013年12月31日

功能化石墨烯量子点合成与荧光传感

国家自然科学基金

0+阅读 · 2012年12月31日

基于噻吩并[3,4-c]吡咯[4,6]二酮的全受体共轭聚合物的合成及性能研究

国家自然科学基金

0+阅读 · 2012年12月31日

全外显子组测序确定下颌前突致病基因

国家自然科学基金

0+阅读 · 2011年12月31日

维吾尔语文本情感倾向性分析技术研究

国家自然科学基金

1+阅读 · 2009年12月31日

基于Sparse-Land模型的SAR图像噪声抑制与分割

国家自然科学基金

0+阅读 · 2009年12月31日

水溶性“#37329;—#27688;基酸”#32476;合物的合成及其应用基础研究

国家自然科学基金

0+阅读 · 2009年12月31日

Galectin-3对肝星状细胞激活及凋亡的影响

国家自然科学基金

0+阅读 · 2008年12月31日

Arxiv

0+阅读 · 2022年10月24日

Hyper-X: A Unified Hypernetwork for Multi-Task Multilingual Transfer

Arxiv

0+阅读 · 2022年10月24日

Few-shot Learning with Multilingual Language Models

Arxiv

0+阅读 · 2022年10月24日

Learning Vector-Quantized Item Representation for Transferable Sequential Recommenders

Arxiv

0+阅读 · 2022年10月22日

Audio-to-Intent Using Acoustic-Textual Subword Representations from End-to-End ASR

Audio-to-Intent Using Acoustic-Textual Subword Representations from End-to-End ASR

Arxiv

0+阅读 · 2022年10月21日

Maestro-U: Leveraging joint speech-text representation learning for zero supervised speech ASR

Maestro-U: Leveraging joint speech-text representation learning for zero supervised speech ASR

Arxiv

0+阅读 · 2022年10月21日

SMaLL-100: Introducing Shallow Multilingual Machine Translation Model for Low-Resource Languages

Arxiv

0+阅读 · 2022年10月20日

What Do Compressed Multilingual Machine Translation Models Forget?

Arxiv

0+阅读 · 2022年10月20日

MuRAG: Multimodal Retrieval-Augmented Generator for Open Question Answering over Images and Text

MuRAG: Multimodal Retrieval-Augmented Generator for Open Question Answering over Images and Text

Arxiv

1+阅读 · 2022年10月20日

A Survey of Model Compression and Acceleration for Deep Neural Networks

Arxiv

66+阅读 · 2019年9月8日

VIP会员

文章信息

相关主题

相关VIP内容

【图机器学习进展与趋势@ICML2022】Graph Machine Learning @ ICML 2022

【图机器学习进展与趋势@ICML2022】Graph Machine Learning @ ICML 2022

专知会员服务

40+阅读 · 2022年7月25日

不可错过！《机器学习100讲》课程，UBC Mark Schmidt讲授

不可错过！《机器学习100讲》课程，UBC Mark Schmidt讲授

专知会员服务

75+阅读 · 2022年6月28日

纽约大学最新《语音识别Speech Recognition》2020课程，不可错过！

纽约大学最新《语音识别Speech Recognition》2020课程，不可错过！

专知会员服务

44+阅读 · 2020年11月2日

50+篇《神经架构搜索NAS》2020论文合集

专知会员服务

61+阅读 · 2020年3月19日

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

专知会员服务

165+阅读 · 2020年3月18日

【跨语言BERT模型大集合】Transfer learning is increasingly going multilingual with language-specific BERT models

专知会员服务

54+阅读 · 2020年1月30日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

开源书：PyTorch深度学习起步

开源书：PyTorch深度学习起步

专知会员服务

51+阅读 · 2019年10月11日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

操作系统智能体：基于多模态大模型（MLLM）的通用计算设备智能体综述

《美国太空军系统全生命周期建模、仿真与分析效能提升方案》最新84页报告

【博士论文】推进数据高效的深度学习：非参数 Transformer、主动测试与上下文学习

自主人工智能：未来战争是否将是自主化的？

相关资讯

RoBERTa for Chinese：大规模中文预训练RoBERTa模型

RoBERTa for Chinese：大规模中文预训练RoBERTa模型

AINLP

30+阅读 · 2019年9月8日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

ICLR2019最佳论文出炉

ICLR2019最佳论文出炉

专知

12+阅读 · 2019年5月6日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

【论文推荐】最新七篇图像分割相关论文—Attention U-Net、对抗结构匹配损失、卷积CRFs、对抗样本、弱监督分割

【论文推荐】最新七篇图像分割相关论文—Attention U-Net、对抗结构匹配损失、卷积CRFs、对抗样本、弱监督分割

专知

19+阅读 · 2018年5月31日

【论文推荐】最新5篇度量学习（Metric Learning）相关论文—人脸验证、BIER、自适应图卷积、注意力机制、单次学习

【论文推荐】最新5篇度量学习（Metric Learning）相关论文—人脸验证、BIER、自适应图卷积、注意力机制、单次学习

专知

17+阅读 · 2018年2月11日

【论文推荐】最新5篇语音识别（ASR）相关论文—音频对抗样本、对抗性语音识别系统、声学模型、序列到序列、口语可理解性矫正

【论文推荐】最新5篇语音识别（ASR）相关论文—音频对抗样本、对抗性语音识别系统、声学模型、序列到序列、口语可理解性矫正

专知

14+阅读 · 2018年2月4日

【论文】变分推断（Variational inference)的总结

【论文】变分推断（Variational inference)的总结

机器学习研究会

39+阅读 · 2017年11月16日

相关论文

Arxiv

0+阅读 · 2022年10月24日

Hyper-X: A Unified Hypernetwork for Multi-Task Multilingual Transfer

Arxiv

0+阅读 · 2022年10月24日

Few-shot Learning with Multilingual Language Models

Arxiv

0+阅读 · 2022年10月24日

Learning Vector-Quantized Item Representation for Transferable Sequential Recommenders

Arxiv

0+阅读 · 2022年10月22日

Audio-to-Intent Using Acoustic-Textual Subword Representations from End-to-End ASR

Audio-to-Intent Using Acoustic-Textual Subword Representations from End-to-End ASR

Arxiv

0+阅读 · 2022年10月21日

Maestro-U: Leveraging joint speech-text representation learning for zero supervised speech ASR

Maestro-U: Leveraging joint speech-text representation learning for zero supervised speech ASR

Arxiv

0+阅读 · 2022年10月21日

SMaLL-100: Introducing Shallow Multilingual Machine Translation Model for Low-Resource Languages

Arxiv

0+阅读 · 2022年10月20日

What Do Compressed Multilingual Machine Translation Models Forget?

Arxiv

0+阅读 · 2022年10月20日

MuRAG: Multimodal Retrieval-Augmented Generator for Open Question Answering over Images and Text

MuRAG: Multimodal Retrieval-Augmented Generator for Open Question Answering over Images and Text

Arxiv

1+阅读 · 2022年10月20日

A Survey of Model Compression and Acceleration for Deep Neural Networks

Arxiv

66+阅读 · 2019年9月8日

相关基金

基于1,2,4-三唑-5-硫酮有机小分子的合成与光伏性能研究

国家自然科学基金

0+阅读 · 2015年12月31日

激光加载对NEA半导体光阴极激活层的破坏机理研究

国家自然科学基金

0+阅读 · 2014年12月31日

基于四嗪为受体单元共轭聚合物的设计、合成及光伏性能研究

国家自然科学基金

0+阅读 · 2013年12月31日

功能化石墨烯量子点合成与荧光传感

国家自然科学基金

0+阅读 · 2012年12月31日

基于噻吩并[3,4-c]吡咯[4,6]二酮的全受体共轭聚合物的合成及性能研究

国家自然科学基金

0+阅读 · 2012年12月31日

全外显子组测序确定下颌前突致病基因

国家自然科学基金

0+阅读 · 2011年12月31日

维吾尔语文本情感倾向性分析技术研究

国家自然科学基金

1+阅读 · 2009年12月31日

基于Sparse-Land模型的SAR图像噪声抑制与分割

国家自然科学基金

0+阅读 · 2009年12月31日

水溶性“#37329;—#27688;基酸”#32476;合物的合成及其应用基础研究

国家自然科学基金

0+阅读 · 2009年12月31日

Galectin-3对肝星状细胞激活及凋亡的影响

国家自然科学基金

0+阅读 · 2008年12月31日

微信扫码咨询专知VIP会员