语音识别的深碎分解前导器 (Deep Sparse Conformer for Speech Recognition) - 专知论文

会员服务 ·

0

Conformer · 稀疏 · 残差连接 · 语音识别 · 层 ·

2022 年 9 月 1 日

Deep Sparse Conformer for Speech Recognition

翻译：语音识别的深碎分解前导器

from arxiv, 5 pages, 1 figure

Conformer has achieved impressive results in Automatic Speech Recognition (ASR) by leveraging transformer's capturing of content-based global interactions and convolutional neural network's exploiting of local features. In Conformer, two macaron-like feed-forward layers with half-step residual connections sandwich the multi-head self-attention and convolution modules followed by a post layer normalization. We improve Conformer's long-sequence representation ability in two directions, \emph{sparser} and \emph{deeper}. We adapt a sparse self-attention mechanism with $\mathcal{O}(L\text{log}L)$ in time complexity and memory usage. A deep normalization strategy is utilized when performing residual connections to ensure our training of hundred-level Conformer blocks. On the Japanese CSJ-500h dataset, this deep sparse Conformer achieves respectively CERs of 5.52\%, 4.03\% and 4.50\% on the three evaluation sets and 4.16\%, 2.84\% and 3.20\% when ensembling five deep sparse Conformer variants from 12 to 16, 17, 50, and finally 100 encoder layers.

翻译：通过利用变压器捕捉基于内容的全球互动和进化神经网络利用当地特点,自动语音识别取得了令人印象深刻的成果。在变压器中,两层马卡龙式的进化向前层,配有半阶段剩余连接,配有多头自留和进化模块,后加一层正常化。我们提高了变压器在两个方向,即\emph{sparser} 和\emph{diter}的长期序列代表能力。我们用时间复杂性和记忆用美元调整一个稀少的自留机制(L\text{log}L),在进行剩余连接时使用了一种深度的正常化战略,以确保我们培训100级的组合块。在日本的CSJ-500h数据集中,这种深度稀薄的组合在三个评价组和4.16{O}、2.84}和3.20之间分别实现了5.52 ⁇ 、4.03 ⁇ 和4.50的核证的排减量。当将五种深度分散的变式从12至16级、17级、50级和50级。

0

相关内容

Conformer

NeurlPS 2022 | 自然语言处理相关论文分类整理

NeurlPS 2022 | 自然语言处理相关论文分类整理

专知会员服务

51+阅读 · 2022年10月2日

纽约大学最新《语音识别Speech Recognition》2020课程，不可错过！

纽约大学最新《语音识别Speech Recognition》2020课程，不可错过！

专知会员服务

44+阅读 · 2020年11月2日

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

专知会员服务

166+阅读 · 2020年3月18日

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

专知会员服务

95+阅读 · 2020年3月12日

【深度学习架构、模型和技巧集合(TensorFlow/PyTorch)】’Deep Learning Models - A collection of various deep learning architectures, models, and tips'

【深度学习架构、模型和技巧集合(TensorFlow/PyTorch)】’Deep Learning Models - A collection of various deep learning architectures, models, and tips'

专知会员服务

58+阅读 · 2020年1月25日

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

专知会员服务

19+阅读 · 2019年10月22日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日

VCIP 2022 Call for Special Session Proposals

VCIP 2022 Call for Special Session Proposals

CCF多媒体专委会

1+阅读 · 2022年4月1日

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

ICLR2019最佳论文出炉

ICLR2019最佳论文出炉

专知

12+阅读 · 2019年5月6日

disentangled-representation-papers

disentangled-representation-papers

CreateAMind

26+阅读 · 2018年9月12日

Capsule Networks解析

Capsule Networks解析

机器学习研究会

11+阅读 · 2017年11月12日

可解释的CNN

可解释的CNN

CreateAMind

17+阅读 · 2017年10月5日

【推荐】深度学习目标检测全面综述

【推荐】深度学习目标检测全面综述

机器学习研究会

21+阅读 · 2017年9月13日

【推荐】深度学习目标检测概览

【推荐】深度学习目标检测概览

机器学习研究会

10+阅读 · 2017年9月1日

TMS1基因响应高温胁迫和ER Stress的分子机制

国家自然科学基金

0+阅读 · 2014年12月31日

黄河中游地区突发性大暴雨MCC结构特征研究

国家自然科学基金

0+阅读 · 2014年12月31日

微纳尺度混合填充增强复合材料热导率的协同效应研究

国家自然科学基金

0+阅读 · 2013年12月31日

脉冲电流条件下纳米晶金属箔材超塑性微成形机理与尺寸效应

国家自然科学基金

0+阅读 · 2013年12月31日

基于 HSS 迭代方法的加性 Schwarz 算法

国家自然科学基金

0+阅读 · 2013年12月31日

废水处理用类荷叶表面微纳米结构超疏水微孔膜的制备及膜蒸馏机理研究

国家自然科学基金

0+阅读 · 2012年12月31日

中红外波段石墨烯锁模级联脉冲光纤激光器的机制分析与实验研究

国家自然科学基金

0+阅读 · 2012年12月31日

粤西海域CTW（Coastal Trapped Wave）特征分析与数值模拟研究

国家自然科学基金

0+阅读 · 2009年12月31日

基于本体的Deep Web搜索技术

国家自然科学基金

2+阅读 · 2009年12月31日

强脉冲光照射对人类皮肤光老化的影响研究

国家自然科学基金

0+阅读 · 2008年12月31日

Anti-Symmetric DGN: a stable architecture for Deep Graph Networks

Arxiv

0+阅读 · 2022年10月18日

Object Recognition in Different Lighting Conditions at Various Angles by Deep Learning Method

Arxiv

0+阅读 · 2022年10月18日

Multi-Level Modeling Units for End-to-End Mandarin Speech Recognition

Arxiv

0+阅读 · 2022年10月18日

LR-Net: A Block-based Convolutional Neural Network for Low-Resolution Image Classification

Arxiv

0+阅读 · 2022年10月17日

Squeezeformer: An Efficient Transformer for Automatic Speech Recognition

Arxiv

0+阅读 · 2022年10月15日

Mechanical features based object recognition

Arxiv

0+阅读 · 2022年10月14日

STAR-Transformer: A Spatio-temporal Cross Attention Transformer for Human Action Recognition

Arxiv

0+阅读 · 2022年10月14日

Look-into-Object: Self-supervised Structure Modeling for Object Recognition

Look-into-Object: Self-supervised Structure Modeling for Object Recognition

Arxiv

15+阅读 · 2020年3月31日

Meta Learning for End-to-End Low-Resource Speech Recognition

Meta Learning for End-to-End Low-Resource Speech Recognition

Arxiv

20+阅读 · 2019年10月26日

A Survey on Deep Learning for Named Entity Recognition

A Survey on Deep Learning for Named Entity Recognition

Arxiv

73+阅读 · 2018年12月22日

VIP会员

文章信息

相关主题

相关VIP内容

NeurlPS 2022 | 自然语言处理相关论文分类整理

NeurlPS 2022 | 自然语言处理相关论文分类整理

专知会员服务

51+阅读 · 2022年10月2日

纽约大学最新《语音识别Speech Recognition》2020课程，不可错过！

纽约大学最新《语音识别Speech Recognition》2020课程，不可错过！

专知会员服务

44+阅读 · 2020年11月2日

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

专知会员服务

166+阅读 · 2020年3月18日

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

专知会员服务

95+阅读 · 2020年3月12日

【深度学习架构、模型和技巧集合(TensorFlow/PyTorch)】’Deep Learning Models - A collection of various deep learning architectures, models, and tips'

【深度学习架构、模型和技巧集合(TensorFlow/PyTorch)】’Deep Learning Models - A collection of various deep learning architectures, models, and tips'

专知会员服务

58+阅读 · 2020年1月25日

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

专知会员服务

19+阅读 · 2019年10月22日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

【CMU博士论文】数据驱动决策中的激励、信息与不确定性

DGP双粒度提示框架：图增强大模型助力欺诈检测

【ICCV2025】ESSENTIAL：用于视频类增量学习的情景记忆与语义记忆整合

唯快不破：大型语言模型高效架构综述

相关资讯

VCIP 2022 Call for Special Session Proposals

VCIP 2022 Call for Special Session Proposals

CCF多媒体专委会

1+阅读 · 2022年4月1日

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

ICLR2019最佳论文出炉

ICLR2019最佳论文出炉

专知

12+阅读 · 2019年5月6日

disentangled-representation-papers

disentangled-representation-papers

CreateAMind

26+阅读 · 2018年9月12日

Capsule Networks解析

Capsule Networks解析

机器学习研究会

11+阅读 · 2017年11月12日

可解释的CNN

可解释的CNN

CreateAMind

17+阅读 · 2017年10月5日

【推荐】深度学习目标检测全面综述

【推荐】深度学习目标检测全面综述

机器学习研究会

21+阅读 · 2017年9月13日

【推荐】深度学习目标检测概览

【推荐】深度学习目标检测概览

机器学习研究会

10+阅读 · 2017年9月1日

相关论文

Anti-Symmetric DGN: a stable architecture for Deep Graph Networks

Arxiv

0+阅读 · 2022年10月18日

Object Recognition in Different Lighting Conditions at Various Angles by Deep Learning Method

Arxiv

0+阅读 · 2022年10月18日

Multi-Level Modeling Units for End-to-End Mandarin Speech Recognition

Arxiv

0+阅读 · 2022年10月18日

LR-Net: A Block-based Convolutional Neural Network for Low-Resolution Image Classification

Arxiv

0+阅读 · 2022年10月17日

Squeezeformer: An Efficient Transformer for Automatic Speech Recognition

Arxiv

0+阅读 · 2022年10月15日

Mechanical features based object recognition

Arxiv

0+阅读 · 2022年10月14日

STAR-Transformer: A Spatio-temporal Cross Attention Transformer for Human Action Recognition

Arxiv

0+阅读 · 2022年10月14日

Look-into-Object: Self-supervised Structure Modeling for Object Recognition

Look-into-Object: Self-supervised Structure Modeling for Object Recognition

Arxiv

15+阅读 · 2020年3月31日

Meta Learning for End-to-End Low-Resource Speech Recognition

Meta Learning for End-to-End Low-Resource Speech Recognition

Arxiv

20+阅读 · 2019年10月26日

A Survey on Deep Learning for Named Entity Recognition

A Survey on Deep Learning for Named Entity Recognition

Arxiv

73+阅读 · 2018年12月22日

相关基金

TMS1基因响应高温胁迫和ER Stress的分子机制

国家自然科学基金

0+阅读 · 2014年12月31日

黄河中游地区突发性大暴雨MCC结构特征研究

国家自然科学基金

0+阅读 · 2014年12月31日

微纳尺度混合填充增强复合材料热导率的协同效应研究

国家自然科学基金

0+阅读 · 2013年12月31日

脉冲电流条件下纳米晶金属箔材超塑性微成形机理与尺寸效应

国家自然科学基金

0+阅读 · 2013年12月31日

基于 HSS 迭代方法的加性 Schwarz 算法

国家自然科学基金

0+阅读 · 2013年12月31日

废水处理用类荷叶表面微纳米结构超疏水微孔膜的制备及膜蒸馏机理研究

国家自然科学基金

0+阅读 · 2012年12月31日

中红外波段石墨烯锁模级联脉冲光纤激光器的机制分析与实验研究

国家自然科学基金

0+阅读 · 2012年12月31日

粤西海域CTW（Coastal Trapped Wave）特征分析与数值模拟研究

国家自然科学基金

0+阅读 · 2009年12月31日

基于本体的Deep Web搜索技术

国家自然科学基金

2+阅读 · 2009年12月31日

强脉冲光照射对人类皮肤光老化的影响研究

国家自然科学基金

0+阅读 · 2008年12月31日

微信扫码咨询专知VIP会员