使用端对端语音翻译分部分两种语言发言单位优化发言分部分 (Speech Segmentation Optimization using Segmented Bilingual Speech Corpus for End-to-end Speech Translation) - 专知论文

会员服务 ·

0

语音翻译 · 端到端 · 优化器 · WebRTC · 二分类 ·

2022 年 7 月 13 日

Speech Segmentation Optimization using Segmented Bilingual Speech Corpus for End-to-end Speech Translation

翻译：使用端对端语音翻译分部分两种语言发言单位优化发言分部分

Ryo Fukuda,Katsuhito Sudoh,Satoshi Nakamura

from arxiv, Accepted to INTERSPEECH 2022

Speech segmentation, which splits long speech into short segments, is essential for speech translation (ST). Popular VAD tools like WebRTC VAD have generally relied on pause-based segmentation. Unfortunately, pauses in speech do not necessarily match sentence boundaries, and sentences can be connected by a very short pause that is difficult to detect by VAD. In this study, we propose a speech segmentation method using a binary classification model trained using a segmented bilingual speech corpus. We also propose a hybrid method that combines VAD and the above speech segmentation method. Experimental results revealed that the proposed method is more suitable for cascade and end-to-end ST systems than conventional segmentation methods. The hybrid approach further improved the translation performance.

翻译：将长篇发言分成短篇部分,对于语言翻译至关重要。WebRTC VAD等流行性VAD工具一般都依赖暂停式分割。不幸的是,暂停语句不一定与句号界限相符,而句子可以通过极短的暂停连接,而这种暂停很难被VAD发现。在本研究报告中,我们建议使用使用使用分段双语语言材料培训的二进制分类模式来使用语言分割法。我们还提议一种混合法,将VAD和上述语言分割法结合起来。实验结果表明,拟议的方法比常规分割法更适合级联和端至端ST系统。混合法进一步提高了翻译的性能。

0

相关内容

语音翻译

通过计算机进行不同语言之间的直接语音翻译，辅助不同语言背景的人们进行沟通已经成为世界各国研究的重点。和一般的文本翻译不同，语音翻译需要把语音识别、机器翻译和语音合成三大技术进行集成，具有很大的挑战性。

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

专知会员服务

60+阅读 · 2022年4月22日

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

专知会员服务

96+阅读 · 2020年3月12日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

50+阅读 · 2019年10月17日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

VCIP 2022 Call for Special Session Proposals

VCIP 2022 Call for Special Session Proposals

CCF多媒体专委会

1+阅读 · 2022年4月1日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

中国图象图形学学会CSIG

0+阅读 · 2021年11月16日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium4

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium4

中国图象图形学学会CSIG

0+阅读 · 2021年11月10日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

【推荐】全卷积语义分割综述

【推荐】全卷积语义分割综述

机器学习研究会

19+阅读 · 2017年8月31日

考虑天气过程随机性的风电场群概率预测及系统优化调度方法

国家自然科学基金

0+阅读 · 2014年12月31日

禾谷镰孢菌Fusarium graminearum CYP51与DMIs类杀菌剂结合的分子机理研究

国家自然科学基金

0+阅读 · 2013年12月31日

薄势垒增强型AlGaN/GaN HEMT及可靠性研究

国家自然科学基金

0+阅读 · 2013年12月31日

超临界二氧化碳钻井井筒变质量流动的相态控制理论与方法研究

国家自然科学基金

0+阅读 · 2013年12月31日

退化抛物方程的可控性

国家自然科学基金

0+阅读 · 2013年12月31日

A repository of automatic GUI test patterns in Android applications: Specification and Analysis using Alloy modeling language

Arxiv

0+阅读 · 2022年9月5日

Multimodal Neural Machine Translation with Search Engine Based Image Retrieval

Arxiv

0+阅读 · 2022年9月3日

Exploiting Pretrained Biochemical Language Models for Targeted Drug Design

Arxiv

0+阅读 · 2022年9月2日

UniInst: Unique Representation for End-to-End Instance Segmentation

Arxiv

0+阅读 · 2022年9月2日

Generative Adversarial Networks and Probabilistic Graph Models for Hyperspectral Image Classification

Arxiv

11+阅读 · 2018年2月10日

VIP会员

文章信息

相关主题

相关VIP内容

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

专知会员服务

60+阅读 · 2022年4月22日

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

专知会员服务

96+阅读 · 2020年3月12日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

50+阅读 · 2019年10月17日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

【博士论文】面向真实世界音视联合语音识别的可扩展框架

《通过仿真与开源数据提升战略决策：机遇与局限》最新报告

【AAAI2026】善始则事半功倍：基于前缀优化的大语言模型推理强化学习

评估大语言模型在科学发现中的作用

相关资讯

VCIP 2022 Call for Special Session Proposals

VCIP 2022 Call for Special Session Proposals

CCF多媒体专委会

1+阅读 · 2022年4月1日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

中国图象图形学学会CSIG

0+阅读 · 2021年11月16日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium4

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium4

中国图象图形学学会CSIG

0+阅读 · 2021年11月10日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

【推荐】全卷积语义分割综述

【推荐】全卷积语义分割综述

机器学习研究会

19+阅读 · 2017年8月31日

相关论文

A repository of automatic GUI test patterns in Android applications: Specification and Analysis using Alloy modeling language

Arxiv

0+阅读 · 2022年9月5日

Multimodal Neural Machine Translation with Search Engine Based Image Retrieval

Arxiv

0+阅读 · 2022年9月3日

Exploiting Pretrained Biochemical Language Models for Targeted Drug Design

Arxiv

0+阅读 · 2022年9月2日

UniInst: Unique Representation for End-to-End Instance Segmentation

Arxiv

0+阅读 · 2022年9月2日

Generative Adversarial Networks and Probabilistic Graph Models for Hyperspectral Image Classification

Arxiv

11+阅读 · 2018年2月10日

相关基金

考虑天气过程随机性的风电场群概率预测及系统优化调度方法

国家自然科学基金

0+阅读 · 2014年12月31日

禾谷镰孢菌Fusarium graminearum CYP51与DMIs类杀菌剂结合的分子机理研究

国家自然科学基金

0+阅读 · 2013年12月31日

薄势垒增强型AlGaN/GaN HEMT及可靠性研究

国家自然科学基金

0+阅读 · 2013年12月31日

超临界二氧化碳钻井井筒变质量流动的相态控制理论与方法研究

国家自然科学基金

0+阅读 · 2013年12月31日

退化抛物方程的可控性

国家自然科学基金

0+阅读 · 2013年12月31日

微信扫码咨询专知VIP会员