交叉的语音识别重新评分方法：基于图形的标签传播 (Cross-utterance ASR Rescoring with Graph-based Label Propagation) - 专知论文

会员服务 ·

0

语音识别 · 标签传播 · 基于图形 · 神经语言模型 · 识别 ·

2023 年 3 月 27 日

Cross-utterance ASR Rescoring with Graph-based Label Propagation

翻译：交叉的语音识别重新评分方法：基于图形的标签传播

Srinath Tankasala,Long Chen,Andreas Stolcke,Anirudh Raju,Qianli Deng,Chander Chandak,Aparna Khare,Roland Maas,Venkatesh Ravichandran

from arxiv, To appear in IEEE ICASSP 2023

We propose a novel approach for ASR N-best hypothesis rescoring with graph-based label propagation by leveraging cross-utterance acoustic similarity. In contrast to conventional neural language model (LM) based ASR rescoring/reranking models, our approach focuses on acoustic information and conducts the rescoring collaboratively among utterances, instead of individually. Experiments on the VCTK dataset demonstrate that our approach consistently improves ASR performance, as well as fairness across speaker groups with different accents. Our approach provides a low-cost solution for mitigating the majoritarian bias of ASR systems, without the need to train new domain- or accent-specific models.

翻译：我们提出了一种基于图形的标签传播的新方法来利用交叉语音相似性对ASR N个最佳假设进行重新评分。与传统的基于神经语言模型（LM）的ASR重新评分/重排模型不同，我们的方法专注于声学信息，并在不同的对话中协同地进行重新评分，而不是在每个对话中单独进行。在VCTK数据集上进行的实验表明，我们的方法始终可以提高ASR的性能，并且可以在具有不同口音的讲话者组之间提供公平性。我们的方法为解决ASR系统的大多数偏见提供了一种低成本的解决方案，无需训练新的特定领域或口音的模型。

0

相关内容

语音识别

语音识别是计算机科学和计算语言学的一个跨学科子领域，它发展了一些方法和技术，使计算机可以将口语识别和翻译成文本。它也被称为自动语音识别（ASR），计算机语音识别或语音转文本（STT）。它整合了计算机科学，语言学和计算机工程领域的知识和研究。

【RecSys22教程】多阶段推荐系统的神经重排序，90页ppt

【RecSys22教程】多阶段推荐系统的神经重排序，90页ppt

专知会员服务

27+阅读 · 2022年9月30日

【KDD2022教程】图算法公平性：方法与趋势，200页ppt

【KDD2022教程】图算法公平性：方法与趋势，200页ppt

专知会员服务

42+阅读 · 2022年8月20日

【2022新书】高效深度学习，Efficient Deep Learning Book

【2022新书】高效深度学习，Efficient Deep Learning Book

专知会员服务

125+阅读 · 2022年4月21日

【AAAI 2022】用于文本摘要任务的序列级对比学习模型

【AAAI 2022】用于文本摘要任务的序列级对比学习模型

专知会员服务

25+阅读 · 2022年1月11日

【KDD2021】检索交互机的表格数据预测

专知会员服务

16+阅读 · 2021年8月13日

【KDD2021】基于因果反事实Shapley的MARL信度分配

专知会员服务

19+阅读 · 2021年7月11日

因果关联学习，Causal Relational Learning

因果关联学习，Causal Relational Learning

专知会员服务

185+阅读 · 2020年4月21日

【北京大学】动态异构图神经网络建模情感，Jointly Modeling Aspect and Sentiment with Dynamic Heterogeneous Graph Neural Networks

【北京大学】动态异构图神经网络建模情感，Jointly Modeling Aspect and Sentiment with Dynamic Heterogeneous Graph Neural Networks

专知会员服务

55+阅读 · 2020年4月15日

【AAAI2020】多模态注意力语义图嵌入多标签分类（Cross-Modality Attention with Semantic Graph Embedding for Multi-Label Classification）

【AAAI2020】多模态注意力语义图嵌入多标签分类（Cross-Modality Attention with Semantic Graph Embedding for Multi-Label Classification）

专知会员服务

92+阅读 · 2019年12月22日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

【KDD2022教程】图算法公平性：方法与趋势，200页ppt

【KDD2022教程】图算法公平性：方法与趋势，200页ppt

专知

1+阅读 · 2022年8月21日

浅聊对比学习（Contrastive Learning）第一弹

浅聊对比学习（Contrastive Learning）第一弹

PaperWeekly

0+阅读 · 2022年6月10日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

逆强化学习-学习人先验的动机

逆强化学习-学习人先验的动机

CreateAMind

16+阅读 · 2019年1月18日

强化学习的Unsupervised Meta-Learning

强化学习的Unsupervised Meta-Learning

CreateAMind

18+阅读 · 2019年1月7日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

disentangled-representation-papers

disentangled-representation-papers

CreateAMind

26+阅读 · 2018年9月12日

笔记 | Deep active learning for named entity recognition

笔记 | Deep active learning for named entity recognition

黑龙江大学自然语言处理实验室

24+阅读 · 2018年5月27日

【论文】图上的表示学习综述

【论文】图上的表示学习综述

机器学习研究会

15+阅读 · 2017年9月24日

高维回归模型的预测稳定性研究

国家自然科学基金

3+阅读 · 2015年12月31日

基于共享变量的多核并发程序模型检测

国家自然科学基金

0+阅读 · 2012年12月31日

跨汉斯拉夫蒙古文的信息检索关键技术研究

国家自然科学基金

0+阅读 · 2012年12月31日

转录因子ZNF580 在缺血-再灌注脑损伤中的作用及分子机制

国家自然科学基金

0+阅读 · 2012年12月31日

模糊Domain中的一些范畴之间的对偶等价

国家自然科学基金

0+阅读 · 2012年12月31日

孤独症动物模型（Fmr1 KO mice）脑功能网络的时空特性研究

国家自然科学基金

0+阅读 · 2011年12月31日

Notch信号通路负性调控哮喘小鼠气道杯状细胞MUC5AC的合成及其机制的研究

国家自然科学基金

0+阅读 · 2009年12月31日

磁性金属纳米催化剂的制备及水热催化碳碳键形成研究

国家自然科学基金

0+阅读 · 2009年12月31日

均衡释放的中药复方缓释制剂药代动力学研究

国家自然科学基金

0+阅读 · 2009年12月31日

瞬时随机光照下的自拼接快速三维轮廓测量

国家自然科学基金

0+阅读 · 2008年12月31日

RelationMatch: Matching In-batch Relationships for Semi-supervised Learning

RelationMatch: Matching In-batch Relationships for Semi-supervised Learning

Arxiv

0+阅读 · 2023年5月17日

Iterated learning and communication jointly explain efficient color naming systems

Arxiv

0+阅读 · 2023年5月17日

A Dictionary-based approach to Time Series Ordinal Classification

Arxiv

0+阅读 · 2023年5月16日

A Novel Framework for Multimodal Named Entity Recognition with Multi-level Alignments

Arxiv

0+阅读 · 2023年5月15日

Self-supervised Neural Factor Analysis for Disentangling Utterance-level Speech Representations

Arxiv

0+阅读 · 2023年5月14日

A Survey of Adversarial Learning on Graphs

Arxiv

38+阅读 · 2020年3月10日

Knowledge Graph Transfer Network for Few-Shot Recognition

Arxiv

15+阅读 · 2019年11月21日

Enhanced Meta-Learning for Cross-lingual Named Entity Recognition with Minimal Resources

Arxiv

13+阅读 · 2019年11月14日

Representation Learning with Ordered Relation Paths for Knowledge Graph Completion

Representation Learning with Ordered Relation Paths for Knowledge Graph Completion

Arxiv

12+阅读 · 2019年9月26日

Learning to Propagate for Graph Meta-Learning

Arxiv

14+阅读 · 2019年9月11日

VIP会员

文章信息

相关主题

神经语言模型

相关VIP内容

【RecSys22教程】多阶段推荐系统的神经重排序，90页ppt

【RecSys22教程】多阶段推荐系统的神经重排序，90页ppt

专知会员服务

27+阅读 · 2022年9月30日

【KDD2022教程】图算法公平性：方法与趋势，200页ppt

【KDD2022教程】图算法公平性：方法与趋势，200页ppt

专知会员服务

42+阅读 · 2022年8月20日

【2022新书】高效深度学习，Efficient Deep Learning Book

【2022新书】高效深度学习，Efficient Deep Learning Book

专知会员服务

125+阅读 · 2022年4月21日

【AAAI 2022】用于文本摘要任务的序列级对比学习模型

【AAAI 2022】用于文本摘要任务的序列级对比学习模型

专知会员服务

25+阅读 · 2022年1月11日

【KDD2021】检索交互机的表格数据预测

专知会员服务

16+阅读 · 2021年8月13日

【KDD2021】基于因果反事实Shapley的MARL信度分配

专知会员服务

19+阅读 · 2021年7月11日

因果关联学习，Causal Relational Learning

因果关联学习，Causal Relational Learning

专知会员服务

185+阅读 · 2020年4月21日

【北京大学】动态异构图神经网络建模情感，Jointly Modeling Aspect and Sentiment with Dynamic Heterogeneous Graph Neural Networks

【北京大学】动态异构图神经网络建模情感，Jointly Modeling Aspect and Sentiment with Dynamic Heterogeneous Graph Neural Networks

专知会员服务

55+阅读 · 2020年4月15日

【AAAI2020】多模态注意力语义图嵌入多标签分类（Cross-Modality Attention with Semantic Graph Embedding for Multi-Label Classification）

【AAAI2020】多模态注意力语义图嵌入多标签分类（Cross-Modality Attention with Semantic Graph Embedding for Multi-Label Classification）

专知会员服务

92+阅读 · 2019年12月22日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

热门VIP内容

开通专知VIP会员享更多权益服务

《复杂工程系统模型驱动设计决策支持系统：早期设计阶段挑战》最新138页

《日本陆上自卫队2040年作战方式与未来作战研究》最新23页slides

人工智能作为战争武器

《后勤保障》最新23页

相关资讯

【KDD2022教程】图算法公平性：方法与趋势，200页ppt

【KDD2022教程】图算法公平性：方法与趋势，200页ppt

专知

1+阅读 · 2022年8月21日

浅聊对比学习（Contrastive Learning）第一弹

浅聊对比学习（Contrastive Learning）第一弹

PaperWeekly

0+阅读 · 2022年6月10日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

逆强化学习-学习人先验的动机

逆强化学习-学习人先验的动机

CreateAMind

16+阅读 · 2019年1月18日

强化学习的Unsupervised Meta-Learning

强化学习的Unsupervised Meta-Learning

CreateAMind

18+阅读 · 2019年1月7日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

disentangled-representation-papers

disentangled-representation-papers

CreateAMind

26+阅读 · 2018年9月12日

笔记 | Deep active learning for named entity recognition

笔记 | Deep active learning for named entity recognition

黑龙江大学自然语言处理实验室

24+阅读 · 2018年5月27日

【论文】图上的表示学习综述

【论文】图上的表示学习综述

机器学习研究会

15+阅读 · 2017年9月24日

相关论文

RelationMatch: Matching In-batch Relationships for Semi-supervised Learning

RelationMatch: Matching In-batch Relationships for Semi-supervised Learning

Arxiv

0+阅读 · 2023年5月17日

Iterated learning and communication jointly explain efficient color naming systems

Arxiv

0+阅读 · 2023年5月17日

A Dictionary-based approach to Time Series Ordinal Classification

Arxiv

0+阅读 · 2023年5月16日

A Novel Framework for Multimodal Named Entity Recognition with Multi-level Alignments

Arxiv

0+阅读 · 2023年5月15日

Self-supervised Neural Factor Analysis for Disentangling Utterance-level Speech Representations

Arxiv

0+阅读 · 2023年5月14日

A Survey of Adversarial Learning on Graphs

Arxiv

38+阅读 · 2020年3月10日

Knowledge Graph Transfer Network for Few-Shot Recognition

Arxiv

15+阅读 · 2019年11月21日

Enhanced Meta-Learning for Cross-lingual Named Entity Recognition with Minimal Resources

Arxiv

13+阅读 · 2019年11月14日

Representation Learning with Ordered Relation Paths for Knowledge Graph Completion

Representation Learning with Ordered Relation Paths for Knowledge Graph Completion

Arxiv

12+阅读 · 2019年9月26日

Learning to Propagate for Graph Meta-Learning

Arxiv

14+阅读 · 2019年9月11日

相关基金

高维回归模型的预测稳定性研究

国家自然科学基金

3+阅读 · 2015年12月31日

基于共享变量的多核并发程序模型检测

国家自然科学基金

0+阅读 · 2012年12月31日

跨汉斯拉夫蒙古文的信息检索关键技术研究

国家自然科学基金

0+阅读 · 2012年12月31日

转录因子ZNF580 在缺血-再灌注脑损伤中的作用及分子机制

国家自然科学基金

0+阅读 · 2012年12月31日

模糊Domain中的一些范畴之间的对偶等价

国家自然科学基金

0+阅读 · 2012年12月31日

孤独症动物模型（Fmr1 KO mice）脑功能网络的时空特性研究

国家自然科学基金

0+阅读 · 2011年12月31日

Notch信号通路负性调控哮喘小鼠气道杯状细胞MUC5AC的合成及其机制的研究

国家自然科学基金

0+阅读 · 2009年12月31日

磁性金属纳米催化剂的制备及水热催化碳碳键形成研究

国家自然科学基金

0+阅读 · 2009年12月31日

均衡释放的中药复方缓释制剂药代动力学研究

国家自然科学基金

0+阅读 · 2009年12月31日

瞬时随机光照下的自拼接快速三维轮廓测量

国家自然科学基金

0+阅读 · 2008年12月31日

微信扫码咨询专知VIP会员