基于知识蒸馏的中文语法错误校正 (Chinese grammatical error correction based on knowledge distillation) - 专知论文

会员服务 ·

0

知识 (knowledge) · 蒸馏 · MoDELS · 稳健性 · 情景 ·

2022 年 8 月 28 日

Chinese grammatical error correction based on knowledge distillation

翻译：基于知识蒸馏的中文语法错误校正

Peng Xia,Yuechi Zhou,Ziyan Zhang,Zecheng Tang,Juntao Li

from arxiv, 10 pages, 4 figures, 5 tables

In view of the poor robustness of existing Chinese grammatical error correction models on attack test sets and large model parameters, this paper uses the method of knowledge distillation to compress model parameters and improve the anti-attack ability of the model. In terms of data, the attack test set is constructed by integrating the disturbance into the standard evaluation data set, and the model robustness is evaluated by the attack test set. The experimental results show that the distilled small model can ensure the performance and improve the training speed under the condition of reducing the number of model parameters, and achieve the optimal effect on the attack test set, and the robustness is significantly improved. Code is available at \url{https://github.com/Richard88888/KD-CGEC}.

翻译：鉴于现有中国攻击试验机组和大型模型参数的语法误差校正模型不够坚固,本文使用知识蒸馏法压缩模型参数,提高模型的反攻击能力,在数据方面,攻击试验组通过将扰动纳入标准评价数据集构建,模型坚固度由攻击试验组进行评估,实验结果表明,蒸馏的小模型可以在减少模型参数数目的条件下确保性能,提高培训速度,实现对攻击试验组的最佳效果,并大大改进强健性,代码可在以下网址查阅:\url{https://github.com/Richard88888/KD-CGEC}。

0

相关内容

知识 (knowledge)

知识 (knowledge)

通过学习、实践或探索所获得的认识、判断或技能。

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

专知会员服务

60+阅读 · 2022年4月22日

NLP必读经典文献100篇

专知会员服务

124+阅读 · 2020年9月8日

史上最全！358篇机器学习&自然语言处理综述论文！都这儿了

专知会员服务

129+阅读 · 2020年7月18日

【新书】数字图像(影像)处理手第二版，2176pdf，Mathematical Methods in Imaging

【新书】数字图像(影像)处理手第二版，2176pdf，Mathematical Methods in Imaging

专知会员服务

93+阅读 · 2020年2月12日

【跨语言BERT模型大集合】Transfer learning is increasingly going multilingual with language-specific BERT models

专知会员服务

54+阅读 · 2020年1月30日

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

专知会员服务

19+阅读 · 2019年10月22日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

50+阅读 · 2019年10月17日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

专知会员服务

65+阅读 · 2019年10月9日

VCIP 2022 Call for Demos

VCIP 2022 Call for Demos

CCF多媒体专委会

1+阅读 · 2022年6月6日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

【ICIG2021】Latest News & Announcements of the Workshop

【ICIG2021】Latest News & Announcements of the Workshop

中国图象图形学学会CSIG

0+阅读 · 2021年12月20日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium6

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium6

中国图象图形学学会CSIG

2+阅读 · 2021年11月12日

【ICIG2021】Latest News & Announcements of the Plenary Talk2

【ICIG2021】Latest News & Announcements of the Plenary Talk2

中国图象图形学学会CSIG

0+阅读 · 2021年11月2日

【ICIG2021】Latest News & Announcements of the Plenary Talk1

【ICIG2021】Latest News & Announcements of the Plenary Talk1

中国图象图形学学会CSIG

0+阅读 · 2021年11月1日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

【论文】变分推断（Variational inference)的总结

【论文】变分推断（Variational inference)的总结

机器学习研究会

39+阅读 · 2017年11月16日

基于BIM的建筑生命周期环境与经济评价及优化设计方法研究

国家自然科学基金

3+阅读 · 2014年12月31日

MiR-155/β-arrestin 2/GSK3β通路在Sca-1+心脏干细胞向心肌分化中的功能研究

国家自然科学基金

0+阅读 · 2012年12月31日

microRNA调节肿瘤抑制因子Caliban应答DNA损伤的机制

国家自然科学基金

1+阅读 · 2012年12月31日

SDIR1互作蛋白ECA1在植物应对干旱胁迫过程中的功能分析

国家自然科学基金

0+阅读 · 2012年12月31日

miR-491通过调控T细胞的增殖和凋亡在诱导T细胞衰竭中的作用机制研究

国家自然科学基金

0+阅读 · 2012年12月31日

缠结凝聚态对淀粉回生调控机制的研究

国家自然科学基金

0+阅读 · 2012年12月31日

干旱诱导表达的苹果AsA转运蛋白功能和在抗逆中的作用分析

国家自然科学基金

0+阅读 · 2011年12月31日

Intermedin-53在心肌肥厚中的作用和机制

国家自然科学基金

0+阅读 · 2011年12月31日

mir-341、mir-1188、mir-370在小鼠发育中的表达与调控机制的研究

国家自然科学基金

0+阅读 · 2009年12月31日

磁性Pickering乳液界面流变学研究

国家自然科学基金

0+阅读 · 2008年12月31日

Sparse Teachers Can Be Dense with Knowledge

Arxiv

0+阅读 · 2022年10月17日

Supervised Prototypical Contrastive Learning for Emotion Recognition in Conversation

Arxiv

0+阅读 · 2022年10月17日

Zero-Shot Learners for Natural Language Understanding via a Unified Multiple Choice Perspective

Arxiv

0+阅读 · 2022年10月16日

On Model Selection Consistency of Lasso for High-Dimensional Ising Models

Arxiv

0+阅读 · 2022年10月15日

Numerically Stable Sparse Gaussian Processes via Minimum Separation using Cover Trees

Numerically Stable Sparse Gaussian Processes via Minimum Separation using Cover Trees

Arxiv

0+阅读 · 2022年10月14日

Joint Reasoning on Hybrid-knowledge sources for Task-Oriented Dialog

Arxiv

0+阅读 · 2022年10月13日

Multilingual Zero Resource Speech Recognition Base on Self-Supervise Pre-Trained Acoustic Models

Arxiv

0+阅读 · 2022年10月13日

Prediction can be safely used as a proxy for explanation in causally consistent Bayesian generalized linear models

Arxiv

0+阅读 · 2022年10月13日

The Unreasonable Effectiveness of Fully-Connected Layers for Low-Data Regimes

Arxiv

0+阅读 · 2022年10月13日

Expectation-Maximizing Network Reconstruction and MostApplicable Network Types Based on Binary Time Series Data

Arxiv

0+阅读 · 2022年10月13日

VIP会员

文章信息

相关主题

知识 (knowledge)

相关VIP内容

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

专知会员服务

60+阅读 · 2022年4月22日

NLP必读经典文献100篇

专知会员服务

124+阅读 · 2020年9月8日

史上最全！358篇机器学习&自然语言处理综述论文！都这儿了

专知会员服务

129+阅读 · 2020年7月18日

【新书】数字图像(影像)处理手第二版，2176pdf，Mathematical Methods in Imaging

【新书】数字图像(影像)处理手第二版，2176pdf，Mathematical Methods in Imaging

专知会员服务

93+阅读 · 2020年2月12日

【跨语言BERT模型大集合】Transfer learning is increasingly going multilingual with language-specific BERT models

专知会员服务

54+阅读 · 2020年1月30日

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

专知会员服务

19+阅读 · 2019年10月22日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

50+阅读 · 2019年10月17日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

专知会员服务

65+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

《俄乌战争背景下俄罗斯的战略性海军分析（2022-2025年）》最新100页报告

【斯坦福博士论文】数据、决策与依赖：构建可信人工智能的挑战

人工智能时代背景下的未来海战

接触战中的无人机优势：美军旅级部队面临的小型无人机系统挑战与调整

相关资讯

VCIP 2022 Call for Demos

VCIP 2022 Call for Demos

CCF多媒体专委会

1+阅读 · 2022年6月6日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

【ICIG2021】Latest News & Announcements of the Workshop

【ICIG2021】Latest News & Announcements of the Workshop

中国图象图形学学会CSIG

0+阅读 · 2021年12月20日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium6

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium6

中国图象图形学学会CSIG

2+阅读 · 2021年11月12日

【ICIG2021】Latest News & Announcements of the Plenary Talk2

【ICIG2021】Latest News & Announcements of the Plenary Talk2

中国图象图形学学会CSIG

0+阅读 · 2021年11月2日

【ICIG2021】Latest News & Announcements of the Plenary Talk1

【ICIG2021】Latest News & Announcements of the Plenary Talk1

中国图象图形学学会CSIG

0+阅读 · 2021年11月1日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

【论文】变分推断（Variational inference)的总结

【论文】变分推断（Variational inference)的总结

机器学习研究会

39+阅读 · 2017年11月16日

相关论文

Sparse Teachers Can Be Dense with Knowledge

Arxiv

0+阅读 · 2022年10月17日

Supervised Prototypical Contrastive Learning for Emotion Recognition in Conversation

Arxiv

0+阅读 · 2022年10月17日

Zero-Shot Learners for Natural Language Understanding via a Unified Multiple Choice Perspective

Arxiv

0+阅读 · 2022年10月16日

On Model Selection Consistency of Lasso for High-Dimensional Ising Models

Arxiv

0+阅读 · 2022年10月15日

Numerically Stable Sparse Gaussian Processes via Minimum Separation using Cover Trees

Numerically Stable Sparse Gaussian Processes via Minimum Separation using Cover Trees

Arxiv

0+阅读 · 2022年10月14日

Joint Reasoning on Hybrid-knowledge sources for Task-Oriented Dialog

Arxiv

0+阅读 · 2022年10月13日

Multilingual Zero Resource Speech Recognition Base on Self-Supervise Pre-Trained Acoustic Models

Arxiv

0+阅读 · 2022年10月13日

Prediction can be safely used as a proxy for explanation in causally consistent Bayesian generalized linear models

Arxiv

0+阅读 · 2022年10月13日

The Unreasonable Effectiveness of Fully-Connected Layers for Low-Data Regimes

Arxiv

0+阅读 · 2022年10月13日

Expectation-Maximizing Network Reconstruction and MostApplicable Network Types Based on Binary Time Series Data

Arxiv

0+阅读 · 2022年10月13日

相关基金

基于BIM的建筑生命周期环境与经济评价及优化设计方法研究

国家自然科学基金

3+阅读 · 2014年12月31日

MiR-155/β-arrestin 2/GSK3β通路在Sca-1+心脏干细胞向心肌分化中的功能研究

国家自然科学基金

0+阅读 · 2012年12月31日

microRNA调节肿瘤抑制因子Caliban应答DNA损伤的机制

国家自然科学基金

1+阅读 · 2012年12月31日

SDIR1互作蛋白ECA1在植物应对干旱胁迫过程中的功能分析

国家自然科学基金

0+阅读 · 2012年12月31日

miR-491通过调控T细胞的增殖和凋亡在诱导T细胞衰竭中的作用机制研究

国家自然科学基金

0+阅读 · 2012年12月31日

缠结凝聚态对淀粉回生调控机制的研究

国家自然科学基金

0+阅读 · 2012年12月31日

干旱诱导表达的苹果AsA转运蛋白功能和在抗逆中的作用分析

国家自然科学基金

0+阅读 · 2011年12月31日

Intermedin-53在心肌肥厚中的作用和机制

国家自然科学基金

0+阅读 · 2011年12月31日

mir-341、mir-1188、mir-370在小鼠发育中的表达与调控机制的研究

国家自然科学基金

0+阅读 · 2009年12月31日

磁性Pickering乳液界面流变学研究

国家自然科学基金

0+阅读 · 2008年12月31日

微信扫码咨询专知VIP会员