单一常态表达式的噪音容忍有区别的学习方法,并进行插接 (A Noise-tolerant Differentiable Learning Approach for Single Occurrence Regular Expression with Interleaving) - 专知论文

会员服务 ·

0

正则化项 · 正则表达式 · Learning · Networking · Neural Networks ·

2023 年 1 月 11 日

A Noise-tolerant Differentiable Learning Approach for Single Occurrence Regular Expression with Interleaving

翻译：单一常态表达式的噪音容忍有区别的学习方法,并进行插接

Rongzhen Ye,Tianqu Zhuang,Hai Wan,Jianfeng Du,Weilin Luo,Pingjia Liang

We study the problem of learning a single occurrence regular expression with interleaving (SOIRE) from a set of text strings possibly with noise. SOIRE fully supports interleaving and covers a large portion of regular expressions used in practice. Learning SOIREs is challenging because it requires heavy computation and text strings usually contain noise in practice. Most of the previous studies only learn restricted SOIREs and are not robust on noisy data. To tackle these issues, we propose a noise-tolerant differentiable learning approach SOIREDL for SOIRE. We design a neural network to simulate SOIRE matching and theoretically prove that certain assignments of the set of parameters learnt by the neural network, called faithful encodings, are one-to-one corresponding to SOIREs for a bounded size. Based on this correspondence, we interpret the target SOIRE from an assignment of the set of parameters of the neural network by exploring the nearest faithful encodings. Experimental results show that SOIREDL outperforms the state-of-the-art approaches, especially on noisy data.

翻译：我们研究从一套可能带有噪音的文本字符串中学习一个单一的定期表达式的问题。 SOIRE 完全支持插入并覆盖实践中使用的很大一部分常规表达式。学习 SOIRE 具有挑战性, 因为它需要大量计算, 文本字符串通常含有实际中的噪音。以往的研究大多只学习有限的 SOIRE, 并且对吵闹的数据不强。为了解决这些问题, 我们为SOIRE 设计了一个不动的可异学习方法 SOIREDL 。我们设计了一个神经网络, 模拟SOIRE 匹配, 并在理论上证明神经网络所学的一组参数( 称为忠实编码) 的某些分配是一对一的, 与SOIRE 相对应, 其尺寸受约束。基于这一通信, 我们从神经网络一系列参数的指定中解释 SOIRE, 探索最近的可靠编码。实验结果显示 SOIREDL 超越了最先进的方法, 特别是热调数据。

0

相关内容

正则化项

高效可扩展图神经网络的研究进展，Recent Advances in Efficient and Scalable Graph Neural Networks

高效可扩展图神经网络的研究进展，Recent Advances in Efficient and Scalable Graph Neural Networks

专知会员服务

78+阅读 · 2022年3月15日

【深度学习表格检测、信息提取和结构化】《Table Detection, Information Extraction and Structuring using Deep Learning》by Vihar Kurama

专知会员服务

38+阅读 · 2020年1月23日

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

专知会员服务

19+阅读 · 2019年10月22日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

Stabilizing Transformers for Reinforcement Learning

Stabilizing Transformers for Reinforcement Learning

专知会员服务

60+阅读 · 2019年10月17日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

专知会员服务

160+阅读 · 2019年10月12日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

VCIP 2022 Call for Demos

VCIP 2022 Call for Demos

CCF多媒体专委会

1+阅读 · 2022年6月6日

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

IEEE TII Call For Papers

IEEE TII Call For Papers

CCF多媒体专委会

3+阅读 · 2022年3月24日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

强化学习的Unsupervised Meta-Learning

强化学习的Unsupervised Meta-Learning

CreateAMind

18+阅读 · 2019年1月7日

无监督元学习表示学习

无监督元学习表示学习

CreateAMind

27+阅读 · 2019年1月4日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

disentangled-representation-papers

disentangled-representation-papers

CreateAMind

26+阅读 · 2018年9月12日

一类稳态Schödinger-Poisson-Slater方程标准化解的研究

国家自然科学基金

1+阅读 · 2015年12月31日

Kahler 曲面中特殊曲面的研究

国家自然科学基金

0+阅读 · 2014年12月31日

可变剪切基因REST调控SRRM3的表达在前列腺癌神经内分泌分化中的作用机制研究

国家自然科学基金

0+阅读 · 2013年12月31日

Schrodinger-Poisson方程的若干问题研究

国家自然科学基金

1+阅读 · 2012年12月31日

Cldn-7调控Intergrin/FAK信号通路参与大肠癌发生、侵袭转移的机制研究

国家自然科学基金

0+阅读 · 2012年12月31日

LaBr3晶体快定时性能的研究

国家自然科学基金

0+阅读 · 2012年12月31日

基于垂直取向的锌基杂化本体异质结薄膜的设计制备、界面调控及光伏器件研究

国家自然科学基金

0+阅读 · 2011年12月31日

广义Kloosterman和的均值估计

国家自然科学基金

1+阅读 · 2011年12月31日

UGT基因簇进化及调控研究

国家自然科学基金

0+阅读 · 2009年12月31日

肾上腺源性及原发性高血压线粒体tRNAIle、tRNALeu(UUR)和tRNAlys基因突变的差异对比研究

国家自然科学基金

0+阅读 · 2009年12月31日

Aquarium: A Fully Differentiable Fluid-Structure Interaction Solver for Robotics Applications

Arxiv

0+阅读 · 2023年3月7日

Pattern recovery by SLOPE

Arxiv

0+阅读 · 2023年3月6日

Inference in Spatial Experiments with Interference using the SpatialEffect Package

Arxiv

0+阅读 · 2023年3月6日

A combinatorial proof for the secretary problem with multiple choices

Arxiv

0+阅读 · 2023年3月4日

Interpretable reduced-order modeling with time-scale separation

Arxiv

0+阅读 · 2023年3月3日

Render unto Numerics: Orthogonal Polynomial Neural Operator for PDEs with Non-periodic Boundary Conditions

Arxiv

0+阅读 · 2023年3月3日

Eryn : A multi-purpose sampler for Bayesian inference

Arxiv

0+阅读 · 2023年3月3日

Queue Scheduling with Adversarial Bandit Learning

Arxiv

0+阅读 · 2023年3月3日

On the Use of Neural Networks for Full Waveform Inversion

Arxiv

0+阅读 · 2023年1月30日

Self-correcting Q-Learning

Arxiv

11+阅读 · 2020年12月2日

VIP会员

文章信息

相关主题

正则表达式

Neural Networks

相关VIP内容

高效可扩展图神经网络的研究进展，Recent Advances in Efficient and Scalable Graph Neural Networks

高效可扩展图神经网络的研究进展，Recent Advances in Efficient and Scalable Graph Neural Networks

专知会员服务

78+阅读 · 2022年3月15日

【深度学习表格检测、信息提取和结构化】《Table Detection, Information Extraction and Structuring using Deep Learning》by Vihar Kurama

专知会员服务

38+阅读 · 2020年1月23日

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

专知会员服务

19+阅读 · 2019年10月22日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

Stabilizing Transformers for Reinforcement Learning

Stabilizing Transformers for Reinforcement Learning

专知会员服务

60+阅读 · 2019年10月17日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

专知会员服务

160+阅读 · 2019年10月12日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

【NeurIPS2025】大型语言模型中关系解码线性算子的结构

《大模型一体机应用研究报告（2025年）》，48页pdf

语言模型如何重塑实体对齐？语言模型驱动实体对齐的进展、基准与未来

【CMU博士论文】迈向具备基础先验的四维感知

相关资讯

VCIP 2022 Call for Demos

VCIP 2022 Call for Demos

CCF多媒体专委会

1+阅读 · 2022年6月6日

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

IEEE TII Call For Papers

IEEE TII Call For Papers

CCF多媒体专委会

3+阅读 · 2022年3月24日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

强化学习的Unsupervised Meta-Learning

强化学习的Unsupervised Meta-Learning

CreateAMind

18+阅读 · 2019年1月7日

无监督元学习表示学习

无监督元学习表示学习

CreateAMind

27+阅读 · 2019年1月4日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

disentangled-representation-papers

disentangled-representation-papers

CreateAMind

26+阅读 · 2018年9月12日

相关论文

Aquarium: A Fully Differentiable Fluid-Structure Interaction Solver for Robotics Applications

Arxiv

0+阅读 · 2023年3月7日

Pattern recovery by SLOPE

Arxiv

0+阅读 · 2023年3月6日

Inference in Spatial Experiments with Interference using the SpatialEffect Package

Arxiv

0+阅读 · 2023年3月6日

A combinatorial proof for the secretary problem with multiple choices

Arxiv

0+阅读 · 2023年3月4日

Interpretable reduced-order modeling with time-scale separation

Arxiv

0+阅读 · 2023年3月3日

Render unto Numerics: Orthogonal Polynomial Neural Operator for PDEs with Non-periodic Boundary Conditions

Arxiv

0+阅读 · 2023年3月3日

Eryn : A multi-purpose sampler for Bayesian inference

Arxiv

0+阅读 · 2023年3月3日

Queue Scheduling with Adversarial Bandit Learning

Arxiv

0+阅读 · 2023年3月3日

On the Use of Neural Networks for Full Waveform Inversion

Arxiv

0+阅读 · 2023年1月30日

Self-correcting Q-Learning

Arxiv

11+阅读 · 2020年12月2日

相关基金

一类稳态Schödinger-Poisson-Slater方程标准化解的研究

国家自然科学基金

1+阅读 · 2015年12月31日

Kahler 曲面中特殊曲面的研究

国家自然科学基金

0+阅读 · 2014年12月31日

可变剪切基因REST调控SRRM3的表达在前列腺癌神经内分泌分化中的作用机制研究

国家自然科学基金

0+阅读 · 2013年12月31日

Schrodinger-Poisson方程的若干问题研究

国家自然科学基金

1+阅读 · 2012年12月31日

Cldn-7调控Intergrin/FAK信号通路参与大肠癌发生、侵袭转移的机制研究

国家自然科学基金

0+阅读 · 2012年12月31日

LaBr3晶体快定时性能的研究

国家自然科学基金

0+阅读 · 2012年12月31日

基于垂直取向的锌基杂化本体异质结薄膜的设计制备、界面调控及光伏器件研究

国家自然科学基金

0+阅读 · 2011年12月31日

广义Kloosterman和的均值估计

国家自然科学基金

1+阅读 · 2011年12月31日

UGT基因簇进化及调控研究

国家自然科学基金

0+阅读 · 2009年12月31日

肾上腺源性及原发性高血压线粒体tRNAIle、tRNALeu(UUR)和tRNAlys基因突变的差异对比研究

国家自然科学基金

0+阅读 · 2009年12月31日

微信扫码咨询专知VIP会员