LinCQA:以线性时间担保更快一致的查询 (LinCQA: Faster Consistent Query Answering with Linear Time Guarantees) - 专知论文

会员服务 ·

0

线性的 · CASES · Extensibility · 可辨认的 · SQL ·

2022 年 8 月 25 日

LinCQA: Faster Consistent Query Answering with Linear Time Guarantees

翻译：LinCQA:以线性时间担保更快一致的查询

Zhiwei Fan,Paraschos Koutris,Xiating Ouyang,Jef Wijsen

Most data analytical pipelines often encounter the problem of querying inconsistent data that violate pre-determined integrity constraints. Data cleaning is an extensively studied paradigm that singles out a consistent repair of the inconsistent data. Consistent query answering (CQA) is an alternative approach to data cleaning that asks for all tuples guaranteed to be returned by a given query on all (in most cases, exponentially many) repairs of the inconsistent data. This paper identifies a class of acyclic select-project-join (SPJ) queries for which CQA can be solved via SQL rewriting with a linear time guarantee. Our rewriting method can be viewed as a generalization of Yannakakis's algorithm for acyclic joins to the inconsistent setting. We present LinCQA, a system that can output rewritings in both SQL and non-recursive Datalog rules for every query in this class. We show that LinCQA often outperforms the existing CQA systems on both synthetic and real-world workloads, and in some cases, by orders of magnitude.

翻译：大多数数据分析管道常常遇到质疑不一致数据的问题,这违反了预先确定的完整限制。数据清理是一个广泛研究的范例,它挑选出对不一致数据进行一致的修复。一致的查询回答(CQA)是数据清理的替代方法,它要求通过对不一致数据的所有(多数情况下是指数性多的)修复进行特定查询,以所有(大多数情况下是指数性的)修复数据来保证归还所有图例。本文确定了一种周期性选择项目-join(SPJ)查询,可以通过SQL以线性时间保证重写CQA(SPJ)来解决这个问题。我们的重写方法可以被看作是Yannakakis的循环计算法与不一致的设置相结合的概括。我们介绍了LinCQA,这个系统可以在SQL和不精确的数据记录规则中输出该类每项查询的重写内容。我们显示,LincQA常常在合成和现实世界工作量方面超越现有的CQA系统,有些情况下,以数量顺序。

0

相关内容

线性的

不可错过！700+ppt《因果推理》课程！杜克大学Fan Li教程

不可错过！700+ppt《因果推理》课程！杜克大学Fan Li教程

专知会员服务

72+阅读 · 2022年7月11日

不可错过！《机器学习100讲》课程，UBC Mark Schmidt讲授

不可错过！《机器学习100讲》课程，UBC Mark Schmidt讲授

专知会员服务

76+阅读 · 2022年6月28日

不可错过！UIUC最新《统计强化学习》课程！

专知会员服务

54+阅读 · 2020年9月7日

2020数据工程师成长路线图

专知会员服务

41+阅读 · 2020年9月6日

50+篇《神经架构搜索NAS》2020论文合集

专知会员服务

61+阅读 · 2020年3月19日

UC.Berkeley CS189讲义教材:《机器学习全面指南》，185页pdf

专知会员服务

162+阅读 · 2020年1月16日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

《DeepGCNs: Making GCNs Go as Deep as CNNs》

《DeepGCNs: Making GCNs Go as Deep as CNNs》

专知会员服务

31+阅读 · 2019年10月17日

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

专知会员服务

160+阅读 · 2019年10月12日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

详解PyTorch中的ModuleList和Sequential

详解PyTorch中的ModuleList和Sequential

极市平台

0+阅读 · 2022年1月28日

讲座报名丨 ICML专场

讲座报名丨 ICML专场

THU数据派

0+阅读 · 2021年9月15日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

无监督元学习表示学习

无监督元学习表示学习

CreateAMind

27+阅读 · 2019年1月4日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

自然语言处理顶会EMNLP2018接受论文列表！

自然语言处理顶会EMNLP2018接受论文列表！

专知

87+阅读 · 2018年8月26日

紧区间上保向微分同胚的光滑嵌入流

国家自然科学基金

0+阅读 · 2015年12月31日

阿司匹林乙酰化修饰ALDH1选择性清除大肠癌干细胞的分子机制研究

国家自然科学基金

0+阅读 · 2014年12月31日

Kronheimer-Nakajima quiver 模空间与有理曲面

国家自然科学基金

1+阅读 · 2013年12月31日

CD147-CD98复合体参与RA患者CD4+CD161+T细胞活化相关功能的新机制

国家自然科学基金

0+阅读 · 2013年12月31日

胶质瘤中长链非编码RNA HOTAIR功能网络的解析及其分子机制的研究

国家自然科学基金

0+阅读 · 2012年12月31日

RegIII信号通路与SOCS3甲基化协同调控胰腺炎症恶性转化的分子机制研究

国家自然科学基金

0+阅读 · 2012年12月31日

MiR-155/β-arrestin 2/GSK3β通路在Sca-1+心脏干细胞向心肌分化中的功能研究

国家自然科学基金

0+阅读 · 2012年12月31日

可压缩Navier-Stokes方程的一些数学问题

国家自然科学基金

0+阅读 · 2012年12月31日

随机泛函微分方程的渐近行为

国家自然科学基金

0+阅读 · 2012年12月31日

一类四阶MEMS方程的解集结构与解的渐近性态

国家自然科学基金

0+阅读 · 2011年12月31日

Expander Graph Propagation

Arxiv

0+阅读 · 2022年10月6日

Fault-tolerant Coding for Entanglement-Assisted Communication

Arxiv

0+阅读 · 2022年10月6日

Locate before Answering: Answer Guided Question Localization for Video Question Answering

Arxiv

0+阅读 · 2022年10月5日

Adversarial Attack on Attackers: Post-Process to Mitigate Black-Box Score-Based Query Attacks

Adversarial Attack on Attackers: Post-Process to Mitigate Black-Box Score-Based Query Attacks

Arxiv

0+阅读 · 2022年10月4日

Scheduling with Many Shared Resources

Arxiv

0+阅读 · 2022年10月4日

Universal Mini-Batch Consistency for Set Encoding Functions

Arxiv

0+阅读 · 2022年10月4日

Unsupervised Model Selection for Time-series Anomaly Detection

Arxiv

0+阅读 · 2022年10月3日

Offline Reinforcement Learning with Differentiable Function Approximation is Provably Efficient

Arxiv

0+阅读 · 2022年10月3日

Offset-value coding in database query processing

Arxiv

0+阅读 · 2022年9月30日

Deterministic Performance Guarantees for Bidirectional BFS on Real-World Networks

Arxiv

0+阅读 · 2022年9月30日

VIP会员

文章信息

相关主题

相关VIP内容

不可错过！700+ppt《因果推理》课程！杜克大学Fan Li教程

不可错过！700+ppt《因果推理》课程！杜克大学Fan Li教程

专知会员服务

72+阅读 · 2022年7月11日

不可错过！《机器学习100讲》课程，UBC Mark Schmidt讲授

不可错过！《机器学习100讲》课程，UBC Mark Schmidt讲授

专知会员服务

76+阅读 · 2022年6月28日

不可错过！UIUC最新《统计强化学习》课程！

专知会员服务

54+阅读 · 2020年9月7日

2020数据工程师成长路线图

专知会员服务

41+阅读 · 2020年9月6日

50+篇《神经架构搜索NAS》2020论文合集

专知会员服务

61+阅读 · 2020年3月19日

UC.Berkeley CS189讲义教材:《机器学习全面指南》，185页pdf

专知会员服务

162+阅读 · 2020年1月16日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

《DeepGCNs: Making GCNs Go as Deep as CNNs》

《DeepGCNs: Making GCNs Go as Deep as CNNs》

专知会员服务

31+阅读 · 2019年10月17日

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

专知会员服务

160+阅读 · 2019年10月12日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

美陆军五大转型方向

一种Agent自主性风险评估框架 | 最新文献

实时无人机指令处理：一种面向无人机系统的大语言模型方法

基于动态知识图谱的人工智能代理自主研究周期 | 文献

相关资讯

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

详解PyTorch中的ModuleList和Sequential

详解PyTorch中的ModuleList和Sequential

极市平台

0+阅读 · 2022年1月28日

讲座报名丨 ICML专场

讲座报名丨 ICML专场

THU数据派

0+阅读 · 2021年9月15日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

无监督元学习表示学习

无监督元学习表示学习

CreateAMind

27+阅读 · 2019年1月4日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

自然语言处理顶会EMNLP2018接受论文列表！

自然语言处理顶会EMNLP2018接受论文列表！

专知

87+阅读 · 2018年8月26日

相关论文

Expander Graph Propagation

Arxiv

0+阅读 · 2022年10月6日

Fault-tolerant Coding for Entanglement-Assisted Communication

Arxiv

0+阅读 · 2022年10月6日

Locate before Answering: Answer Guided Question Localization for Video Question Answering

Arxiv

0+阅读 · 2022年10月5日

Adversarial Attack on Attackers: Post-Process to Mitigate Black-Box Score-Based Query Attacks

Adversarial Attack on Attackers: Post-Process to Mitigate Black-Box Score-Based Query Attacks

Arxiv

0+阅读 · 2022年10月4日

Scheduling with Many Shared Resources

Arxiv

0+阅读 · 2022年10月4日

Universal Mini-Batch Consistency for Set Encoding Functions

Arxiv

0+阅读 · 2022年10月4日

Unsupervised Model Selection for Time-series Anomaly Detection

Arxiv

0+阅读 · 2022年10月3日

Offline Reinforcement Learning with Differentiable Function Approximation is Provably Efficient

Arxiv

0+阅读 · 2022年10月3日

Offset-value coding in database query processing

Arxiv

0+阅读 · 2022年9月30日

Deterministic Performance Guarantees for Bidirectional BFS on Real-World Networks

Arxiv

0+阅读 · 2022年9月30日

相关基金

紧区间上保向微分同胚的光滑嵌入流

国家自然科学基金

0+阅读 · 2015年12月31日

阿司匹林乙酰化修饰ALDH1选择性清除大肠癌干细胞的分子机制研究

国家自然科学基金

0+阅读 · 2014年12月31日

Kronheimer-Nakajima quiver 模空间与有理曲面

国家自然科学基金

1+阅读 · 2013年12月31日

CD147-CD98复合体参与RA患者CD4+CD161+T细胞活化相关功能的新机制

国家自然科学基金

0+阅读 · 2013年12月31日

胶质瘤中长链非编码RNA HOTAIR功能网络的解析及其分子机制的研究

国家自然科学基金

0+阅读 · 2012年12月31日

RegIII信号通路与SOCS3甲基化协同调控胰腺炎症恶性转化的分子机制研究

国家自然科学基金

0+阅读 · 2012年12月31日

MiR-155/β-arrestin 2/GSK3β通路在Sca-1+心脏干细胞向心肌分化中的功能研究

国家自然科学基金

0+阅读 · 2012年12月31日

可压缩Navier-Stokes方程的一些数学问题

国家自然科学基金

0+阅读 · 2012年12月31日

随机泛函微分方程的渐近行为

国家自然科学基金

0+阅读 · 2012年12月31日

一类四阶MEMS方程的解集结构与解的渐近性态

国家自然科学基金

0+阅读 · 2011年12月31日

微信扫码咨询专知VIP会员