Distinct Elements in Streams: An Algorithm for the (Text) Book - 专知论文

会员服务 ·

0

相互独立的 · state-of-the-art · 成对型 · 估计/估计量 · SimPLe ·

2023 年 5 月 24 日

Distinct Elements in Streams: An Algorithm for the (Text) Book

翻译：暂无翻译

Sourav Chakraborty,N. V. Vinodchandran,Kuldeep S. Meel

from arxiv, The version of the paper, as published in ESA-22, contained an error in the proof of Claim 4. The current revised version fixes the error as well as several other errors pointed by Donald E. Knuth. The main theorem and algorithm remain unchanged. The authors decided to forgo the old convention of alphabetical ordering of authors in favor of a randomized ordering, denoted by \textcircled{r}

Given a data stream $\mathcal{A} = \langle a_1, a_2, \ldots, a_m \rangle$ of $m$ elements where each $a_i \in [n]$, the Distinct Elements problem is to estimate the number of distinct elements in $\mathcal{A}$.Distinct Elements has been a subject of theoretical and empirical investigations over the past four decades resulting in space optimal algorithms for it.All the current state-of-the-art algorithms are, however, beyond the reach of an undergraduate textbook owing to their reliance on the usage of notions such as pairwise independence and universal hash functions. We present a simple, intuitive, sampling-based space-efficient algorithm whose description and the proof are accessible to undergraduates with the knowledge of basic probability theory.

翻译：暂无翻译

0

相关内容

相互独立的

相互独立的

不可错过！《机器学习100讲》课程，UBC Mark Schmidt讲授

不可错过！《机器学习100讲》课程，UBC Mark Schmidt讲授

专知会员服务

76+阅读 · 2022年6月28日

Meta最新WWW2022《联邦计算导论》教程，附77页ppt

Meta最新WWW2022《联邦计算导论》教程，附77页ppt

专知会员服务

60+阅读 · 2022年5月5日

【2022新书】高效深度学习，Efficient Deep Learning Book

【2022新书】高效深度学习，Efficient Deep Learning Book

专知会员服务

125+阅读 · 2022年4月21日

【ETH】最新《几何数据分析》2020课程，附PPT下载

专知会员服务

44+阅读 · 2020年12月18日

Linux导论，Introduction to Linux，96页ppt

Linux导论，Introduction to Linux，96页ppt

专知会员服务

81+阅读 · 2020年7月26日

UC.Berkeley CS189讲义教材:《机器学习全面指南》，185页pdf

专知会员服务

162+阅读 · 2020年1月16日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

机器学习入门的经验与建议

机器学习入门的经验与建议

专知会员服务

94+阅读 · 2019年10月10日

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

专知会员服务

79+阅读 · 2019年10月10日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

IEEE | DSC 2019诚邀稿件 (EI检索)

IEEE | DSC 2019诚邀稿件 (EI检索)

Call4Papers

10+阅读 · 2019年2月25日

强化学习的Unsupervised Meta-Learning

强化学习的Unsupervised Meta-Learning

CreateAMind

18+阅读 · 2019年1月7日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

《模式识别与机器学习(PRML)》正式开放免费下载

《模式识别与机器学习(PRML)》正式开放免费下载

AINLP

27+阅读 · 2018年11月27日

AI实战圣经《Machine Learning Yearning》第1-52章中英文版pdf分享

AI实战圣经《Machine Learning Yearning》第1-52章中英文版pdf分享

深度学习与NLP

15+阅读 · 2018年9月8日

【论文】变分推断（Variational inference)的总结

【论文】变分推断（Variational inference)的总结

机器学习研究会

39+阅读 · 2017年11月16日

【推荐】SVM实例教程

【推荐】SVM实例教程

机器学习研究会

17+阅读 · 2017年8月26日

RnCoX3n+2(R=Y,Sc,Zr,Hf,Sm Pr,Ce等,n=1,2,∞,X=Ga,In)化合物中的新超导体探索

国家自然科学基金

0+阅读 · 2014年12月31日

混凝土Weibull统计尺寸效应理论模型改进研究

国家自然科学基金

0+阅读 · 2013年12月31日

脉冲激励氢原子钟伺服系统设计

国家自然科学基金

0+阅读 · 2013年12月31日

Cyber体系脆弱性仿真分析方法研究

国家自然科学基金

3+阅读 · 2013年12月31日

Lee偏差在试验设计中的应用研究

国家自然科学基金

0+阅读 · 2013年12月31日

机电集成压电谐波传动系统

国家自然科学基金

0+阅读 · 2012年12月31日

基于盲源分离的扩谱通信抗干扰方法研究

国家自然科学基金

0+阅读 · 2011年12月31日

改进Max-SAT算法的关键技术研究

国家自然科学基金

0+阅读 · 2009年12月31日

一种适用于高维问题的Co-kriging代理模型新方法研究

国家自然科学基金

0+阅读 · 2009年12月31日

高频PWM多机逆变微电网的稳定性分析和能量优化管理

国家自然科学基金

0+阅读 · 2009年12月31日

Exploring Model Misspecification in Statistical Finite Elements via Shallow Water Equations

Arxiv

0+阅读 · 2023年7月11日

Generalization Error of First-Order Methods for Statistical Learning with Generic Oracles

Arxiv

0+阅读 · 2023年7月11日

Selective Sampling and Imitation Learning via Online Regression

Arxiv

0+阅读 · 2023年7月11日

Beyond the Two-Trials Rule

Arxiv

0+阅读 · 2023年7月10日

An Algorithm with Optimal Dimension-Dependence for Zero-Order Nonsmooth Nonconvex Stochastic Optimization

Arxiv

0+阅读 · 2023年7月10日

The WQN algorithm for EEG artifact removal in the absence of scale invariance

Arxiv

0+阅读 · 2023年7月9日

What is the meaning of proofs? A Fregean distinction in proof-theoretic semantics

Arxiv

0+阅读 · 2023年7月8日

k-strip: A novel segmentation algorithm in k-space for the application of skull stripping

Arxiv

0+阅读 · 2023年7月7日

Variational quantum regression algorithm with encoded data structure

Arxiv

0+阅读 · 2023年7月7日

A Machine-Learned Ranking Algorithm for Dynamic and Personalised Car Pooling Services

Arxiv

0+阅读 · 2023年7月6日

VIP会员

文章信息

相关主题

相互独立的

state-of-the-art

估计/估计量

相关VIP内容

不可错过！《机器学习100讲》课程，UBC Mark Schmidt讲授

不可错过！《机器学习100讲》课程，UBC Mark Schmidt讲授

专知会员服务

76+阅读 · 2022年6月28日

Meta最新WWW2022《联邦计算导论》教程，附77页ppt

Meta最新WWW2022《联邦计算导论》教程，附77页ppt

专知会员服务

60+阅读 · 2022年5月5日

【2022新书】高效深度学习，Efficient Deep Learning Book

【2022新书】高效深度学习，Efficient Deep Learning Book

专知会员服务

125+阅读 · 2022年4月21日

【ETH】最新《几何数据分析》2020课程，附PPT下载

专知会员服务

44+阅读 · 2020年12月18日

Linux导论，Introduction to Linux，96页ppt

Linux导论，Introduction to Linux，96页ppt

专知会员服务

81+阅读 · 2020年7月26日

UC.Berkeley CS189讲义教材:《机器学习全面指南》，185页pdf

专知会员服务

162+阅读 · 2020年1月16日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

机器学习入门的经验与建议

机器学习入门的经验与建议

专知会员服务

94+阅读 · 2019年10月10日

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

专知会员服务

79+阅读 · 2019年10月10日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

综述：面向移动端大语言模型的隐私与安全

运用小型语言模型解锁战术边缘人工智能优势

【博士论文】半结构化表格数据上的信息检索

《无人机飞行控制中的人工智能：基于深度强化学习的固定翼无人机高度保持策略》

相关资讯

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

IEEE | DSC 2019诚邀稿件 (EI检索)

IEEE | DSC 2019诚邀稿件 (EI检索)

Call4Papers

10+阅读 · 2019年2月25日

强化学习的Unsupervised Meta-Learning

强化学习的Unsupervised Meta-Learning

CreateAMind

18+阅读 · 2019年1月7日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

《模式识别与机器学习(PRML)》正式开放免费下载

《模式识别与机器学习(PRML)》正式开放免费下载

AINLP

27+阅读 · 2018年11月27日

AI实战圣经《Machine Learning Yearning》第1-52章中英文版pdf分享

AI实战圣经《Machine Learning Yearning》第1-52章中英文版pdf分享

深度学习与NLP

15+阅读 · 2018年9月8日

【论文】变分推断（Variational inference)的总结

【论文】变分推断（Variational inference)的总结

机器学习研究会

39+阅读 · 2017年11月16日

【推荐】SVM实例教程

【推荐】SVM实例教程

机器学习研究会

17+阅读 · 2017年8月26日

相关论文

Exploring Model Misspecification in Statistical Finite Elements via Shallow Water Equations

Arxiv

0+阅读 · 2023年7月11日

Generalization Error of First-Order Methods for Statistical Learning with Generic Oracles

Arxiv

0+阅读 · 2023年7月11日

Selective Sampling and Imitation Learning via Online Regression

Arxiv

0+阅读 · 2023年7月11日

Beyond the Two-Trials Rule

Arxiv

0+阅读 · 2023年7月10日

An Algorithm with Optimal Dimension-Dependence for Zero-Order Nonsmooth Nonconvex Stochastic Optimization

Arxiv

0+阅读 · 2023年7月10日

The WQN algorithm for EEG artifact removal in the absence of scale invariance

Arxiv

0+阅读 · 2023年7月9日

What is the meaning of proofs? A Fregean distinction in proof-theoretic semantics

Arxiv

0+阅读 · 2023年7月8日

k-strip: A novel segmentation algorithm in k-space for the application of skull stripping

Arxiv

0+阅读 · 2023年7月7日

Variational quantum regression algorithm with encoded data structure

Arxiv

0+阅读 · 2023年7月7日

A Machine-Learned Ranking Algorithm for Dynamic and Personalised Car Pooling Services

Arxiv

0+阅读 · 2023年7月6日

相关基金

RnCoX3n+2(R=Y,Sc,Zr,Hf,Sm Pr,Ce等,n=1,2,∞,X=Ga,In)化合物中的新超导体探索

国家自然科学基金

0+阅读 · 2014年12月31日

混凝土Weibull统计尺寸效应理论模型改进研究

国家自然科学基金

0+阅读 · 2013年12月31日

脉冲激励氢原子钟伺服系统设计

国家自然科学基金

0+阅读 · 2013年12月31日

Cyber体系脆弱性仿真分析方法研究

国家自然科学基金

3+阅读 · 2013年12月31日

Lee偏差在试验设计中的应用研究

国家自然科学基金

0+阅读 · 2013年12月31日

机电集成压电谐波传动系统

国家自然科学基金

0+阅读 · 2012年12月31日

基于盲源分离的扩谱通信抗干扰方法研究

国家自然科学基金

0+阅读 · 2011年12月31日

改进Max-SAT算法的关键技术研究

国家自然科学基金

0+阅读 · 2009年12月31日

一种适用于高维问题的Co-kriging代理模型新方法研究

国家自然科学基金

0+阅读 · 2009年12月31日

高频PWM多机逆变微电网的稳定性分析和能量优化管理

国家自然科学基金

0+阅读 · 2009年12月31日

微信扫码咨询专知VIP会员