HyrLogLooLooLolog: 更多单日志的红心估计 (HyperLogLogLog: Cardinality Estimation With One Log More) - 专知论文

会员服务 ·

0

估计/估计量 · 讲稿 · 近似 · Apache · Google ·

2022 年 5 月 23 日

HyperLogLogLog: Cardinality Estimation With One Log More

翻译：HyrLogLooLooLolog: 更多单日志的红心估计

Matti Karppa,Rasmus Pagh

from arxiv, 10 pages, 7 figures, KDD '22

We present HyperLogLogLog, a practical compression of the HyperLogLog sketch that compresses the sketch from $O(m\log\log n)$ bits down to $m \log_2\log_2\log_2 m + O(m+\log\log n)$ bits for estimating the number of distinct elements~$n$ using $m$~registers. The algorithm works as a drop-in replacement that preserves all estimation properties of the HyperLogLog sketch, it is possible to convert back and forth between the compressed and uncompressed representations, and the compressed sketch maintains mergeability in the compressed domain. The compressed sketch can be updated in amortized constant time, assuming $n$ is sufficiently larger than $m$. We provide a C++ implementation of the sketch, and show by experimental evaluation against well-known implementations by Google and Apache that our implementation provides small sketches while maintaining competitive update and merge times. Concretely, we observed approximately a 40% reduction in the sketch size. Furthermore, we obtain as a corollary a theoretical algorithm that compresses the sketch down to $m\log_2\log_2\log_2\log_2 m+O(m\log\log\log m/\log\log m+\log\log n)$ bits.

翻译：我们展示了超LogLogLogLog的超LogLog 素描, 将素描从$O( m\log\log n) bits 压缩到$$( log_ 2\log_ 2 m + O( log\ log n) bits), 用于估算不同元素的数量 ~ 美元 ~ 美元, 使用 ~ 注册者。算法作为一个滴入替换工具, 保存超LogLog素描的所有估计属性, 压缩和未压缩的面貌之间可以互换, 压缩的素描在压缩的域中保持合并性。压缩的素描可以以折现常数时间更新, 假设美元大于 $( log_ log\ mlog\ m) 。我们提供素描图的C+ 执行情况, 并用实验性评估显示, 我们的实施在保持竞争性更新和合并时提供了小的草图。具体地说, 我们观察到了绘图大小约减少40% 。此外, 我们得到了一个必然的理论算算法, 缩略图2\ praglog\ $_ 。

0

相关内容

估计/估计量

估计/估计量

不可错过！《机器学习100讲》课程，UBC Mark Schmidt讲授

不可错过！《机器学习100讲》课程，UBC Mark Schmidt讲授

专知会员服务

76+阅读 · 2022年6月28日

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

专知会员服务

60+阅读 · 2022年4月22日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

专知会员服务

163+阅读 · 2019年10月12日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium5

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium5

中国图象图形学学会CSIG

1+阅读 · 2021年11月11日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium4

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium4

中国图象图形学学会CSIG

0+阅读 · 2021年11月10日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium3

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium3

中国图象图形学学会CSIG

0+阅读 · 2021年11月9日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium2

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium2

中国图象图形学学会CSIG

0+阅读 · 2021年11月8日

Delta-Sarcoglycan基因的两个新突变在东亚人遗传性心肌病中的致病作用及其机理

国家自然科学基金

0+阅读 · 2014年12月31日

猪肺部感染组织中中性粒细胞的募集、浸润及SHP-2介导ICAM-1信号调控机制

国家自然科学基金

0+阅读 · 2012年12月31日

PI3Kγ对血管平滑肌细胞凋亡和移植物动脉硬化的调控作用及分子机制

国家自然科学基金

0+阅读 · 2011年12月31日

复形范畴中的Gorenstein同调维数

国家自然科学基金

0+阅读 · 2009年12月31日

甘薯AGPase基因TRAP分子标记筛选及高淀粉育种新策略研究

国家自然科学基金

0+阅读 · 2008年12月31日

LIDL: Local Intrinsic Dimension Estimation Using Approximate Likelihood

Arxiv

0+阅读 · 2022年7月11日

Doubly Optimal No-Regret Online Learning in Strongly Monotone Games with Bandit Feedback

Arxiv

0+阅读 · 2022年7月10日

High-frequency Estimation of the Lévy-driven Graph Ornstein-Uhlenbeck process

Arxiv

0+阅读 · 2022年7月9日

ABC for model selection and parameter estimation of drill-string bit-rock interaction models and stochastic stability

ABC for model selection and parameter estimation of drill-string bit-rock interaction models and stochastic stability

Arxiv

0+阅读 · 2022年7月8日

Towards Open World Object Detection

Arxiv

13+阅读 · 2021年3月3日

VIP会员

文章信息

相关主题

估计/估计量

相关VIP内容

不可错过！《机器学习100讲》课程，UBC Mark Schmidt讲授

不可错过！《机器学习100讲》课程，UBC Mark Schmidt讲授

专知会员服务

76+阅读 · 2022年6月28日

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

专知会员服务

60+阅读 · 2022年4月22日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

专知会员服务

163+阅读 · 2019年10月12日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

热门VIP内容

开通专知VIP会员享更多权益服务

前沿人工智能趋势报告（Frontier AI Trends Report）

【AAAI2026】善始则事半功倍：基于前缀优化的大语言模型推理强化学习

Andrej Karpathy：2025 年 LLM 年度回顾（2025 LLM Year in Review）

音退化问题：基于输入操控的鲁棒语音转换综述

相关资讯

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium5

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium5

中国图象图形学学会CSIG

1+阅读 · 2021年11月11日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium4

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium4

中国图象图形学学会CSIG

0+阅读 · 2021年11月10日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium3

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium3

中国图象图形学学会CSIG

0+阅读 · 2021年11月9日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium2

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium2

中国图象图形学学会CSIG

0+阅读 · 2021年11月8日

相关论文

LIDL: Local Intrinsic Dimension Estimation Using Approximate Likelihood

Arxiv

0+阅读 · 2022年7月11日

Doubly Optimal No-Regret Online Learning in Strongly Monotone Games with Bandit Feedback

Arxiv

0+阅读 · 2022年7月10日

High-frequency Estimation of the Lévy-driven Graph Ornstein-Uhlenbeck process

Arxiv

0+阅读 · 2022年7月9日

ABC for model selection and parameter estimation of drill-string bit-rock interaction models and stochastic stability

ABC for model selection and parameter estimation of drill-string bit-rock interaction models and stochastic stability

Arxiv

0+阅读 · 2022年7月8日

Towards Open World Object Detection

Arxiv

13+阅读 · 2021年3月3日

相关基金

Delta-Sarcoglycan基因的两个新突变在东亚人遗传性心肌病中的致病作用及其机理

国家自然科学基金

0+阅读 · 2014年12月31日

猪肺部感染组织中中性粒细胞的募集、浸润及SHP-2介导ICAM-1信号调控机制

国家自然科学基金

0+阅读 · 2012年12月31日

PI3Kγ对血管平滑肌细胞凋亡和移植物动脉硬化的调控作用及分子机制

国家自然科学基金

0+阅读 · 2011年12月31日

复形范畴中的Gorenstein同调维数

国家自然科学基金

0+阅读 · 2009年12月31日

甘薯AGPase基因TRAP分子标记筛选及高淀粉育种新策略研究

国家自然科学基金

0+阅读 · 2008年12月31日

微信扫码咨询专知VIP会员