加快五氯苯甲醚的生长速度 (Speeding up PCA with priming)

We introduce primed-PCA (pPCA), a two-step algorithm for speeding up the approximation of principal components. This algorithm first runs any approximate-PCA method to get an initial estimate of the principal components (priming), and then applies an exact PCA in the subspace they span. Since this subspace is of small dimension in any practical use, the second step is extremely cheap computationally. Nonetheless, it improves accuracy significantly for a given computational budget across datasets. In this setup, the purpose of the priming is to narrow down the search space, and prepare the data for the second step, an exact calculation. We show formally that pPCA improves upon the priming algorithm under very mild conditions, and we provide experimental validation on both synthetic and real large-scale datasets showing that it systematically translates to improved performance. In our experiments we prime pPCA by several approximate algorithms and report an average speedup by a factor of 7.2 over Oja's rule, and a factor of 10.5 over EigenGame.

翻译：我们引入了原始- PCA( PPCA), 这是加速主元件近似的两步算法。这个算法首先运行任何近似- PCA 方法, 以初步估计主要元件( 最优), 然后在它们横跨的子空间中应用精确的 CPE 。由于这个子空间在任何实际用途中都很小, 第二步是极廉价的计算。尽管如此, 它还是大大提高了跨数据集的计算预算的准确性。在此设置中, 缩小范围的目的是缩小搜索空间, 为第二步准备数据, 精确的计算。我们正式显示 PPCA 在非常温和的条件下改进了边际算法, 我们提供合成和真实的大型数据集的实验性验证, 表明它系统地转换为改进了性能。在我们的实验中, 我们通过几个近似算法, 并报告平均速度为7.2 超过 Oja 规则的系数, 和 EigenGame 10.5 的系数。

相关内容

PCA

关注 3

在统计中，主成分分析（PCA）是一种通过最大化每个维度的方差来将较高维度空间中的数据投影到较低维度空间中的方法。给定二维，三维或更高维空间中的点集合，可以将“最佳拟合”线定义为最小化从点到线的平均平方距离的线。可以从垂直于第一条直线的方向类似地选择下一条最佳拟合线。重复此过程会产生一个正交的基础，其中数据的不同单个维度是不相关的。这些基向量称为主成分。

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

专知会员服务

60+阅读 · 2022年4月22日

INRIA 最新《机器学习理论》课程笔记，176页pdf

专知会员服务

51+阅读 · 2020年12月14日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日