低质询预算制度中的简单而高效的硬标签黑盒 (Simple and Efficient Hard Label Black-box Adversarial Attacks in Low Query Budget Regimes)

We focus on the problem of black-box adversarial attacks, where the aim is to generate adversarial examples for deep learning models solely based on information limited to output label~(hard label) to a queried data input. We propose a simple and efficient Bayesian Optimization~(BO) based approach for developing black-box adversarial attacks. Issues with BO's performance in high dimensions are avoided by searching for adversarial examples in a structured low-dimensional subspace. We demonstrate the efficacy of our proposed attack method by evaluating both $\ell_\infty$ and $\ell_2$ norm constrained untargeted and targeted hard label black-box attacks on three standard datasets - MNIST, CIFAR-10 and ImageNet. Our proposed approach consistently achieves 2x to 10x higher attack success rate while requiring 10x to 20x fewer queries compared to the current state-of-the-art black-box adversarial attacks.

翻译：我们的重点是黑盒对抗性攻击问题,目的是仅仅根据限于输出标签~(硬标签)到查询数据输入的信息,为深层次学习模式生成对抗性例子。我们提出了一种简单而高效的巴伊西亚优化~(BO)基于方法来发展黑盒对抗性攻击。通过在结构化低维次空间中寻找对立实例,可以避免BO高层面的性能问题。我们通过对三种标准数据集----MNIST、CIFAR-10和图像网络----的非针对性和有针对性的硬标签黑盒攻击进行评估,显示了我们拟议攻击方法的有效性。我们提出的方法一贯达到2x至10倍高攻击性攻击成功率,同时比目前最先进的黑盒对抗性攻击少10x20x查询。

相关内容

黑盒

关注 1

在科学，计算和工程学中，黑盒是一种设备，系统或对象，可以根据其输入和输出（或传输特性）对其进行查看，而无需对其内部工作有任何了解。它的实现是“不透明的”（黑色）。几乎任何事物都可以被称为黑盒：晶体管，引擎，算法，人脑，机构或政府。为了使用典型的“黑匣子方法”来分析建模为开放系统的事物，仅考虑刺激/响应的行为，以推断（未知）盒子。该黑匣子系统的通常表示形式是在该方框中居中的数据流程图。黑盒的对立面是一个内部组件或逻辑可用于检查的系统，通常将其称为白盒（有时也称为“透明盒”或“玻璃盒”）。

【CVPR2021】坐标注意力的高效移动网络设计

专知会员服务

23+阅读 · 2021年3月9日

近期必读的六篇AAAI 2021【对抗攻击（Adversarial Attack）】相关论文和代码

专知会员服务

55+阅读 · 2021年2月17日

【Google】深度学习对抗鲁棒性，43页ppt

专知会员服务

45+阅读 · 2020年10月31日