Debias 后门:减少以后门攻击为基础的人造比重为模式的比重</s> (Backdoor for Debias: Mitigating Model Bias with Backdoor Attack-based Artificial Bias) - 专知论文

会员服务 ·

0

有偏 · MoDELS · 可约的 · state-of-the-art · motivation ·

2023 年 3 月 1 日

Backdoor for Debias: Mitigating Model Bias with Backdoor Attack-based Artificial Bias

翻译：Debias 后门:减少以后门攻击为基础的人造比重为模式的比重

Shangxi Wu,Qiuyang He,Fangzhao Wu,Jitao Sang,Yaowei Wang,Changsheng Xu

With the swift advancement of deep learning, state-of-the-art algorithms have been utilized in various social situations. Nonetheless, some algorithms have been discovered to exhibit biases and provide unequal results. The current debiasing methods face challenges such as poor utilization of data or intricate training requirements. In this work, we found that the backdoor attack can construct an artificial bias similar to the model bias derived in standard training. Considering the strong adjustability of backdoor triggers, we are motivated to mitigate the model bias by carefully designing reverse artificial bias created from backdoor attack. Based on this, we propose a backdoor debiasing framework based on knowledge distillation, which effectively reduces the model bias from original data and minimizes security risks from the backdoor attack. The proposed solution is validated on both image and structured datasets, showing promising results. This work advances the understanding of backdoor attacks and highlights its potential for beneficial applications. The code for the study can be found at \url{https://anonymous.4open.science/r/DwB-BC07/}.

翻译：随着深层次学习的迅速发展,在各种社会情况中采用了最先进的算法,然而,还是发现了一些算法,以显示偏见和提供不平等的结果。目前的贬低方法面临着数据利用不善或训练要求复杂等挑战。在这项工作中,我们发现后门攻击可以形成类似于标准培训模式偏见的人工偏见。考虑到后门触发器的强大可调整性,我们通过仔细设计后门攻击产生的反向人为偏差来减少模型偏差。在此基础上,我们提议了一个基于知识蒸馏的后门偏差框架,有效地减少原始数据的模型偏差,并尽量减少后门攻击产生的安全风险。拟议的解决办法在图像和结构数据集上都得到验证,显示出有希望的结果。这项工作提高了对后门攻击的理解,并突出了其有利应用的潜力。研究的代码可以在\url{https://anonimous4.open.science/r/DwB-BC07/}找到。</s>

0

相关内容

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

专知会员服务

60+阅读 · 2022年4月22日

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

专知会员服务

166+阅读 · 2020年3月18日

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

专知会员服务

95+阅读 · 2020年3月12日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

专知会员服务

79+阅读 · 2019年10月10日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

专知会员服务

65+阅读 · 2019年10月9日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

VCIP 2022 Call for Demos

VCIP 2022 Call for Demos

CCF多媒体专委会

1+阅读 · 2022年6月6日

强化学习三篇论文避免遗忘等

强化学习三篇论文避免遗忘等

CreateAMind

20+阅读 · 2019年5月24日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

逆强化学习-学习人先验的动机

逆强化学习-学习人先验的动机

CreateAMind

16+阅读 · 2019年1月18日

强化学习的Unsupervised Meta-Learning

强化学习的Unsupervised Meta-Learning

CreateAMind

18+阅读 · 2019年1月7日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

disentangled-representation-papers

disentangled-representation-papers

CreateAMind

26+阅读 · 2018年9月12日

【论文推荐】最新六篇自动问答相关论文—无监督迁移学习、综述、生成式问答、QDEE、可扩展文档理解

【论文推荐】最新六篇自动问答相关论文—无监督迁移学习、综述、生成式问答、QDEE、可扩展文档理解

专知

12+阅读 · 2018年5月9日

亚稳态分子间复合物（MIC）反应特性和微观点火机理的研究

国家自然科学基金

0+阅读 · 2015年12月31日

深部充填开采留巷球应力壳与偏应力场演化及协同控制

国家自然科学基金

0+阅读 · 2015年12月31日

基于SWAN模式的卫星遥感海浪方向谱集合同化研究

国家自然科学基金

0+阅读 · 2013年12月31日

Intraflagellar Transport运输纤毛蛋白的分子机理

国家自然科学基金

0+阅读 · 2012年12月31日

Ni-M(M=Cu, Ag, Au)双金属催化剂催化甲烷水蒸气重整制氢的理论研究

国家自然科学基金

0+阅读 · 2012年12月31日

热喷涂羟基磷灰石涂层的应力-组织协同控制研究

国家自然科学基金

0+阅读 · 2012年12月31日

AB2O4(B=Al、Ga、In)基尖晶石型可见光催化剂结构和性能的理论与实验研究

国家自然科学基金

0+阅读 · 2011年12月31日

冷等离子体作用下离子液体催化甲烷转化气-液反应机理研究

国家自然科学基金

0+阅读 · 2009年12月31日

TR3相互作用新蛋白机理研究

国家自然科学基金

1+阅读 · 2008年12月31日

动力扰动下深部高应力巷道围岩分区破裂机理研究

国家自然科学基金

0+阅读 · 2008年12月31日

BackCache: Mitigating Contention-Based Cache Timing Attacks by Hiding Cache Line Evictions

Arxiv

0+阅读 · 2023年4月25日

Synthpop++: A Hybrid Framework for Generating A Country-scale Synthetic Population

Arxiv

0+阅读 · 2023年4月24日

Enhancing Fine-Tuning Based Backdoor Defense with Sharpness-Aware Minimization

Arxiv

0+阅读 · 2023年4月24日

Policy Learning under Biased Sample Selection

Arxiv

0+阅读 · 2023年4月23日

Launching a Robust Backdoor Attack under Capability Constrained Scenarios

Launching a Robust Backdoor Attack under Capability Constrained Scenarios

Arxiv

0+阅读 · 2023年4月21日

A Survey of Learning on Small Data

Arxiv

19+阅读 · 2022年7月29日

Advances in adversarial attacks and defenses in computer vision: A survey

Arxiv

22+阅读 · 2021年9月2日

Backdoor Learning: A Survey

Arxiv

14+阅读 · 2020年10月26日

Subgraph Neural Networks

Arxiv

27+阅读 · 2020年6月19日

HyperGCN: A New Method of Training Graph Convolutional Networks on Hypergraphs

HyperGCN: A New Method of Training Graph Convolutional Networks on Hypergraphs

Arxiv

13+阅读 · 2019年5月22日

VIP会员

文章信息

相关主题

state-of-the-art

相关VIP内容

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

专知会员服务

60+阅读 · 2022年4月22日

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

专知会员服务

166+阅读 · 2020年3月18日

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

专知会员服务

95+阅读 · 2020年3月12日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

专知会员服务

79+阅读 · 2019年10月10日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

专知会员服务

65+阅读 · 2019年10月9日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

【CMU博士论文】数据驱动决策中的激励、信息与不确定性

DGP双粒度提示框架：图增强大模型助力欺诈检测

【ICCV2025】ESSENTIAL：用于视频类增量学习的情景记忆与语义记忆整合

唯快不破：大型语言模型高效架构综述

相关资讯

VCIP 2022 Call for Demos

VCIP 2022 Call for Demos

CCF多媒体专委会

1+阅读 · 2022年6月6日

强化学习三篇论文避免遗忘等

强化学习三篇论文避免遗忘等

CreateAMind

20+阅读 · 2019年5月24日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

逆强化学习-学习人先验的动机

逆强化学习-学习人先验的动机

CreateAMind

16+阅读 · 2019年1月18日

强化学习的Unsupervised Meta-Learning

强化学习的Unsupervised Meta-Learning

CreateAMind

18+阅读 · 2019年1月7日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

disentangled-representation-papers

disentangled-representation-papers

CreateAMind

26+阅读 · 2018年9月12日

【论文推荐】最新六篇自动问答相关论文—无监督迁移学习、综述、生成式问答、QDEE、可扩展文档理解

【论文推荐】最新六篇自动问答相关论文—无监督迁移学习、综述、生成式问答、QDEE、可扩展文档理解

专知

12+阅读 · 2018年5月9日

相关论文

BackCache: Mitigating Contention-Based Cache Timing Attacks by Hiding Cache Line Evictions

Arxiv

0+阅读 · 2023年4月25日

Synthpop++: A Hybrid Framework for Generating A Country-scale Synthetic Population

Arxiv

0+阅读 · 2023年4月24日

Enhancing Fine-Tuning Based Backdoor Defense with Sharpness-Aware Minimization

Arxiv

0+阅读 · 2023年4月24日

Policy Learning under Biased Sample Selection

Arxiv

0+阅读 · 2023年4月23日

Launching a Robust Backdoor Attack under Capability Constrained Scenarios

Launching a Robust Backdoor Attack under Capability Constrained Scenarios

Arxiv

0+阅读 · 2023年4月21日

A Survey of Learning on Small Data

Arxiv

19+阅读 · 2022年7月29日

Advances in adversarial attacks and defenses in computer vision: A survey

Arxiv

22+阅读 · 2021年9月2日

Backdoor Learning: A Survey

Arxiv

14+阅读 · 2020年10月26日

Subgraph Neural Networks

Arxiv

27+阅读 · 2020年6月19日

HyperGCN: A New Method of Training Graph Convolutional Networks on Hypergraphs

HyperGCN: A New Method of Training Graph Convolutional Networks on Hypergraphs

Arxiv

13+阅读 · 2019年5月22日

相关基金

亚稳态分子间复合物（MIC）反应特性和微观点火机理的研究

国家自然科学基金

0+阅读 · 2015年12月31日

深部充填开采留巷球应力壳与偏应力场演化及协同控制

国家自然科学基金

0+阅读 · 2015年12月31日

基于SWAN模式的卫星遥感海浪方向谱集合同化研究

国家自然科学基金

0+阅读 · 2013年12月31日

Intraflagellar Transport运输纤毛蛋白的分子机理

国家自然科学基金

0+阅读 · 2012年12月31日

Ni-M(M=Cu, Ag, Au)双金属催化剂催化甲烷水蒸气重整制氢的理论研究

国家自然科学基金

0+阅读 · 2012年12月31日

热喷涂羟基磷灰石涂层的应力-组织协同控制研究

国家自然科学基金

0+阅读 · 2012年12月31日

AB2O4(B=Al、Ga、In)基尖晶石型可见光催化剂结构和性能的理论与实验研究

国家自然科学基金

0+阅读 · 2011年12月31日

冷等离子体作用下离子液体催化甲烷转化气-液反应机理研究

国家自然科学基金

0+阅读 · 2009年12月31日

TR3相互作用新蛋白机理研究

国家自然科学基金

1+阅读 · 2008年12月31日

动力扰动下深部高应力巷道围岩分区破裂机理研究

国家自然科学基金

0+阅读 · 2008年12月31日

微信扫码咨询专知VIP会员