FLEA: 从不可靠的培训数据中学习的公平、公平、多种来源 (FLEA: Provably Robust Fair Multisource Learning from Unreliable Training Data) - 专知论文

会员服务 ·

0

Facebook AI Research · Learning · 训练数据 · 稳健性 · 可辨认的 ·

2023 年 1 月 11 日

FLEA: Provably Robust Fair Multisource Learning from Unreliable Training Data

翻译：FLEA: 从不可靠的培训数据中学习的公平、公平、多种来源

Eugenia Iofinova,Nikola Konstantinov,Christoph H. Lampert

from arxiv, 10 pages in main text; 42 pages including bibliography and appendix. Published in Transactions of Machine Learning Research (TMLR), 2022, https://openreview.net/forum?id=XsPopigZX; project website at https://github.com/ISTAustria-CVML/FLEA

Fairness-aware learning aims at constructing classifiers that not only make accurate predictions, but also do not discriminate against specific groups. It is a fast-growing area of machine learning with far-reaching societal impact. However, existing fair learning methods are vulnerable to accidental or malicious artifacts in the training data, which can cause them to unknowingly produce unfair classifiers. In this work we address the problem of fair learning from unreliable training data in the robust multisource setting, where the available training data comes from multiple sources, a fraction of which might not be representative of the true data distribution. We introduce FLEA, a filtering-based algorithm that identifies and suppresses those data sources that would have a negative impact on fairness or accuracy if they were used for training. As such, FLEA is not a replacement of prior fairness-aware learning methods but rather an augmentation that makes any of them robust against unreliable training data. We show the effectiveness of our approach by a diverse range of experiments on multiple datasets. Additionally, we prove formally that -- given enough data -- FLEA protects the learner against corruptions as long as the fraction of affected data sources is less than half. Our source code and documentation are available at https://github.com/ISTAustria-CVML/FLEA.

翻译：公平认识的学习旨在构建不仅作出准确预测,而且不歧视特定群体的分类,这是一个快速增长的机器学习领域,具有深远的社会影响;然而,现有的公平学习方法容易在培训数据中出现意外或恶意的文物,从而导致他们不知情地产生不公平的分类者;在这项工作中,我们处理从强有力的多来源环境中不可靠的培训数据中公平学习的问题,因为现有培训数据来自多个来源,其中一小部分可能无法代表真实的数据分配。我们引入了基于过滤的算法,即查明和压制那些如果用于培训会对公平或准确性产生消极影响的数据源。因此,公平学习方法不是取代先前的公平意识学习方法,而是扩大方法,使之在不可靠的培训数据方面变得强大。我们通过在多个数据集上进行多种多样的实验来展示我们的方法的有效性。此外,我们正式证明,只要有足够的数据,FLEA将保护学习者免受腐败,只要受影响的数据源的分数为MALM/MLA。我们的源码和AFLA/MLA/MOLS/MOLs。

0

相关内容

Facebook AI Research

Facebook AI Research

Facebook AI Research

不可错过！杜克大学《因果推断》课程，全面讲述因果推理

不可错过！杜克大学《因果推断》课程，全面讲述因果推理

专知会员服务

52+阅读 · 2022年10月22日

不可错过！《机器学习100讲》课程，UBC Mark Schmidt讲授

不可错过！《机器学习100讲》课程，UBC Mark Schmidt讲授

专知会员服务

76+阅读 · 2022年6月28日

【干货书】机器学习速查手册，135页pdf

【干货书】机器学习速查手册，135页pdf

专知会员服务

127+阅读 · 2020年11月20日

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

专知会员服务

166+阅读 · 2020年3月18日

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

专知会员服务

19+阅读 · 2019年10月22日

【机器学习基础最新版】（Mathematics for Machine Learning），417页pdf

【机器学习基础最新版】（Mathematics for Machine Learning），417页pdf

专知会员服务

244+阅读 · 2019年10月21日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

IEEE TII Call For Papers

IEEE TII Call For Papers

CCF多媒体专委会

3+阅读 · 2022年3月24日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

强化学习的Unsupervised Meta-Learning

强化学习的Unsupervised Meta-Learning

CreateAMind

18+阅读 · 2019年1月7日

无监督元学习表示学习

无监督元学习表示学习

CreateAMind

27+阅读 · 2019年1月4日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

disentangled-representation-papers

disentangled-representation-papers

CreateAMind

26+阅读 · 2018年9月12日

【论文推荐】最新7篇聊天机器人（Chatbot）相关论文—触动你的心、DeepProbe、饮食推荐、知识学习、交互、挑战、管理

【论文推荐】最新7篇聊天机器人（Chatbot）相关论文—触动你的心、DeepProbe、饮食推荐、知识学习、交互、挑战、管理

专知

12+阅读 · 2018年3月15日

圆柱壳体振动陀螺品质特征的飞秒激光精密修调机理研究

国家自然科学基金

0+阅读 · 2014年12月31日

miRNA-126/SDF-1/CXCR7通路在内皮祖细胞移植促进卒中后血管新生中的作用机制研究

国家自然科学基金

0+阅读 · 2013年12月31日

RIP2调控CD40-NF-кB信号通路在血管内皮细胞损伤中的作用机制研究

国家自然科学基金

0+阅读 · 2013年12月31日

可压缩Navier-Stokes方程和Boltzmann方程解的渐近行为

国家自然科学基金

0+阅读 · 2013年12月31日

YB-1介导血管内皮细胞凋亡的分子机制研究

国家自然科学基金

0+阅读 · 2013年12月31日

抑郁症认知偏差的神经环路特征与临床意义

国家自然科学基金

0+阅读 · 2012年12月31日

有限维Banach几何与关于凸体覆盖的Hadwiger猜想

国家自然科学基金

0+阅读 · 2012年12月31日

概率并发理论

国家自然科学基金

1+阅读 · 2011年12月31日

多重耦合非线性偏微分方程组的奇性解

国家自然科学基金

0+阅读 · 2011年12月31日

脂肪因子adiponutrin在肥胖、胰岛素抵抗和2型糖尿病发病机制中的作用

国家自然科学基金

0+阅读 · 2009年12月31日

Prior and Posterior Networks: A Survey on Evidential Deep Learning Methods For Uncertainty Estimation

Arxiv

0+阅读 · 2023年3月7日

Positive unlabeled learning with tensor networks

Arxiv

0+阅读 · 2023年3月7日

Data Valuation Without Training of a Model

Arxiv

0+阅读 · 2023年3月7日

Masked Images Are Counterfactual Samples for Robust Fine-tuning

Arxiv

0+阅读 · 2023年3月6日

PRECISION: Decentralized Constrained Min-Max Learning with Low Communication and Sample Complexities

Arxiv

0+阅读 · 2023年3月5日

VRA: Out-of-Distribution Detection with variational rectified activations

Arxiv

0+阅读 · 2023年3月3日

Understanding the Role of Nonlinearity in Training Dynamics of Contrastive Learning

Arxiv

0+阅读 · 2023年3月3日

Deep Class-Incremental Learning: A Survey

Arxiv

13+阅读 · 2023年2月7日

Active Learning for Domain Adaptation: An Energy-based Approach

Arxiv

13+阅读 · 2021年12月2日

Neural Architecture Search without Training

Neural Architecture Search without Training

Arxiv

10+阅读 · 2021年6月11日

VIP会员

文章信息

相关主题

Facebook AI Research

相关VIP内容

不可错过！杜克大学《因果推断》课程，全面讲述因果推理

不可错过！杜克大学《因果推断》课程，全面讲述因果推理

专知会员服务

52+阅读 · 2022年10月22日

不可错过！《机器学习100讲》课程，UBC Mark Schmidt讲授

不可错过！《机器学习100讲》课程，UBC Mark Schmidt讲授

专知会员服务

76+阅读 · 2022年6月28日

【干货书】机器学习速查手册，135页pdf

【干货书】机器学习速查手册，135页pdf

专知会员服务

127+阅读 · 2020年11月20日

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

专知会员服务

166+阅读 · 2020年3月18日

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

专知会员服务

19+阅读 · 2019年10月22日

【机器学习基础最新版】（Mathematics for Machine Learning），417页pdf

【机器学习基础最新版】（Mathematics for Machine Learning），417页pdf

专知会员服务

244+阅读 · 2019年10月21日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

【普林斯顿博士论文】在线学习：优化、控制与学习理论

不确定环境下无人机三维路径规划研究 | 221页

【NeurIPS2025】《LeapFactual：基于条件流匹配的可靠视觉反事实解释》

大语言模型将如何改变军事指挥结构

相关资讯

IEEE TII Call For Papers

IEEE TII Call For Papers

CCF多媒体专委会

3+阅读 · 2022年3月24日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

强化学习的Unsupervised Meta-Learning

强化学习的Unsupervised Meta-Learning

CreateAMind

18+阅读 · 2019年1月7日

无监督元学习表示学习

无监督元学习表示学习

CreateAMind

27+阅读 · 2019年1月4日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

disentangled-representation-papers

disentangled-representation-papers

CreateAMind

26+阅读 · 2018年9月12日

【论文推荐】最新7篇聊天机器人（Chatbot）相关论文—触动你的心、DeepProbe、饮食推荐、知识学习、交互、挑战、管理

【论文推荐】最新7篇聊天机器人（Chatbot）相关论文—触动你的心、DeepProbe、饮食推荐、知识学习、交互、挑战、管理

专知

12+阅读 · 2018年3月15日

相关论文

Prior and Posterior Networks: A Survey on Evidential Deep Learning Methods For Uncertainty Estimation

Arxiv

0+阅读 · 2023年3月7日

Positive unlabeled learning with tensor networks

Arxiv

0+阅读 · 2023年3月7日

Data Valuation Without Training of a Model

Arxiv

0+阅读 · 2023年3月7日

Masked Images Are Counterfactual Samples for Robust Fine-tuning

Arxiv

0+阅读 · 2023年3月6日

PRECISION: Decentralized Constrained Min-Max Learning with Low Communication and Sample Complexities

Arxiv

0+阅读 · 2023年3月5日

VRA: Out-of-Distribution Detection with variational rectified activations

Arxiv

0+阅读 · 2023年3月3日

Understanding the Role of Nonlinearity in Training Dynamics of Contrastive Learning

Arxiv

0+阅读 · 2023年3月3日

Deep Class-Incremental Learning: A Survey

Arxiv

13+阅读 · 2023年2月7日

Active Learning for Domain Adaptation: An Energy-based Approach

Arxiv

13+阅读 · 2021年12月2日

Neural Architecture Search without Training

Neural Architecture Search without Training

Arxiv

10+阅读 · 2021年6月11日

相关基金

圆柱壳体振动陀螺品质特征的飞秒激光精密修调机理研究

国家自然科学基金

0+阅读 · 2014年12月31日

miRNA-126/SDF-1/CXCR7通路在内皮祖细胞移植促进卒中后血管新生中的作用机制研究

国家自然科学基金

0+阅读 · 2013年12月31日

RIP2调控CD40-NF-кB信号通路在血管内皮细胞损伤中的作用机制研究

国家自然科学基金

0+阅读 · 2013年12月31日

可压缩Navier-Stokes方程和Boltzmann方程解的渐近行为

国家自然科学基金

0+阅读 · 2013年12月31日

YB-1介导血管内皮细胞凋亡的分子机制研究

国家自然科学基金

0+阅读 · 2013年12月31日

抑郁症认知偏差的神经环路特征与临床意义

国家自然科学基金

0+阅读 · 2012年12月31日

有限维Banach几何与关于凸体覆盖的Hadwiger猜想

国家自然科学基金

0+阅读 · 2012年12月31日

概率并发理论

国家自然科学基金

1+阅读 · 2011年12月31日

多重耦合非线性偏微分方程组的奇性解

国家自然科学基金

0+阅读 · 2011年12月31日

脂肪因子adiponutrin在肥胖、胰岛素抵抗和2型糖尿病发病机制中的作用

国家自然科学基金

0+阅读 · 2009年12月31日

微信扫码咨询专知VIP会员