公平、多代理、多武装、低 Regret 的公平多代理多武装强盗高效算法 (An Efficient Algorithm for Fair Multi-Agent Multi-Armed Bandit with Low Regret) - 专知论文

会员服务 ·

0

赌博机/老虎机 · Facebook AI Research · Learning · 优化器 · 在线 ·

2022 年 9 月 23 日

An Efficient Algorithm for Fair Multi-Agent Multi-Armed Bandit with Low Regret

翻译：公平、多代理、多武装、低 Regret 的公平多代理多武装强盗高效算法

Matthew Jones,Huy Lê Nguyen,Thy Nguyen

Recently a multi-agent variant of the classical multi-armed bandit was proposed to tackle fairness issues in online learning. Inspired by a long line of work in social choice and economics, the goal is to optimize the Nash social welfare instead of the total utility. Unfortunately previous algorithms either are not efficient or achieve sub-optimal regret in terms of the number of rounds $T$. We propose a new efficient algorithm with lower regret than even previous inefficient ones. For $N$ agents, $K$ arms, and $T$ rounds, our approach has a regret bound of $\tilde{O}(\sqrt{NKT} + NK)$. This is an improvement to the previous approach, which has regret bound of $\tilde{O}( \min(NK, \sqrt{N} K^{3/2})\sqrt{T})$. We also complement our efficient algorithm with an inefficient approach with $\tilde{O}(\sqrt{KT} + N^2K)$ regret. The experimental findings confirm the effectiveness of our efficient algorithm compared to the previous approaches.

翻译：最近提出了经典多武装土匪的多剂变体,以解决网上学习中的公平问题。在社会选择和经济学方面一长串工作的启发下,目标是优化纳什的社会福利,而不是总效用。不幸的是,以前的算法不是效率不高,就是在回合数方面没有达到最优的遗憾。我们提出的新的高效算法比以往低遗憾,甚至比以往低效率的算法要低。对于美元代理法、美元军火和美元回合,我们的方法对美元tilde{O}(\qrt{NKT}+NK)有遗憾。这是对前一种方法的改进,因为前一种方法对美元tilde{O}(\qrt{NQ2K}(\qrt{K}+N%2K)有遗憾。实验结果证实了我们有效算法与以往方法相比的有效性。

0

相关内容

赌博机/老虎机

赌博机/老虎机

高效可扩展图神经网络的研究进展，Recent Advances in Efficient and Scalable Graph Neural Networks

高效可扩展图神经网络的研究进展，Recent Advances in Efficient and Scalable Graph Neural Networks

专知会员服务

78+阅读 · 2022年3月15日

【跨语言BERT模型大集合】Transfer learning is increasingly going multilingual with language-specific BERT models

专知会员服务

54+阅读 · 2020年1月30日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

Stabilizing Transformers for Reinforcement Learning

Stabilizing Transformers for Reinforcement Learning

专知会员服务

60+阅读 · 2019年10月17日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

专知会员服务

65+阅读 · 2019年10月9日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

VCIP 2022 Call for Special Session Proposals

VCIP 2022 Call for Special Session Proposals

CCF多媒体专委会

1+阅读 · 2022年4月1日

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

ACM TOMM Call for Papers

ACM TOMM Call for Papers

CCF多媒体专委会

2+阅读 · 2022年3月23日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

无监督元学习表示学习

无监督元学习表示学习

CreateAMind

27+阅读 · 2019年1月4日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

disentangled-representation-papers

disentangled-representation-papers

CreateAMind

26+阅读 · 2018年9月12日

再生核希尔伯特空间图像稀疏表达算法研究

国家自然科学基金

1+阅读 · 2013年12月31日

基于时空域模型分解策略的流程企业级协同优化方法研究

国家自然科学基金

0+阅读 · 2013年12月31日

基于SURE/PURE准则的图像盲反卷积算法研究

国家自然科学基金

3+阅读 · 2013年12月31日

基于合成光学孔径的延长光学相干层析成像焦深的方法研究

国家自然科学基金

0+阅读 · 2013年12月31日

基于CS算法的数字信号压缩和高效数字系统设计的研究

国家自然科学基金

0+阅读 · 2012年12月31日

维生素D和维生素D受体基因多态性在2型糖尿病发病中的作用研究

国家自然科学基金

0+阅读 · 2012年12月31日

海量、动态、嘈杂语义数据集上的递增随时推理方法研究

国家自然科学基金

0+阅读 · 2012年12月31日

地下管线磁异常三层分量联合反演成像探测新方法研究

国家自然科学基金

0+阅读 · 2012年12月31日

基于list-mode数据的快速SART真3D PET断层重建算法的研究

国家自然科学基金

0+阅读 · 2011年12月31日

矩阵分解的低延迟并行算法

国家自然科学基金

0+阅读 · 2009年12月31日

Quasi-Newton Steps for Efficient Online Exp-Concave Optimization

Arxiv

0+阅读 · 2022年11月2日

An efficient algorithm for the $\ell_{p}$ norm based metric nearness problem

Arxiv

0+阅读 · 2022年11月2日

ProtoBandit: Efficient Prototype Selection via Multi-Armed Bandits

Arxiv

0+阅读 · 2022年11月1日

MARS: A second-order reduction algorithm for high-dimensional sparse precision matrices estimation

Arxiv

0+阅读 · 2022年11月1日

L-GreCo: An Efficient and General Framework for Layerwise-Adaptive Gradient Compression

Arxiv

0+阅读 · 2022年10月31日

Linear regression with partially mismatched data: local search with theoretical guarantees

Arxiv

0+阅读 · 2022年10月31日

Learning to Compare Nodes in Branch and Bound with Graph Neural Networks

Arxiv

0+阅读 · 2022年10月30日

On the Efficient Implementation of the Matrix Exponentiated Gradient Algorithm for Low-Rank Matrix Optimization

Arxiv

0+阅读 · 2022年10月30日

Federated X-Armed Bandit

Arxiv

0+阅读 · 2022年10月28日

Dynamic Bandits with an Auto-Regressive Temporal Structure

Arxiv

0+阅读 · 2022年10月28日

VIP会员

文章信息

相关主题

赌博机/老虎机

Facebook AI Research

相关VIP内容

高效可扩展图神经网络的研究进展，Recent Advances in Efficient and Scalable Graph Neural Networks

高效可扩展图神经网络的研究进展，Recent Advances in Efficient and Scalable Graph Neural Networks

专知会员服务

78+阅读 · 2022年3月15日

【跨语言BERT模型大集合】Transfer learning is increasingly going multilingual with language-specific BERT models

专知会员服务

54+阅读 · 2020年1月30日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

Stabilizing Transformers for Reinforcement Learning

Stabilizing Transformers for Reinforcement Learning

专知会员服务

60+阅读 · 2019年10月17日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

专知会员服务

65+阅读 · 2019年10月9日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

《物联网（IoT）中的无人机通信高效控制》135页

《在GNSS信号降级环境中利用共识实现无人机集群稳健协调》

中程单向攻击无人机的战略意义：俄乌战争启示

《面向无人机集群的避障动态传感器覆盖算法》最新38页

相关资讯

VCIP 2022 Call for Special Session Proposals

VCIP 2022 Call for Special Session Proposals

CCF多媒体专委会

1+阅读 · 2022年4月1日

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

ACM TOMM Call for Papers

ACM TOMM Call for Papers

CCF多媒体专委会

2+阅读 · 2022年3月23日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

无监督元学习表示学习

无监督元学习表示学习

CreateAMind

27+阅读 · 2019年1月4日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

disentangled-representation-papers

disentangled-representation-papers

CreateAMind

26+阅读 · 2018年9月12日

相关论文

Quasi-Newton Steps for Efficient Online Exp-Concave Optimization

Arxiv

0+阅读 · 2022年11月2日

An efficient algorithm for the $\ell_{p}$ norm based metric nearness problem

Arxiv

0+阅读 · 2022年11月2日

ProtoBandit: Efficient Prototype Selection via Multi-Armed Bandits

Arxiv

0+阅读 · 2022年11月1日

MARS: A second-order reduction algorithm for high-dimensional sparse precision matrices estimation

Arxiv

0+阅读 · 2022年11月1日

L-GreCo: An Efficient and General Framework for Layerwise-Adaptive Gradient Compression

Arxiv

0+阅读 · 2022年10月31日

Linear regression with partially mismatched data: local search with theoretical guarantees

Arxiv

0+阅读 · 2022年10月31日

Learning to Compare Nodes in Branch and Bound with Graph Neural Networks

Arxiv

0+阅读 · 2022年10月30日

On the Efficient Implementation of the Matrix Exponentiated Gradient Algorithm for Low-Rank Matrix Optimization

Arxiv

0+阅读 · 2022年10月30日

Federated X-Armed Bandit

Arxiv

0+阅读 · 2022年10月28日

Dynamic Bandits with an Auto-Regressive Temporal Structure

Arxiv

0+阅读 · 2022年10月28日

相关基金

再生核希尔伯特空间图像稀疏表达算法研究

国家自然科学基金

1+阅读 · 2013年12月31日

基于时空域模型分解策略的流程企业级协同优化方法研究

国家自然科学基金

0+阅读 · 2013年12月31日

基于SURE/PURE准则的图像盲反卷积算法研究

国家自然科学基金

3+阅读 · 2013年12月31日

基于合成光学孔径的延长光学相干层析成像焦深的方法研究

国家自然科学基金

0+阅读 · 2013年12月31日

基于CS算法的数字信号压缩和高效数字系统设计的研究

国家自然科学基金

0+阅读 · 2012年12月31日

维生素D和维生素D受体基因多态性在2型糖尿病发病中的作用研究

国家自然科学基金

0+阅读 · 2012年12月31日

海量、动态、嘈杂语义数据集上的递增随时推理方法研究

国家自然科学基金

0+阅读 · 2012年12月31日

地下管线磁异常三层分量联合反演成像探测新方法研究

国家自然科学基金

0+阅读 · 2012年12月31日

基于list-mode数据的快速SART真3D PET断层重建算法的研究

国家自然科学基金

0+阅读 · 2011年12月31日

矩阵分解的低延迟并行算法

国家自然科学基金

0+阅读 · 2009年12月31日

微信扫码咨询专知VIP会员