具有背景信息的最佳武器标识 (Semiparametric Best Arm Identification with Contextual Information) - 专知论文

会员服务 ·

0

INFORMS · ARM · 可辨认的 · 赌博机/老虎机 · 随机采样 ·

2022 年 9 月 15 日

Semiparametric Best Arm Identification with Contextual Information

翻译：具有背景信息的最佳武器标识

Masahiro Kato,Masaaki Imaizumi,Takuya Ishihara,Toru Kitagawa

We study best-arm identification with a fixed budget and contextual (covariate) information in stochastic multi-armed bandit problems. In each round, after observing contextual information, we choose a treatment arm using past observations and current context. Our goal is to identify the best treatment arm, a treatment arm with the maximal expected reward marginalized over the contextual distribution, with a minimal probability of misidentification. First, we derive semiparametric lower bounds for this problem, where we regard the gaps between the expected rewards of the best and suboptimal treatment arms as parameters of interest, and all other parameters, such as the expected rewards conditioned on contexts, as the nuisance parameters. We then develop the "Contextual RS-AIPW strategy," which consists of the random sampling (RS) rule tracking a target allocation ratio and the recommendation rule using the augmented inverse probability weighting (AIPW) estimator. Our proposed Contextual RS-AIPW strategy is optimal because the upper bound for the probability of misidentification matches the semiparametric lower bound when the budget goes to infinity, and the gaps converge to zero.

翻译：我们用固定预算和背景(共变)信息研究固定预算和多武装土匪问题的最佳武器识别信息。每回合,在观察背景信息后,我们使用以往的观察和当前背景选择一个处理臂。我们的目标是确定最佳处理臂,这是在背景分布中处于边缘地位的最大预期奖赏的处理臂,其误判概率最小。首先,我们从这一问题中得出半对称下限,将最佳和次最佳处理臂的预期奖赏视为利益参数,以及所有其他参数,例如环境条件下的预期奖赏,作为骚扰参数。然后我们制定“原始的RS-AIPW战略 ”, 其中包括随机抽样(RS) 规则, 跟踪目标分配比率, 以及建议规则, 使用增加的反概率加权(AIPW) 。我们提议的RS-AIPW 框架战略是最佳的, 因为错误识别概率的上限在预算走向无限性时与半偏差的较低界限相匹配, 差距接近于零。

0

相关内容

INFORMS

《计算机信息》杂志发表高质量的论文，扩大了运筹学和计算的范围，寻求有关理论、方法、实验、系统和应用方面的原创研究论文、新颖的调查和教程论文，以及描述新的和有用的软件工具的论文。官网链接：https://pubsonline.informs.org/journal/ijoc

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

专知会员服务

60+阅读 · 2022年4月22日

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

专知会员服务

19+阅读 · 2019年10月22日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

机器学习入门的经验与建议

机器学习入门的经验与建议

专知会员服务

94+阅读 · 2019年10月10日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

专知会员服务

65+阅读 · 2019年10月9日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

VCIP 2022 Call for Demos

VCIP 2022 Call for Demos

CCF多媒体专委会

1+阅读 · 2022年6月6日

VCIP 2022 Call for Special Session Proposals

VCIP 2022 Call for Special Session Proposals

CCF多媒体专委会

1+阅读 · 2022年4月1日

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

ACM TOMM Call for Papers

ACM TOMM Call for Papers

CCF多媒体专委会

2+阅读 · 2022年3月23日

【ICIG2021】Latest News & Announcements of the Tutorial

【ICIG2021】Latest News & Announcements of the Tutorial

中国图象图形学学会CSIG

3+阅读 · 2021年12月20日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

disentangled-representation-papers

disentangled-representation-papers

CreateAMind

26+阅读 · 2018年9月12日

基于天然产物Drimenal的新型杀菌剂分子设计、合成及构效关系研究

国家自然科学基金

0+阅读 · 2013年12月31日

基于Cre/loxP系统的肝特异性表达REGγ转基因小鼠的建立及脂质代谢分析

国家自然科学基金

0+阅读 · 2013年12月31日

柑橘黄龙病亚洲种病原( Cadidatus Liberibacter assiaticus)重组抗体的研究

国家自然科学基金

0+阅读 · 2012年12月31日

MeCP2-PTEN调控神经干细胞增殖分化影响孤独症发生的分子机制

国家自然科学基金

0+阅读 · 2012年12月31日

基于Steger-Warming FVS 的长管道气液两相瞬变流计算及其水锤的气阀防护研究

国家自然科学基金

0+阅读 · 2012年12月31日

基于模态灵敏度分析的网状可展天线形面主动控制机理及实验研究

国家自然科学基金

0+阅读 · 2012年12月31日

考虑班轮公司和货代公司委托代理关系的集装箱调度管理

国家自然科学基金

0+阅读 · 2011年12月31日

面向属性的CPN建模及On the Fly辅助的测试生成方法研究

国家自然科学基金

0+阅读 · 2011年12月31日

多维高次有限元超收敛后处理研究

国家自然科学基金

0+阅读 · 2011年12月31日

宫颈癌干细胞的特异基因表达分析

国家自然科学基金

0+阅读 · 2009年12月31日

Region of Interest focused MRI to Synthetic CT Translation using Regression and Classification Multi-task Network

Arxiv

0+阅读 · 2022年10月25日

Confidence-Calibrated Face and Kinship Verification

Arxiv

0+阅读 · 2022年10月25日

Cost-Effective Online Contextual Model Selection

Arxiv

0+阅读 · 2022年10月24日

Regularized Nonlinear Regression with Dependent Errors and its Application to a Biomechanical Model

Arxiv

0+阅读 · 2022年10月24日

Local Metric Learning for Off-Policy Evaluation in Contextual Bandits with Continuous Actions

Arxiv

0+阅读 · 2022年10月24日

On Elimination Strategies for Bandit Fixed-Confidence Identification

Arxiv

0+阅读 · 2022年10月24日

Static Information Flow Control Made Simpler

Arxiv

0+阅读 · 2022年10月24日

Covariate adjustment in multi-armed, possibly factorial experiments

Arxiv

0+阅读 · 2022年10月24日

Fast Beam Alignment via Pure Exploration in Multi-armed Bandits

Arxiv

0+阅读 · 2022年10月23日

Tackling cyclicity in causal models with cross-sectional data using a partial least square approach. Implication for the sequential model on internet appropriation

Arxiv

0+阅读 · 2022年10月22日

VIP会员

文章信息

相关主题

赌博机/老虎机

相关VIP内容

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

专知会员服务

60+阅读 · 2022年4月22日

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

专知会员服务

19+阅读 · 2019年10月22日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

机器学习入门的经验与建议

机器学习入门的经验与建议

专知会员服务

94+阅读 · 2019年10月10日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

专知会员服务

65+阅读 · 2019年10月9日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

《物联网（IoT）中的无人机通信高效控制》135页

《在GNSS信号降级环境中利用共识实现无人机集群稳健协调》

中程单向攻击无人机的战略意义：俄乌战争启示

《面向无人机集群的避障动态传感器覆盖算法》最新38页

相关资讯

VCIP 2022 Call for Demos

VCIP 2022 Call for Demos

CCF多媒体专委会

1+阅读 · 2022年6月6日

VCIP 2022 Call for Special Session Proposals

VCIP 2022 Call for Special Session Proposals

CCF多媒体专委会

1+阅读 · 2022年4月1日

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

ACM TOMM Call for Papers

ACM TOMM Call for Papers

CCF多媒体专委会

2+阅读 · 2022年3月23日

【ICIG2021】Latest News & Announcements of the Tutorial

【ICIG2021】Latest News & Announcements of the Tutorial

中国图象图形学学会CSIG

3+阅读 · 2021年12月20日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

disentangled-representation-papers

disentangled-representation-papers

CreateAMind

26+阅读 · 2018年9月12日

相关论文

Region of Interest focused MRI to Synthetic CT Translation using Regression and Classification Multi-task Network

Arxiv

0+阅读 · 2022年10月25日

Confidence-Calibrated Face and Kinship Verification

Arxiv

0+阅读 · 2022年10月25日

Cost-Effective Online Contextual Model Selection

Arxiv

0+阅读 · 2022年10月24日

Regularized Nonlinear Regression with Dependent Errors and its Application to a Biomechanical Model

Arxiv

0+阅读 · 2022年10月24日

Local Metric Learning for Off-Policy Evaluation in Contextual Bandits with Continuous Actions

Arxiv

0+阅读 · 2022年10月24日

On Elimination Strategies for Bandit Fixed-Confidence Identification

Arxiv

0+阅读 · 2022年10月24日

Static Information Flow Control Made Simpler

Arxiv

0+阅读 · 2022年10月24日

Covariate adjustment in multi-armed, possibly factorial experiments

Arxiv

0+阅读 · 2022年10月24日

Fast Beam Alignment via Pure Exploration in Multi-armed Bandits

Arxiv

0+阅读 · 2022年10月23日

Tackling cyclicity in causal models with cross-sectional data using a partial least square approach. Implication for the sequential model on internet appropriation

Arxiv

0+阅读 · 2022年10月22日

相关基金

基于天然产物Drimenal的新型杀菌剂分子设计、合成及构效关系研究

国家自然科学基金

0+阅读 · 2013年12月31日

基于Cre/loxP系统的肝特异性表达REGγ转基因小鼠的建立及脂质代谢分析

国家自然科学基金

0+阅读 · 2013年12月31日

柑橘黄龙病亚洲种病原( Cadidatus Liberibacter assiaticus)重组抗体的研究

国家自然科学基金

0+阅读 · 2012年12月31日

MeCP2-PTEN调控神经干细胞增殖分化影响孤独症发生的分子机制

国家自然科学基金

0+阅读 · 2012年12月31日

基于Steger-Warming FVS 的长管道气液两相瞬变流计算及其水锤的气阀防护研究

国家自然科学基金

0+阅读 · 2012年12月31日

基于模态灵敏度分析的网状可展天线形面主动控制机理及实验研究

国家自然科学基金

0+阅读 · 2012年12月31日

考虑班轮公司和货代公司委托代理关系的集装箱调度管理

国家自然科学基金

0+阅读 · 2011年12月31日

面向属性的CPN建模及On the Fly辅助的测试生成方法研究

国家自然科学基金

0+阅读 · 2011年12月31日

多维高次有限元超收敛后处理研究

国家自然科学基金

0+阅读 · 2011年12月31日

宫颈癌干细胞的特异基因表达分析

国家自然科学基金

0+阅读 · 2009年12月31日

微信扫码咨询专知VIP会员