序列决策的线性部分监测: 等级、遗憾边界和应用 (Linear Partial Monitoring for Sequential Decision-Making: Algorithms, Regret Bounds and Applications) - 专知论文

会员服务 ·

0

线性的 · 情景 · 赌博机/老虎机 · MoDELS · Bandits ·

2023 年 2 月 7 日

Linear Partial Monitoring for Sequential Decision-Making: Algorithms, Regret Bounds and Applications

翻译：序列决策的线性部分监测: 等级、遗憾边界和应用

Johannes Kirschner,Tor Lattimore,Andreas Krause

Partial monitoring is an expressive framework for sequential decision-making with an abundance of applications, including graph-structured and dueling bandits, dynamic pricing and transductive feedback models. We survey and extend recent results on the linear formulation of partial monitoring that naturally generalizes the standard linear bandit setting. The main result is that a single algorithm, information-directed sampling (IDS), is (nearly) worst-case rate optimal in all finite-action games. We present a simple and unified analysis of stochastic partial monitoring, and further extend the model to the contextual and kernelized setting.

翻译：部分监测是连续决策的一个明确框架,其应用范围很广,包括图表结构化和决断式强盗、动态定价和传输反馈模型。我们调查并推广关于部分监测线性表述的最新结果,这种监测自然地概括了标准的线性强盗设置。主要结果是,单一算法、信息导向抽样(IDS)在所有有限行动游戏中(几乎)最差的速率是最佳的。我们对随机性部分监测进行简单和统一的分析,并将模型进一步扩展至背景和内嵌式环境。

0

相关内容

线性的

高效可扩展图神经网络的研究进展，Recent Advances in Efficient and Scalable Graph Neural Networks

高效可扩展图神经网络的研究进展，Recent Advances in Efficient and Scalable Graph Neural Networks

专知会员服务

78+阅读 · 2022年3月15日

INRIA 最新《机器学习理论》课程笔记，176页pdf

专知会员服务

51+阅读 · 2020年12月14日

【新书】人工智能Python代码，227页pdf，Python code for Artificial Intelligence: Foundations of Computational Agents

【新书】人工智能Python代码，227页pdf，Python code for Artificial Intelligence: Foundations of Computational Agents

专知会员服务

103+阅读 · 2020年6月21日

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

专知会员服务

96+阅读 · 2020年3月12日

深度学习金融应用综述论文，52页pdf，Deep Learning for Financial Applications

深度学习金融应用综述论文，52页pdf，Deep Learning for Financial Applications

专知会员服务

83+阅读 · 2020年2月18日

【新书】数字图像(影像)处理手第二版，2176pdf，Mathematical Methods in Imaging

【新书】数字图像(影像)处理手第二版，2176pdf，Mathematical Methods in Imaging

专知会员服务

93+阅读 · 2020年2月12日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

直播 | Interpretable and Trustworthy Graph Geometric Deep Learning

直播 | Interpretable and Trustworthy Graph Geometric Deep Learning

图与推荐

2+阅读 · 2022年11月2日

VCIP 2022 Call for Demos

VCIP 2022 Call for Demos

CCF多媒体专委会

1+阅读 · 2022年6月6日

征稿 | CFP：Special Issue of NLP and KG(JCR Q2，IF2.67)

征稿 | CFP：Special Issue of NLP and KG(JCR Q2，IF2.67)

开放知识图谱

1+阅读 · 2022年4月4日

局部学习的特征选择：Local-Learning-Based Feature Selection

局部学习的特征选择：Local-Learning-Based Feature Selection

我爱读PAMI

14+阅读 · 2019年9月20日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

强化学习的Unsupervised Meta-Learning

强化学习的Unsupervised Meta-Learning

CreateAMind

18+阅读 · 2019年1月7日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

强化学习族谱

强化学习族谱

CreateAMind

26+阅读 · 2017年8月2日

近似最优径向基函数插值的理论与算法研究

国家自然科学基金

0+阅读 · 2013年12月31日

Partial Spread Bent函数与Bent-Negabent函数的构造及密码学性质研究

国家自然科学基金

0+阅读 · 2013年12月31日

滑坡碎屑流持速效应机理研究

国家自然科学基金

0+阅读 · 2012年12月31日

预应力淬硬磨削复合加工新工艺机理及全参数表面完整性研究

国家自然科学基金

0+阅读 · 2012年12月31日

人肝细胞特异性分泌III型干扰素的调控机制研究

国家自然科学基金

0+阅读 · 2012年12月31日

基于贝叶斯推理的模糊逻辑强化学习模型研究

国家自然科学基金

18+阅读 · 2012年12月31日

多元代数插值的计算机数学方法

国家自然科学基金

1+阅读 · 2011年12月31日

《计算机研究与发展》学术期刊

国家自然科学基金

1+阅读 · 2011年12月31日

基于Decorin基因甲基化调控的非小细胞肺癌转移的分子机制

国家自然科学基金

0+阅读 · 2011年12月31日

高速干切削（准干切削）加工过程智能监控

国家自然科学基金

0+阅读 · 2008年12月31日

An inexact linearized proximal algorithm for a class of DC composite optimization problems and applications

Arxiv

0+阅读 · 2023年3月29日

Mixtures of All Trees

Arxiv

0+阅读 · 2023年3月29日

Towards Quantifying Calibrated Uncertainty via Deep Ensembles in Multi-output Regression Task

Arxiv

0+阅读 · 2023年3月28日

Learning linear dynamical systems under convex constraints

Arxiv

0+阅读 · 2023年3月27日

A Survey on Causal Discovery Methods for Temporal and Non-Temporal Data

Arxiv

0+阅读 · 2023年3月27日

Sequential Knockoffs for Variable Selection in Reinforcement Learning

Arxiv

0+阅读 · 2023年3月24日

Greedy Training Algorithms for Neural Networks and Applications to PDEs

Arxiv

0+阅读 · 2023年3月24日

On the convergence and sampling of randomized primal-dual algorithms and their application to parallel MRI reconstruction

Arxiv

0+阅读 · 2023年3月24日

Lower Bounds on the Bayesian Risk via Information Measures

Arxiv

0+阅读 · 2023年3月24日

Enable Deep Learning on Mobile Devices: Methods, Systems, and Applications

Arxiv

36+阅读 · 2022年4月25日

VIP会员

文章信息

相关主题

赌博机/老虎机

相关VIP内容

高效可扩展图神经网络的研究进展，Recent Advances in Efficient and Scalable Graph Neural Networks

高效可扩展图神经网络的研究进展，Recent Advances in Efficient and Scalable Graph Neural Networks

专知会员服务

78+阅读 · 2022年3月15日

INRIA 最新《机器学习理论》课程笔记，176页pdf

专知会员服务

51+阅读 · 2020年12月14日

【新书】人工智能Python代码，227页pdf，Python code for Artificial Intelligence: Foundations of Computational Agents

【新书】人工智能Python代码，227页pdf，Python code for Artificial Intelligence: Foundations of Computational Agents

专知会员服务

103+阅读 · 2020年6月21日

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

专知会员服务

96+阅读 · 2020年3月12日

深度学习金融应用综述论文，52页pdf，Deep Learning for Financial Applications

深度学习金融应用综述论文，52页pdf，Deep Learning for Financial Applications

专知会员服务

83+阅读 · 2020年2月18日

【新书】数字图像(影像)处理手第二版，2176pdf，Mathematical Methods in Imaging

【新书】数字图像(影像)处理手第二版，2176pdf，Mathematical Methods in Imaging

专知会员服务

93+阅读 · 2020年2月12日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

大语言模型智能体强化学习：全景综述

《城市滨海地区：理解复杂多变环境下的指挥控制框架》50页报告

【伯克利博士论文】从推理服务到训练：面向大规模 LLM 智能体的高效系统

美空军“顶点2025”实验：推进AI在C2、动态目标锁定与联盟集成中的应用

相关资讯

直播 | Interpretable and Trustworthy Graph Geometric Deep Learning

直播 | Interpretable and Trustworthy Graph Geometric Deep Learning

图与推荐

2+阅读 · 2022年11月2日

VCIP 2022 Call for Demos

VCIP 2022 Call for Demos

CCF多媒体专委会

1+阅读 · 2022年6月6日

征稿 | CFP：Special Issue of NLP and KG(JCR Q2，IF2.67)

征稿 | CFP：Special Issue of NLP and KG(JCR Q2，IF2.67)

开放知识图谱

1+阅读 · 2022年4月4日

局部学习的特征选择：Local-Learning-Based Feature Selection

局部学习的特征选择：Local-Learning-Based Feature Selection

我爱读PAMI

14+阅读 · 2019年9月20日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

强化学习的Unsupervised Meta-Learning

强化学习的Unsupervised Meta-Learning

CreateAMind

18+阅读 · 2019年1月7日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

强化学习族谱

强化学习族谱

CreateAMind

26+阅读 · 2017年8月2日

相关论文

An inexact linearized proximal algorithm for a class of DC composite optimization problems and applications

Arxiv

0+阅读 · 2023年3月29日

Mixtures of All Trees

Arxiv

0+阅读 · 2023年3月29日

Towards Quantifying Calibrated Uncertainty via Deep Ensembles in Multi-output Regression Task

Arxiv

0+阅读 · 2023年3月28日

Learning linear dynamical systems under convex constraints

Arxiv

0+阅读 · 2023年3月27日

A Survey on Causal Discovery Methods for Temporal and Non-Temporal Data

Arxiv

0+阅读 · 2023年3月27日

Sequential Knockoffs for Variable Selection in Reinforcement Learning

Arxiv

0+阅读 · 2023年3月24日

Greedy Training Algorithms for Neural Networks and Applications to PDEs

Arxiv

0+阅读 · 2023年3月24日

On the convergence and sampling of randomized primal-dual algorithms and their application to parallel MRI reconstruction

Arxiv

0+阅读 · 2023年3月24日

Lower Bounds on the Bayesian Risk via Information Measures

Arxiv

0+阅读 · 2023年3月24日

Enable Deep Learning on Mobile Devices: Methods, Systems, and Applications

Arxiv

36+阅读 · 2022年4月25日

相关基金

近似最优径向基函数插值的理论与算法研究

国家自然科学基金

0+阅读 · 2013年12月31日

Partial Spread Bent函数与Bent-Negabent函数的构造及密码学性质研究

国家自然科学基金

0+阅读 · 2013年12月31日

滑坡碎屑流持速效应机理研究

国家自然科学基金

0+阅读 · 2012年12月31日

预应力淬硬磨削复合加工新工艺机理及全参数表面完整性研究

国家自然科学基金

0+阅读 · 2012年12月31日

人肝细胞特异性分泌III型干扰素的调控机制研究

国家自然科学基金

0+阅读 · 2012年12月31日

基于贝叶斯推理的模糊逻辑强化学习模型研究

国家自然科学基金

18+阅读 · 2012年12月31日

多元代数插值的计算机数学方法

国家自然科学基金

1+阅读 · 2011年12月31日

《计算机研究与发展》学术期刊

国家自然科学基金

1+阅读 · 2011年12月31日

基于Decorin基因甲基化调控的非小细胞肺癌转移的分子机制

国家自然科学基金

0+阅读 · 2011年12月31日

高速干切削（准干切削）加工过程智能监控

国家自然科学基金

0+阅读 · 2008年12月31日

微信扫码咨询专知VIP会员