Bayesian优化和序列估计的推算法和近似动态方案 (Rollout Algorithms and Approximate Dynamic Programming for Bayesian Optimization and Sequential Estimation) - 专知论文

会员服务 ·

0

dynamic programming · 近似动态规划 · 估计/估计量 · 近似 · CASE ·

2022 年 12 月 29 日

Rollout Algorithms and Approximate Dynamic Programming for Bayesian Optimization and Sequential Estimation

翻译：Bayesian优化和序列估计的推算法和近似动态方案

Dimitri Bertsekas

We provide a unifying approximate dynamic programming framework that applies to a broad variety of problems involving sequential estimation. We consider first the construction of surrogate cost functions for the purposes of optimization, and we focus on the special case of Bayesian optimization, using the rollout algorithm and some of its variations. We then discuss the more general case of sequential estimation of a random vector using optimal measurement selection, and its application to problems of stochastic and adaptive control. We distinguish between adaptive control of deterministic and stochastic systems: the former are better suited for the use of rollout, while the latter are well suited for the use of rollout with certainty equivalence approximations. As an example of the deterministic case, we discuss sequential decoding problems, and a rollout algorithm for the approximate solution of the Wordle and Mastermind puzzles, recently developed in the paper [BBB22].

翻译：我们首先考虑为优化目的构建代用成本功能,我们侧重于贝叶斯优化的特例,使用推出算法及其某些变异。然后我们讨论使用最佳计量选择对随机矢量进行顺序估算的更一般性案例,以及将其应用于随机矢量控制的问题。我们区分了确定性和随机系统的适应性控制:前者更适合使用推出,而后者则非常适合使用确定性等效近似值的推出。作为确定性案例的一个实例,我们讨论了顺序解码问题,并讨论了最近在论文[BBB22]中开发的关于“Wordle”和“Mastermind”拼图近似解决方案的推出算法。

0

相关内容

dynamic programming

dynamic programming

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

专知会员服务

60+阅读 · 2022年4月22日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

IEEE TII Call For Papers

IEEE TII Call For Papers

CCF多媒体专委会

3+阅读 · 2022年3月24日

【ICIG2021】Latest News & Announcements of the Tutorial

【ICIG2021】Latest News & Announcements of the Tutorial

中国图象图形学学会CSIG

3+阅读 · 2021年12月20日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

强化学习的Unsupervised Meta-Learning

强化学习的Unsupervised Meta-Learning

CreateAMind

18+阅读 · 2019年1月7日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

MDM2介导的有丝分裂灾难- - -糖尿病肾病足细胞损伤的新机制

国家自然科学基金

0+阅读 · 2014年12月31日

基于氧化石墨烯的新型界面材料与有机太阳能电池

国家自然科学基金

0+阅读 · 2013年12月31日

microRNA调节肿瘤抑制因子Caliban应答DNA损伤的机制

国家自然科学基金

1+阅读 · 2012年12月31日

基于GH/IGF-1轴糖尿病肾病大鼠Snail 1通路及TEMT的研究

国家自然科学基金

0+阅读 · 2012年12月31日

一类非线性薄板的建模与控制

国家自然科学基金

1+阅读 · 2011年12月31日

Learning to Estimate Two Dense Depths from LiDAR and Event Data

Arxiv

0+阅读 · 2023年2月28日

Estimation-of-Distribution Algorithms for Multi-Valued Decision Variables

Arxiv

0+阅读 · 2023年2月28日

Optimistic Planning by Regularized Dynamic Programming

Arxiv

0+阅读 · 2023年2月27日

Online Black-Box Confidence Estimation of Deep Neural Networks

Arxiv

0+阅读 · 2023年2月27日

A Numerical Approach to Optimal Sequential Multi-Hypothesis Testing

Arxiv

0+阅读 · 2023年2月26日

VIP会员

文章信息

相关主题

dynamic programming

近似动态规划

估计/估计量

相关VIP内容

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

专知会员服务

60+阅读 · 2022年4月22日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

【伯克利博士论文】从推理服务到模型训练：面向大规模 LLM 智能体的高效系统构建

面向作战人员负责任地寻求生成式人工智能

《Hello-Agents》项目正式发布，一起从零学习智能体！

智能体 AI (Agentic AI) 的新进展：回归初心，预见未来

相关资讯

IEEE TII Call For Papers

IEEE TII Call For Papers

CCF多媒体专委会

3+阅读 · 2022年3月24日

【ICIG2021】Latest News & Announcements of the Tutorial

【ICIG2021】Latest News & Announcements of the Tutorial

中国图象图形学学会CSIG

3+阅读 · 2021年12月20日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

强化学习的Unsupervised Meta-Learning

强化学习的Unsupervised Meta-Learning

CreateAMind

18+阅读 · 2019年1月7日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

相关论文

Learning to Estimate Two Dense Depths from LiDAR and Event Data

Arxiv

0+阅读 · 2023年2月28日

Estimation-of-Distribution Algorithms for Multi-Valued Decision Variables

Arxiv

0+阅读 · 2023年2月28日

Optimistic Planning by Regularized Dynamic Programming

Arxiv

0+阅读 · 2023年2月27日

Online Black-Box Confidence Estimation of Deep Neural Networks

Arxiv

0+阅读 · 2023年2月27日

A Numerical Approach to Optimal Sequential Multi-Hypothesis Testing

Arxiv

0+阅读 · 2023年2月26日

相关基金

MDM2介导的有丝分裂灾难- - -糖尿病肾病足细胞损伤的新机制

国家自然科学基金

0+阅读 · 2014年12月31日

基于氧化石墨烯的新型界面材料与有机太阳能电池

国家自然科学基金

0+阅读 · 2013年12月31日

microRNA调节肿瘤抑制因子Caliban应答DNA损伤的机制

国家自然科学基金

1+阅读 · 2012年12月31日

基于GH/IGF-1轴糖尿病肾病大鼠Snail 1通路及TEMT的研究

国家自然科学基金

0+阅读 · 2012年12月31日

一类非线性薄板的建模与控制

国家自然科学基金

1+阅读 · 2011年12月31日

微信扫码咨询专知VIP会员