Stochastic Window Mean-payoff Games - 专知论文

会员服务 ·

0

Microsoft Windows · 回合 · MoDELS · 讲稿 · 滑动窗口 ·

2023 年 4 月 23 日

Stochastic Window Mean-payoff Games

翻译：暂无翻译

Laurent Doyen,Pranshu Gaba,Shibashis Guha

from arxiv, 39 pages

Stochastic two-player games model systems with both adversarial and stochastic environment. The adversarial environment is modeled by a player (Player 2) who tries to prevent the system (Player 1) from achieving its objective. We consider finitary versions of the traditional mean-payoff objective, replacing the long-run average of the payoffs by payoff average computed over a finite sliding window. Two variants have been considered; in one variant, the maximum window length is fixed and given, while in the other, it is not fixed but is required to be bounded. For both variants, we present complexity bounds and algorithmic solutions for computing strategies for Player 1 to ensure that the objective is satisfied with positive probability, with probability 1, or with a probability at least $p$. The solution crucially relies on a reduction to the special case of nonstochastic two-player games. We give a general characterization of prefix-independent objectives for which this reduction holds. The positive and almost-sure decision problems are in ${\sf PTIME}$ for the fixed variant and in ${\sf NP \cap coNP}$ for the bounded variant. For arbitrary $p$, the decision problem is in ${\sf NP \cap coNP}$ for both variants, thus matching the bounds for simple stochastic games. The memory requirements for both players in stochastic games are also the same as for nonstochastic games by our reduction. Further, for nonstochastic games, we improve upon the upper bound on the memory requirement of Player 1 and the lower bound on the memory requirement of Player 2. To the best of our knowledge, this is the first work to consider stochastic games with finitary quantitative objectives.

翻译：暂无翻译

0

相关内容

Microsoft Windows

Microsoft Windows

Microsoft Windows（视窗操作系统）是微软公司推出的一系列操作系统。它问世于1985年，当时是DOS之下的操作环境，而后其后续版本作逐渐发展成为个人电脑和服务器用户设计的操作系统。

INRIA 最新《机器学习理论》课程笔记，176页pdf

专知会员服务

51+阅读 · 2020年12月14日

Linux导论，Introduction to Linux，96页ppt

Linux导论，Introduction to Linux，96页ppt

专知会员服务

82+阅读 · 2020年7月26日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

50+阅读 · 2019年10月17日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

Stabilizing Transformers for Reinforcement Learning

Stabilizing Transformers for Reinforcement Learning

专知会员服务

60+阅读 · 2019年10月17日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

机器学习入门的经验与建议

机器学习入门的经验与建议

专知会员服务

94+阅读 · 2019年10月10日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

局部学习的特征选择：Local-Learning-Based Feature Selection

局部学习的特征选择：Local-Learning-Based Feature Selection

我爱读PAMI

14+阅读 · 2019年9月20日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

深度自进化聚类：Deep Self-Evolution Clustering

深度自进化聚类：Deep Self-Evolution Clustering

我爱读PAMI

15+阅读 · 2019年4月13日

逆强化学习-学习人先验的动机

逆强化学习-学习人先验的动机

CreateAMind

16+阅读 · 2019年1月18日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

【论文】变分推断（Variational inference)的总结

【论文】变分推断（Variational inference)的总结

机器学习研究会

39+阅读 · 2017年11月16日

【推荐】YOLO实时目标检测(6fps)

【推荐】YOLO实时目标检测(6fps)

机器学习研究会

20+阅读 · 2017年11月5日

【推荐】RNN/LSTM时序预测

【推荐】RNN/LSTM时序预测

机器学习研究会

25+阅读 · 2017年9月8日

最优控制的快速算法

国家自然科学基金

0+阅读 · 2014年12月31日

糖肾方对糖尿病肾病蛋白聚糖介导肾脏脂质沉积的机制研究

国家自然科学基金

0+阅读 · 2014年12月31日

流形上的Bakry-Emery曲率，泛函不等式和热核分析

国家自然科学基金

0+阅读 · 2012年12月31日

基于压缩感知的机载激光扫描数据完好性检验及特征级融合

国家自然科学基金

0+阅读 · 2012年12月31日

内燃机核基动态监测诊断方法研究

国家自然科学基金

0+阅读 · 2012年12月31日

TGF-β1通路调控MET在滑膜肉瘤双相分化和侵袭转移中作用及机制

国家自然科学基金

0+阅读 · 2012年12月31日

机载InSAR区域网平差方法研究

国家自然科学基金

0+阅读 · 2012年12月31日

压缩感知框架下多视光学遥感影像超分辨率重建方法

国家自然科学基金

0+阅读 · 2011年12月31日

赋值理论与几何不等式的研究

国家自然科学基金

1+阅读 · 2011年12月31日

亚椭圆算子的泛函不等式和热核分析

国家自然科学基金

0+阅读 · 2011年12月31日

Stochastic noise can be helpful for variational quantum algorithms

Arxiv

0+阅读 · 2023年6月8日

Unconstrained Online Learning with Unbounded Losses

Arxiv

0+阅读 · 2023年6月8日

Fisher information of correlated stochastic processes

Arxiv

0+阅读 · 2023年6月7日

Temporal Difference Learning with Continuous Time and State in the Stochastic Setting

Arxiv

0+阅读 · 2023年6月7日

Matroid-Constrained Vertex Cover

Arxiv

0+阅读 · 2023年6月7日

Stochastic Collapse: How Gradient Noise Attracts SGD Dynamics Towards Simpler Subnetworks

Arxiv

0+阅读 · 2023年6月7日

Finding Counterfactually Optimal Action Sequences in Continuous State Spaces

Arxiv

0+阅读 · 2023年6月6日

Correlated Pseudorandomness from the Hardness of Quasi-Abelian Decoding

Arxiv

0+阅读 · 2023年6月6日

Nonlinear Distributionally Robust Optimization

Arxiv

0+阅读 · 2023年6月5日

Sketching low-rank matrices with a shared column space by convex programming

Arxiv

0+阅读 · 2023年6月5日

VIP会员

文章信息

相关主题

Microsoft Windows

相关VIP内容

INRIA 最新《机器学习理论》课程笔记，176页pdf

专知会员服务

51+阅读 · 2020年12月14日

Linux导论，Introduction to Linux，96页ppt

Linux导论，Introduction to Linux，96页ppt

专知会员服务

82+阅读 · 2020年7月26日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

50+阅读 · 2019年10月17日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

Stabilizing Transformers for Reinforcement Learning

Stabilizing Transformers for Reinforcement Learning

专知会员服务

60+阅读 · 2019年10月17日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

机器学习入门的经验与建议

机器学习入门的经验与建议

专知会员服务

94+阅读 · 2019年10月10日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

《在单一作战合成环境（SSE）中运用人工智能与大型语言模型以提供灵活人文地形及可信角色组》报告

《俄罗斯的未来战争方式第二部分：核威慑》报告

《提示战争：大语言模型如何决定军事干预》报告

《俄罗斯的未来战争方式第三部分：军事改革》报告

相关资讯

局部学习的特征选择：Local-Learning-Based Feature Selection

局部学习的特征选择：Local-Learning-Based Feature Selection

我爱读PAMI

14+阅读 · 2019年9月20日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

深度自进化聚类：Deep Self-Evolution Clustering

深度自进化聚类：Deep Self-Evolution Clustering

我爱读PAMI

15+阅读 · 2019年4月13日

逆强化学习-学习人先验的动机

逆强化学习-学习人先验的动机

CreateAMind

16+阅读 · 2019年1月18日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

【论文】变分推断（Variational inference)的总结

【论文】变分推断（Variational inference)的总结

机器学习研究会

39+阅读 · 2017年11月16日

【推荐】YOLO实时目标检测(6fps)

【推荐】YOLO实时目标检测(6fps)

机器学习研究会

20+阅读 · 2017年11月5日

【推荐】RNN/LSTM时序预测

【推荐】RNN/LSTM时序预测

机器学习研究会

25+阅读 · 2017年9月8日

相关论文

Stochastic noise can be helpful for variational quantum algorithms

Arxiv

0+阅读 · 2023年6月8日

Unconstrained Online Learning with Unbounded Losses

Arxiv

0+阅读 · 2023年6月8日

Fisher information of correlated stochastic processes

Arxiv

0+阅读 · 2023年6月7日

Temporal Difference Learning with Continuous Time and State in the Stochastic Setting

Arxiv

0+阅读 · 2023年6月7日

Matroid-Constrained Vertex Cover

Arxiv

0+阅读 · 2023年6月7日

Stochastic Collapse: How Gradient Noise Attracts SGD Dynamics Towards Simpler Subnetworks

Arxiv

0+阅读 · 2023年6月7日

Finding Counterfactually Optimal Action Sequences in Continuous State Spaces

Arxiv

0+阅读 · 2023年6月6日

Correlated Pseudorandomness from the Hardness of Quasi-Abelian Decoding

Arxiv

0+阅读 · 2023年6月6日

Nonlinear Distributionally Robust Optimization

Arxiv

0+阅读 · 2023年6月5日

Sketching low-rank matrices with a shared column space by convex programming

Arxiv

0+阅读 · 2023年6月5日

相关基金

最优控制的快速算法

国家自然科学基金

0+阅读 · 2014年12月31日

糖肾方对糖尿病肾病蛋白聚糖介导肾脏脂质沉积的机制研究

国家自然科学基金

0+阅读 · 2014年12月31日

流形上的Bakry-Emery曲率，泛函不等式和热核分析

国家自然科学基金

0+阅读 · 2012年12月31日

基于压缩感知的机载激光扫描数据完好性检验及特征级融合

国家自然科学基金

0+阅读 · 2012年12月31日

内燃机核基动态监测诊断方法研究

国家自然科学基金

0+阅读 · 2012年12月31日

TGF-β1通路调控MET在滑膜肉瘤双相分化和侵袭转移中作用及机制

国家自然科学基金

0+阅读 · 2012年12月31日

机载InSAR区域网平差方法研究

国家自然科学基金

0+阅读 · 2012年12月31日

压缩感知框架下多视光学遥感影像超分辨率重建方法

国家自然科学基金

0+阅读 · 2011年12月31日

赋值理论与几何不等式的研究

国家自然科学基金

1+阅读 · 2011年12月31日

亚椭圆算子的泛函不等式和热核分析

国家自然科学基金

0+阅读 · 2011年12月31日

微信扫码咨询专知VIP会员