通过 " 决定-估计估计效率 " 计算出,每平方美美美元-区域</s> (Lower Bounds for $γ$-Regret via the Decision-Estimation Coefficient) - 专知论文

会员服务 ·

0

赌博机/老虎机 · 分解的 · 确切的 · 偏移量 · 原点 ·

2023 年 3 月 6 日

Lower Bounds for $γ$-Regret via the Decision-Estimation Coefficient

翻译：通过 " 决定-估计估计效率 " 计算出,每平方美美美元-区域

Margalit Glasgow,Alexander Rakhlin

In this note, we give a new lower bound for the $\gamma$-regret in bandit problems, the regret which arises when comparing against a benchmark that is $\gamma$ times the optimal solution, i.e., $\mathsf{Reg}_{\gamma}(T) = \sum_{t = 1}^T \gamma \max_{\pi} f(\pi) - f(\pi_t)$. The $\gamma$-regret arises in structured bandit problems where finding an exact optimum of $f$ is intractable. Our lower bound is given in terms of a modification of the constrained Decision-Estimation Coefficient (DEC) of~\citet{foster2023tight} (and closely related to the original offset DEC of \citet{foster2021statistical}), which we term the $\gamma$-DEC. When restricted to the traditional regret setting where $\gamma = 1$, our result removes the logarithmic factors in the lower bound of \citet{foster2023tight}.

翻译：在本说明中,我们给土匪问题中的$gamma$-regret提供了一个新的较低约束值。美元gamma$- regret 出现在结构化的土匪问题中, 找到精确最佳的美元是难以解决的。我们的较低约束值是修改限制的决定- 估计系数(DEC) 的“citet{foster2023tight}”(与我们称之为$gama$2021statistical}的原始抵消DEC密切相关) 。当我们限制在$\gamma$=1美元的传统悔恨状态中时, 我们的结果消除了低约束范围(DEC) 的对数系数。 {citet{foster2023t} (与我们称之为$gamma$2021statistical} 的原始抵消值DEC密切相关) 。当我们限制在$\gamma =1美元20tstestrate 设置时, 我们的结果消除了下层的对数系数。</s>

0

相关内容

赌博机/老虎机

赌博机/老虎机

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

专知会员服务

96+阅读 · 2020年3月12日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

50+阅读 · 2019年10月17日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

机器学习入门的经验与建议

机器学习入门的经验与建议

专知会员服务

94+阅读 · 2019年10月10日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

强化学习的Unsupervised Meta-Learning

强化学习的Unsupervised Meta-Learning

CreateAMind

18+阅读 · 2019年1月7日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

【推荐】SVM实例教程

【推荐】SVM实例教程

机器学习研究会

17+阅读 · 2017年8月26日

强化学习族谱

强化学习族谱

CreateAMind

26+阅读 · 2017年8月2日

Vlasov-Poisson-Boltzmann方程研究

国家自然科学基金

0+阅读 · 2013年12月31日

关于具有奇异参数的偏微分方程边值问题与带双边反射的随机偏微分方程的研究

国家自然科学基金

0+阅读 · 2013年12月31日

大黄鱼抗氧化酶Peroxiredoxin IV调控炎症反应的机理研究

国家自然科学基金

0+阅读 · 2012年12月31日

LIMK1：罗格列酮抑制人胃癌细胞增殖、迁移及侵袭的作用靶点

国家自然科学基金

0+阅读 · 2012年12月31日

小电导Ca2+激活K+通道与ryanodine受体功能性偶联的研究

国家自然科学基金

0+阅读 · 2008年12月31日

On the Order of Power Series and the Sum of Square Roots Problem

Arxiv

0+阅读 · 2023年4月26日

On the simultanenous identification of the nonlinearity coefficient and the sound speed in the Westervelt equation

Arxiv

0+阅读 · 2023年4月26日

Post-processing and improved error estimates of numerical methods for evolutionary systems

Arxiv

0+阅读 · 2023年4月25日

IMUPoser: Full-Body Pose Estimation using IMUs in Phones, Watches, and Earbuds

Arxiv

0+阅读 · 2023年4月25日

Theory of Posterior Concentration for Generalized Bayesian Additive Regression Trees

Arxiv

0+阅读 · 2023年4月25日

VIP会员

文章信息

相关主题

赌博机/老虎机

相关VIP内容

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

专知会员服务

96+阅读 · 2020年3月12日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

50+阅读 · 2019年10月17日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

机器学习入门的经验与建议

机器学习入门的经验与建议

专知会员服务

94+阅读 · 2019年10月10日

热门VIP内容

开通专知VIP会员享更多权益服务

《基于AI的动态任务分配策略实现多智能体系统有意义人类控制》报告

《超越连接：AI驱动网络未来愿景》最新报告

人工智能赋能多域作战：能力与挑战

《战场空间决策优势：AI基础与应用研究》总结报告

相关资讯

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

强化学习的Unsupervised Meta-Learning

强化学习的Unsupervised Meta-Learning

CreateAMind

18+阅读 · 2019年1月7日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

【推荐】SVM实例教程

【推荐】SVM实例教程

机器学习研究会

17+阅读 · 2017年8月26日

强化学习族谱

强化学习族谱

CreateAMind

26+阅读 · 2017年8月2日

相关论文

On the Order of Power Series and the Sum of Square Roots Problem

Arxiv

0+阅读 · 2023年4月26日

On the simultanenous identification of the nonlinearity coefficient and the sound speed in the Westervelt equation

Arxiv

0+阅读 · 2023年4月26日

Post-processing and improved error estimates of numerical methods for evolutionary systems

Arxiv

0+阅读 · 2023年4月25日

IMUPoser: Full-Body Pose Estimation using IMUs in Phones, Watches, and Earbuds

Arxiv

0+阅读 · 2023年4月25日

Theory of Posterior Concentration for Generalized Bayesian Additive Regression Trees

Arxiv

0+阅读 · 2023年4月25日

相关基金

Vlasov-Poisson-Boltzmann方程研究

国家自然科学基金

0+阅读 · 2013年12月31日

关于具有奇异参数的偏微分方程边值问题与带双边反射的随机偏微分方程的研究

国家自然科学基金

0+阅读 · 2013年12月31日

大黄鱼抗氧化酶Peroxiredoxin IV调控炎症反应的机理研究

国家自然科学基金

0+阅读 · 2012年12月31日

LIMK1：罗格列酮抑制人胃癌细胞增殖、迁移及侵袭的作用靶点

国家自然科学基金

0+阅读 · 2012年12月31日

小电导Ca2+激活K+通道与ryanodine受体功能性偶联的研究

国家自然科学基金

0+阅读 · 2008年12月31日

微信扫码咨询专知VIP会员