随机实用模型学习 (Learning in Random Utility Models Via Online Decision Problems)

This paper studies the Random Utility Model (RUM) in environments where the decision maker is imperfectly informed about the payoffs associated to each of the alternatives he faces. By embedding the RUM into an online decision problem, we make four contributions. First, we propose a gradient-based learning algorithm and show that a large class of RUMs are Hannan consistent (\citet{Hahn1957}); that is, the average difference between the expected payoffs generated by a RUM and that of the best fixed policy in hindsight goes to zero as the number of periods increase. Second, we show that the class of Generalized Extreme Value (GEV) models can be implemented with our learning algorithm. Examples in the GEV class include the Nested Logit, Ordered, and Product Differentiation models among many others. Third, we show that our gradient-based algorithm is the dual, in a convex analysis sense, of the Follow the Regularized Leader (FTRL) algorithm, which is widely used in the Machine Learning literature. Finally, we discuss how our approach can incorporate recency bias and be used to implement prediction markets in general environments.

翻译：本文在决策者不完全了解与他所面临的各种替代方案相关的回报的环境下研究随机实用模型(RUM) 。通过将 RUM 嵌入在线决策问题, 我们做出四项贡献。首先, 我们提出基于梯度的学习算法, 并显示一大批RUM都是汉南一致的(\citet{Hahn1957}); 也就是说, 由RUM 产生的预期收益与后视最佳固定政策的预期回报之间的平均差异随着时间的增加而变为零。其次, 我们展示了普通极端值(GEV) 模型的类别可以与我们的学习算法一起实施。 GEV 类中的例子包括Nested Logit、有秩序的和产品差异模型等等。第三, 我们显示,我们的梯度算法是遵循正规领导人(FTRL) 算法的双重性, 后者在机器学习文献中广泛使用。最后, 我们讨论我们的方法如何纳入耐久性偏差, 并用于在一般环境中实施预测市场。

相关内容

MoDELS

关注 43

ACM/IEEE第23届模型驱动工程语言和系统国际会议，是模型驱动软件和系统工程的首要会议系列，由ACM-SIGSOFT和IEEE-TCSE支持组织。自1998年以来，模型涵盖了建模的各个方面，从语言和方法到工具和应用程序。模特的参加者来自不同的背景，包括研究人员、学者、工程师和工业专业人士。MODELS 2019是一个论坛，参与者可以围绕建模和模型驱动的软件和系统交流前沿研究成果和创新实践经验。今年的版本将为建模社区提供进一步推进建模基础的机会，并在网络物理系统、嵌入式系统、社会技术系统、云计算、大数据、机器学习、安全、开源等新兴领域提出建模的创新应用以及可持续性。官网链接：http://www.modelsconference.org/