具有选择程序经济性的序贯抽样

Sequential Sampling with Economics of Selection Procedures

Management Science · 2011

被引 102

人大 A+FT50UTD24ABS 4*

Stephen E. Chick · 欧洲工商管理学院
Peter I. Frazier · 康奈尔大学

中文导读

提出了经济驱动的序贯抽样程序，通过动态规划将序贯抽样建模为选择前的学习期权，在数值实验中优于现有方法，适用于随机模拟等场景。

Abstract

Sequential sampling problems arise in stochastic simulation and many other applications. Sampling is used to infer the unknown performance of several alternatives before one alternative is selected as best. This paper presents new economically motivated fully sequential sampling procedures to solve such problems, called economics of selection procedures. The optimal procedure is derived for comparing a known standard with one alternative whose unknown reward is inferred with sampling. That result motivates heuristics when multiple alternatives have unknown rewards. The resulting procedures are more effective in numerical experiments than any previously proposed procedure of which we are aware and are easily implemented. The key driver of the improvement is the use of dynamic programming to model sequential sampling as an option to learn before selecting an alternative. It accounts for the expected benefit of adaptive stopping policies for sampling, rather than of one-stage policies, as is common in the literature. This paper was accepted by Assaf Zeevi, stochastic models and simulation.

序贯抽样经济选择程序动态规划自适应停止策略

阅读原文 ↗