情境动态定价:算法、最优性与局部差分隐私约束

Contextual Dynamic Pricing: Algorithms, Optimality, and Local Differential Privacy Constraints

Journal of the American Statistical Association · 2026
被引 1 · 同刊同年前 2%
ABS 4

中文导读

研究了企业向顺序到达的消费者销售产品时的情境动态定价问题,提出了探索后承诺算法,并分析了在局部差分隐私约束下的最优遗憾界,通过数值实验验证了算法效率。

Abstract

We study contextual dynamic pricing problems where a firm sells products to T sequentially-arriving consumers, behaving according to an unknown demand model. The firm aims to minimize its regret over a clairvoyant that knows the model in advance. The demand follows a generalized linear model (GLM), allowing for stochastic feature vectors in Rd encoding product and consumer information. We first show the optimal regret is of order dT, up to logarithmic factors, improving existing upper bounds by a d factor, achieved by an explore-then-commit (ETC) algorithm.We further study contextual dynamic pricing under local differential privacy (LDP) constraints. We propose a stochastic gradient descent-based ETC algorithm achieving regret upper bounds of order dT/ϵ, up to logarithmic factors, where is the privacy parameter. The upper bounds with and without LDP constraints are matched by newly constructed minimax lower bounds, characterizing costs of privacy. Moreover, we extend our study to dynamic pricing under mixed privacy constraints, which naturally bridges private and non-private dynamic pricing. We propose a two-stage ETC algorithm and show that it improves the privacy-utility tradeoff by efficiently leveraging public data. Its optimality is further established via a newly-derived minimax lower bound. To our knowledge, this is the first time such setting is studied in the dynamic pricing literature. Extensive numerical experiments and real data applications are conducted to illustrate the efficiency and practical value of our algorithms.

动态定价差分隐私情境信息广义线性模型遗憾最小化