Pith. sign in

REVIEW 2 cited by

On Dynamic Pricing with Covariates

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2112.13254 v3 pith:ZVO2NT3L submitted 2021-12-25 cs.LG stat.ML

classification cs.LGstat.ML
keywords regretboundcovariatespricingdynamicmodelachievablealgorithms
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
abstract

We consider dynamic pricing with covariates under a generalized linear demand model: a seller can dynamically adjust the price of a product over a horizon of $T$ time periods, and at each time period $t$, the demand of the product is jointly determined by the price and an observable covariate vector $x_t\in\mathbb{R}^d$ through a generalized linear model with unknown co-efficients. Most of the existing literature assumes the covariate vectors $x_t$'s are independently and identically distributed (i.i.d.); the few papers that relax this assumption either sacrifice model generality or yield sub-optimal regret bounds. In this paper, we show that UCB and Thompson sampling-based pricing algorithms can achieve an $O(d\sqrt{T}\log T)$ regret upper bound without assuming any statistical structure on the covariates $x_t$. Our upper bound on the regret matches the lower bound up to logarithmic factors. We thus show that (i) the i.i.d. assumption is not necessary for obtaining low regret, and (ii) the regret bound can be independent of the (inverse) minimum eigenvalue of the covariance matrix of the $x_t$'s, a quantity present in previous bounds. Moreover, we consider a constrained setting of the dynamic pricing problem where there is a limited and unreplenishable inventory and we develop theoretical results that relate the best achievable algorithm performance to a variation measure with respect to the temporal distribution shift of the covariates. We also discuss conditions under which a better regret is achievable and demonstrate the proposed algorithms' performance with numerical experiments.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Minimax-Optimal Semiparametric Contextual Dynamic Pricing with Multimodal Revenue

    stat.ML 2026-08 accept novelty 7.0 of 10

    For semiparametric contextual pricing with arbitrary covariates and bounded quantity feedback, a pilot-corrected layered policy achieves the minimax regret exponent (beta+1)/(2beta+1) without concavity, unimodality, o...

  2. Transfer Learning for Nonparametric Contextual Dynamic Pricing

    cs.LG 2025-01 conditional novelty 6.0 of 10

    TLDP is a nonparametric contextual dynamic pricing algorithm with provably minimax-optimal regret when source-domain data are available under covariate shift.

Pith tools