REVIEW 3 cited by
Degeneracy is OK: Logarithmic Regret for Network Revenue Management with Indiscrete Distributions
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
abstract
We study the classical Network Revenue Management (NRM) problem with accept/reject decisions and $T$ IID arrivals. We consider a distributional form where each arrival must fall under a finite number of possible categories, each with a deterministic resource consumption vector, but a random value distributed continuously over an interval. We develop an online algorithm that achieves $O(\log^2 T)$ regret under this model, with the only (necessary) assumption being that the probability densities are bounded away from 0. We derive a second result that achieves $O(\log T)$ regret under an additional assumption of second-order growth. To our knowledge, these are the first results achieving logarithmic-level regret in an NRM model with continuous values that do not require any kind of "non-degeneracy" assumptions. Our results are achieved via new techniques including a new method of bounding myopic regret, a "semi-fluid" relaxation of the offline allocation, and an improved bound on the "dual convergence".
Forward citations
Cited by 3 Pith papers
-
Beyond Non-Degeneracy: Revisiting Certainty Equivalent Heuristic for Online Linear Programming
The Certainty Equivalent heuristic achieves near-optimal hindsight regret for online linear programming under mild distributional assumptions, without requiring non-degeneracy or second-order growth conditions.
-
Online Pricing and Allocation with Demand Learning and Fulfillment Cost
An online pricing-and-allocation algorithm with lower-confidence-bound agent selection achieves O~(sqrt(T) mn) regret, but the proof rests on a false convexity lemma.
-
Learning to Price with Resource Constraints: From Full Information to Machine-Learned Prices
Claims logarithmic and square-root regret bounds for dynamic pricing with inventory constraints across three information settings; key proof steps in the no-information and informed-price results are invalid as written.
Discussion (0). Continue with ORCID to comment.