Pith. sign in

REVIEW 1 cited by

Score Function Gradient Estimation to Widen the Applicability of Decision-Focused Learning

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2307.05213 v3 pith:EAHOR7CO submitted 2023-07-11 cs.LG cs.AI

classification cs.LGcs.AI
keywords problemsmethodsparametersfunctionlearningassumptionsconstraintsdecision-focused
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Many real-world optimization problems contain parameters that are unknown before deployment time, either due to stochasticity or to lack of information (e.g., demand or travel times in delivery problems). A common strategy in such cases is to estimate said parameters via machine learning (ML) models trained to minimize the prediction error, which however is not necessarily aligned with the downstream task-level error. The decision-focused learning (DFL) paradigm overcomes this limitation by training to directly minimize a task loss, e.g. regret. Since the latter has non-informative gradients for combinatorial problems, state-of-the-art DFL methods introduce surrogates and approximations that enable training. But these methods exploit specific assumptions about the problem structures (e.g., convex or linear problems, unknown parameters only in the objective function). We propose an alternative method that makes no such assumptions, it combines stochastic smoothing with score function gradient estimation which works on any task loss. This opens up the use of DFL methods to nonlinear objectives, uncertain parameters in the problem constraints, and even two-stage stochastic optimization. Experiments show that it typically requires more epochs, but that it is on par with specialized methods and performs especially well for the difficult case of problems with uncertainty in the constraints, in terms of solution quality, scalability, or both.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Decision-Focused Learning for Complex System Identification: HVAC Management System Application

    eess.SY 2025-01 conditional novelty 5.0 of 10

    Decision-focused learning identifies RC model parameters inside a convex HVAC scheduling policy, reducing day-ahead cost prediction error from 389 to 16 euros in a 15-zone building simulation.

Pith tools