Pith. sign in

REVIEW 1 cited by

UniCO: Towards a Unified Model for Combinatorial Optimization Problems

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2505.06290 v1 pith:JZ5ZJ65X submitted 2025-05-07 cs.LG cs.DM

classification cs.LGcs.DM
keywords problemsmodelunicounifiedapproachcombinatorialdatamethods
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Combinatorial Optimization (CO) encompasses a wide range of problems that arise in many real-world scenarios. While significant progress has been made in developing learning-based methods for specialized CO problems, a unified model with a single architecture and parameter set for diverse CO problems remains elusive. Such a model would offer substantial advantages in terms of efficiency and convenience. In this paper, we introduce UniCO, a unified model for solving various CO problems. Inspired by the success of next-token prediction, we frame each problem-solving process as a Markov Decision Process (MDP), tokenize the corresponding sequential trajectory data, and train the model using a transformer backbone. To reduce token length in the trajectory data, we propose a CO-prefix design that aggregates static problem features. To address the heterogeneity of state and action tokens within the MDP, we employ a two-stage self-supervised learning approach. In this approach, a dynamic prediction model is first trained and then serves as a pre-trained model for subsequent policy generation. Experiments across 10 CO problems showcase the versatility of UniCO, emphasizing its ability to generalize to new, unseen problems with minimal fine-tuning, achieving even few-shot or zero-shot performance. Our framework offers a valuable complement to existing neural CO methods that focus on optimizing performance for individual problems.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. LaT: LLM-as-Trainer for Multi-Task Vehicle Routing Solvers

    cs.AI 2026-07 conditional novelty 6.0 of 10

    A pretrained LLM converts cross-task validation gaps into a five-number guidance signal injected into each encoder layer, improving several multi-task VRP solvers by roughly 0.1–0.5 percentage points on trained and un...

Pith tools