Pith. sign in

REVIEW 1 cited by

Scaling TabPFN: Sketching and Feature Selection for Tabular Prior-Data Fitted Networks

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2311.10609 v1 pith:DA5MSYVE submitted 2023-11-17 cs.LG cs.DB

classification cs.LGcs.DB
keywords tabulardatamodeltrainingfittedsamplestabpfnclassify
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Tabular classification has traditionally relied on supervised algorithms, which estimate the parameters of a prediction model using its training data. Recently, Prior-Data Fitted Networks (PFNs) such as TabPFN have successfully learned to classify tabular data in-context: the model parameters are designed to classify new samples based on labelled training samples given after the model training. While such models show great promise, their applicability to real-world data remains limited due to the computational scale needed. Here we study the following question: given a pre-trained PFN for tabular data, what is the best way to summarize the labelled training samples before feeding them to the model? We conduct an initial investigation of sketching and feature-selection methods for TabPFN, and note certain key differences between it and conventionally fitted tabular models.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. TabPFN Unleashed: A Scalable and Effective Solution to Tabular Classification Problems

    cs.LG 2025-02 conditional novelty 6.0 of 10

    BETA augments TabPFN with encoder fine-tuning and bagging to reduce both bias and variance, achieving SOTA accuracy on 200+ tabular benchmarks while scaling to larger and higher-dimensional data.

Pith tools