Pith. sign in

REVIEW 2 cited by

Generalization and Overfitting in Matrix Product State Machine Learning Architectures

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2208.04372 v1 pith:GTR7WQXK submitted 2022-08-08 cs.LG quant-ph

classification cs.LGquant-ph
keywords dataoverfittinggeneralizationpropertieswhilearchitecturescomplexexactly
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

While overfitting and, more generally, double descent are ubiquitous in machine learning, increasing the number of parameters of the most widely used tensor network, the matrix product state (MPS), has generally lead to monotonic improvement of test performance in previous studies. To better understand the generalization properties of architectures parameterized by MPS, we construct artificial data which can be exactly modeled by an MPS and train the models with different number of parameters. We observe model overfitting for one-dimensional data, but also find that for more complex data overfitting is less significant, while with MNIST image data we do not find any signatures of overfitting. We speculate that generalization properties of MPS depend on the properties of data: with one-dimensional data (for which the MPS ansatz is the most suitable) MPS is prone to overfitting, while with more complex data which cannot be fit by MPS exactly, overfitting may be much less significant.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Regularized second-order optimization of tensor-network Born machines

    cs.LG 2025-01 conditional novelty 6.0 of 10

    A regularized constrained Newton optimizer on the sphere of normalized tensors reaches lower negative log-likelihood faster than gradient descent for tensor-network Born machines.

  2. Bayesian perspectives for quantum states and application to ab initio quantum chemistry

    cond-mat.str-el 2025-08 conditional novelty 3.0 of 10

    A review of Bayesian Gaussian Process States for ab initio quantum chemistry, with new MNIST digit classification results reaching about 1.6% test error.

Pith tools