Pith. sign in

REVIEW 2 cited by

From Points to Functions: Infinite-dimensional Representations in Diffusion Models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2210.13774 v1 pith:UG5V5DYX submitted 2022-10-25 cs.LG

classification cs.LG
keywords informationmodelsdistributiondownstreamgenerativecontentdiffusionlearn
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Diffusion-based generative models learn to iteratively transfer unstructured noise to a complex target distribution as opposed to Generative Adversarial Networks (GANs) or the decoder of Variational Autoencoders (VAEs) which produce samples from the target distribution in a single step. Thus, in diffusion models every sample is naturally connected to a random trajectory which is a solution to a learned stochastic differential equation (SDE). Generative models are only concerned with the final state of this trajectory that delivers samples from the desired distribution. Abstreiter et. al showed that these stochastic trajectories can be seen as continuous filters that wash out information along the way. Consequently, it is reasonable to ask if there is an intermediate time step at which the preserved information is optimal for a given downstream task. In this work, we show that a combination of information content from different time steps gives a strictly better representation for the downstream task. We introduce an attention and recurrence based modules that ``learn to mix'' information content of various time-steps such that the resultant representation leads to superior performance in downstream tasks.

Discussion (0). Sign in to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Real-Time 3D Vision-Language Embedding Mapping

    cs.RO 2025-08 unverdicted novelty 4.0 of 10

    Combining local embedding masking with confidence-weighted 3D integration yields, the paper claims, a real-time metric-accurate 3D map of vision-language embeddings for language-guided object localization.

  2. Growth Rate Analysis in $f(R,L_m)$ Gravity: A Comparative Study with \boldmath$\Lambda$CDM Cosmology

    gr-qc 2025-07 reject novelty 4.0 of 10

    The paper claims f(R,L_m) gravity with matter-curvature coupling fits f_sigma8 measurements better than LambdaCDM, but the supporting comparison table contains unphysical negative LambdaCDM growth rates and no statist...

Pith tools