Pith. sign in

Paper Citation Record · LEDGER

Scaling 4D Representations

As of 6 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2412.15212.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.15212 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T18:43:43.563351Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T00:26:39.628379Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 3f95764d-646c-414b-9d2c-358331e7599a · inbound

V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning cites this paper.

V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning Scaling 4D Representations

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:33:50.730654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-11T00:33:50.471804Z digest=sha256:6b7e2dc4d17d29631de84da7bf92475401c3044900dbccc7cb2fc1408aec5e3d

Observation 987e4f4f-0610-45b9-a150-53cfdb66061e · inbound

Frozen Forecasting: A Unified Evaluation cites this paper.

Frozen Forecasting: A Unified Evaluation Scaling 4D Representations

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-19T03:42:57.307979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-19T03:42:54.620069Z digest=sha256:896a045cafcc1c3cdcc6159405cddccd456050c38d023077442903d691cf4f09

Observation b38649bd-72f1-4353-8d6a-7c3454b70b61 · inbound

Unique Lives, Shared World: Learning from Single-Life Videos cites this paper.

Unique Lives, Shared World: Learning from Single-Life Videos Scaling 4D Representations

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-03T18:43:43.563351Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:43:43.563351Z digest=sha256:829af3f11ffce3a19cc17b984f739e0a86beda617165fbdb4039babe8d665474

Observation 17bdcb93-8b3e-47f5-953b-7f19a69424b2 · inbound

LA-Pose: Latent Action Pretraining Meets Pose Estimation cites this paper.

LA-Pose: Latent Action Pretraining Meets Pose Estimation Scaling 4D Representations

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:51:28.839842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-07T08:57:27.503349Z digest=sha256:958d443790d35fe3293912725543549c160f053bd4573453e93e96f1383b2003

Observation 7f655eea-b5fc-4541-8a3f-5952a9598b38 · inbound

Towards Data-Efficient Video Pre-training with Frozen Image Foundation Models cites this paper.

Towards Data-Efficient Video Pre-training with Frozen Image Foundation Models Scaling 4D Representations

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:33:13.031523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-20T10:28:42.027773Z digest=sha256:1348388c80a30cdcfde4440ca8cc9a607b95fb5a882930ae6b1f59b04b9d0ccc

Observation 394c74aa-a514-4f79-96b3-0293b0ebf52c · inbound

Where Do We (Not) Need Temporal Context in Low-Resource Video Task Adaptation? cites this paper.

Where Do We (Not) Need Temporal Context in Low-Resource Video Task Adaptation? Scaling 4D Representations

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T02:56:28.965993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T10:31:10.409674Z digest=sha256:9ee92fd0ac9381d0d5669bf29efab51c362a36645cab4bebd13578387f03c2c0

Observation a3458577-a4a7-4f4c-a660-a63981166db6 · inbound

Gen4U: Unifying Video Generation and Understanding via Diffusion cites this paper.

Gen4U: Unifying Video Generation and Understanding via Diffusion Scaling 4D Representations

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-07-10T00:26:39.630147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-10T00:16:43.190961Z digest=sha256:96a3db23069d2c418abd5a0662e15029c0d157c18890f615f54e758c62782d56

Observation 50b5a864-d3ae-4dc9-a564-937d49a0c200 · inbound

SeeSE3: Emergence of 3D Space in Vision Features cites this paper.

SeeSE3: Emergence of 3D Space in Vision Features Scaling 4D Representations

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-02T02:49:58.712379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T02:49:58.712379Z digest=sha256:5106a61ff49dca2cdc4e5906276007cc427ccafadd2d89ce377f14dc0c01a449

Observation 9b3a923d-c9c3-4986-9f24-86606157a485 · inbound

Self-Supervised Learning of Structured Dynamics from Videos cites this paper.

Self-Supervised Learning of Structured Dynamics from Videos Scaling 4D Representations

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T07:07:43.307749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T07:07:43.307749Z digest=sha256:c21250eeda4c842d08bd8cd3eb36630eb7ab710595fc8818caa6d20c20daa189