Pith. sign in

Paper Citation Record · LEDGER

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion

As of 4 August 2026, this Paper Citation Record lists 87 of 87 outbound references and 0 inbound Pith citation observations for arXiv:2605.12957.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.12957 v1

Coverage vector

measured 87 of 87 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-14T19:39:16.410139Z

measured 87 of 87 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

87 of 87 outbound references displayed

  • verified exact4
  • verified fuzzy61
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch22

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9097fda5-42bb-4359-af7e-b8276ceaa69b · outbound

This paper cites In: Proceedings of the First International Conference on Computer Vision Theory and Applications, pp.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion In: Proceedings of the First International Conference on Computer Vision Theory and Applications, pp

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.272121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:d64c570cd12af9885f247fae9da01c80d2830c4b86fcc2623c25cb98a0e5d981

Observation 65450e5d-9451-43de-afc8-e4e324b11398 · outbound

This paper cites Advances in 3D Generation: A Survey.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion Advances in 3D Generation: A Survey

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T19:39:23.553976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:01e935afd48b39fc12fd07dbbc44cc044043cc75a1ec5fb896d7cd372090db99

Observation 06ca3168-e64d-4149-be64-38d63c261fed · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.298514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:c46b3e724725b955116af131e3f46207f1fd2eeb6bf88094886d59f2241c4318

Observation 052c6c09-3393-4180-8b0f-33dc6f73c2f4 · outbound

This paper cites Artificial Intelligence Review56(9), 9175–9219 (2023).

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion Artificial Intelligence Review56(9), 9175–9219 (2023)

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.293156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:aa52b84c6413fc6db70c924f0feeabf9a2e9c32ffde440d096c7fb2081c9be96

Observation 1704447d-816a-4324-82e2-d265a41bbead · outbound

This paper cites 3D Scene Generation: A Survey.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion 3D Scene Generation: A Survey

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T19:39:23.560248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:56774675f0263e57662e626cdfa363ea9c6d5ece57feb030f2d0936b25659c7f

Observation 1a7c67d5-3008-4742-b831-68fe255ef658 · outbound

This paper cites International Journal of Computer Vision112(2), 188–203 (2015).

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion International Journal of Computer Vision112(2), 188–203 (2015)

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.282793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:a120d13052711218e197abce5fbf1bf9982abffd24549bc5a188afa0261be1ea

Observation 25a56f9a-3fcc-4ec2-b755-33154d3a8db5 · outbound

This paper cites International Journal of Computer Vision132(10), 4456–4472 (2024) 20.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion International Journal of Computer Vision132(10), 4456–4472 (2024) 20

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.267996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:a8500e0253c6f7c3616cfbde5af6022e7d760f114960a01f4ef29482998acb33

Observation 0c421605-51fa-4de0-8830-576dbd0efd26 · outbound

This paper cites Interna- tional Journal of Computer Vision128(10), 2534–2551 (2020).

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion Interna- tional Journal of Computer Vision128(10), 2534–2551 (2020)

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.272471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:636d496ea49035a86c8eebe4187da3e9dd447e825c030092dd51bc8cd0a40d04

Observation 85645e86-5d9d-43e5-ab9f-48351ecfe0ef · outbound

This paper cites Interna- tional Journal of Computer Vision133(5), 2886–2909 (2025).

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion Interna- tional Journal of Computer Vision133(5), 2886–2909 (2025)

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.288210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:b846116e7efc0dc07087b6d1a02c005ad26d1c1158a23c02dabcd3e8f96770c0

Observation df555f3d-6710-45aa-9ec7-87e27f45619b · outbound

This paper cites Interna- tional Journal of Computer Vision131(11), 2816–2844 (2023).

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion Interna- tional Journal of Computer Vision131(11), 2816–2844 (2023)

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.278123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:4f6bbe22f8cc90b1bb5d6913807f3d998645d3016f426ef5379d21ac1bedabb5

Observation 624661b5-5921-4b40-9785-b95ca8f11018 · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.179207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:5114c332f53174fea372d3484bb809371205f54ac47ae159cd16fa6df9e1e4d0

Observation 55d670ec-060a-4262-9cbc-aa13f973463e · outbound

This paper cites EmbodiedGen: Towards a Generative 3D World Engine for Embodied Intelligence.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion EmbodiedGen: Towards a Generative 3D World Engine for Embodied Intelligence

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T19:39:23.438683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:ec401bd134c6fd4c934bf98c1554e029f1831c8867fe5fcae5d083397294936f

Observation 409a2041-7df7-4def-9a2a-2643d5f783b5 · outbound

This paper cites 3D and 4D World Modeling: A Survey.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion 3D and 4D World Modeling: A Survey

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-07-21T03:22:29.733321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:0154fc5d446bc736fcc8e716c1bbf218269ca35835c4d401e884a9ac8ed1588b

Observation 069373a7-eee4-4c32-84a6-4660f0bd751a · outbound

This paper cites ACM Computing Surveys 58(3), 1–38 (2025).

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion ACM Computing Surveys 58(3), 1–38 (2025)

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.142871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:889d455ee9a51bb7d3d53e304e33be31412c29c34128ee81b93a06b559baa87e

Observation 71b703ce-75a5-493b-84f8-3c8a35ae3038 · outbound

This paper cites In: European Conference on Computer Vision, pp.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion In: European Conference on Computer Vision, pp

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.073928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:28e3d2158fcc28983847266d7c3c1fcbeac2f3111e4c1f061b0c22727e1fef92

Observation 98331e12-8812-498b-90ce-fbee626d1396 · outbound

This paper cites DreamFusion: Text-to-3D using 2D Diffusion.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion DreamFusion: Text-to-3D using 2D Diffusion

Reference 16

Resolution
metadata mismatch
local_arxiv, observed 2026-05-14T19:39:23.496328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:0ac3c4b5ababfa384e01eb74f8a4431082c761662e20a9584232130365cf9dea

Observation f3c09bed-7def-48bf-8407-238588d09afc · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.120072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:be3f1968f4a49437dd9ce7e9a7bc2f1bc81993b420690249761391ad6cd4c939

Observation daad5753-6eb6-4459-8037-02cf1f756b20 · outbound

This paper cites In: Proceedings of the IEEE/CVF Inter- national Conference on Computer Vision, pp.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion In: Proceedings of the IEEE/CVF Inter- national Conference on Computer Vision, pp

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.118293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:40973ad4c0c90ca52787a5ef198a4bc0119713c9570c49028fd90b14ed58a4e4

Observation d740ed3b-0095-4b91-8f7b-5c0403547217 · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.101907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:aa37ffa293d9d39dd83e13fca4fd6c7a5477eac0eba4c345a53a40780f567598

Observation c8d5072a-e192-4abf-822d-64474e47295e · outbound

This paper cites CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers

Reference 20

Resolution
metadata mismatch
local_arxiv, observed 2026-05-14T19:39:23.463338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:91e70b3369b24f352bf0928275d8ec106b21803b231f944aa475adb0e88b999c

Observation 232f7f3a-7221-4567-87a9-a01407d4305e · outbound

This paper cites In: European Conference on Com- puter Vision, pp.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion In: European Conference on Com- puter Vision, pp

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.192778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:aa9db163c23b18b96345847b199eb238344e5c9716ad3d09b133e207d8672c3c

Observation 5a369e40-a1df-4c39-b18c-204556f9c5eb · outbound

This paper cites In: SIG- GRAPH Asia 2024 Conference Papers, pp.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion In: SIG- GRAPH Asia 2024 Conference Papers, pp

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.157882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:7c3a0a83ca3137ae757d89cfc99d63e30651ddd2d169bec113f4c4b30c9511e9

Observation e847931f-721f-4a66-8405-5d49e270f879 · outbound

This paper cites OpenVid-1M: A Large-Scale High-Quality Dataset for Text-to-video Generation.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion OpenVid-1M: A Large-Scale High-Quality Dataset for Text-to-video Generation

Reference 23

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T20:34:53.267725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:8464496c8076f89396440d2efb3b79cbbbcba00ca874c018966e3fa5a75eb1a2

Observation 8dc8fe02-ae02-4b84-9d97-9f945aa585c6 · outbound

This paper cites InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T06:30:22.925421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:5537b93481e9ecbac17d4145e53ab978419626040817b3d76b9325379e90524a

Observation f1cd2688-4729-4590-987f-92b7fa53a40d · outbound

This paper cites DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 25

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T19:39:23.538276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:8ebaaed3e66ed3272d5332c7cb7f4ff863689e2d817820de5edf645ba6ef9adb

Observation e40401fc-5561-4a8f-9789-9592c690a073 · outbound

This paper cites In: Proceedings of the Computer Vision and Pattern Recognition Conference, pp.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion In: Proceedings of the Computer Vision and Pattern Recognition Conference, pp

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.197536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:0525a6059003ecebc352076804cc5310c6a917e6e796096995c43e66c7dd8855

Observation 24225c09-ab84-49bb-9b75-a9f357a87b15 · outbound

This paper cites In: Proceedings of the Computer Vision and Pattern Recognition Conference, pp.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion In: Proceedings of the Computer Vision and Pattern Recognition Conference, pp

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.174576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:3d816136bce77ee3cfb7885ed0c8b38abbdb55137dcd06a8977c1554de641f09

Observation 159930a4-16ea-437c-9bf1-c4f85200aae5 · outbound

This paper cites ViewCrafter: Taming Video Diffusion Models for High-fidelity Novel View Synthesis.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion ViewCrafter: Taming Video Diffusion Models for High-fidelity Novel View Synthesis

Reference 28

Resolution
metadata mismatch
local_arxiv, observed 2026-05-14T19:39:23.435450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:32bb22a72ecb6f88ed8d399875849c2092f092c2b33df7422d8412e175d738da

Observation d2637e77-5c3d-4c00-be24-a5d83bb10a93 · outbound

This paper cites In: Proceedings of the IEEE/CVF International Conference on Computer Vision, pp.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion In: Proceedings of the IEEE/CVF International Conference on Computer Vision, pp

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.221238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:8b8286f0681eaabf2bd399a8d0073ee57dae460a11b805e12d062ce0b0525a26

Observation f4ba53a4-b482-47f3-8c98-bf2a07674512 · outbound

This paper cites CameraCtrl: Enabling Camera Control for Text-to-Video Generation.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion CameraCtrl: Enabling Camera Control for Text-to-Video Generation

Reference 30

Resolution
metadata mismatch
local_arxiv, observed 2026-05-14T19:39:23.508614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:971b9a502a122811eb38e2ca7a00dcd585c6884d5f3e6fc4cb15e194717e2971

Observation 7b07cd39-d014-4019-b99d-bd13a56dadea · outbound

This paper cites In: ACM SIGGRAPH 2024 Conference Papers, pp.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion In: ACM SIGGRAPH 2024 Conference Papers, pp

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.026598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:46effae5b49bd65f6f36cf47629d89676a83fe0df43bca85277827e009c00e13

Observation d58e9eac-855f-48b2-bae4-70164d0dbf36 · outbound

This paper cites ACM Transactions on Graphics (TOG)44(6), 1–15 (2025).

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion ACM Transactions on Graphics (TOG)44(6), 1–15 (2025)

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.216890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:11ac30b33687f70bebd1152886c84fbfadc08af62e50a260d2ff5aaea3224186

Observation 0a4a4ff6-71ef-4077-8fb5-69aff102573c · outbound

This paper cites Fantasyworld: Geometry-consistent world modeling via unified video and 3d prediction.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion Fantasyworld: Geometry-consistent world modeling via unified video and 3d prediction

Reference 33

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T19:39:23.546069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:708b1fec201537868d73a9933fe14eec29efb79b696935169a478966550422f3

Observation a4a506e7-dca3-4dd3-a5e5-e4881bf0a366 · outbound

This paper cites arXiv preprint arXiv:2511.23127 (2025).

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion arXiv preprint arXiv:2511.23127 (2025)

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-14T19:39:23.513976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:dfca3663768eac065f029ccc579257d04518b94c56d4fe66feffd1bdae3caa10

Observation 891d9fd4-0fa8-45e5-b829-8556a02d0012 · outbound

This paper cites Cognitive psychology9(3), 353–383 (1977).

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion Cognitive psychology9(3), 353–383 (1977)

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.212279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:e51de365733272c120a1fe21806730fd819448ef0543390a1e1dcab2b7ef5b7a

Observation 59db2d8e-ca12-4848-a36e-89cbb5157bb9 · outbound

This paper cites Journal of the Opti- cal Society of America A4(10), 2006–2021 (1987).

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion Journal of the Opti- cal Society of America A4(10), 2006–2021 (1987)

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.167009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:50320e451e263c703cab51bf4b9b0be72d9fc5091d38c7326b28dc4a8d548c66

Observation 0f4bfa7b-b8ed-4cf0-9787-79f173687fd3 · outbound

This paper cites ACM Computing Surveys57(2), 1–42 (2024).

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion ACM Computing Surveys57(2), 1–42 (2024)

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.209820Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:e9395c943d74b1804b09e0265b3280ff2f5f2f8266c1474c7169e1144abfbe5f

Observation b755346f-c6a4-4e65-b5a8-2277860dffba · outbound

This paper cites ACM Computing Surveys58(6), 1–35 (2025).

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion ACM Computing Surveys58(6), 1–35 (2025)

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.213575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:d655d7dad28bb26eb1e87f86279eb0ca5e27e18b53ae986498a8ac964a6f5469

Observation 83a0d08d-0db3-4b84-9caf-73d2a6ccd386 · outbound

This paper cites Artificial Intelligence Review58(11), 338 (2025).

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion Artificial Intelligence Review58(11), 338 (2025)

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.170658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:71122b0157ac1594d9c52cf8c7f3278302609bc18b080f0a5426eef492e90609

Observation 9c54046c-2a93-41ff-a86e-c124c7fbaacf · outbound

This paper cites In: Proceedings of the IEEE Conference on Computer Vision and 22 Pattern Recognition, pp.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion In: Proceedings of the IEEE Conference on Computer Vision and 22 Pattern Recognition, pp

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.255200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:9b4cf2e40e18d9e34de121c5f99823bb55621ea6b46efa37bfb8ea056c054fe2

Observation 2eb69874-94d0-4f5f-824a-d835d19bd6f2 · outbound

This paper cites In: Proceed- ings of the AAAI Conference on Artificial Intelligence, vol.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion In: Proceed- ings of the AAAI Conference on Artificial Intelligence, vol

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.078414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:049d5dde5d3110df43dc36ed33e19635274e2e96d465e160ef6a05b3d2258c53

Observation 16c55b4a-7a2e-4cbd-a63f-072df8152b39 · outbound

This paper cites Make-A-Video: Text-to-Video Generation without Text-Video Data.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion Make-A-Video: Text-to-Video Generation without Text-Video Data

Reference 42

Resolution
metadata mismatch
local_arxiv, observed 2026-05-14T19:39:23.507759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:e887bd3d5b39c81c02010af250659480e120521100a465b30359c29befc60f50

Observation d4a38439-ebca-4dce-a452-acc1bb52c67e · outbound

This paper cites Advances in Neural Information Processing Systems37, 29489– 29513 (2024).

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion Advances in Neural Information Processing Systems37, 29489– 29513 (2024)

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.162611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:f4c8de7cce4171d2896fa13e2267d878f286925ed9f67bc73490b76f397ea81a

Observation f1754eb4-2005-42ed-9731-7a555825a920 · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.069737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:b6dc8db60ae903d88861ce7dcf5e86ee6004522d6555bfc79770c4d0fa44076d

Observation a0314f0a-dc1f-4b86-8a77-a0e4dc4a23bb · outbound

This paper cites International Journal of Computer Vision 133(5), 3059–3078 (2025).

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion International Journal of Computer Vision 133(5), 3059–3078 (2025)

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.230919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:c7b7e59ef21306b230971edde29960723b477119767d5588da052fcf13c868ee

Observation 9014f510-a537-41dd-a104-b721a7019620 · outbound

This paper cites In: Proceed- ings of the IEEE/CVF Conference on Com- puter Vision and Pattern Recognition, pp.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion In: Proceed- ings of the IEEE/CVF Conference on Com- puter Vision and Pattern Recognition, pp

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.183210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:e8c4ab483983e84b014964d179d02103844fd1cb771f6c73d4bbc80a2d84197f

Observation c71e59d8-75b2-4efe-8708-988136f03760 · outbound

This paper cites Controllable video generation: A survey.arXiv preprint arXiv:2507.16869.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion Controllable video generation: A survey.arXiv preprint arXiv:2507.16869

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-14T19:39:23.532668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:8443adcaba4ffda8ff42a4adb70859d3ae3fd4d88290da1d89a753e9302707a7

Observation 239b3c76-25d1-465f-a236-3fb64c6a57d3 · outbound

This paper cites In: Proceedings of the IEEE Conference on Com- puter Vision and Pattern Recognition, pp.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion In: Proceedings of the IEEE Conference on Com- puter Vision and Pattern Recognition, pp

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.087029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:581c41019ad7dff5cc8845da24b9f9da5852321f4d8285e838cb6e1e0519f86d

Observation 5ea7c983-16ee-4626-88a2-86e7050d1cbc · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.082892Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:2636f1f94dc60b93ce518ce35e11359fdeda31b034be24b4c56c7a42d52cce19

Observation 5e07adf8-aac8-4193-b2f9-ed5854db3abb · outbound

This paper cites In: ACM SIG- GRAPH 2024 Conference Papers, pp.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion In: ACM SIG- GRAPH 2024 Conference Papers, pp

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.259620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:2d5b5e125b462ecf4b0c3d60ada89a172bd8dc649d35f991d1ee6cf1aa99c497

Observation fb6c872f-e6c3-45be-b0df-db303aec2f41 · outbound

This paper cites A Survey on Long Video Generation: Challenges, Methods, and Prospects.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion A Survey on Long Video Generation: Challenges, Methods, and Prospects

Reference 51

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T19:39:23.448664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:50eaa750979beefae2e4b54c8ba21d33f01996f923e361228e74cad82d3fc6d0

Observation 6aaa0ea7-a7bd-45f8-8eb0-9811eab2add7 · outbound

This paper cites In: European Conference on Computer Vision, pp.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion In: European Conference on Computer Vision, pp

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.172187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:d2bf0685fe150eb60eaaaf7c035591ef98cbe19b51b00b2172da6a02f75097b1

Observation 722ac9b7-7939-4205-bf52-069f75d96fda · outbound

This paper cites In: Proceedings of the Computer Vision and Pattern Recog- nition Conference, pp.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion In: Proceedings of the Computer Vision and Pattern Recog- nition Conference, pp

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.168533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:86c90ad32e005503609289753ac887a302310d7b106850c94882359280451bed

Observation 625a29d2-17d4-452e-830c-5708b9cdf476 · outbound

This paper cites Advances in Neural Information Processing Systems 37, 131434–131455 (2024).

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion Advances in Neural Information Processing Systems 37, 131434–131455 (2024)

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.044997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:88bf7968ea1161dcdcae9c549c032afbf507bbe1a51a6efebefc81c826c9b3b6

Observation e9764e44-67f5-4bf6-a93f-d75319927a02 · outbound

This paper cites Bidirectional sparse attention for faster video diffusion training.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion Bidirectional sparse attention for faster video diffusion training

Reference 55

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T19:39:23.519903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:a818edde108426db7fe4ee19d35924815cc21be27eb51f4c2b6d40148eb40c58

Observation 4c538df8-94ec-4909-9fb4-f6a7c57a54d4 · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer 23 Vision and Pattern Recognition, pp.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion In: Proceedings of the IEEE/CVF Conference on Computer 23 Vision and Pattern Recognition, pp

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.175765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:7703e3ea7f73cea0a514d7e7d895e2f89f332acd3d282883843019112f5628ba

Observation edc118c8-2bb9-4c2d-a369-0991112a71ce · outbound

This paper cites VideoControlNet: A Motion-Guided Video-to-Video Translation Framework by Using Diffusion Model with ControlNet.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion VideoControlNet: A Motion-Guided Video-to-Video Translation Framework by Using Diffusion Model with ControlNet

Reference 57

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T19:39:23.526496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:df1d582bffc0fee77556f7f8e59c524c7a7fd3ec23a2183092992fd36dc97290

Observation c34c8fa8-94e6-4f73-8d15-50911a53ec82 · outbound

This paper cites Mixture of contexts for long video generation.arXiv preprint arXiv:2508.21058.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion Mixture of contexts for long video generation.arXiv preprint arXiv:2508.21058

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-05-14T19:39:23.526166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:9386c0fc8d8b2786173653f097ce141dc4c78f7733d6424298beafb84911c522

Observation a5deb092-a0b5-4cad-b2da-7db5cecf58b6 · outbound

This paper cites Advances in neural information processing systems36, 8406– 8441 (2023).

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion Advances in neural information processing systems36, 8406– 8441 (2023)

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.205715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:b78748ccb41272cecb60cdb404297fed311793cf9fcc9745b930f6c4478bc498

Observation 97e65404-5853-4693-9d4a-5bfb3541d5cd · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.088216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:283f9027456c37d6aa07612c7cd398d6625ab8aa7164e70a2dadf22f296e2702

Observation cff6a2c0-ca38-4403-9c3f-d0570a53bafc · outbound

This paper cites In: Proceedings of the IEEE/CVF International Conference on Computer Vision, pp.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion In: Proceedings of the IEEE/CVF International Conference on Computer Vision, pp

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.111168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:92e25e10bf7f42efd31511ba80009560225c3f0e0be434406f71f81a2f850584

Observation 6825fffa-b70e-4da2-a3ff-2f63b3f2c0c3 · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.231195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:cf027c68f29e419a3b2e4b33f8e1f28d3df8c208770c8d77226b6aed606f0ac9

Observation baa21025-5de8-400e-9c81-3b11e94f9ca3 · outbound

This paper cites In: Proceedings of the Computer Vision and Pattern Recognition Conference, pp.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion In: Proceedings of the Computer Vision and Pattern Recognition Conference, pp

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.155087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:9f9e1566c9f5d0beff1b091d6619eca4bf4289880317884c2419bbf0bc8a0615

Observation d845de5a-dd61-4817-8d6c-fbdb73567b87 · outbound

This paper cites In: Proceedings of the IEEE/CVF International Conference on Computer Vision, pp.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion In: Proceedings of the IEEE/CVF International Conference on Computer Vision, pp

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.149562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:a54a84abf7b7abc62823becb72363c0c9dbc213ecf4378b78119c4526fe6f5d8

Observation c503edb9-aed1-4769-a540-8d931e764e99 · outbound

This paper cites Advances in Neural Information Processing Systems36, 22226– 22246 (2023).

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion Advances in Neural Information Processing Systems36, 22226– 22246 (2023)

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.236451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:daa743147bbce9ec9456d3824576dad6e9cfafb732a1d002a660d47fdd10c287

Observation 868051ea-bc84-4d31-b4e6-f9f10ef9a18d · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.164386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:b0d20f5ddade0c6d9ab63dcd41fb1855efc19b306a566bb8d012b5059086d6e4

Observation c86e20f7-edb9-4f7a-8a0d-efe3742abc0e · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.187490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:af5bdb8a27e1d08669a171da4292509b45d1f8a0fcb544a108f21070127e230a

Observation 866b06ad-891d-4b11-be46-e36d231c922c · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recogni- tion, pp.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recogni- tion, pp

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.217406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:6ddd4596b183a241fb945c4c1488073f35c2fd38b535174132d261fb3e6d0e23

Observation b9c925e8-78d0-4aec-a0c3-574ee6d9d27e · outbound

This paper cites LucidDreamer: Domain-free Generation of 3D Gaussian Splatting Scenes.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion LucidDreamer: Domain-free Generation of 3D Gaussian Splatting Scenes

Reference 69

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T19:39:23.533443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:0515e9d1aa60408aaa2ae315eba04d8f02f91d7ed8ac83b2d08d8d5eae1ef924

Observation 27fe4bca-ec18-49ac-9a84-611a753a5106 · outbound

This paper cites In: Proceedings of the Computer Vision and Pat- tern Recognition Conference, pp.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion In: Proceedings of the Computer Vision and Pat- tern Recognition Conference, pp

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.147267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:218a69badbddb7679b0f9cf844b82ad88c7f6c4e75e985298b3bd9160ebb6dbd

Observation 071c8dd0-49c9-4c04-8bd8-04a933641be8 · outbound

This paper cites WonderTurbo: Generating Interactive 3D World in 0.72 Seconds.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion WonderTurbo: Generating Interactive 3D World in 0.72 Seconds

Reference 71

Resolution
verified exact
arxiv_id, observed 2026-05-14T19:39:23.539721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:b60e6a8f8d31ba9d21f5886393fc605768dca540d22a06f080b7eca89b1db513

Observation ba25342a-341f-4e50-acf6-9ab676a9698b · outbound

This paper cites FlexWorld: Progressively Expanding 3D Scenes for Flexiable-View Synthesis.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion FlexWorld: Progressively Expanding 3D Scenes for Flexiable-View Synthesis

Reference 72

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T19:39:23.514524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:45e68ba1899af0a818c132219d7fbe2d99c8d6fe9cb50249e287c533694267cf

Observation 263f65e2-ebca-4bd2-bc7b-800e5eee77be · outbound

This paper cites In: 2025 International Conference on 3D Vision (3DV), pp.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion In: 2025 International Conference on 3D Vision (3DV), pp

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.131053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:b56a6b568964afdefca994bc3d96ff6ba5c45f865c5f232f03a18f3129e8badf

Observation cba3e222-1dc6-4c4a-a796-3b4fd37766e7 · outbound

This paper cites In: Proceedings of the Computer Vision and Pattern Recognition Conference, pp.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion In: Proceedings of the Computer Vision and Pattern Recognition Conference, pp

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.115517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:362cba43c91ad7d8e970a9904feb7722f13bcbf41502e4a5db74b64b161c7dca

Observation b1a206d3-4809-4c67-8c0d-eccf0fa97eff · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.226229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:8db80108a9db825e05e5c4ea42c3eb70cabe8958a388ec625ef9df8fe6e7d94c

Observation ac1d1ebc-9ab0-4aaa-962c-4d489eb220c9 · outbound

This paper cites Advances in neural information processing systems33, 6840– 6851 (2020).

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion Advances in neural information processing systems33, 6840– 6851 (2020)

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.201728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:2245ab189f77478c70f9c0285f53e0369ce2945d0a9b3178f87fc63a81aec5de

Observation 03dbf02f-a22c-402a-b8af-e72e63866bfe · outbound

This paper cites Advances in neural information processing systems35, 8633–8646 (2022).

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion Advances in neural information processing systems35, 8633–8646 (2022)

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.222530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:b843107356874a4feb43453c21f5581dfc1923093a238fa641a10b6f3d6628b5

Observation 37580a44-85d5-4124-8b87-cf3f39e4c183 · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.245821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:965165e8adcc7ec2d48eade071c556004294c34db4dd0b7d41523a38cdb0a005

Observation c84702a4-db52-41ca-bdad-fe65dd1e82c1 · outbound

This paper cites In: Proceedings of the Computer Vision and Pattern Recognition Conference, pp.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion In: Proceedings of the Computer Vision and Pattern Recognition Conference, pp

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.187886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:1b7eeebbbac69f02eedb5b6a46d224ce3d7c83fba89f1ad0617d2fa1909fad6e

Observation b82f5a7f-884e-446e-b0c9-b5d40419692c · outbound

This paper cites In: Pro- ceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion In: Pro- ceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.236113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:d0aece8193a98b4e9f4da6fc73eb2e6e61425fd8870745f52c8a7a352ee005f2

Observation e0cb7f61-0de9-4ccf-88bc-2b99a0c87a16 · outbound

This paper cites Stereo Magnification: Learning View Synthesis using Multiplane Images.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion Stereo Magnification: Learning View Synthesis using Multiplane Images

Reference 81

Resolution
metadata mismatch
local_arxiv, observed 2026-05-14T19:39:23.476389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:f284a92d2e8a7976e11db757a847659b9c67fe5fa4864a58e94f718f62c5571a

Observation 1420a960-5bec-4c51-b6c5-9bd44991e114 · outbound

This paper cites TrajectoryCrafter: Redirecting Camera Trajectory for Monocular Videos via Diffusion Models.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion TrajectoryCrafter: Redirecting Camera Trajectory for Monocular Videos via Diffusion Models

Reference 82

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T19:39:23.410882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:60eae62519fc9e453560e36c8c81ef7f22c73ad0f81d5221cbbeeedfd9e8aa4d

Observation d076b35b-2140-466e-8a82-2cc3d5a20eaa · outbound

This paper cites Q-Align: Teaching LMMs for Visual Scoring via Discrete Text-Defined Levels.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion Q-Align: Teaching LMMs for Visual Scoring via Discrete Text-Defined Levels

Reference 83

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T16:35:48.845096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:c9752a63fe67e9f136b16c278851e7839f9f95e567ab5cd196a198026e06d9c4

Observation e447fa58-de48-483e-83e8-d1c6926b74ad · outbound

This paper cites CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer

Reference 84

Resolution
metadata mismatch
local_arxiv, observed 2026-05-14T19:39:23.495415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:94031b317a58eef86835818f26a92821969d3ba92dc2ddf2550d3ccb62ac6af5

Observation 44bf464e-cf88-45b4-909c-b3c5f5f16b51 · outbound

This paper cites In: Pro- ceedings of the Computer Vision and Pat- tern Recognition Conference, pp.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion In: Pro- ceedings of the Computer Vision and Pat- tern Recognition Conference, pp

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.241584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:084ecb8e8e10ae694f44de44c65e9b4fa4727ba32d4b2154a754d0d6f3b50175

Observation 84020691-92e4-4043-bcd7-b95a329097af · outbound

This paper cites In: Proceedings of the Computer Vision and Pattern Recognition Conference, pp.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion In: Proceedings of the Computer Vision and Pattern Recognition Conference, pp

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.250842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:9d7a5c9f7acdac986918548fe91b0d0821aa83a74c84df29debff359620ec033

Observation 5d45e4bd-c57e-423e-8e7a-b13e1a2961b1 · outbound

This paper cites Advances in Neural Information Processing Systems37, 21875–21911 (2024) 25.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion Advances in Neural Information Processing Systems37, 21875–21911 (2024) 25

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T19:39:24.226881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:fc44beeb0a8f8993a8372af5b8cb89f42620883a16c6c45507866a216d124638

Pith citing papers

No inbound Pith citation observations are available.