Pith. sign in

Paper Citation Record · LEDGER

Geometric Context Transformer for Streaming 3D Reconstruction

As of 5 August 2026, this Paper Citation Record lists 100 of 105 outbound references and 20 inbound Pith citation observations for arXiv:2604.14141.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2604.14141 v2

Coverage vector

measured 100 of 105 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-10T13:54:13.644019Z

measured 120 of 120 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 20 of 20 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T13:07:07.298312Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-08T23:55:43.003851Z

Reference resolution

100 of 105 outbound references displayed

  • verified exact13
  • verified fuzzy85
  • unresolved1
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation fa5d0d2f-8f4b-4b78-b594-80c793228afb · outbound

This paper cites Map-free visual relocalization: Metric pose relative to a single image.

Geometric Context Transformer for Streaming 3D Reconstruction Map-free visual relocalization: Metric pose relative to a single image

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.685083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:cf7b02cd45da6a8ff134101fbad5950da0d3178f739b5efb1747c40aa0c7668a

Observation 60ed0500-5dd8-4736-9078-1da28a2b96ca · outbound

This paper cites Neural rgb-d surface reconstruction.

Geometric Context Transformer for Streaming 3D Reconstruction Neural rgb-d surface reconstruction

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.688844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:a4106e6091839758a5564bd1b7c6ad58023760eaba0811154c128412581fcadb

Observation 3e4607d5-0cb8-417e-a185-5e880795e0d3 · outbound

This paper cites Virtual KITTI 2.

Geometric Context Transformer for Streaming 3D Reconstruction Virtual KITTI 2

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-13T16:00:33.810478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:f5fcca0cf45fe883574680998515e8ae86eafc44869a26ceac8b5bfb94e7500a

Observation a4b55268-d88b-4faa-8aaa-290385b0a1a7 · outbound

This paper cites Orb-slam3: An accurate open-source library for visual, visual–inertial, and multimap slam.IEEE transactions on robotics, 37(6):1874–1890.

Geometric Context Transformer for Streaming 3D Reconstruction Orb-slam3: An accurate open-source library for visual, visual–inertial, and multimap slam.IEEE transactions on robotics, 37(6):1874–1890

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.492148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:155803548f3a8aac8771a83bd4bdcfb9c51921b033b7ed81b4ceab79bb5dc599

Observation e1c3eea8-7ed9-4e6d-9394-e7e7adc3d0ca · outbound

This paper cites Matterport3d: Learning from RGB-D data in indoor environments.

Geometric Context Transformer for Streaming 3D Reconstruction Matterport3d: Learning from RGB-D data in indoor environments

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.500287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:d1340831cd4d8d8f427f2255b2ea8b5cbccdd00522bcdbc62e2770246620ee45

Observation bd7aa65e-aa96-4f00-834f-2041f13ca329 · outbound

This paper cites Easi3r: Estimating disentangled motion from dust3r without training.

Geometric Context Transformer for Streaming 3D Reconstruction Easi3r: Estimating disentangled motion from dust3r without training

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.498950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:a82c671c9dc99be338cad592bca95f80ea397599cb4c9be903bfb1de354c2406

Observation 14107e27-2321-49c1-8bfc-4a5f76918a3b · outbound

This paper cites Ttt3r: 3d reconstruction as test-time training.Int.

Geometric Context Transformer for Streaming 3D Reconstruction Ttt3r: 3d reconstruction as test-time training.Int

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.509012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:ecb2609ea208c52231ab777cd468bc3a25a005240cd7d2df5a1bdb267af32071

Observation ae4f50c5-dc6b-41ef-a128-f87521db5f18 · outbound

This paper cites Long3r: Long sequence streaming 3d reconstruction.

Geometric Context Transformer for Streaming 3D Reconstruction Long3r: Long sequence streaming 3d reconstruction

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.547610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:98122c72c868a2adc6ed9a856c4d8161d95dd4dd3853b93fb4190297ff0027e3

Observation fbeb52d5-8550-4dae-99b7-ea816221b9c6 · outbound

This paper cites Scannet: Richly- annotated 3d reconstructions of indoor scenes.

Geometric Context Transformer for Streaming 3D Reconstruction Scannet: Richly- annotated 3d reconstructions of indoor scenes

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.540327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:0ac88fe85230411f75e8b3008e66d8cc273dcfbe13b3db44745c9d697a66b5f4

Observation f4eacae2-4795-48bb-b2cd-286e36ecb009 · outbound

This paper cites Objaverse: A universe of annotated 3d objects.

Geometric Context Transformer for Streaming 3D Reconstruction Objaverse: A universe of annotated 3d objects

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.353120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:ae8d5462fc77a1f6ed31b03548b655f9af3b5ebf63217a1751750d729e51333b

Observation 86902775-bced-450f-ab51-207304980f4c · outbound

This paper cites VGGT-Long: Chunk it, Loop it, Align it -- Pushing VGGT's Limits on Kilometer-scale Long RGB Sequences.

Geometric Context Transformer for Streaming 3D Reconstruction VGGT-Long: Chunk it, Loop it, Align it -- Pushing VGGT's Limits on Kilometer-scale Long RGB Sequences

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-17T08:55:00.688421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:eed876d5e0211ceabe1f5e53b9dd73bf6de16e065acd48c7d7c2c1d48ae55da4

Observation 4dc69e25-2979-49f3-b5df-c6c86964bca5 · outbound

This paper cites Superpoint: Self-supervised interest point detection and description.

Geometric Context Transformer for Streaming 3D Reconstruction Superpoint: Self-supervised interest point detection and description

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.280075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:ed32a95ff211fde0262b721e82e1e20871c0c5293b4cf28c9e3575359db845f8

Observation 890173ad-5d9d-4b4d-9313-73ce744b68ca · outbound

This paper cites St4rtrack: Simultaneous 4d reconstruction and tracking in the world.

Geometric Context Transformer for Streaming 3D Reconstruction St4rtrack: Simultaneous 4d reconstruction and tracking in the world

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.366534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:f075ff6a9d2014bac7ef62155d17c5c1096b1faa42242c482143a3ef7209c84c

Observation 1940f9a7-3cf2-4124-830a-61bcd9b69169 · outbound

This paper cites Mid-air: A multi-modal dataset for extremely low altitude drone flights.

Geometric Context Transformer for Streaming 3D Reconstruction Mid-air: A multi-modal dataset for extremely low altitude drone flights

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.253847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:b2d8e73d9348e2815fff433014756f4099c93c86ddf1470712f4213653b1d33e

Observation 071fa672-c65d-43b8-a5ab-2cfaa74204a1 · outbound

This paper cites Accurate, dense, and robust multiview stereopsis.IEEE transactions on pattern analysis and machine intelligence, 32(8):1362–1376.

Geometric Context Transformer for Streaming 3D Reconstruction Accurate, dense, and robust multiview stereopsis.IEEE transactions on pattern analysis and machine intelligence, 32(8):1362–1376

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.371771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:28cb84ae3a3f3cc429f66c4d258bbd245dfed8deb4648c53d48e81f114730a9f

Observation e2f749e9-ca34-45aa-9c83-2b9253fef0d3 · outbound

This paper cites Kubric: A scalable dataset generator.

Geometric Context Transformer for Streaming 3D Reconstruction Kubric: A scalable dataset generator

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.511023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:5d86904d70d53600beba57adf712c9f922e0bba7763df3a119d1e5cb3718f398

Observation b418b36c-28a0-408a-876f-fd02119ba3b3 · outbound

This paper cites LRM: Large reconstruction model for single image to 3D.

Geometric Context Transformer for Streaming 3D Reconstruction LRM: Large reconstruction model for single image to 3D

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.522539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:cbb4ec1aeec21498ae5c4e9d02ccedb133265753b40646bd907c58ca2a5a159a

Observation 14a050be-e210-4603-9078-99cfea916c78 · outbound

This paper cites ViPE: Video Pose Engine for 3D Geometric Perception.

Geometric Context Transformer for Streaming 3D Reconstruction ViPE: Video Pose Engine for 3D Geometric Perception

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.910540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:6cb1c92854881dc04fb96bebda7298f5155d142ac7ebe4fddd7cd2ddf8f38604

Observation ef7889d7-c1dd-4110-bc29-f53ae81957cb · outbound

This paper cites Deepmvs: Learning multi-view stereopsis.

Geometric Context Transformer for Streaming 3D Reconstruction Deepmvs: Learning multi-view stereopsis

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.681024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:b03af1b7351bf17cce5ea0c3cf6cb0a0b78f05c17f035e4ab8d44a7bf38f9048

Observation 99b750b4-ed1f-4641-8eab-75a80d4fcb13 · outbound

This paper cites DeepSpeed Ulysses: System Optimizations for Enabling Training of Extreme Long Sequence Transformer Models.

Geometric Context Transformer for Streaming 3D Reconstruction DeepSpeed Ulysses: System Optimizations for Enabling Training of Extreme Long Sequence Transformer Models

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:07:22.650116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:e3727517efc20871b9701dad06af9918a8fb5e60417971434a478efb2ced6521

Observation 49193a2f-5cb3-4786-91a9-41a9ac91f502 · outbound

This paper cites Pow3r: Empowering unconstrained 3d reconstruction with camera and scene priors.

Geometric Context Transformer for Streaming 3D Reconstruction Pow3r: Empowering unconstrained 3d reconstruction with camera and scene priors

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.362855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:bfe177e5ed751a423ddbe4e358e1df0a99f294593768d2b4a23ce019ff2b8100

Observation bd29b91c-8012-43ea-acdd-54f582dd39e9 · outbound

This paper cites Anysplat: Feed-forward 3d gaussian splatting from unconstrained views.ACM Trans.

Geometric Context Transformer for Streaming 3D Reconstruction Anysplat: Feed-forward 3d gaussian splatting from unconstrained views.ACM Trans

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.265720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:63e35b9adb82383f97a729b3641bb657774462e6aca07771f6c7b05f16e2fa3d

Observation af0b075e-5846-4cc4-b11d-d728ec872970 · outbound

This paper cites Barron, Noah Snavely, and Aleksander Hoły´nski.

Geometric Context Transformer for Streaming 3D Reconstruction Barron, Noah Snavely, and Aleksander Hoły´nski

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.514746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:7b6f8dedceb8664d2a8dbf893569ee42a612060973f0d039787800b99c658368

Observation 94fee29c-de4c-4ac2-bf44-77cef8699614 · outbound

This paper cites Stereo4D: Learning How Things Move in 3D from Internet Stereo Videos.

Geometric Context Transformer for Streaming 3D Reconstruction Stereo4D: Learning How Things Move in 3D from Internet Stereo Videos

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.293105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:377849ef790eff5035434cf996f847c595f2a4934f3c95e299d67df76b464579

Observation e5c26eaa-33bc-4e8f-a0f6-4423c39397c4 · outbound

This paper cites MapAnything: Universal Feed-Forward Metric 3D Reconstruction.

Geometric Context Transformer for Streaming 3D Reconstruction MapAnything: Universal Feed-Forward Metric 3D Reconstruction

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-12T11:06:15.374425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:4bdb27b57c4a8bc85f064cb0e1faf7efda247e5a60de83412ec2d76d9801ee45

Observation 3a8f90fc-d3f8-4f99-b6de-90b9695350a9 · outbound

This paper cites Tanks and temples: Benchmarking large-scale scene reconstruction.ACM Trans.

Geometric Context Transformer for Streaming 3D Reconstruction Tanks and temples: Benchmarking large-scale scene reconstruction.ACM Trans

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.338792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:fae99b359e365673a3061651f5d8630096f20a20f1ed7a20f20a458c130a01f2

Observation 14c024c1-641e-4754-aae1-9f273e816e96 · outbound

This paper cites Gonzalez, Hao Zhang, and Ion Stoica.

Geometric Context Transformer for Streaming 3D Reconstruction Gonzalez, Hao Zhang, and Ion Stoica

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.366121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:6eec914c8effe7f3b0a138a19221a9e2d59171c70628c31132440f3ec15da1cd

Observation b5ddc2f3-e762-4dc4-995f-821cbe604469 · outbound

This paper cites STream3R: Scalable sequential 3D reconstruction with causal transformer.

Geometric Context Transformer for Streaming 3D Reconstruction STream3R: Scalable sequential 3D reconstruction with causal transformer

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.369148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:e0c5e4d1fff0d86bb9de01358edfa008e164f0839aab1d75095bb4289a08f76b

Observation 472facde-5fb5-41c3-ae0b-f86539a3932d · outbound

This paper cites Instant3d: Fast text-to-3D with sparse-view generation and large reconstruction model.

Geometric Context Transformer for Streaming 3D Reconstruction Instant3d: Fast text-to-3D with sparse-view generation and large reconstruction model

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.544307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:6c3f86ea9e56fba168a116114c6bd1c1f6d8294abf51afbd863b2fe87a82d9ed

Observation 437d4197-d327-4028-b8bf-fb2332227875 · outbound

This paper cites Matrixcity: A large-scale city dataset for city-scale neural rendering and beyond.

Geometric Context Transformer for Streaming 3D Reconstruction Matrixcity: A large-scale city dataset for city-scale neural rendering and beyond

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.493934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:114955b97a6ab0feb3ae45a571b19f15846d9c2445a5b892dc8915f4d2477101

Observation 0584824d-3e7c-4514-9ce8-2b12453346d6 · outbound

This paper cites Megadepth: Learning single-view depth prediction from internet photos.

Geometric Context Transformer for Streaming 3D Reconstruction Megadepth: Learning single-view depth prediction from internet photos

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.314296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:60e2e988b617dc7ddfa569a9b6e28e19c1fcf0d4bbcdc1cacab0a6113bb1f87f

Observation 7e8ec462-55e3-4237-8af3-ca0bd2fac4c0 · outbound

This paper cites Megasam: Accurate, fast and robust structure and motion from casual dynamic videos.

Geometric Context Transformer for Streaming 3D Reconstruction Megasam: Accurate, fast and robust structure and motion from casual dynamic videos

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.543936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:cd6702bdc4613f666b449fb936306e166dadf3879856632d4abd389ce0e16240

Observation e5f3ce6e-1af7-4dc1-8b9e-cb140ec7e3b2 · outbound

This paper cites Wint3r: Window-based streaming reconstruction with camera token pool.

Geometric Context Transformer for Streaming 3D Reconstruction Wint3r: Window-based streaming reconstruction with camera token pool

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.622058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:1b1558448da59b0cf506a5fa997d1f5f4a46de78f58888399bf932cfbb945498

Observation 115ef8c4-3a48-4d1d-9bf7-2070a2dd7077 · outbound

This paper cites Torchtitan: One-stop pytorch native solution for production ready LLM pretraining.

Geometric Context Transformer for Streaming 3D Reconstruction Torchtitan: One-stop pytorch native solution for production ready LLM pretraining

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.667594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:f38852eee2cfc3e142b4e2ee29f20c4214f571661a399e6f84ed2be19310faa5

Observation 93644598-0195-4e50-a79f-055db9809c9d · outbound

This paper cites Kitti-360: A novel dataset and benchmarks for urban scene understanding in 2d and 3d.IEEE Trans.

Geometric Context Transformer for Streaming 3D Reconstruction Kitti-360: A novel dataset and benchmarks for urban scene understanding in 2d and 3d.IEEE Trans

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.671146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:407fcec6c18be9971720b853272f90e57fbe1b98c2612a7a0749d64567c309cf

Observation 17e7db54-78ab-4f23-a116-3033904c1f78 · outbound

This paper cites Longsplat: Robust unposed 3d gaussian splatting for casual long videos.

Geometric Context Transformer for Streaming 3D Reconstruction Longsplat: Robust unposed 3d gaussian splatting for casual long videos

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.551558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:81d84d609014981c27c829a468b75481af4acabfc27d0dc3f617dfce582e2b50

Observation b2f72a62-b84b-412b-bdb8-94c781572ca9 · outbound

This paper cites Depth Anything 3: Recovering the Visual Space from Any Views.

Geometric Context Transformer for Streaming 3D Reconstruction Depth Anything 3: Recovering the Visual Space from Any Views

Reference 37

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T02:07:59.495600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:e58c2cc5409ce333840255878d5cb95fbdd16795c92b610df6d01cf638535db7

Observation 70c0bf23-4991-4eb1-8e61-fe7765cbc2dd · outbound

This paper cites Dl3dv-10k: A large-scale scene dataset for deep learning-based 3d vision.

Geometric Context Transformer for Streaming 3D Reconstruction Dl3dv-10k: A large-scale scene dataset for deep learning-based 3d vision

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.454947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:ac22b67c10f2e898acc042a37e415d156ea54cfe32ffdf64ca184f94ed7fa832

Observation 9f9278f9-a2c4-429c-b47e-1b9bf8fe9942 · outbound

This paper cites Slam3r: Real-time dense scene reconstruction from monocular rgb videos.

Geometric Context Transformer for Streaming 3D Reconstruction Slam3r: Real-time dense scene reconstruction from monocular rgb videos

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.540743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:57d1ea821cd2487aaf5f233494d5f76c2f31dddc788537ecb3b14e1bba26eb70

Observation 536c4bec-9984-48e0-b655-8141915db963 · outbound

This paper cites Align3r: Aligned monocular depth estimation for dynamic videos.

Geometric Context Transformer for Streaming 3D Reconstruction Align3r: Aligned monocular depth estimation for dynamic videos

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.554984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:67153b90f4985c60da86a0f4e0b4778292f8275f8950acb32c4204d3aa47385f

Observation 5eea0d83-2315-465b-91e5-6f21572d7d4b · outbound

This paper cites Matrix3d: Large photogrammetry model all-in-one.

Geometric Context Transformer for Streaming 3D Reconstruction Matrix3d: Large photogrammetry model all-in-one

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.536650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:6c664d284c95263f4132f57435770e41e2bc5cd4c1061ec61bbe657cbc10d2b3

Observation c1c234d9-e0b4-4043-89b1-f28d311b2399 · outbound

This paper cites Vggt-slam: Dense rgb slam optimized on the sl (4) manifold.Adv.

Geometric Context Transformer for Streaming 3D Reconstruction Vggt-slam: Dense rgb slam optimized on the sl (4) manifold.Adv

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.407826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:afb20a09cb21282718132d4064de8bda297f98fba9b03ff7c2fd2f760f029fdb

Observation 612da0d6-35a0-40b8-9bec-95849f764d78 · outbound

This paper cites Scenenet rgb-d: Can 5m synthetic images beat generic imagenet pre-training on indoor segmentation? InInt.

Geometric Context Transformer for Streaming 3D Reconstruction Scenenet rgb-d: Can 5m synthetic images beat generic imagenet pre-training on indoor segmentation? InInt

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.360555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:0aefffb7bf4538e7fa8c2d83b64c02e5ff5cd86199d372540c5be78a1eba75db

Observation a2d4d9b0-81b6-4414-844c-795528f57de6 · outbound

This paper cites Orb-slam: A versatile and accurate monocular slam system.

Geometric Context Transformer for Streaming 3D Reconstruction Orb-slam: A versatile and accurate monocular slam system

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.645134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:e5b39b95ef4dcaa0751b4b7d723d2ff91b7bc47235e1bb298a7daecdb8d4dc5b

Observation d32c8725-923e-48d4-802b-b3b6b4c93796 · outbound

This paper cites Orb-slam2: An open-source slam system for monocular, stereo, and rgb-d cameras.IEEE transactions on robotics, 33(5):1255–1262.

Geometric Context Transformer for Streaming 3D Reconstruction Orb-slam2: An open-source slam system for monocular, stereo, and rgb-d cameras.IEEE transactions on robotics, 33(5):1255–1262

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.650748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:c880d3e137bdb10b6f0600eeea2978b07dd2e2170ea5d1794e3420db71a767a0

Observation e349d7b0-8c82-4de1-aa2a-a0de5bdbeace · outbound

This paper cites Mast3r-slam: Real-time dense slam with 3d reconstruction priors.

Geometric Context Transformer for Streaming 3D Reconstruction Mast3r-slam: Real-time dense slam with 3d reconstruction priors

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.632199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:e2f4f985990b0ebf33cb8a5faaf537bf57817c197897f600cc9fe60568f5c2d2

Observation 00155a2a-767b-4d69-a8a4-46c57573844b · outbound

This paper cites an unresolved cited work.

Geometric Context Transformer for Streaming 3D Reconstruction Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-05-18T20:22:52.662044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:e20290bc43a6c77066bb1247f7dceb489f3037cf59636a19e382f79cbb1bb867

Observation e52ac131-f8b0-4a28-9886-c7f94ea04ab2 · outbound

This paper cites Global structure-from-motion revisited.

Geometric Context Transformer for Streaming 3D Reconstruction Global structure-from-motion revisited

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.395159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:310de9eb327d7b8e4ceb693ddd6e4c38eeb0dbecc863d5d786cbdadfba833808

Observation bbbc1551-7a87-4f30-a0ca-5afa1d229c31 · outbound

This paper cites Aria digital twin: A new benchmark dataset for egocentric 3d machine perception.

Geometric Context Transformer for Streaming 3D Reconstruction Aria digital twin: A new benchmark dataset for egocentric 3d machine perception

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.475962Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:0bf1c624e2080095584e522bdfa6babfdac4ee1231cc8984ae43227b0826add5

Observation 76f175ef-ef9b-45b1-a59f-bd965913b604 · outbound

This paper cites Tartanground: A large-scale dataset for ground robot perception and navigation.

Geometric Context Transformer for Streaming 3D Reconstruction Tartanground: A large-scale dataset for ground robot perception and navigation

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.450885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:97925f904eb9d1b8c2ebcca912f3f7078da2cb716b795451f7440b055a9a42f8

Observation 492d794c-0423-4d05-9471-f426a4a8c1aa · outbound

This paper cites Chang, Manolis Savva, Yili Zhao, and Dhruv Batra.

Geometric Context Transformer for Streaming 3D Reconstruction Chang, Manolis Savva, Yili Zhao, and Dhruv Batra

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.525564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:12abc32dea1152e1b6f376e636beaace7d032ce532540e841050dd662cd77299

Observation 09b2add1-59b6-4812-93a7-5dbc9faf6997 · outbound

This paper cites Common objects in 3d: Large-scale learning and evaluation of real-life 3d category reconstruction.

Geometric Context Transformer for Streaming 3D Reconstruction Common objects in 3d: Large-scale learning and evaluation of real-life 3d category reconstruction

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.516085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:63f3f52da22bc3e33c2f078e3bc472d2025bf9eae0237fbb55ba26bd354fa584

Observation 9f934fbe-4edc-4dff-bc94-e40bbc1cafde · outbound

This paper cites Hypersim: A photorealistic synthetic dataset for holistic indoor scene understanding.

Geometric Context Transformer for Streaming 3D Reconstruction Hypersim: A photorealistic synthetic dataset for holistic indoor scene understanding

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.524744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:dc952d64d67e0318eaefddace2bcda963b792dbf47a68c5915cf047af1857185

Observation 4f89c0d1-2ace-4acd-b04f-bb9557673013 · outbound

This paper cites Superglue: Learning feature matching with graph neural networks.

Geometric Context Transformer for Streaming 3D Reconstruction Superglue: Learning feature matching with graph neural networks

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.528875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:aff6b62f79026405dcc34a5fcdc424fdde7d7ab0f03635f77ab06c8de511d7e5

Observation 242784e4-874a-4f22-b321-47b50837bfcb · outbound

This paper cites Habitat: A platform for embodied ai research.

Geometric Context Transformer for Streaming 3D Reconstruction Habitat: A platform for embodied ai research

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.537129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:0a8ef980e25738e810bb0e45abd16786cb3b903884c9d96b5f4f0309c1329224

Observation 3c53b42f-644f-428f-8f58-cee208ef1aa1 · outbound

This paper cites Structure-from-motion revisited.

Geometric Context Transformer for Streaming 3D Reconstruction Structure-from-motion revisited

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.502431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:47cdb1d9d5cf9c277085eb13c781a001c4e2719e947d8bbc72e87e6932e10271

Observation 8072b2b3-ed17-4e01-a7ba-e4e449c72e9f · outbound

This paper cites Pixelwise view selection for unstructured multi-view stereo.

Geometric Context Transformer for Streaming 3D Reconstruction Pixelwise view selection for unstructured multi-view stereo

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.505716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:c4cd880d66456e6dd130afad343a46f2211be4f337406b4c578844231eb4e5af

Observation 9aca7d24-df72-4a03-971d-679f3ce819cd · outbound

This paper cites A multi-view stereo benchmark with high-resolution images and multi-camera videos.

Geometric Context Transformer for Streaming 3D Reconstruction A multi-view stereo benchmark with high-resolution images and multi-camera videos

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.656576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:1333801efe8398b20adf29646a3d5f0ef68dd4f4b99f5342500cf4f424e64828

Observation baf39580-8086-4ceb-b39c-9f80a2f9e267 · outbound

This paper cites FastVGGT: Training-Free Acceleration of Visual Geometry Transformer.

Geometric Context Transformer for Streaming 3D Reconstruction FastVGGT: Training-Free Acceleration of Visual Geometry Transformer

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-05-15T23:36:07.097348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:13ea5ba1d3495e5ad893a56ff45342970a42e0281b11c0e4b58cd5821ef203f1

Observation 134d49cd-41b8-4425-ad1d-8d806b4703e4 · outbound

This paper cites Scene coordinate regression forests for camera relocalization in rgb-d images.

Geometric Context Transformer for Streaming 3D Reconstruction Scene coordinate regression forests for camera relocalization in rgb-d images

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.548041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:f274488d43f1606f2be77a02e51bae00042d1a1e9cc6683762fe5560dc7652f4

Observation d45ae004-589b-4f8b-aebc-72603eb84df2 · outbound

This paper cites Splatt3R: Zero-shot Gaussian Splatting from Uncalibrated Image Pairs.

Geometric Context Transformer for Streaming 3D Reconstruction Splatt3R: Zero-shot Gaussian Splatting from Uncalibrated Image Pairs

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-05-17T01:16:06.916611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:e0bcd4f65ddfa52a2e37548e5f631fb9775a2fb35d93813fea679b6ed5cb1ac2

Observation 1ac3dbe7-13f8-40da-8cc5-2ad12c32c6a5 · outbound

This paper cites Seitz, and Richard Szeliski.

Geometric Context Transformer for Streaming 3D Reconstruction Seitz, and Richard Szeliski

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.600552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:0125649025ee0b0816e626324034f10c1ccc4ca6d227aa592d139ae2679cde2b

Observation f37778c6-1b97-4f64-8d86-165250774dd9 · outbound

This paper cites The Replica Dataset: A Digital Replica of Indoor Spaces.

Geometric Context Transformer for Streaming 3D Reconstruction The Replica Dataset: A Digital Replica of Indoor Spaces

Reference 63

Resolution
verified exact
arxiv_id, observed 2026-05-12T16:34:29.536218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:259b55f7d2ffa1aa0fe52cbcc5d62aece413bdc4e978ea9822b2f121732125fb

Observation bc526bdf-8111-426e-abdf-0b1e90f06eda · outbound

This paper cites Dynamic point maps: A versatile representation for dynamic 3d reconstruction.

Geometric Context Transformer for Streaming 3D Reconstruction Dynamic point maps: A versatile representation for dynamic 3d reconstruction

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.503765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:94e04407fe84689a32dfe59eafc216427e4ee527ca3d0e5a3ab77996b1bf092a

Observation 9495bfd1-229c-4fd1-8f75-5c192bd69e5e · outbound

This paper cites Loftr: Detector-free local feature matching with transformers.

Geometric Context Transformer for Streaming 3D Reconstruction Loftr: Detector-free local feature matching with transformers

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.507260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:0a663d8d6e08ab278f486b70bfc78cb2c417c7d3a0d4577c5cb572d6cb1fb11c

Observation f762456f-f508-45b4-9cfa-be47dd98366c · outbound

This paper cites Scalability in perception for autonomous driving: Waymo open dataset.

Geometric Context Transformer for Streaming 3D Reconstruction Scalability in perception for autonomous driving: Waymo open dataset

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.479846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:05b3d15929cf6a0b0a12b948748815f309923ef02907c9b0b397ef0b8e834670

Observation 2390b199-0064-4b9a-97c6-404cb81f6156 · outbound

This paper cites LGM: Large multi-view gaussian model for high-resolution 3D content creation.

Geometric Context Transformer for Streaming 3D Reconstruction LGM: Large multi-view gaussian model for high-resolution 3D content creation

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.518147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:b941411a2b2055b1617ee631bf07b9f4682ff10267c35c802e5f36f0e1f70c8f

Observation 88624e12-a687-4d07-a6fe-c3edfae61a44 · outbound

This paper cites The oxford spires dataset: Benchmarking large-scale lidar-visual localisation, reconstruction and radiance field methods.International Journal of Robotics Research.

Geometric Context Transformer for Streaming 3D Reconstruction The oxford spires dataset: Benchmarking large-scale lidar-visual localisation, reconstruction and radiance field methods.International Journal of Robotics Research

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.674789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:95b704a8033fabfaf42c4b638892261e5e301c88d41f1575b12482d033547829

Observation a9300d73-0121-4391-9480-f01e1c4809d6 · outbound

This paper cites Droid-slam: Deep visual slam for monocular, stereo, and rgb-d cameras.Adv.

Geometric Context Transformer for Streaming 3D Reconstruction Droid-slam: Deep visual slam for monocular, stereo, and rgb-d cameras.Adv

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.445921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:bc70a0602c7f3aaa9f5d5edf2270bd262169b7636092046542fa94f4fb7badb5

Observation d037963a-2697-40a1-98e6-47d3748b6d78 · outbound

This paper cites Smd-nets: Stereo mixture density networks.

Geometric Context Transformer for Streaming 3D Reconstruction Smd-nets: Stereo mixture density networks

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.455434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:4803b5e864d4514f3747e1bf4605c90e46abaacde115969b7915abd6004a9f35

Observation 2aaa984e-687c-4b19-8475-01d1aa63a182 · outbound

This paper cites Least-squares estimation of transformation parameters between two point patterns.IEEE Trans.

Geometric Context Transformer for Streaming 3D Reconstruction Least-squares estimation of transformation parameters between two point patterns.IEEE Trans

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.617837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:c9037e6c2c137ce036cf3ad96c016a8d790bf15cb38895559de982725eb2a8c3

Observation 210673f1-0c4d-4927-a0b5-a5fa9c504e54 · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

Geometric Context Transformer for Streaming 3D Reconstruction Wan: Open and Advanced Large-Scale Video Generative Models

Reference 72

Resolution
verified exact
local_arxiv, observed 2026-05-10T13:55:28.766457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:f7cef75ec0e40b107b261e8c89ee5a41b9856fafc6e027db32701297693ade2b

Observation 59020b42-f109-4cdd-9755-349625214790 · outbound

This paper cites 3d reconstruction with spatial memory.

Geometric Context Transformer for Streaming 3D Reconstruction 3d reconstruction with spatial memory

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.485397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:aabf894527f6380da7085814152d16ed4bf3bcd2b60169755474425edc17d147

Observation f4251df6-839d-4d57-b87a-39c63cfcb98c · outbound

This paper cites Spatialvid: A large-scale video dataset with spatial annotations.

Geometric Context Transformer for Streaming 3D Reconstruction Spatialvid: A large-scale video dataset with spatial annotations

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.462854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:6a7b74c6f93814e2a06d20b6578eadb219f810d738d4d80534cb5f9dae2a3160

Observation 94dc25d2-6655-44c7-9c41-b19456d41da7 · outbound

This paper cites Vggt: Visual geometry grounded transformer.

Geometric Context Transformer for Streaming 3D Reconstruction Vggt: Visual geometry grounded transformer

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.420781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:64781781fb91431380857fe36304185288414062f0a2c0efacd784f6c5d09a6a

Observation 51775605-29a5-460b-a7fe-db97863a716d · outbound

This paper cites Vggsfm: Visual geometry grounded deep structure from motion.

Geometric Context Transformer for Streaming 3D Reconstruction Vggsfm: Visual geometry grounded deep structure from motion

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.425421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:5cbc6171477231175ff719f189dc306c7b7572023505b5abf74b49369d378aaa

Observation 9035b9ee-3c10-4e73-b4c4-2df251e9ee23 · outbound

This paper cites Flow-motion and depth network for monocular stereo and beyond.IEEE Robotics and Automation Letters, 5(2):3307–3314.

Geometric Context Transformer for Streaming 3D Reconstruction Flow-motion and depth network for monocular stereo and beyond.IEEE Robotics and Automation Letters, 5(2):3307–3314

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.414026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:ce6524a0336f800d37d11fde7658e543fd684be4d8d28029fc8170e0ce45e341

Observation 3567a891-aa06-48a0-a997-5c10e84a98e5 · outbound

This paper cites PF-LRM: Pose-free large reconstruction model for joint pose and shape prediction.

Geometric Context Transformer for Streaming 3D Reconstruction PF-LRM: Pose-free large reconstruction model for joint pose and shape prediction

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.607342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:8cd2a6cd36cfa2a2b7a49dc887c72163805188937cd650ce22b73c04f2e6ec70

Observation 68488113-5386-436b-8cdf-6f4291ed10e2 · outbound

This paper cites Efros, and Angjoo Kanazawa.

Geometric Context Transformer for Streaming 3D Reconstruction Efros, and Angjoo Kanazawa

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.611105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:049dd6afe0a67b376a50726c41f89f16d9e0e2281e17b95562181450259fa56f

Observation d1960e2b-743e-44c0-8c82-dec43d3bbbe4 · outbound

This paper cites Continuous 3d perception model with persistent state.

Geometric Context Transformer for Streaming 3D Reconstruction Continuous 3d perception model with persistent state

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.628785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:94605ea60cfc6ae90fae742783c1002a9db5ca330ea98c9b4c81d9a16c1900f0

Observation ab53510c-e8cc-472f-af04-93fc5f57e35d · outbound

This paper cites Dust3r: Geometric 3d vision made easy.

Geometric Context Transformer for Streaming 3D Reconstruction Dust3r: Geometric 3d vision made easy

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.641951Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:0eaf1a240983c000e90e198b63146d25c38a1a6890b039892e17032778b755a3

Observation 418d1a20-5439-4b2a-89ae-12f20a0d1bc8 · outbound

This paper cites Tartanair: A dataset to push the limits of visual slam.

Geometric Context Transformer for Streaming 3D Reconstruction Tartanair: A dataset to push the limits of visual slam

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.397181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:b64dcee5ae22bf2fbcf1f94e9effeb4b82c80c56c8ea31326b824de169b661fb

Observation 208934b4-9125-48d6-bd46-6102b51a06cb · outbound

This paper cites $\pi^3$: Permutation-Equivariant Visual Geometry Learning.

Geometric Context Transformer for Streaming 3D Reconstruction $\pi^3$: Permutation-Equivariant Visual Geometry Learning

Reference 83

Resolution
verified exact
arxiv_id, observed 2026-05-12T17:14:22.853753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:50c95c8736016670576b9cab088cb2b68c0c6c2b7f396d153ad5bef03df73da3

Observation 70c95b65-352f-4507-97a2-65d2647dde87 · outbound

This paper cites Zamir, Zhiyang He, Alexander Sax, Jitendra Malik, and Silvio Savarese.

Geometric Context Transformer for Streaming 3D Reconstruction Zamir, Zhiyang He, Alexander Sax, Jitendra Malik, and Silvio Savarese

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.590027Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:19ac2e9fb485e7fbf1a24d3a9e9cf600e80ccfdc5cf1c817ff7e431a8cc3ec6a

Observation 2e13e931-1bdc-4df5-b041-2414af00f1c4 · outbound

This paper cites Rgbd objects in the wild: Scaling real-world 3d object learning from rgb-d videos.

Geometric Context Transformer for Streaming 3D Reconstruction Rgbd objects in the wild: Scaling real-world 3d object learning from rgb-d videos

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.593883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:7731b0bc42a74411ce574b57247374477f5bba47821927fca673c5bd3bd5d9d2

Observation 281a1098-a2f1-4742-9ec9-23fc85d9d61b · outbound

This paper cites Spatialtrackerv2: 3d point tracking made easy.

Geometric Context Transformer for Streaming 3D Reconstruction Spatialtrackerv2: 3d point tracking made easy

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.432014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:a9d846ea1a3f8b60a5b7065c16dc99ab2507b3512a6d73e6c2f5035b42839388

Observation d928b326-d492-4b2b-a6cc-69957c8e8d57 · outbound

This paper cites Scal3r: Scalable test-time training for large-scale 3d reconstruction.

Geometric Context Transformer for Streaming 3D Reconstruction Scal3r: Scalable test-time training for large-scale 3d reconstruction

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.581976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:ef33230e76fd5a5427f55328b7b27a2e30e7b64b7c7ea9b73276117130e112b8

Observation f75b2d54-32d6-4ec6-95cd-dcf1cb18d0bf · outbound

This paper cites GRM: Large gaussian reconstruction model for efficient 3d reconstruction and generation.

Geometric Context Transformer for Streaming 3D Reconstruction GRM: Large gaussian reconstruction model for efficient 3d reconstruction and generation

Reference 88

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.550850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:fb83d6f4b5f379f28938d21ddbaef9466ac2f2dbe37a2e4c826488f6197771d5

Observation 06268585-ffa8-483d-978b-1f5fb9e3ab5e · outbound

This paper cites DMV3D: Denoising multi-view diffusion using 3D large reconstruction model.

Geometric Context Transformer for Streaming 3D Reconstruction DMV3D: Denoising multi-view diffusion using 3D large reconstruction model

Reference 89

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.578058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:381017084af29ca00aff84b5fd857a02049284567ed9b3e5f124df5becac90ad

Observation 0940ae0e-e58c-47e9-813e-03f46a52e41f · outbound

This paper cites Fast3r: Towards 3d reconstruction of 1000+ images in one forward pass.

Geometric Context Transformer for Streaming 3D Reconstruction Fast3r: Towards 3d reconstruction of 1000+ images in one forward pass

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.597482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:c6f11bc15cab4b3ec58c411fa057a1ff86b0f41d83fa5c8ec0cd3ddcc03f8151

Observation 9f49bd1c-1b67-4e9c-abce-49415d7056e1 · outbound

This paper cites Mvsnet: Depth inference for unstructured multi-view stereo.

Geometric Context Transformer for Streaming 3D Reconstruction Mvsnet: Depth inference for unstructured multi-view stereo

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.570497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:b2d8a45e40cae2b3b16442e281ef2009505187b535ed73a6f1d4461f490a35e1

Observation 4c5742a9-a0dd-4f63-920e-a840aab4c3ef · outbound

This paper cites Recurrent mvsnet for high-resolution multi-view stereo depth inference.

Geometric Context Transformer for Streaming 3D Reconstruction Recurrent mvsnet for high-resolution multi-view stereo depth inference

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.404023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:5e8e10c708ae864be171d75b998a815f328fadb802b2c0e227053e8eaf975b20

Observation b3b2e18d-955f-4290-b411-6134acdb6590 · outbound

This paper cites Blendedmvs: A large-scale dataset for generalized multi-view stereo networks.

Geometric Context Transformer for Streaming 3D Reconstruction Blendedmvs: A large-scale dataset for generalized multi-view stereo networks

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.416599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:3daddc62f1c3cc037a682a88aea8553787a33f5e963f6fe315e5f7d205fa3815

Observation 92ce4e80-d1f1-49a0-a6bf-068f693bb944 · outbound

This paper cites No pose, no problem: Surprisingly simple 3d gaussian splats from sparse unposed images.

Geometric Context Transformer for Streaming 3D Reconstruction No pose, no problem: Surprisingly simple 3d gaussian splats from sparse unposed images

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.562873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:88db4585befea3a3c85efad6f595aff20e4ec595eb9bd7ceef0dd0138308e6f6

Observation 7ce282e1-d508-4f0f-97e1-20ef00eac9ab · outbound

This paper cites FlashInfer: Efficient and Customizable Attention Engine for LLM Inference Serving.

Geometric Context Transformer for Streaming 3D Reconstruction FlashInfer: Efficient and Customizable Attention Engine for LLM Inference Serving

Reference 95

Resolution
verified exact
arxiv_id, observed 2026-05-16T13:26:34.840430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:1092543d5224942c33be19da222d0dccbc8c7cf22f0d0b0e79484b9c10ad2164

Observation f09e2938-85c2-4550-9818-4f2265268b20 · outbound

This paper cites Scannet++: A high-fidelity dataset of 3d indoor scenes.

Geometric Context Transformer for Streaming 3D Reconstruction Scannet++: A high-fidelity dataset of 3d indoor scenes

Reference 96

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.467183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:90b6111595c2906c1966f076b54d960236ecc19ed93d19961a089b3d67ac458f

Observation 911ff603-7502-41f3-9029-6548f5db47e3 · outbound

This paper cites InfiniteVGGT: Visual geometry grounded transformer for endless streams.

Geometric Context Transformer for Streaming 3D Reconstruction InfiniteVGGT: Visual geometry grounded transformer for endless streams

Reference 97

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:55:28.739497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:bf6fb24012438171e7d3eb5f53d2688749d528a15746cb0c44ee697baf1398be

Observation dda3a2d2-fddd-416a-aec0-52c171821dc0 · outbound

This paper cites Monst3r: A simple approach for estimating geometry in the presence of motion.

Geometric Context Transformer for Streaming 3D Reconstruction Monst3r: A simple approach for estimating geometry in the presence of motion

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.443649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:074ebcdfdf6f5e91fae179b59348e1d06c04a9c04aa41f7de55dfb7271010797

Observation 1b4820e7-8fad-4ec5-a02e-41a3b4f96c6d · outbound

This paper cites LoGeR: Long-Context Geometric Reconstruction with Hybrid Memory.

Geometric Context Transformer for Streaming 3D Reconstruction LoGeR: Long-Context Geometric Reconstruction with Hybrid Memory

Reference 99

Resolution
verified exact
local_arxiv, observed 2026-05-10T13:55:28.743882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:5cad3a7139389bf8a71bede6749c0188bd687297af3c9509f0f8e4161a454251

Observation 40b5cd4a-ece8-4812-a00f-8c53b79be863 · outbound

This paper cites Flare: Feed-forward geometry, appearance and camera estimation from uncalibrated sparse views.

Geometric Context Transformer for Streaming 3D Reconstruction Flare: Feed-forward geometry, appearance and camera estimation from uncalibrated sparse views

Reference 100

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T20:22:52.559337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:cf096d09fc138478053d33a3f5dd6af06edbdadaaec83da2f38b3ee6a9b59ffe

Pith citing papers

Observation 763518fe-b8df-4b12-8d0b-4d4acb764d4d · inbound

NavOne: One-Step Global Planning for Vision-Language Navigation on Top-Down Maps cites this paper.

NavOne: One-Step Global Planning for Vision-Language Navigation on Top-Down Maps Geometric Context Transformer for Streaming 3D Reconstruction

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-05-11T18:51:08.593187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T13:33:55.535856Z digest=sha256:0943275be917e745f2acdd9fc8b500c052c19b094f7cce5b477b5e36ecf9ff90

Observation 0a217dcf-b29a-41e3-9cd7-4ccf2e4b1380 · inbound

NavOne: One-Step Global Planning for Vision-Language Navigation on Top-Down Maps cites this paper.

NavOne: One-Step Global Planning for Vision-Language Navigation on Top-Down Maps Geometric Context Transformer for Streaming 3D Reconstruction

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-05-11T00:50:49.552353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-11T00:50:11.709217Z digest=sha256:81b12fc5ca7e24191e84bceac2ec4023f474d3dee748b52895c813549716f251

Observation 64250608-a83b-43eb-9809-a148d3db096f · inbound

NavOne: One-Step Global Planning for Vision-Language Navigation on Top-Down Maps cites this paper.

NavOne: One-Step Global Planning for Vision-Language Navigation on Top-Down Maps Geometric Context Transformer for Streaming 3D Reconstruction

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-05-20T23:13:50.266966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-20T23:13:41.784895Z digest=sha256:59c5852401b0d3cb74511079a3590ec9887a7a5e65f865f4321cea04df78c1d9

Observation 210b51cd-3de6-4433-a790-d66644c1c223 · inbound

NavOne: One-Step Global Planning for Vision-Language Navigation on Top-Down Maps cites this paper.

NavOne: One-Step Global Planning for Vision-Language Navigation on Top-Down Maps Geometric Context Transformer for Streaming 3D Reconstruction

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-06-30T23:25:07.737429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T23:22:52.586386Z digest=sha256:7d31172a73827a8defdea450f9e39df5f3cdfae157c2173f3738786cf6746d33

Observation 113334b0-d1e3-405c-ac8e-e97eac779621 · inbound

SEDualVLN: A Spatially-Enhanced Dual-System for Vision-Language Navigation cites this paper.

SEDualVLN: A Spatially-Enhanced Dual-System for Vision-Language Navigation Geometric Context Transformer for Streaming 3D Reconstruction

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-05-20T13:33:19.178681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-20T13:31:16.012419Z digest=sha256:a7bc90b499aa4e9c344a899150f0311d1f42fd56bfbb39173aa19172937286fa

Observation baf80952-7111-4f94-9b2c-568da5c2dfce · inbound

SEDualVLN: A Spatially-Enhanced Dual-System for Vision-Language Navigation cites this paper.

SEDualVLN: A Spatially-Enhanced Dual-System for Vision-Language Navigation Geometric Context Transformer for Streaming 3D Reconstruction

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-06-30T19:45:00.681245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T19:42:41.238072Z digest=sha256:c2e24eb2955459c856fce8c9c485ccfb81d7e5aa5beeb460882c5ff07bd01ecf

Observation 7e1b4d02-0052-404a-b899-f29cea6e22c7 · inbound

LongDPM: Overlap-Aware 4D Reconstruction from Long Monocular Videos cites this paper.

LongDPM: Overlap-Aware 4D Reconstruction from Long Monocular Videos Geometric Context Transformer for Streaming 3D Reconstruction

Reference 3

Resolution
metadata mismatch
local_arxiv, observed 2026-05-20T14:38:21.656244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-20T14:34:54.528395Z digest=sha256:ee6e9edb90eb069da70693814f942afd0e546cd0c10ac16ffc5d5fd7c50f1823

Observation f18cd508-a9f0-449d-8deb-5c7d1652b163 · inbound

Mamba-VGGT: Persistent Long-Sequence Video Geometry Grounded Transformer via External Sliding Window Mamba Memory cites this paper.

Mamba-VGGT: Persistent Long-Sequence Video Geometry Grounded Transformer via External Sliding Window Mamba Memory Geometric Context Transformer for Streaming 3D Reconstruction

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-05-20T14:23:21.726832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-20T14:19:03.850707Z digest=sha256:cb71b106c78d3c3a532d4fe0cf6dc524d06bcecf627b7f3ccbc46ecd551c04b9

Observation 9a017dcb-d91b-4318-ad12-93d38d1f2ab9 · inbound

HorizonStream: Long-Horizon Attention for Streaming 3D Reconstruction cites this paper.

HorizonStream: Long-Horizon Attention for Streaming 3D Reconstruction Geometric Context Transformer for Streaming 3D Reconstruction

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-05-25T04:35:20.857053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-25T04:33:48.127707Z digest=sha256:7535eab88ce9fbc14ffd69f21ab46598b5ba43251a746fbe51d76f87920b55b2

Observation ebdd55c1-679b-405a-97c7-fb6340aeec87 · inbound

Good Token Hunting: A Hitchhiker's Guide to Token Selection for Visual Geometry Transformers cites this paper.

Good Token Hunting: A Hitchhiker's Guide to Token Selection for Visual Geometry Transformers Geometric Context Transformer for Streaming 3D Reconstruction

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-05-25T04:30:20.363482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-25T04:27:19.553379Z digest=sha256:9157e715de43759d12739035aee1d5bc7930ec682137c64b4315e432f8fed385

Observation e3b18c78-4020-4d98-bf9f-b9f5d27549e3 · inbound

Nano World Models: A Minimalist Implementation of Future Video Prediction cites this paper.

Nano World Models: A Minimalist Implementation of Future Video Prediction Geometric Context Transformer for Streaming 3D Reconstruction

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-06-30T18:55:00.272355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T18:52:29.675477Z digest=sha256:92de1e453a38e0039368008f41d6b968f2afc171654f9e5170ac389212faea9f

Observation 4661bad1-3c51-405d-977d-b84ac88d9e8e · inbound

$R^3$: 3D Reconstruction via Relative Regression cites this paper.

$R^3$: 3D Reconstruction via Relative Regression Geometric Context Transformer for Streaming 3D Reconstruction

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-06-29T18:03:48.159649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-29T17:56:05.704750Z digest=sha256:d61719098a4c968f780d777e4bffe3094385d0228a59b1dfc2851907c5f2909d

Observation 53242d8b-565d-43ec-a959-bb2b823d0c1d · inbound

SpatialBench: Is Your Spatial Foundation Model an All-Round Player? cites this paper.

SpatialBench: Is Your Spatial Foundation Model an All-Round Player? Geometric Context Transformer for Streaming 3D Reconstruction

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-06-29T17:53:47.256777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-29T17:49:58.532910Z digest=sha256:36d95de4dfe03e866d244139eb164417c2137bdd9ca9f4830c63afe3a4900ab4

Observation 5f82d311-9c4d-4447-8406-7b8dc860deba · inbound

MemoryWAM: Efficient World Action Modeling with Persistent Memory cites this paper.

MemoryWAM: Efficient World Action Modeling with Persistent Memory Geometric Context Transformer for Streaming 3D Reconstruction

Reference 50

Resolution
verified exact
local_arxiv, observed 2026-07-04T04:29:34.838395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-26T17:00:43.138441Z digest=sha256:5815be33cb999be2662b2e4d424ce778002797fb8ce145d22d69f862020b983b

Observation fbd891ea-b21f-4608-b512-80f09aab4082 · inbound

The Turning Point of 3D Plant Phenotyping: 3D Foundation Models Enable Minute-to-Second Cross-Crop Reconstruction and Beyond cites this paper.

The Turning Point of 3D Plant Phenotyping: 3D Foundation Models Enable Minute-to-Second Cross-Crop Reconstruction and Beyond Geometric Context Transformer for Streaming 3D Reconstruction

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-07-03T16:38:39.694613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-07-03T16:34:17.398245Z digest=sha256:ffa20d62253a945febcb56e7c5e04777f900e41d0c93e5e851e49052722cf392

Observation 5b35a6ff-1232-4ed8-a8f9-277634bac96b · inbound

TRIG: Trajectory-Rig Decoupled Metric Geometry Learning cites this paper.

TRIG: Trajectory-Rig Decoupled Metric Geometry Learning Geometric Context Transformer for Streaming 3D Reconstruction

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-07-08T23:55:43.005778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-08T23:54:53.463703Z digest=sha256:3086fac3f52243d2aa5b41752d2ebf47eab68ba05fae3f95fc2350552111dd47

Observation 4e816f02-9bad-47c1-bb5d-4601d6502214 · inbound

Glob3R: Global Structure-from-Motion with 3D Foundation Models cites this paper.

Glob3R: Global Structure-from-Motion with 3D Foundation Models Geometric Context Transformer for Streaming 3D Reconstruction

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-13T01:21:15.228734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T01:21:15.228734Z digest=sha256:e135f1a8f425f29f9fda52da59c449c569c629354b5dbffd7f3dea117a63230c

Observation af5e3c28-3f2e-4c9d-9cb6-426ab42c9670 · inbound

Context by Distinct Information: An Auditable Dirichlet-Process Working Memory for Long, Redundant Context Streams cites this paper.

Context by Distinct Information: An Auditable Dirichlet-Process Working Memory for Long, Redundant Context Streams Geometric Context Transformer for Streaming 3D Reconstruction

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-14T11:45:09.424370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T11:45:09.424370Z digest=sha256:cbafe368659c64748acc6632afcae0799ebaed6e46c6153e7a9134f694ef3f43

Observation ead7adfa-cd99-40d0-af7d-67c21abb7d98 · inbound

IGGT4D: Streaming 4D Instance-Grounded Geometry Transformer cites this paper.

IGGT4D: Streaming 4D Instance-Grounded Geometry Transformer Geometric Context Transformer for Streaming 3D Reconstruction

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-01T13:07:07.298312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:07:07.298312Z digest=sha256:ff00d12e2bdc9769981ccc1d50d14bf65b9d2d6ea873c4150664163fc5d7de76

Observation 77378d27-065b-465f-a0d0-1a196f949afb · inbound

VidMap: Exploiting Temporal Structure for Video-Based Structure-from-Motion cites this paper.

VidMap: Exploiting Temporal Structure for Video-Based Structure-from-Motion Geometric Context Transformer for Streaming 3D Reconstruction

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-30T13:50:35.806186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T13:50:35.806186Z digest=sha256:9543613d208ef95362b46bc5cc304e2c529ad22d975f4c3970229ba8858bb176