Pith. sign in

Paper Citation Record · LEDGER

Ross3D: Reconstructive Visual Instruction Tuning with 3D-Awareness

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2504.01901.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.01901 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:46:16.379978Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T16:39:57.809937Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 8bfcbaab-93bb-47aa-a2c7-679f23797e2e · inbound

VLM-3R: Vision-Language Models Augmented with Instruction-Aligned 3D Reconstruction cites this paper.

VLM-3R: Vision-Language Models Augmented with Instruction-Aligned 3D Reconstruction Ross3D: Reconstructive Visual Instruction Tuning with 3D-Awareness

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-05-19T12:57:17.982684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T12:54:01.013242Z digest=sha256:f31b1a62ef6cec74fe386a6eace9b73ed50e5511cb0b680fda1ed39c8409d0c6

Observation 8504db6e-dd20-4c52-a550-76c5b5c5afae · inbound

Uni3D-MoE: Scalable Multimodal 3D Scene Understanding via Mixture of Experts cites this paper.

Uni3D-MoE: Scalable Multimodal 3D Scene Understanding via Mixture of Experts Ross3D: Reconstructive Visual Instruction Tuning with 3D-Awareness

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T13:46:16.379978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:46:16.379978Z digest=sha256:9c1f21d45236f85636341ec25623657be7bcd4f691d20fce1cdbfd61d8f4ce8e

Observation 2d112118-348a-46f6-8d42-b5cb8416732c · inbound

VGR: Visual Grounded Reasoning cites this paper.

VGR: Visual Grounded Reasoning Ross3D: Reconstructive Visual Instruction Tuning with 3D-Awareness

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-19T09:12:14.381105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T09:11:00.295700Z digest=sha256:4abf5eb78d66ff1972fe32f77700872b3524e08bc891db92f1c81ec26fabc5bd

Observation bba36ad4-258d-42bf-8e49-8f14d29010fb · inbound

SPAZER: Spatial-Semantic Progressive Reasoning Agent for Zero-shot 3D Visual Grounding cites this paper.

SPAZER: Spatial-Semantic Progressive Reasoning Agent for Zero-shot 3D Visual Grounding Ross3D: Reconstructive Visual Instruction Tuning with 3D-Awareness

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T22:21:10.825875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:21:10.825875Z digest=sha256:9d457e701cd3c4067712706c72386a9a918c9ce104c2a72b67443627d72cef94

Observation 423e956a-1019-49d8-bea8-56dd9f10e73a · inbound

Reconstruction Alignment Improves Unified Multimodal Models cites this paper.

Reconstruction Alignment Improves Unified Multimodal Models Ross3D: Reconstructive Visual Instruction Tuning with 3D-Awareness

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-04T22:36:08.186645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T22:36:08.186645Z digest=sha256:064bf5e5b9b2afead3ccfc0bc3617488e951d83701b33a448615e63f34abdeee

Observation 11043a64-5a12-42f2-a59c-3c7d03c05e93 · inbound

POMA-3D: The Point Map Way to 3D Scene Understanding cites this paper.

POMA-3D: The Point Map Way to 3D Scene Understanding Ross3D: Reconstructive Visual Instruction Tuning with 3D-Awareness

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-17T20:30:11.566090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T20:27:27.347592Z digest=sha256:2ea144387bb92d9eddc56c87599d30b53dba4c221b59ccd4427838ad34117553

Observation f695d0f4-77d6-47e9-bce7-c4279037e67c · inbound

Vision-Language Memory for Spatial Reasoning cites this paper.

Vision-Language Memory for Spatial Reasoning Ross3D: Reconstructive Visual Instruction Tuning with 3D-Awareness

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-03T20:15:34.849814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T20:15:34.849814Z digest=sha256:85b25a1f6cbcb8f2250a53e575284716a1afca89ddc71d0bc369cc20c24578df

Observation 6fefb318-0208-43d5-bfdb-06927cceb23c · inbound

Thinking with Geometry: Active Geometry Integration for Spatial Reasoning cites this paper.

Thinking with Geometry: Active Geometry Integration for Spatial Reasoning Ross3D: Reconstructive Visual Instruction Tuning with 3D-Awareness

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-16T06:40:42.220268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T06:39:29.010937Z digest=sha256:5cd1fd49d207e5034159331d19c815d6aed33bbca82bc7e0874b59432e4020a1

Observation c66d1156-c2e6-4c16-acab-ca984e3d0aec · inbound

MAG-3D: Multi-Agent Grounded Reasoning for 3D Understanding cites this paper.

MAG-3D: Multi-Agent Grounded Reasoning for 3D Understanding Ross3D: Reconstructive Visual Instruction Tuning with 3D-Awareness

Reference 41

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T06:51:18.870345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T17:25:31.097385Z digest=sha256:acf31b8b4a8a2968fba54e4ed726cec3a597a2f95c56759a6aa923552ea3417f

Observation 69f14586-43f4-4ac8-8045-a4d963e96b63 · inbound

Thinking with Imagination: Agentic Visual Spatial Reasoning with World Simulators cites this paper.

Thinking with Imagination: Agentic Visual Spatial Reasoning with World Simulators Ross3D: Reconstructive Visual Instruction Tuning with 3D-Awareness

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-02T12:06:56.220175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T02:25:30.998989Z digest=sha256:05950c0a6a2a135bdf097d836ec6ba0c4a6ce27f922eab78c1e110bf670e46a8

Observation f71a0c4a-b0ed-4f15-a422-d0e4fed69ce3 · inbound

Stream3D-VLM: Online 3D Spatial Understanding with Incremental Geometry Priors cites this paper.

Stream3D-VLM: Online 3D Spatial Understanding with Incremental Geometry Priors Ross3D: Reconstructive Visual Instruction Tuning with 3D-Awareness

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T16:17:09.634782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T22:49:02.846428Z digest=sha256:8e9d5fc04670ec8c7acb7a116ae61fe0ff1dac374befc2356d028042710b5f74

Observation 3584a925-bd43-427d-9c83-ae92ad24c96a · inbound

SpatialSV: Internalizing Interpretable 3D Spatial Awareness in MLLMs via Task-Oriented Visual Supervision cites this paper.

SpatialSV: Internalizing Interpretable 3D Spatial Awareness in MLLMs via Task-Oriented Visual Supervision Ross3D: Reconstructive Visual Instruction Tuning with 3D-Awareness

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-07-04T03:29:30.841479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T17:59:21.032539Z digest=sha256:4b873074bd96707158414698bd252cb96db7cfa9849561cc98e5fb5e08836f20

Observation af7e5b41-52ad-4616-a92b-5c4952271999 · inbound

Agentic Collaborative Cognition for Zero-Shot 3D Understanding cites this paper.

Agentic Collaborative Cognition for Zero-Shot 3D Understanding Ross3D: Reconstructive Visual Instruction Tuning with 3D-Awareness

Reference 42

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T16:39:57.811475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T00:22:06.082183Z digest=sha256:ed457995a59fdf1d527f33c954a16102e9eedc5b66e60565b1d20456b210e948

Observation a95ecf52-e902-4210-be1c-49a24557df22 · inbound

Agentic Collaborative Cognition for Zero-Shot 3D Understanding cites this paper.

Agentic Collaborative Cognition for Zero-Shot 3D Understanding Ross3D: Reconstructive Visual Instruction Tuning with 3D-Awareness

Reference 42

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T12:59:52.660709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T05:37:41.407624Z digest=sha256:ae256c1b1c5c001fe911a2ac12a9089dd3ac175d19c5c56e6ef73bfab223a125