Pith. sign in

Paper Citation Record · LEDGER

ViewMind3D: Modular View-Aware Inference for Training-Free 3D-QA

As of 20 August 2026, this Paper Citation Record lists 29 of 29 outbound references and 0 inbound Pith citation observations for arXiv:2607.28442.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.28442 v1

Coverage vector

measured 29 of 29 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-31T07:16:36.490043Z

measured 29 of 29 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

29 of 29 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved29
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation fdc5db96-c1a8-45f1-a0b1-3f8df6e9c6fe · outbound

This paper cites 3D-LLM: Injecting the 3D World into Large Language Models,.

ViewMind3D: Modular View-Aware Inference for Training-Free 3D-QA 3D-LLM: Injecting the 3D World into Large Language Models,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-31T07:16:36.414135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T07:16:36.414135Z digest=sha256:cbcee5ce1bf1be0ce0dbcf732441f23c83fa5d91b8fab66774586944b8f51a8e

Observation 88b0b9db-5622-4b0d-b4ed-7fcacb0f6319 · outbound

This paper cites PointLLM: Empowering Large Language Models to Understand Point Clouds,.

ViewMind3D: Modular View-Aware Inference for Training-Free 3D-QA PointLLM: Empowering Large Language Models to Understand Point Clouds,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-31T07:16:36.417640Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T07:16:36.417640Z digest=sha256:133f96acde1871c2d84c0b20d3cda9cee6b0955f1a5ef67abd163f302aa8cc65

Observation d76ad42b-cb8d-4a04-9737-9d67cd18dee9 · outbound

This paper cites ShapeLLM: Universal 3D Object Understanding for Embodied Interaction,.

ViewMind3D: Modular View-Aware Inference for Training-Free 3D-QA ShapeLLM: Universal 3D Object Understanding for Embodied Interaction,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-31T07:16:36.420724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T07:16:36.420724Z digest=sha256:83c4da18f635097753fa4e34c147bf4ef68e47a709afc6b9f6f2317b6f59e524

Observation 0b92832a-cea1-49e8-b611-9ce276a2513a · outbound

This paper cites GPT-4V(ision) System Card,.

ViewMind3D: Modular View-Aware Inference for Training-Free 3D-QA GPT-4V(ision) System Card,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-31T07:16:36.423857Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T07:16:36.423857Z digest=sha256:38cba63a7b624fd6534590f0afc7e520cec3d48e6906712a56c4439dce974e5b

Observation a34b8c7f-814d-438c-bc80-9b34296a8e53 · outbound

This paper cites ScanQA: 3D Question Answering for Spatial Scene Understanding,.

ViewMind3D: Modular View-Aware Inference for Training-Free 3D-QA ScanQA: 3D Question Answering for Spatial Scene Understanding,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-31T07:16:36.426702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T07:16:36.426702Z digest=sha256:7ad7542c6ebf022ab7ce5e845b5e57ef22b2262e5701ab8afa7740602fd0bb5f

Observation 8dc45514-3dd8-407d-b4c5-880828bc34ac · outbound

This paper cites SQA3D: Situated Question Answering in 3D Scenes,.

ViewMind3D: Modular View-Aware Inference for Training-Free 3D-QA SQA3D: Situated Question Answering in 3D Scenes,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-31T07:16:36.429923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T07:16:36.429923Z digest=sha256:0b3267205966c9acf99239ee2a5da653bad436740171d2059981b9e5c6970571

Observation 95c20697-b62d-45ae-b9b0-ff02f06b0e36 · outbound

This paper cites Grounded 3D-LLM with Referent Tokens.

ViewMind3D: Modular View-Aware Inference for Training-Free 3D-QA Grounded 3D-LLM with Referent Tokens

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-31T07:16:36.432881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T07:16:36.432881Z digest=sha256:ced1ebe9c51ca576b4c180a88b6ce71e86c26b3252ec8b31209d8b31792de46e

Observation bc13bb8c-8937-452c-a868-977507b00562 · outbound

This paper cites Scene-LLM: Extending Language Model for 3D Visual Reasoning,.

ViewMind3D: Modular View-Aware Inference for Training-Free 3D-QA Scene-LLM: Extending Language Model for 3D Visual Reasoning,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-31T07:16:36.435995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T07:16:36.435995Z digest=sha256:0589ea9d551bad6bdee8e8417f4159c0cba124495e6084b0c0ddfc168c8551c2

Observation 4f72c9d9-a190-48ce-8cee-9d7193ae3242 · outbound

This paper cites GPT4Scene: Understand 3D Scenes from Videos with Vision-Language Models.

ViewMind3D: Modular View-Aware Inference for Training-Free 3D-QA GPT4Scene: Understand 3D Scenes from Videos with Vision-Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-31T07:16:36.438659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T07:16:36.438659Z digest=sha256:f7b69458f8c523b808a9cc04f3481649db998a8fc25c637b1bb39978b51cda30

Observation 3a766ae1-7e20-475f-848b-2cc52af3df51 · outbound

This paper cites Chat-scene: Bridging 3d scene and large language models with object identifiers,.

ViewMind3D: Modular View-Aware Inference for Training-Free 3D-QA Chat-scene: Bridging 3d scene and large language models with object identifiers,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-31T07:16:36.441509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T07:16:36.441509Z digest=sha256:abc2864cb573604b9d782afd51ff18337db2721403ccf7a6ff77e2ca75c5a251

Observation 0cf9400b-0a9b-4d0e-b5c3-a7af00965c2e · outbound

This paper cites Chat-Scene++: Exploiting Context-Rich Object Identification for 3D LLM.

ViewMind3D: Modular View-Aware Inference for Training-Free 3D-QA Chat-Scene++: Exploiting Context-Rich Object Identification for 3D LLM

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-31T07:16:36.444118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T07:16:36.444118Z digest=sha256:45cb76e9f063271c8860c28098db10d444d5633793c313a7e99be4bfbb9e3ec1

Observation 38b309dd-18c5-4881-b9d1-cafb326d3d02 · outbound

This paper cites LLM-Grounder: Open-V ocabulary 3D Visual Grounding With Large Language Model as an Agent,.

ViewMind3D: Modular View-Aware Inference for Training-Free 3D-QA LLM-Grounder: Open-V ocabulary 3D Visual Grounding With Large Language Model as an Agent,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-31T07:16:36.446927Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T07:16:36.446927Z digest=sha256:3545bb87df49cc99793521f29fedd10f483ddfb8e4eb077716b0f339ac09323c

Observation babff4ad-e4d6-4c9c-92f3-8f44be787153 · outbound

This paper cites [Online].

ViewMind3D: Modular View-Aware Inference for Training-Free 3D-QA [Online]

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-31T07:16:36.449670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T07:16:36.449670Z digest=sha256:e4a554f630fdd3118f84ebb25bc31f40e21aeb450c5949f181abd5f1757dba9f

Observation ade514d7-ae26-4872-95a7-ceaa89fe6acf · outbound

This paper cites Learning Transferable Visual Models From Natural Lan- guage Supervision,.

ViewMind3D: Modular View-Aware Inference for Training-Free 3D-QA Learning Transferable Visual Models From Natural Lan- guage Supervision,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-31T07:16:36.452327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T07:16:36.452327Z digest=sha256:1c9d7db7300f205ccff1285f4d01b6c1a53748e36c562725f8f1d1cf057541f4

Observation 3c2cd300-f63f-4f31-baee-cc1c39dd4e21 · outbound

This paper cites BLIP: Bootstrapping Language-Image Pre-training for Unified Vision-Language Under- standing and Generation,.

ViewMind3D: Modular View-Aware Inference for Training-Free 3D-QA BLIP: Bootstrapping Language-Image Pre-training for Unified Vision-Language Under- standing and Generation,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-31T07:16:36.454768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T07:16:36.454768Z digest=sha256:c7190624af8b8229efc60db3751292e53bfafccf73abb78ebdca841b9e2b2cd5

Observation c53d5e7b-a486-4cc3-ae00-1a801027ffb8 · outbound

This paper cites BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models,.

ViewMind3D: Modular View-Aware Inference for Training-Free 3D-QA BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-31T07:16:36.457117Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T07:16:36.457117Z digest=sha256:14ae8ebac03ea6913228a3a23f37e5167d5256bda5f120b6299fa3cece9ccc84

Observation 255e5c43-0bd4-4ed1-bbfa-a136fa7619f7 · outbound

This paper cites Visual Instruction Tuning,.

ViewMind3D: Modular View-Aware Inference for Training-Free 3D-QA Visual Instruction Tuning,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-31T07:16:36.459557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T07:16:36.459557Z digest=sha256:0d7f7f068a0e71d2f29d2a57984b3f8e5f083594f07437d1ed91cb7e41de3bc8

Observation 5e535469-3726-49c3-8103-fca79e63c0a5 · outbound

This paper cites LL3DA: Visual Interactive Instruction Tuning for Omni-3D Understanding Reasoning and Planning,.

ViewMind3D: Modular View-Aware Inference for Training-Free 3D-QA LL3DA: Visual Interactive Instruction Tuning for Omni-3D Understanding Reasoning and Planning,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-31T07:16:36.461931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T07:16:36.461931Z digest=sha256:fb51005c84cd81a7e51954e5f9dfad7e12c63193e4630ee2ca190f4649731bd7

Observation 221a8b00-5876-485e-9aa3-6c0838fe03ee · outbound

This paper cites DSPNet: Dual-vision Scene Perception for Robust 3D Question Answering,.

ViewMind3D: Modular View-Aware Inference for Training-Free 3D-QA DSPNet: Dual-vision Scene Perception for Robust 3D Question Answering,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-31T07:16:36.464739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T07:16:36.464739Z digest=sha256:4ba234576f86b24b00f3365e0bb59bbef7cc2c97bfe87171db78cbb7f4eafe3c

Observation dec6cf22-360c-472c-8e41-c8578243300b · outbound

This paper cites Advancing 3D Scene Understanding with MV-ScanQA Multi-View Reasoning Evaluation and TripAlign Pre-training Dataset,.

ViewMind3D: Modular View-Aware Inference for Training-Free 3D-QA Advancing 3D Scene Understanding with MV-ScanQA Multi-View Reasoning Evaluation and TripAlign Pre-training Dataset,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-07-31T07:16:36.468255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T07:16:36.468255Z digest=sha256:92d89270eb8483d970d8e5afb828a8820fcf46f9b52d276e6c003059d5be8561

Observation dfb91dd8-5000-4cbe-a24a-7a5b517c4775 · outbound

This paper cites ScanNet: Richly-Annotated 3D Reconstructions of In- door Scenes,.

ViewMind3D: Modular View-Aware Inference for Training-Free 3D-QA ScanNet: Richly-Annotated 3D Reconstructions of In- door Scenes,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-31T07:16:36.470592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T07:16:36.470592Z digest=sha256:27d337d84443c810c32a4d4140f8ec139773f3ba77b076d3233fd9b23bda9742

Observation 78d9931c-e1c2-4acc-96c0-48c7008eabc8 · outbound

This paper cites Florence-2: Advancing a Unified Representation for a Variety of Vision Tasks,.

ViewMind3D: Modular View-Aware Inference for Training-Free 3D-QA Florence-2: Advancing a Unified Representation for a Variety of Vision Tasks,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-31T07:16:36.473184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T07:16:36.473184Z digest=sha256:0778764a974fc8870f408d0e33c5be635454038aab440e2e454af00dcfd6a3b8

Observation defe407c-635a-4942-8cae-3a1fb3e8875b · outbound

This paper cites ScanRefer: 3D Object Localization in RGB-D Scans Using Natural Language,.

ViewMind3D: Modular View-Aware Inference for Training-Free 3D-QA ScanRefer: 3D Object Localization in RGB-D Scans Using Natural Language,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-31T07:16:36.475542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T07:16:36.475542Z digest=sha256:a1fae3f7b9ca09e4807f0596ebda602c1f1c9f2383dae91bdbf92c5d0b48c7b0

Observation 0f25679b-e567-4310-ae6f-4b4280ba7ec6 · outbound

This paper cites Flamingo: A Visual Language Model for Few-Shot Learning,.

ViewMind3D: Modular View-Aware Inference for Training-Free 3D-QA Flamingo: A Visual Language Model for Few-Shot Learning,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-31T07:16:36.478063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T07:16:36.478063Z digest=sha256:5b380c59858bca526b27a064cfd2c87e83831e917489e47cac0911735753b641

Observation d213aa31-bc98-4952-818c-e8e102eaa6aa · outbound

This paper cites BLEU: A Method for Automatic Evaluation of Machine Translation,.

ViewMind3D: Modular View-Aware Inference for Training-Free 3D-QA BLEU: A Method for Automatic Evaluation of Machine Translation,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-31T07:16:36.480349Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T07:16:36.480349Z digest=sha256:69dd84c2840b27df6bfbcc499df88b3c583928672fcd4044f2d0f26ad729b195

Observation 9b4dcfa2-a3ea-42be-8db1-db7ff16d72e7 · outbound

This paper cites METEOR: An Automatic Metric for MT Evaluation with Improved Correlation with Human Judgments,.

ViewMind3D: Modular View-Aware Inference for Training-Free 3D-QA METEOR: An Automatic Metric for MT Evaluation with Improved Correlation with Human Judgments,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-07-31T07:16:36.482777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T07:16:36.482777Z digest=sha256:575a035ace7e6fd93675d6e57fee5f1f14f26a5b40983510ef9353707e27e69d

Observation 753d3161-6b7d-4990-9474-17283febe802 · outbound

This paper cites ROUGE: A Package for Automatic Evaluation of Sum- maries,.

ViewMind3D: Modular View-Aware Inference for Training-Free 3D-QA ROUGE: A Package for Automatic Evaluation of Sum- maries,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-07-31T07:16:36.485178Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T07:16:36.485178Z digest=sha256:90160e5bd314207a1b3e8190093a23a3d6ac2bc8a52da455f59f9f5d554ca257

Observation e4d8f0f2-2636-42e1-97d0-d2da1b6dabdf · outbound

This paper cites CIDEr: Consensus-Based Image Description Evaluation,.

ViewMind3D: Modular View-Aware Inference for Training-Free 3D-QA CIDEr: Consensus-Based Image Description Evaluation,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-31T07:16:36.487662Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T07:16:36.487662Z digest=sha256:45ccab0ee834c3da7c6c6a78252c6490e4933725a9d754a481c2bcc595a3d52e

Observation 9a9363fe-f24f-4094-bf72-bc2ba591f98f · outbound

This paper cites BERTScore: Evaluating Text Generation with BERT.

ViewMind3D: Modular View-Aware Inference for Training-Free 3D-QA BERTScore: Evaluating Text Generation with BERT

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-31T07:16:36.490043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T07:16:36.490043Z digest=sha256:d28755ad0b2e638874d8d93db842d46fc66477b6e495f87418e7ac8d5cd49b1e

Pith citing papers

No inbound Pith citation observations are available.