Pith. sign in

Paper Citation Record · LEDGER

Holo-Captioning: Toward the Text Equivalent of 3D Scenes

As of 23 August 2026, this Paper Citation Record lists 91 of 91 outbound references and 0 inbound Pith citation observations for arXiv:2607.02908.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.02908 v1

Coverage vector

measured 91 of 91 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-12T06:12:48.722467Z

measured 91 of 91 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

91 of 91 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved91
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 829da139-5baf-49bd-b2b1-69ea8e432f23 · outbound

This paper cites In: European Conference on Computer Vision (2020) 4, 5, 6, 10.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: European Conference on Computer Vision (2020) 4, 5, 6, 10

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:f2c49f65d5a197174225f79f028138d248c6fcddd3ac0c8a0d37a391bd68f306

Observation d50d6c80-85c6-4aba-b2d6-b517086105a8 · outbound

This paper cites In: IEEE/CVF International Conference on Computer Vision (2019) 4.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: IEEE/CVF International Conference on Computer Vision (2019) 4

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:1d58f8c5fdcf647ddc3f2e8bd5ec829ae355e31458156f688ead12174478cc35

Observation c4ffb155-c23b-4907-b9fa-fbb9f1e6f21b · outbound

This paper cites In: European Conference on Computer Vision (2024) 1, 2, 5, 6, 10.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: European Conference on Computer Vision (2024) 1, 2, 5, 6, 10

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:22e542ef436ab89aee04a4893d3acc3f28a4a83441cb5e39f13cbdb8834d16bb

Observation 2014ffd6-ff72-4ed6-a07f-3ee4a53d2a24 · outbound

This paper cites In: IEEE/CVF Conference on Computer Vision and Pattern Recognition (2022) 5.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: IEEE/CVF Conference on Computer Vision and Pattern Recognition (2022) 5

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:8081aeb433cc74a9c53ecae887261ff9b975066cb080680a162c494203f5bae0

Observation 42c802a7-0006-4cab-bdd2-26a847a5ea1b · outbound

This paper cites arXiv preprint arXiv:2510.10903 (2025) 1, 5.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes arXiv preprint arXiv:2510.10903 (2025) 1, 5

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:edbf051b61870b2c6da53b8f453f9ebd8eeb7fd5085da2c28e0d1a5383ad899d

Observation 7996503c-3d77-4a51-af29-0adff0710d9b · outbound

This paper cites In: Proceedings of the Workshop on Intrinsic and Extrinsic Evaluation Measures for Machine Translation and/or Summarization@ACL 2005 (2005) 6.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: Proceedings of the Workshop on Intrinsic and Extrinsic Evaluation Measures for Machine Translation and/or Summarization@ACL 2005 (2005) 6

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:c82d4e39d638d7b1b210c634a3bd9d3d62583bca466c52365705520e74ea0f59

Observation 0e27fb37-0e02-4e41-a145-63af416aafd2 · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:bbbfa11632888a94d1891c2d1034a33c471fd6f79981468674fa9bede50f1cef

Observation b0e46359-f5b8-407e-af2a-bd3e03632bca · outbound

This paper cites In: IEEE/CVF Conference on Computer Vision and Pattern Recognition (2023) 5.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: IEEE/CVF Conference on Computer Vision and Pattern Recognition (2023) 5

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:d84f8a49b4c9dde730b6ab350f7099cf9e3adc6dde0f5108bfcd96c6a2bb40ad

Observation fdf0691e-41d1-46d7-acf2-d62d52d71eba · outbound

This paper cites In: IEEE/CVF Conference on Computer Vision and Pattern Recognition.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: IEEE/CVF Conference on Computer Vision and Pattern Recognition

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:29f0bf9bd79f1cadb72f22b800756485ffc6da9c5c5165361109dde3f86f5ff2

Observation 61e9e3cf-2d70-40e4-83fa-3fa33572f338 · outbound

This paper cites Exploiting Scene-specific Features for Object Goal Navigation.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes Exploiting Scene-specific Features for Object Goal Navigation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:1c75c7e6b7cb6c659e822da82bdf518cb3d4efc12969a3b5e515d238610aa877

Observation e6f208cb-ed03-4605-8657-208cc7545df4 · outbound

This paper cites In: International Conference on 3D Vision (2017) 9.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: International Conference on 3D Vision (2017) 9

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:b6a03481e27a6ec1aceff3e2892fd4abb64fe9e76e09d21687b3d8af765156a0

Observation 31eb07b1-21ce-4df1-aa4a-3ec52890eeee · outbound

This paper cites In: Conference on Robot Learning (2023) 6.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: Conference on Robot Learning (2023) 6

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:cbdadc2e47e276c649d79b3952b5717159bc4f17fd7f1ebda04af65cb9fec555

Observation fad97fa5-aea7-493f-bb42-e2b76ea9725c · outbound

This paper cites In: European Conference on Computer Vision (2020) 4, 5, 6, 10, 13.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: European Conference on Computer Vision (2020) 4, 5, 6, 10, 13

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:522391858716b2730bf5dba1a5a7daa82f1a3ac54a0a2a2312dcd9f15b37c1b2

Observation e6bb58fa-fec9-4ee6-b6e3-87f7e9b77f93 · outbound

This paper cites In: IEEE Conference on Computer Vision and Pattern Recognition (2021) 1, 2, 4, 5, 6, 16 26 K.-Y.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: IEEE Conference on Computer Vision and Pattern Recognition (2021) 1, 2, 4, 5, 6, 16 26 K.-Y

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:8d213d6b421ce2f8b3488a967302da4aafdda5470f2022d2a44b434b9e0d69f4

Observation be56e23e-0cbd-4673-a3e7-009cf26ef299 · outbound

This paper cites In: European Confer- ence on Computer Vision.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: European Confer- ence on Computer Vision

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:6d3dca04ada38178640ebc87b9953e8c301cec9dfa61743644220e9fd08cd6a0

Observation 66d60d5f-f247-49de-9ef8-01c55ae13a84 · outbound

This paper cites In: IEEE/CVF Conference on Computer Vision and Pattern Recognition (2024) 3, 5, 12, 13, 20, 23.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: IEEE/CVF Conference on Computer Vision and Pattern Recognition (2024) 3, 5, 12, 13, 20, 23

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:3f5bad7ea3517890f9e23f14653ba6046dd75261244c6437621ddd5f0650708c

Observation f99e3f39-9949-4dd4-b578-715bef9cf573 · outbound

This paper cites In: IEEE/CVF Conference on Computer Vision and Pattern Recognition (2023) 1, 2, 4, 5, 6, 12, 13, 20.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: IEEE/CVF Conference on Computer Vision and Pattern Recognition (2023) 1, 2, 4, 5, 6, 12, 13, 20

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:4509d391c093dc508866207579fe865510b2260437f124f1b6062568eb6bf97c

Observation 51fda539-d9d0-4a4a-931a-83530cdc14a4 · outbound

This paper cites IEEE Transactions on Pattern Analysis and Machine Intelligence 46(2024) 4, 6, 12, 13.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes IEEE Transactions on Pattern Analysis and Machine Intelligence 46(2024) 4, 6, 12, 13

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:4d7a6b928d04b743fe26920b273f1f0b8840b798052328cbd3626ba6deef28a7

Observation b6f23ffa-4aff-478c-b4e3-39027550aa3b · outbound

This paper cites In: IEEE Conference on Computer Vision and Pattern Recognition (2017) 9, 10.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: IEEE Conference on Computer Vision and Pattern Recognition (2017) 9, 10

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:ea9c3005151987284a9b8a42ac39a305de4e0afee61d6638c76cf687035312d6

Observation e64a9cdf-4faa-433d-85be-cd206e4be924 · outbound

This paper cites In: Neural Information Processing Systems Datasets and Benchmarks Track (2021) 5, 9.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: Neural Information Processing Systems Datasets and Benchmarks Track (2021) 5, 9

Reference 20

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:907963a70402347099443846770a362b5ac7e645c03a3d174425b5a3a94a87f8

Observation 89e1f93b-9097-4459-bdd9-b244652788cc · outbound

This paper cites In: IEEE/CVF Conference on Computer Vision and Pattern Recognition (2025) 5.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: IEEE/CVF Conference on Computer Vision and Pattern Recognition (2025) 5

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:8a9211aa50199f37e959ccc915199face20dafb332a4138d73232940a1ba5a39

Observation f4e26d98-e229-4de5-8b26-f342ab24b51f · outbound

This paper cites Benchmarking and Improving Detail Image Caption.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes Benchmarking and Improving Detail Image Caption

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:f1168137bdee6c4d509b087cc858b02fa1f613eb3d27a2539483b01c5d561df6

Observation 9a764f80-878a-4389-ad58-e120f8206fff · outbound

This paper cites The Llama 3 Herd of Models.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes The Llama 3 Herd of Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:2262fcd535edc2b183607c624733ccbe6f8f7fb4285deb1f3df2b441f65580c8

Observation d649baf1-37c3-433a-8f72-8241392fb09a · outbound

This paper cites arXiv preprint arXiv:2509.14981 (2025) 1.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes arXiv preprint arXiv:2509.14981 (2025) 1

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:89c03a6f14f2ef77e2047d8bccdf9e240d55111a6e78da7c6985d949c8a20128

Observation 7a71bddb-50d6-407d-837a-365cbd884f13 · outbound

This paper cites In: IEEE/CVF Conference on Computer Vision and Pattern Recognition (2024) 4.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: IEEE/CVF Conference on Computer Vision and Pattern Recognition (2024) 4

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:9527266f48907a2ef19b9c6cfa4d59fa1b8281c69d9ad611e755e1e800ca7a39

Observation 95aac788-2628-4dc9-b7b0-69e340e579a7 · outbound

This paper cites In: IEEE International Conference on Robotics and Automation (2024) 6.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: IEEE International Conference on Robotics and Automation (2024) 6

Reference 26

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:b08633dd1fe86f7b19cdf9eed067b59317aea5d9def6d06ea9e1912299890b37

Observation 7ab4ca9f-5742-44e4-b1fb-9d2096214d1a · outbound

This paper cites an unresolved cited work.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:eafe36e45e6e57724afc43507ed6466a57fd0477f48936397b1afb9f7315c2cf

Observation 7d0edb07-7c65-4d8c-b3a8-b0d4bb772179 · outbound

This paper cites In: Neural Information Processing Systems (2023) 5.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: Neural Information Processing Systems (2023) 5

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:c7457bb1f9cc4480e59402815f6f5ea45d106e4c750d4b096ce44167b73fc1e5

Observation 63e88fe8-c75c-495e-86f3-0371da425b27 · outbound

This paper cites In: Neural Information Processing Systems (2024) 5 Holo-Captioning 27.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: Neural Information Processing Systems (2024) 5 Holo-Captioning 27

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:01a65633069f24f1ea6d8f036766fe254a801ff46a525a6182e1de19b1f0e491

Observation a74de81b-07b9-4eeb-b1b2-e349229c8462 · outbound

This paper cites In: International Conference on Machine Learning (2024) 5, 12, 13, 23.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: International Conference on Machine Learning (2024) 5, 12, 13, 23

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:5642cb25d14a39e8e0cd27856a75496aeaf8947e036dc094563285b51775c6bb

Observation e147e5dc-eb09-4b09-b902-16ed734977d4 · outbound

This paper cites SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:03722fcbf887aeed6e5e8d07a4b06a848856fd7f28ab24fe348aabbfa393a2ad

Observation b6594511-bf0c-4e67-86e7-8989054b6a2b · outbound

This paper cites 3D-R1: Enhancing Reasoning in 3D VLMs for Unified Scene Understanding.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes 3D-R1: Enhancing Reasoning in 3D VLMs for Unified Scene Understanding

Reference 32

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:5c6155ed08994b7b6f0fc401c3da5c8a7345fadd95969c9aff6d5edb3a5476f0

Observation 741cf6ad-22f5-47d8-b90f-1cf6f79482d3 · outbound

This paper cites In: IEEE Conference on Computer Vision and Pattern Recognition (2026) 5.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: IEEE Conference on Computer Vision and Pattern Recognition (2026) 5

Reference 33

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:7697a933aded1169de3b57b2bf5e98de9725067101b5de0be8787823b950d083

Observation f774a01f-8616-471f-bcb4-7444972bd9af · outbound

This paper cites In: Neural Information Processing Systems (2025) 5.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: Neural Information Processing Systems (2025) 5

Reference 34

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:df8c8b0733381925e69bf252f50b7585af13b5aa075cbdb4356cc2d595ab0888

Observation 04122914-1e17-445a-9542-cd28c1f364a0 · outbound

This paper cites In: European Conference on Computer Vision (2024) 5, 9.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: European Conference on Computer Vision (2024) 5, 9

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:992202ce196a77995dcf89d44257202c68b4b0dbb6ed52a556cb279ada4b56df

Observation f1172411-616b-45ac-aad7-9289564cb49f · outbound

This paper cites In: IEEE/CVF International Conference on Computer Vision (2025) 5.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: IEEE/CVF International Conference on Computer Vision (2025) 5

Reference 36

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:3c3fc0fcf63749583baf66cd42dd8e35e585d5121b2b81a918e7342a3ffcc47f

Observation 2350d330-780e-469b-ac8c-ac9f28f1db3b · outbound

This paper cites In: European Conference on Computer Vision.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: European Conference on Computer Vision

Reference 37

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:df2a8f640e2be5ee085fb83c89c36094fb972238ad4f6147d3e36f4ae77496fa

Observation ce9edb96-a021-4646-90b9-89b8d0ce388e · outbound

This paper cites Gemma 3 Technical Report.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes Gemma 3 Technical Report

Reference 38

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:9c9269dca0ca5a3255c620eebc39b93cbfb28554dd9d3d60fa65a8477c90185c

Observation 75a5ef87-b3db-433b-8778-094abeb1a62c · outbound

This paper cites In: Findings of the Association for Computational Linguistics (2024) 4.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: Findings of the Association for Computational Linguistics (2024) 4

Reference 39

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:b11842091075f7f94cae531e52a1b8897abc8dc0c83a47507daf2fda319f1407

Observation 867294db-1147-4dcd-9cee-8ae12ec083f2 · outbound

This paper cites an unresolved cited work.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:f4d7c556ea0893a559a89a280d66c981a406b3e8ac409bb470166413750bb05e

Observation f869c0df-c551-4320-aa08-c4be945004eb · outbound

This paper cites arXiv preprint arXiv:2509.16721 (2025) 5.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes arXiv preprint arXiv:2509.16721 (2025) 5

Reference 41

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:e1d400f76104ff8d42ab8d122c64e123743b2709cc0055dc10348f124e6c443b

Observation cf47ff04-a51d-4a7d-a12e-14eac959b408 · outbound

This paper cites In: European Conference on Computer Vision.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: European Conference on Computer Vision

Reference 42

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:8f31091b47f2d6b6aba1110788c32ab26be3c07f4cbd9596d767bd1d2377c486

Observation 1c6f3731-5b7b-41d9-984f-eba33afdfdb4 · outbound

This paper cites Code as Policies: Language Model Programs for Embodied Control.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes Code as Policies: Language Model Programs for Embodied Control

Reference 43

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:d839a3e3cc840176fd3baa3d7583bbbe93bb4a787c24306f7e1f3ca2556c978f

Observation 0ccd50ba-e862-4787-b316-29b1cca859c0 · outbound

This paper cites In: Neural Information Processing Systems (2025) 4.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: Neural Information Processing Systems (2025) 4

Reference 44

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:67cc4bcb0f5fab41804e3a0410518640c9207d17d4e9102351e40c1bdb8290b3

Observation 9d3fb531-e082-454e-8ef8-45440efe40ce · outbound

This paper cites In: Neural Information Processing Systems Datasets and Benchmarks Track (2024) 5.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: Neural Information Processing Systems Datasets and Benchmarks Track (2024) 5

Reference 45

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:6d809915e8f66bcf1277a69f6f0b2d2ca0e8544a81cfe509a4bf7e385430fb0d

Observation e36daff3-104a-4e56-b33c-4625d85c35dc · outbound

This paper cites In: Neural Information Processing Systems (2023) 4.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: Neural Information Processing Systems (2023) 4

Reference 46

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:45c9ea5fc6f1691f144f1af4ff0076d537d8a16ca2e88cdf8463ec6f3595db4d

Observation 71ca3713-c23d-407e-b38e-af546263400a · outbound

This paper cites In: European Conference on Computer Vision (2024) 4 28 K.-Y.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: European Conference on Computer Vision (2024) 4 28 K.-Y

Reference 47

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:477161546dcaf39441f4fca2ee3a44edc9d412360ca81e936aa3fca3b4599489

Observation 28d2ff66-47c0-442a-878c-f6a7397ec894 · outbound

This paper cites In: Oh, A., Naumann, T., Globerson, A., Saenko, K., Hardt, M., Levine, S.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: Oh, A., Naumann, T., Globerson, A., Saenko, K., Hardt, M., Levine, S

Reference 48

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:e2fbf05c66f9afc82e933a46deca8b40789bdfa5ae5f207dc0ccef9b67bd1a06

Observation 45f84eab-563c-4fdb-a63d-5b7757d93e4f · outbound

This paper cites In: Neural Information Processing Systems (2024) 5, 9.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: Neural Information Processing Systems (2024) 5, 9

Reference 49

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:af17a27abb4626dc01b0840b9303111ef55b0f4402ae00c12cd2487708f50fc3

Observation 714bb697-39cc-4409-805d-9cd04f2dda0d · outbound

This paper cites In: International Conference on Learning Representations (2023) 5.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: International Conference on Learning Representations (2023) 5

Reference 50

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:2fc03f94dbfa9a17729f5dfc95e773032db0c0d25a590011fdba45502a2036eb

Observation c14def2a-3f46-420f-b322-c726a11c7dd3 · outbound

This paper cites In: Neural Information Processing Systems (2022) 3, 14.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: Neural Information Processing Systems (2022) 3, 14

Reference 51

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:731dbbf86bd649e28c0dded644ec57874c8a9f38457b0561e55aa7d41b50a516

Observation 5dcff752-7ed8-488c-bfc3-b496218f8380 · outbound

This paper cites In: Neural Information Processing Systems (2025) 1, 2, 3, 5, 6, 10, 13, 18, 19, 20, 23.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: Neural Information Processing Systems (2025) 1, 2, 3, 5, 6, 10, 13, 18, 19, 20, 23

Reference 52

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:027f98085cf376402efaf618ff835b2cee1fe8e0726e68b39efbaf3778917b78

Observation a5c3cbcc-1314-45a2-8bdc-1e6df7580e1a · outbound

This paper cites In: IEEE/CVF Conference on Computer Vision and Pattern Recognition (2025) 5.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: IEEE/CVF Conference on Computer Vision and Pattern Recognition (2025) 5

Reference 53

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:4595f90c9501f20bcdc7089d49a1f651593a035a016841c1ffe93342f4a33bed

Observation 2e2ca8eb-c4c5-444f-852c-2bb1c2cf64fc · outbound

This paper cites In: Annual Meeting of the Association for Compu- tational Linguistics (2002) 3, 6, 22.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: Annual Meeting of the Association for Compu- tational Linguistics (2002) 3, 6, 22

Reference 54

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:d75d0fc72df2423ddc54e73b8eb2eed448acafa7581b28c0e67842825cd79214

Observation 94eb3c21-e4a1-4a59-9fde-60aeeb1b28de · outbound

This paper cites GPT4Scene: Understand 3D Scenes from Videos with Vision-Language Models.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes GPT4Scene: Understand 3D Scenes from Videos with Vision-Language Models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:340a0fe7ef6871d875475d3d7c0b46f463f0af3f4b11ff6b337c6df02337da0d

Observation f67b2348-bec8-43e6-bf35-5e03b5c80509 · outbound

This paper cites Qwen3-VL Technical Report.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes Qwen3-VL Technical Report

Reference 56

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:4eafdf87c290cce3abb7dd05240e3cf21c7bf77574f29fdb56e6e42c45aab24d

Observation db048688-e5ab-4fad-b3d4-9f8eb300ff5b · outbound

This paper cites In: Conference on Robot Learning (2023) 6.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: Conference on Robot Learning (2023) 6

Reference 57

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:8315a9b624904dc2cfa594c1e8617d2a8f2e0dd0d81e37cb949c06f88d72560c

Observation a913066b-f9fe-456a-a9cd-2cc06c854199 · outbound

This paper cites In: IEEE International Conference on Robotics and Automation (2023) 4.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: IEEE International Conference on Robotics and Automation (2023) 4

Reference 58

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:7ac0397ebe887df61c0f62c5ec3294f9a9b5ab514339d038a90729d88c25c7ac

Observation 98946e16-b693-44aa-bddc-5331367eaed1 · outbound

This paper cites Hunyuan3D 2.1: From Images to High-Fidelity 3D Assets with Production-Ready PBR Material.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes Hunyuan3D 2.1: From Images to High-Fidelity 3D Assets with Production-Ready PBR Material

Reference 59

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:631380aeebfc0a4273cbd084a0133934bc8a1ab6cbb807a2f3eeb38b7928fad6

Observation d72a4de0-a73c-43b9-b2cc-83081bc2919e · outbound

This paper cites Journal of Mech- anisms and Robotics15(2), 020801 (2022) 21.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes Journal of Mech- anisms and Robotics15(2), 020801 (2022) 21

Reference 60

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:1e8746c11522b0d86a1bd84897bc3d8db238800f7e9ada2568f3af1d8308f600

Observation e34bc011-7df3-4b30-9ffe-e313892487c8 · outbound

This paper cites In: IEEE Conference on Computer Vision and Pattern Recognition (2015) 3, 6, 22.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: IEEE Conference on Computer Vision and Pattern Recognition (2015) 3, 6, 22

Reference 61

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:74d2ddcd128cb5f9f8096da38165eb72e7b43cf2fbf727db681095051c58af40

Observation be7f900d-0aa3-4c87-a851-0e2093a0c33e · outbound

This paper cites In: IEEE/CVF Interna- tional Conference on Computer Vision (2019) 9.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: IEEE/CVF Interna- tional Conference on Computer Vision (2019) 9

Reference 62

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:c6a899da1b07aff1e565c7132b314c959ae0934c49ebd4d8483b258f4481a330

Observation e2002e46-0685-467d-b36f-45c098115478 · outbound

This paper cites In: IEEE/CVF Conference on Computer Vision and Pattern Recognition (2020) 4 Holo-Captioning 29.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: IEEE/CVF Conference on Computer Vision and Pattern Recognition (2020) 4 Holo-Captioning 29

Reference 63

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:e7bb8eb7edc5b688e065f571037ab326373186934b9ce2a9cfbffc6dcb2c9e47

Observation 22f52b19-7a67-4be8-b823-4acc4895b8db · outbound

This paper cites In: AAAI (2025) 4.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: AAAI (2025) 4

Reference 64

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:3e6346c98b0f71412821790ec29dae0c2c5daf8d83ed09d945641337ca188917

Observation 2882e1d0-29c5-4af5-8def-5b09e871fc6b · outbound

This paper cites In: IEEE/CVF Conference on Computer Vision and Pattern Recognition (2024) 5, 9.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: IEEE/CVF Conference on Computer Vision and Pattern Recognition (2024) 5, 9

Reference 65

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:942936f497e685e283ea30ff7ad182931837471bb9146a1d46430b7fdaae65cf

Observation 01bb8a30-c8cc-4e9f-b4d4-88022c153ef6 · outbound

This paper cites Chat-3D: Data-efficiently Tuning Large Language Model for Universal Dialogue of 3D Scenes.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes Chat-3D: Data-efficiently Tuning Large Language Model for Universal Dialogue of 3D Scenes

Reference 66

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:02add0582467a204fb07f4b50a3fdc22099a37b350f529718f35a1378f4109fd

Observation 7a25c228-b89a-4da8-99eb-89c20d569d5f · outbound

This paper cites In: Neural Information Processing Systems (2024) 5.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: Neural Information Processing Systems (2024) 5

Reference 67

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:28d610840e40d352e2fa074c73015afd5d5a65ba58e9d77d8d22a42947298825

Observation 55ed052f-8217-47ed-89d4-d756dd23aa1a · outbound

This paper cites In: IEEE/CVF International Conference on Computer Vi- sion.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: IEEE/CVF International Conference on Computer Vi- sion

Reference 68

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:44d8b9a8a28555613d6969624af7357ef8997bf1e6b198dc892fc76eb7a35ebc

Observation 6c0cca11-d749-4a24-8a78-6189041ca2e7 · outbound

This paper cites In: Robotics: Science and Systems (2024) 1, 5.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: Robotics: Science and Systems (2024) 1, 5

Reference 69

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:8dc0ab96917fe5b83aa01fcc1ab678e35ba8a9bc31ac305fc466a98d26b65d6a

Observation a5c70583-ca5d-4a04-8c9d-d1b1723e5b19 · outbound

This paper cites In: IEEE/CVF Con- ference on Computer Vision and Pattern Recognition (2025) 4.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: IEEE/CVF Con- ference on Computer Vision and Pattern Recognition (2025) 4

Reference 70

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:11de56297eeee13b164c9708f0c381958f3318e3982a6dc1b191265937dc6d29

Observation 2c22d3dd-04f1-4d07-9c76-df11de86f607 · outbound

This paper cites Neural Comput.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes Neural Comput

Reference 71

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:df3401a945bb379bcdf284d5b842aa16fe4764e39142eb30a185f052a30dd058

Observation db54430c-b913-4b84-b245-15c42a97f94d · outbound

This paper cites In: European Conference on Computer Vi- sion (2024) 5.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: European Conference on Computer Vi- sion (2024) 5

Reference 72

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:0cf49d8142e9347757ce4cd57ca26c920f1d9640c6b8cd14c7faea1ca3e3b9b3

Observation 3b97bee3-d79b-49fc-9152-b2a8582f36cc · outbound

This paper cites In: International Conference on Machine Learning (2026) 5.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: International Conference on Machine Learning (2026) 5

Reference 73

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:94aee067ade6d09b8afb67c8c48e4fc670cfad51d3409aff44578bf41d866943

Observation 3fb45b70-10a8-4b63-aeae-bf447c39aee4 · outbound

This paper cites In: IEEE/CVF Conference on Computer Vision and Pat- tern Recognition (2025) 13, 18.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: IEEE/CVF Conference on Computer Vision and Pat- tern Recognition (2025) 13, 18

Reference 74

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:04dab3030866576086c026791d81e524e6061fd08d78965385fe175bb3e8300d

Observation 639e746c-19dc-48ea-9133-928191c379c1 · outbound

This paper cites Human-in-the-Loop Local Corrections of 3D Scene Layouts via Infilling.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes Human-in-the-Loop Local Corrections of 3D Scene Layouts via Infilling

Reference 75

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:f8257fd43a161f122378ddcc5b907b4e1bedfd77d7cb626c47258f27fee2705f

Observation 407ca237-5053-4f30-a0ec-b3765727a29d · outbound

This paper cites In: European Conference on Computer Vision (2024) 4.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: European Conference on Computer Vision (2024) 4

Reference 76

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:b9fd00582b679bd33f43e1b97e51091b8c14254b5a1413f6ebc10c81c49f282e

Observation 652cdef4-2f4f-49ce-917a-17bcfba86749 · outbound

This paper cites Qwen3 Technical Report.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes Qwen3 Technical Report

Reference 77

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:2fd8806aa7f8fa823e484175d0c415a73b14541591b40132c06a4bec2ebaeeac

Observation 5a91e304-266f-40d5-a551-5f4159801379 · outbound

This paper cites Qwen2.5 Technical Report.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes Qwen2.5 Technical Report

Reference 78

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:24713672d1f3603c7f3155d85c99d7c6f9020264f13923a4de62c58b89efd6eb

Observation 5f9e2024-7174-4afe-93ed-bbcdb52b1bcb · outbound

This paper cites In: IEEE/CVF Conference on Computer Vision and Pattern Recog- nition.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: IEEE/CVF Conference on Computer Vision and Pattern Recog- nition

Reference 79

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:65614b03c3fc9650b490a79730fc23c690dff91cd3eaaba6bcc1d9f85e96ec73

Observation c273b637-595f-4b68-8196-8f4267cb4212 · outbound

This paper cites In: International Con- ference on Learning Representations (2025) 6 30 K.-Y.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: International Con- ference on Learning Representations (2025) 6 30 K.-Y

Reference 80

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:97736c233fa3b1094450bffb5071d68cfa3dcd23daa370ab7eb998c496916e41

Observation 95fd4445-4c85-4960-b37c-1f7ed9e664fd · outbound

This paper cites In: IEEE/CVF International Conference on Computer Vision (2025) 4, 6, 10.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: IEEE/CVF International Conference on Computer Vision (2025) 4, 6, 10

Reference 81

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:fba721625f1e3288260c234bf7d8a7a038a2d0f59bd5023923fa2b1870877087

Observation bdb69f46-18d2-425b-9a97-11a27c4bbdb0 · outbound

This paper cites In: IEEE/CVF Conference on Computer Vision and Pattern Recognition (2022) 4.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: IEEE/CVF Conference on Computer Vision and Pattern Recognition (2022) 4

Reference 82

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:36d4f3723c014b3b5be43c62ec50975290b4ba039e16042e5e55f2c54e8b2da7

Observation 5f128144-2635-4ef0-ac3d-84bc7e5375a6 · outbound

This paper cites In: International Joint Conference on Artificial Intelligence (2025) 4.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: International Joint Conference on Artificial Intelligence (2025) 4

Reference 83

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:7f08befffe05a76e3958b729fb8a8aac534a49ab7e158ba944f2f86e949ca67e

Observation 42395726-968e-4059-84ae-bdf76ada751a · outbound

This paper cites Qwen3 Embedding: Advancing Text Embedding and Reranking Through Foundation Models.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes Qwen3 Embedding: Advancing Text Embedding and Reranking Through Foundation Models

Reference 84

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:d2607e028f96faebd6192b96d2cf828c864c7347a12f784cf109f9a2975a5362

Observation 740b03b4-7fbc-47df-befa-3b36602b016f · outbound

This paper cites In: Neural Information Processing Systems (2025) 5.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: Neural Information Processing Systems (2025) 5

Reference 85

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:d5a649e90242c99527c1ab87fe1fad3b67f66ea8887dc1723a25a6506465cebb

Observation c400efe8-3cc9-4e76-bef7-42ef2ad4045a · outbound

This paper cites In: IEEE/CVF Conference on Computer Vision and Pattern Recognition (2025) 5.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: IEEE/CVF Conference on Computer Vision and Pattern Recognition (2025) 5

Reference 86

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:04c3c1e447afc0401b44cf4a762f9a0a3f81ff096f21a657c8f8d5542395fd85

Observation d7e1b9f1-f476-4bb0-ba89-3b174180dc60 · outbound

This paper cites In: European Conference on Computer Vision (2020) 9.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: European Conference on Computer Vision (2020) 9

Reference 87

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:ed7ac8276602c060629328d3d2d75b57bac3b8e58e379009dee7217b7d359068

Observation 424e6aab-0cf0-4ffa-9e3d-65adde122029 · outbound

This paper cites In: IEEE/CVF Conference on Computer Vision and Pattern Recognition (2025) 1.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: IEEE/CVF Conference on Computer Vision and Pattern Recognition (2025) 1

Reference 88

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:7d85f2c2748fc6e2c60a960f93b5bf51b241c98156f6cd4b1dcfa5b2137892b8

Observation 9cd1492b-93b6-407d-ae36-dc53e6301976 · outbound

This paper cites In: Neural Information Processing Systems (2025) 1.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: Neural Information Processing Systems (2025) 1

Reference 89

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:20cfc39fea313bbb559f9cf6bf88f4168a8af4d29df7691e71f5062582b80810

Observation d24a58c8-db4a-4f08-8c87-ad22903669f9 · outbound

This paper cites In: IEEE/CVF International Con- ference on Computer Vision (2023) 5.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: IEEE/CVF International Con- ference on Computer Vision (2023) 5

Reference 90

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:644438efa44004eb6e881bf312e41e5c4f996c1893f3ac3c964a308bfcca5773

Observation 4432c857-dd52-4622-ae4b-92d490ef6797 · outbound

This paper cites In: European Conference on Computer Vision (2024) 5.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes In: European Conference on Computer Vision (2024) 5

Reference 91

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:27aeb8a79559e2da5dbbef021775c92e01749d66af6760ed6d15969b715c4787

Pith citing papers

No inbound Pith citation observations are available.