Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-29T22:47:05.747852Z
Paper Citation Record · LEDGER
As of 24 July 2026, this Paper Citation Record lists 32 of 32 outbound references and 0 inbound Pith citation observations for arXiv:2605.25901.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-29T22:47:05.747852Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-07-24T06:31:00.690269+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
32 of 32 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 621ec68a-bcc2-4f28-9af0-710e32d3bff8 · outbound
AgentGrounder: Zero-Shot 3D Visual Pointcloud Grounding using Multimodal Language Models Gˆ 3-lq: Marrying hyperbolic alignment with explicit semantic-geometric modeling for 3d visual grounding,
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c9da0c0c-1e08-4048-a7c0-9b72e9b54f3f · outbound
AgentGrounder: Zero-Shot 3D Visual Pointcloud Grounding using Multimodal Language Models Multi-branch collaborative learning network for 3d visual grounding,
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9ab67e8-a0f2-438c-95ec-1d65b3489d65 · outbound
AgentGrounder: Zero-Shot 3D Visual Pointcloud Grounding using Multimodal Language Models Chat-scene: Bridging 3d scene and large language models with object identifiers,
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51e219cd-d1ab-4265-add6-d9737eae432b · outbound
AgentGrounder: Zero-Shot 3D Visual Pointcloud Grounding using Multimodal Language Models Video-3d llm: Learning position-aware video representation for 3d scene understanding,
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7102f1d4-a20b-4d15-b2d2-ae916c3d3072 · outbound
AgentGrounder: Zero-Shot 3D Visual Pointcloud Grounding using Multimodal Language Models Llm-grounder: Open-vocabulary 3d visual grounding with large language model as an agent,
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb61c667-2b73-4a02-8300-b16713cbaef3 · outbound
AgentGrounder: Zero-Shot 3D Visual Pointcloud Grounding using Multimodal Language Models Visual programming for zero-shot open- vocabulary 3d visual grounding,
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c1e8d07-ce2a-4b31-83dd-007933d773e5 · outbound
AgentGrounder: Zero-Shot 3D Visual Pointcloud Grounding using Multimodal Language Models Vlm-grounder: A vlm agent for zero-shot 3d visual grounding,
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a94e15b2-98c4-4047-a865-3cd6849ea469 · outbound
AgentGrounder: Zero-Shot 3D Visual Pointcloud Grounding using Multimodal Language Models See- ground: See and ground for zero-shot open-vocabulary 3d visual grounding,
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4bd7f0ab-0d5b-43f0-8e52-8c1aa303fa51 · outbound
AgentGrounder: Zero-Shot 3D Visual Pointcloud Grounding using Multimodal Language Models Solving Zero-Shot 3D Visual Grounding as Constraint Satisfaction Problems
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.
Observation bd0b770d-1351-4f08-998d-733a7bf42cf4 · outbound
AgentGrounder: Zero-Shot 3D Visual Pointcloud Grounding using Multimodal Language Models Sort3d: Spatial object-centric rea- soning toolbox for zero-shot 3d grounding using large language models,
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5abcaaca-cd13-4feb-9078-2603b60edad4 · outbound
AgentGrounder: Zero-Shot 3D Visual Pointcloud Grounding using Multimodal Language Models Transcrib3d: 3d referring expression resolution through large language models,
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a07273b-3ba3-4af3-ad29-674b8084aa9e · outbound
AgentGrounder: Zero-Shot 3D Visual Pointcloud Grounding using Multimodal Language Models Scanrefer: 3d object localization in rgb-d scans using natural lan- guage,
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba5ea7fa-6a64-4c12-84f1-d9db03d5b78c · outbound
AgentGrounder: Zero-Shot 3D Visual Pointcloud Grounding using Multimodal Language Models Referit3d: Neural listeners for fine- grained 3d object identification in real-world scenes,
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38b20b93-1502-4d02-b450-39428e94ab7e · outbound
AgentGrounder: Zero-Shot 3D Visual Pointcloud Grounding using Multimodal Language Models Four ways to improve verbo-visual fusion for dense 3d visual grounding,
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb040efe-7a4b-43af-80fa-113c0f2b131c · outbound
AgentGrounder: Zero-Shot 3D Visual Pointcloud Grounding using Multimodal Language Models OpenMask3D: Open-Vocabulary 3D Instance Segmentation
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.
Observation 0ff0fb14-1a73-44f0-b6cc-89ca21138747 · outbound
AgentGrounder: Zero-Shot 3D Visual Pointcloud Grounding using Multimodal Language Models Isbnet: A 3d point cloud instance segmentation network with instance-aware sampling and box-aware dynamic con- volution,
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45358fbe-78bc-4736-8fc6-78d01825bd88 · outbound
AgentGrounder: Zero-Shot 3D Visual Pointcloud Grounding using Multimodal Language Models SAM3D: Segment Anything in 3D Scenes
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.
Observation cf86a4f2-e806-457b-9169-53279f855ad0 · outbound
AgentGrounder: Zero-Shot 3D Visual Pointcloud Grounding using Multimodal Language Models Open3dis: Open-vocabulary 3d instance segmentation with 2d mask guidance,
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 49e580fe-5d5b-4830-9e83-65bb7bb6026e · outbound
AgentGrounder: Zero-Shot 3D Visual Pointcloud Grounding using Multimodal Language Models Any3dis: Class-agnostic 3d instance segmentation by 2d mask tracking,
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e1d08be-0b11-4443-b714-081085a2e5dc · outbound
AgentGrounder: Zero-Shot 3D Visual Pointcloud Grounding using Multimodal Language Models Conceptgraphs: Open-vocabulary 3d scene graphs for perception and planning,
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e75f00bd-bd61-424d-a91c-acb2c5eb3f5f · outbound
AgentGrounder: Zero-Shot 3D Visual Pointcloud Grounding using Multimodal Language Models ConceptFusion: Open-set Multimodal 3D Mapping
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.
Observation 331a537f-d8b2-4772-becc-1fca06193087 · outbound
AgentGrounder: Zero-Shot 3D Visual Pointcloud Grounding using Multimodal Language Models Phygrasp: Generalizing robotic grasp- ing with physics-informed large multimodal models,
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af2ec1cf-a924-498c-8813-9c69a17fe1e3 · outbound
AgentGrounder: Zero-Shot 3D Visual Pointcloud Grounding using Multimodal Language Models Zero-shot object navigation with vision-language models reasoning,
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ec2d34f-b841-446d-87f2-776521f3067b · outbound
AgentGrounder: Zero-Shot 3D Visual Pointcloud Grounding using Multimodal Language Models Iref-vla: A benchmark for interactive referen- tial grounding with imperfect language in 3d scenes,
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 561bfa1f-4a12-4808-94b8-b6571e253da6 · outbound
AgentGrounder: Zero-Shot 3D Visual Pointcloud Grounding using Multimodal Language Models Sceneverse: Scaling 3d vision-language learning for grounded scene understanding,
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5bf9c083-32d1-4bf2-bb20-047d1d535f0c · outbound
AgentGrounder: Zero-Shot 3D Visual Pointcloud Grounding using Multimodal Language Models Mask3d: Mask transformer for 3d semantic instance segmentation,
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f01fbfbe-c7a8-4faf-8764-197323671bb7 · outbound
AgentGrounder: Zero-Shot 3D Visual Pointcloud Grounding using Multimodal Language Models Qwen3 Technical Report
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.
Observation 0c0e2192-34ab-4268-bdf8-e0225399c842 · outbound
AgentGrounder: Zero-Shot 3D Visual Pointcloud Grounding using Multimodal Language Models Using ollama,
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa9cb1c7-2d56-470b-80c0-b7049d59ccc1 · outbound
AgentGrounder: Zero-Shot 3D Visual Pointcloud Grounding using Multimodal Language Models Chase,Langchain,https : / / github
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10454cfd-1cdb-4d81-94e9-768de310711f · outbound
AgentGrounder: Zero-Shot 3D Visual Pointcloud Grounding using Multimodal Language Models GPT4Scene: Understand 3D Scenes from Videos with Vision-Language Models
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.
Observation 27652e4b-f858-40ac-886e-d13de3c027a6 · outbound
AgentGrounder: Zero-Shot 3D Visual Pointcloud Grounding using Multimodal Language Models Mikasa: Multi-key-anchor & scene-aware trans- former for 3d visual grounding,
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab0266f1-8843-4483-994b-352c5dcba901 · outbound
AgentGrounder: Zero-Shot 3D Visual Pointcloud Grounding using Multimodal Language Models Language conditioned spatial relation rea- soning for 3d object grounding,
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.