Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:12:19.747542Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 78 of 78 outbound references and 1 inbound Pith citation observation for arXiv:2505.22429.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:12:19.747542Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-01T06:41:38.649424Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-01T06:45:29.442037Z
78 of 78 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 960747c0-afc7-4659-b22b-4e9984151d05 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Scanrefer: 3d object localization in rgb-d scans using natural language,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8b9a2380-ea31-4728-a03e-027b0f870747 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models RayDF: Neural Ray-surface Distance Fields with Multi-view Consistency
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation dbfea0d5-1991-4263-ac19-3d58190694cb · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Deep view synthesis via self-consistent gen- erative network,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e3b704f0-aab2-4e93-86ab-32d1e3515bac · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models PhysFlow: Unleashing the Potential of Multi-modal Foundation Models and Video Diffusion for 4D Dynamic Physical Scene Simulation
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 57a6e328-f12f-4ab9-af53-7caa8174205f · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models SIR: Multi-view Inverse Rendering with Decomposable Shadow Under Indoor Intense Lighting
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4f6dd310-9524-424d-8221-e830ebea8546 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models An Examination of the Compositionality of Large Generative Vision-Language Models
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation aa17f378-4420-4b33-b61d-de83946f313d · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Think global, act local: Dual-scale graph transformer for vision-and-language navigation,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 330b875b-8571-4c73-ae2c-62fa7c4aaeb1 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Assister: Assistive navigation via condi- tional instruction generation,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9f53beb9-7803-4995-a8ac-6306a5f28797 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models From Cognition to Precognition: A Future-Aware Framework for Social Navigation
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation dd10ce10-0d22-4d89-9acd-dcdaeb9deac4 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Clip2scene: Towards label-efficient 3d scene understanding by clip,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 31d13b7b-2b84-4766-85e6-5592b983dad8 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Robo3d: Towards robust and reliable 3d perception against corruptions,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c1d22c9f-0892-4398-8716-e2d9ccbc6a59 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Rethinking range view representation for lidar segmentation,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2bedc6dc-5cfb-46f0-91a9-202261a576b1 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Xvo: Generalized visual odometry via cross- modal self-training,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 03bfdf04-2234-4fe5-a716-86bba0b57d4b · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models COARSE3D: Class-Prototypes for Contrastive Learning in Weakly-Supervised 3D Point Cloud Segmentation
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa261fc4-6dda-44f3-a79d-88fba013e8b5 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Perception-aware multi-sensor fusion for 3d lidar semantic segmentation,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6b23063d-3182-489e-ae81-83d3d6d434eb · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Epmf: Efficient perception-aware multi- sensor fusion for 3d semantic segmentation,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1040d8e3-f317-4bd5-8225-322ce41fcea5 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Tfnet: Exploiting temporal cues for fast and accurate lidar semantic segmentation,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation dd2fa089-3b28-4518-a393-274e9128ad3c · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Robust 3D Semantic Occupancy Prediction with Calibration-free Spatial Transformation
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6adcb6dd-1a21-4311-bcf5-680846d8444c · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Dhp-mapping: A dense panoptic mapping sys- tem with hierarchical world representation and label opti- mization techniques,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation dd1c1904-35de-4441-b3eb-dd532efb8937 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Multi-modal data-efficient 3d scene un- derstanding for autonomous drivin,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 77efeec2-affe-46a1-8e6c-5590ca719030 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Dynamiccity: Large-scale 4d occu- pancy generation from dynamic scenes,
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21947037-35b4-4109-83a4-c704eadff963 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Calib3d: Calibrating model preferences for reliable 3d scene understanding,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5c6df824-f55a-43df-b72b-d33033b5e4f1 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Bottom up top down detection transform- ers for language grounding in images and point clouds,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bc2adfb0-b5bc-4cec-8761-c20b7c790c46 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models 3d-vista: Pre-trained transformer for 3d vision and text alignment,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7c0380bf-1754-4b5b-91dd-f9bc7301825e · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Eda: Explicit text-decoupling and dense align- ment for 3d visual grounding,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d50e43f0-cf9a-41c7-b7b3-64936739d156 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models 3dvg-transformer: Relation modeling for vi- sual grounding on point clouds,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bd712acb-4ec8-4fd0-bc3f-3c2a28ce6d8f · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Instancerefer: Cooperative holistic under- standing for visual grounding on point clouds through in- stance multi-level contextual referring,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f3b330c4-f015-42f9-86e6-d02a4f2ff0f9 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Multi-branch collaborative learning network for 3d visual grounding,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5f07f699-f8bc-441a-8085-dcc8ec68e53d · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Semantickitti: A dataset for semantic scene understanding of lidar sequences,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2f3a2f58-fa72-4e45-9ff5-3861b4926bab · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Scalability in perception for autonomous driv- ing: Waymo open dataset,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7354791c-6c9e-4f60-956a-af68e73e773c · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Panoptic nuscenes: A large-scale bench- mark for lidar panoptic segmentation and tracking,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 11375c3b-8862-4593-bad2-7761dfc1ac0e · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Visual programming for zero-shot open- vocabulary 3d visual grounding,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a012d746-72fe-44e3-bdeb-0d1aee487572 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Llm-grounder: Open-vocabulary 3d visual grounding with large language model as an agent,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c155db64-fdc3-4454-8b88-6e86c1baba4a · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Training language models to fol- low instructions with human feedback,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1e1b7fb4-f508-42b6-8fab-a9bf6ea0c806 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models GPT-4 Technical Report
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3579b74e-7593-4f22-8ebb-e28f7c17fb6a · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5426ae40-e7e6-4ba6-bb60-6b115f0d2f4a · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models CogVLM2: Visual Language Models for Image and Video Understanding
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63c38e4a-105a-4267-8616-ce139f488a22 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Sceneverse: Scaling 3d vision-language learn- ing for grounded scene understanding,
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 37219787-af06-4101-8af5-57e158d3730c · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models SeeGround: See and Ground for Zero-Shot Open-Vocabulary 3D Visual Grounding
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd2be870-d9b6-4bae-b676-1b4ff513377a · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Referit3d: Neural listeners for fine- grained 3d object identification in real-world scenes,
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1c3d62e2-a284-4093-9485-c07bfd12986b · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Viewrefer: Grasp the multi-view knowledge for 3d visual grounding,
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a70094c3-e7ce-4b6c-8fc8-cb4d1670b286 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Multi-view transformer for 3d visual grounding,
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9229b3cc-9dcb-4273-9701-4f45767296f1 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Look around and refer: 2d synthetic se- mantics knowledge distillation for 3d visual grounding,
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d6ce677b-b1b6-4488-b020-89df092fe040 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Sat: 2d semantics assisted training for 3d visual grounding,
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1ca3c590-6dea-4a27-874d-747b1befea47 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Four ways to improve verbo-visual fusion for dense 3d visual grounding,
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8b9e7394-52c4-4635-90bc-357d17e6ace1 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Distilling coarse-to-fine semantic matching knowledge for weakly supervised 3d visual grounding,
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8ad91c79-521c-4f81-8bfc-9e2577bc4909 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Unifying 3d vision-language understanding via promptable queries,
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f858176d-4543-448f-94e2-5417467643c2 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models OVGaussian: Generalizable 3D Gaussian Segmentation with Open Vocabularies
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b52141f5-02d0-4ff5-8e29-8998abfa1c46 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Multi-space alignments towards universal lidar segmentation,
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 02ec2c3f-ff86-42ff-8ffb-34b28baecacf · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Towards label-free scene understanding by vision foundation models,
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 72357686-fe9d-482c-a614-4c8e5f09b4b6 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models LiMoE: Mixture of LiDAR Representation Learners from Automotive Scenes
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 197f289d-d06e-4527-8ea0-b7333f8cc82a · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models GEAL: Generalizable 3D Affordance Learning with Cross-Modal Consistency
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 030e2204-a25d-4b24-a0b4-3397826df452 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Openscene: 3d scene understanding with open vocabularies,
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 423967c6-2e0e-4c80-bcf4-565da2f9c4d3 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Lerf: Language embedded radiance fields,
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 94103a13-4e1b-43a6-bbc2-ed6e9e18a290 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Ovir-3d: Open-vocabulary 3d instance retrieval without training on 3d data,
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1f8f6e63-38ef-47e1-98da-a0ea5409aeeb · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Agent3D-Zero: An Agent for Zero-shot 3D Understanding
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d2c5493-8c26-42f6-90f4-97d8c35ede01 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Regionplc: Regional point-language con- trastive learning for open-world 3d scene understanding,
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 739b960d-0787-420d-8179-33da110aa4cf · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models OpenMask3D: Open-Vocabulary 3D Instance Segmentation
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9fff50e1-dbfb-428b-bd79-ed7ded1e7bfd · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Openins3d: Snap and lookup for 3d open- vocabulary instance segmentation,
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6b063554-7f08-45ac-8735-ec16104d3204 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Sai3d: Segment any instance in 3d scenes,
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b8e8a98d-4d75-4d7e-859e-cbf1efa418be · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Lasermix for semi-supervised lidar semantic segmentation,
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 88a4132e-37ac-4dc3-aec5-a64dc8004e8f · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Segment any point cloud sequences by distilling vision foundation models,
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c44806ef-9311-4844-9525-3cd2395c8625 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models 4d contrastive superflows are dense 3d repre- sentation learners,
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1c3f78e0-9855-4610-a55d-3cdbde6b628c · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Frnet: Frustum-range networks for scalable lidar segmentation,
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c4b43a23-36bf-481c-ae1a-d60fab1e0d84 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Scene-LLM: Extending Language Model for 3D Visual Understanding and Reasoning
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b469ac04-b778-4bfe-b0ae-85bc252010b2 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Uni3DL: Unified Model for 3D and Language Understanding
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6171bf57-8155-4ef5-bc3d-97e28edb525c · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Conceptfusion: Open-set multi- modal 3d mapping,
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b1e0260d-814e-4607-846a-a2b4f28ac11a · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models GLOVER: Generalizable Open-Vocabulary Affordance Reasoning for Task-Oriented Grasping
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11fb35db-9415-452a-8a43-1742b926fdb6 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Interactive planning using large language models for partially observable robotic tasks,
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fce09665-8e57-4ecf-ac78-7d026fcc53c8 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models 3d-llm: Injecting the 3d world into large language models,
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5fcd2776-fd95-49b2-b879-6e63ed38e8fd · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Is your lidar placement optimized for 3d scene understanding?,
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7343377b-7b76-4d17-b048-d0c15d351101 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models G3-lq: Marrying hyperbolic alignment with explicit semantic-geometric modeling for 3d visual ground- ing,
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3f863d07-94ef-4de9-ad28-d001bf53d84e · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Learning transferable visual models from natural language supervision,
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation edcfabd1-07f9-4cd5-b156-2823a5aa0bc2 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models VLM-Grounder: A VLM Agent for Zero-Shot 3D Visual Grounding
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6bb7d6ca-a924-4237-8f5e-d5f833206cbb · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Text-guided graph neural networks for referring 3d instance segmentation,
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4761a55c-7a2e-4aa9-9ef1-8cf0368ce164 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Mikasa: Multi-key-anchor & scene- aware transformer for 3d visual grounding,
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0de609bf-af3c-4a3d-b24c-fc6057d60b72 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Language conditioned spatial relation rea- soning for 3d object grounding,
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b9c59bea-69c3-4075-8a2c-08946499f3a6 · outbound
Zero-Shot 3D Visual Grounding from Vision-Language Models Mask3d: Mask transformer for 3d semantic instance segmentation,
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7e135835-ad52-43f6-9e78-62ef14a3da16 · inbound
PruneGround: Plug-and-play Spatial Pruning for 3D Visual Grounding Zero-Shot 3D Visual Grounding from Vision-Language Models
Reference 113
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.