Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-08T16:39:53.448383Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 55 of 55 outbound references and 0 inbound Pith citation observations for arXiv:2607.06097.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-08T16:39:53.448383Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
55 of 55 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 770cf900-7afa-4733-9f33-39959fc5b2bf · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet Referit3d: Neural listeners for fine-grained 3d object identification in real-world scenes
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 2c946741-9a0c-4648-be0f-92d6f0977cbb · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet Meteor: An automatic metric for mt evaluation with improved correlation with hu- man judgments
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5a8dc24a-39b3-4a2e-9951-97b85b7a79be · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet 3djcg: A unified framework for joint dense captioning and visual grounding on 3d point clouds
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7995818b-976a-4a7b-92f5-cb1da9b0de65 · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet Scanrefer: 3d object localization in rgb-d scans using natural language
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 968b226e-7f3d-4e69-ad5d-54b99e90cf27 · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet D 3 net: A unified speaker-listener architecture for 3d dense captioning and visual grounding
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 857daf50-1891-4fdf-ba81-64f5b0b19714 · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet End-to-end 3d dense captioning with vote2cap-detr
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b9b56170-c9d7-4ce9-b064-c85786464f7f · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet V ote2cap-detr++: Decoupling localization and describing for end-to-end 3d dense captioning, 2023
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a393026d-9d76-4cfc-929c-073404d8e129 · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet Segment and Select: Vision-Language Segmentation in 3D Scenarios
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation fa03f646-0c25-40ff-8fbb-f3f4c1cec50f · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet Uniter: Universal image-text representation learning
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 46e1c952-63e7-4e4e-ac5e-040a46854904 · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet Scan2cap: Context-aware dense captioning in rgb- d scans
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d517a76e-1d56-4441-8de4-e4494312825f · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet Unit3d: A unified trans- former for 3d dense captioning and visual grounding
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 69dc074a-e70a-4b58-b9a8-68ae88779ae0 · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet Back-tracing representative points for voting- based 3d object detection in point clouds
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6f8b0d8d-a725-405a-ba53-9a6c8f64b670 · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet 4d spatio-temporal convnets: Minkowski convolutional neural networks, 2019
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation bb682023-94e7-4dd0-8de3-e105e7967eae · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet Chang, Manolis Savva, Maciej Hal- ber, Thomas Funkhouser, and Matthias Nießner
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0c37d686-f28e-44b4-88c5-f3e78b0c0fe3 · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet V otenet: A deep learning label fusion method for multi-atlas segmenta- tion
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6c6aa963-c870-4022-8526-6430dc607748 · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet An empirical study of training end-to-end vision-and-language transformers
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5ca690bb-f449-47eb-ad4d-3b06f20a5d27 · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet Learning lightweight lane detection CNNs by self atten- tion distillation
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5d5521b8-bc05-4580-8616-3beeca30e485 · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet Point-to-Voxel Knowledge Distillation for Li- DAR Semantic Segmentation
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation bb32d021-743e-4184-a67d-c1db2e4ac1fe · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet Scaling up vision-language pre-training for image captioning
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c82a5b9a-1dae-4fe3-b802-998d7dd0a511 · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet Nerf-det++: Incorporat- ing semantic cues and perspective-aware depth supervision for indoor multi-view 3d detection.IEEE Transactions on Image Processing, 2025
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 49ebd448-0871-44d6-a417-5ec01004c8ca · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet Perturb, predict & para- phrase: Semi-supervised learning using noisy student for im- age captioning
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 2ba7fa23-cc38-42ff-bc3a-dc6805539583 · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet Recurrent fusion network for image captioning
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ace88239-e12d-494b-aa21-3a6de33b8bda · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet More: Multi-order relation mining for dense captioning in 3d scenes
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 51eada91-8947-4a2a-9920-1917e97ea4c0 · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet Context-aware alignment and mutual masking for 3d- language pre-training
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation dc1c7fd1-f2a1-42b7-8d6c-70c35142d45f · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet DLIP: Distilling Language-Image Pre-training
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ac293cb5-3f5a-4f4e-a4b5-6ada96e398f3 · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet Align before fuse: Vision and language representation learn- ing with momentum distillation.Advances in neural infor- mation processing systems, 34:9694–9705
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ac034ae4-5770-439d-bba8-0701697233c1 · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet Blip: Bootstrapping language-image pre-training for uni- fied vision-language understanding and generation
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation fe1ff3a3-27a5-44c5-9619-369381b78827 · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet Oscar: Object-semantics aligned pre-training for vision-language tasks
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b2ab76ca-055c-4786-be4d-185cb75fef55 · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet MoE3D: Mixture of Experts Meets Multi-Modal 3D Understanding.arXiv 2025
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6db7c481-5339-4174-8308-920c24d3d103 · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet Rouge: A package for automatic evaluation of summaries
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 646f077b-536c-43b1-bd6e-da1d1cc4596b · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet Complete 3d relationships extraction modality align- ment network for 3d dense captioning.IEEE Transactions on Visualization and Computer Graphics, 2023
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7b7d8d4d-9552-4b08-9b80-34fc72754ab8 · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet An end-to- end transformer model for 3d object detection
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9132598e-e2ba-42d3-a77c-d75faeadf435 · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet Bleu: a method for automatic evaluation of machine translation
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b6075cc7-32ed-41c5-b104-b6b5ddb31d35 · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet Pointnet: Deep learning on point sets for 3d classification and segmentation
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e65cc2fb-4474-4c2f-aa3e-7e2f27e373c5 · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet Pointnet++: Deep hierarchical feature learning on point sets in a metric space.Advances in neural information processing systems, 30, 2017
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d9b72fbd-de5f-4b21-b9e9-0ab36724dd4b · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet Qi, Or Litany, Kaiming He, and Leonidas J
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 06a05bf8-35a4-4a7c-9472-b64c4e205fd7 · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet Rennie, Etienne Marcheret, Youssef Mroueh, Jarret Ross, and Vaibhava Goel
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f23972b7-ab42-4b7d-bb5f-85c097badd53 · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet Octnet: Learning deep 3d representations at high resolutions
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 892a1f4d-15d4-49a2-bad3-af402d4079a5 · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet Fcaf3d: Fully convolutional anchor-free 3d object detection
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 76db27ec-5603-4ef8-8737-61444b85fab7 · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet Unresolved cited work
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e438d67d-24ca-44e6-8b77-6e4ecf96f82e · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet VL-BERT: Pre-training of Generic Visual-Linguistic Representations
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation dfcecb9e-7e53-4638-b14d-87e9fff93cdb · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet Lawrence Zitnick, and Devi Parikh
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 65cc72ba-d572-4462-9755-2aaa13a88a01 · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet Cagroup3d: Class- aware grouping for 3d object detection on point clouds
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 843ef6ab-92be-4618-9218-1b0916d1afc5 · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet Rbgnet: Ray-based grouping for 3d object detection
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4f20f37f-a401-4d89-90ef-215406087962 · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet Spatiality-guided Transformer for 3D Dense Captioning on Point Clouds
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 473c21da-dfb2-4285-b3c8-2732efcefc86 · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet Point transformer v2: Grouped vector atten- tion and partition-based pooling, 2022
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3eef39bf-dfb0-4a07-ab9d-3a66fcc82be4 · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet Taseg: Temporal aggregation network for lidar semantic segmentation
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3e809ed1-5a40-47b4-aa28-94c4dbec1535 · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet Second: Sparsely embed- ded convolutional detection.Sensors, 18(10):3337
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 126695a3-1139-4a12-902b-90a0e4eac2bf · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet Vision-language pre-training with triple contrastive learning
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9b088032-702b-491d-b486-7351219afa1b · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet Swin3d: A pretrained transformer backbone for 3d indoor scene understanding, 2023
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 20835cc4-818b-4f34-9c78-5bd570f1c944 · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet X-trans2cap: Cross- modal knowledge transfer using transformer for 3d dense captioning
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6557e455-29be-498b-a661-e0e011d43a58 · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet H3dnet: 3d object detection using hybrid geometric primi- tives
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation fb825369-20a4-4810-b10a-8977896acd63 · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet Contextual modeling for 3d dense captioning on point clouds, 2022
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 01df9144-44bd-4d80-8a35-2f8247dfedd1 · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet V oxelnet: End-to-end learning for point cloud based 3d object detection
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9b2aa759-6b9f-4e44-bba0-7200dcde6860 · outbound
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet 3d-vista: Pre-trained transformer for 3d vision and text alignment
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
No inbound Pith citation observations are available.