Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T19:51:43.669537Z
Paper Citation Record · LEDGER
As of 23 August 2026, this Paper Citation Record lists 65 of 65 outbound references and 1 inbound Pith citation observation for arXiv:2412.06322.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T19:51:43.669537Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-22T08:47:25.712575Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-22T08:51:18.523931Z
65 of 65 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 39497bb6-902b-404c-8fc9-8b2b36052a7d · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 922e3bcb-3b7b-4ba1-9a1e-5a871276ec37 · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Llama 3 model card
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 3b93f7d4-901e-4ece-b0a1-2ded9a7b569a · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Vqa: Visual question answering
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69b3aa84-8263-4ba0-8bde-0a8db499ad73 · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Adabins: Depth estimation using adaptive bins
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 46e608b6-3100-4d2f-8f7e-f0cf91a01ea8 · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Transformerfusion: Monocular rgb scene reconstruction using transformers
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation c99b669d-4cd7-4f08-ad7b-172a98bb1040 · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Language Models are Few-Shot Learners
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab2b56cd-7348-40e8-88b9-1ea909cb40ba · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Sift flow: Dense correspondence across different scenes
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 48438ee1-5f7e-4dd8-b156-f1dc65cc05a8 · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations SpatialVLM: Endowing Vision-Language Models with Spatial Reasoning Capabilities
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0fc116a0-cd53-4837-8dca-9d578651ef64 · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations ShareGPT4V: Improving Large Multi-Modal Models with Better Captions
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 670d3a78-c783-48c4-8554-aedf8d70167b · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Knowledge-embedded routing network for scene graph gen- eration
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 650499dc-1685-49f1-845d-2a0110608a6c · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8d9fd59-c5d3-4f34-9177-bf9d8526f891 · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Internvl: Scaling up vision foundation mod- els and aligning for generic visual-linguistic tasks
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8110f618-5fc2-478b-9e0a-cbfeec625a61 · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations SpatialRGPT: Grounded Spatial Reasoning in Vision Language Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a36950d6-2f13-41f6-9284-96ccef36a17d · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Deep- videomvs: Multi-view stereo on video with recurrent spatio- temporal fusion
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation f2c82a5a-d708-4297-bdfe-78c76707b1fa · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Depth map prediction from a single image using a multi-scale deep net- work
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e79127c0-45e7-494e-9e5e-90f4e4ca4778 · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Making the v in vqa matter: Elevating the role of image understanding in visual question answer- ing
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74396314-a112-445e-939d-6c2433987710 · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Recov- ering surface layout from an image
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation fe669a19-ecd7-4604-b2ac-83923e1cdf11 · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Language is not all you need: Aligning perception with language mod- els
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 7032a545-41f7-45ea-85f4-b89444bfd56f · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Vi- sual prompt tuning
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f2678cb-59a8-4ca0-86f5-de1f36e3533a · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Image retrieval using scene graphs
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 2143d770-f82e-4d3f-8b4b-55ef52194054 · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Poisson surface reconstruction
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 9f8d8a15-300d-4ab4-b22b-2917dd973665 · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Two algorithms for constructing a delaunay triangulation
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 85e8f3ae-b47d-4feb-b8ae-76a2fc8df492 · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Depth and surface normal estimation from monocular images using regression on deep features and hi- erarchical crfs
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 2a7ec07c-d549-41d0-a158-58b9bfeb0ac4 · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d30914f-4e13-40c4-8ecc-5b38f0ae760d · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Align before fuse: Vision and language representation learn- ing with momentum distillation
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 762cab21-7500-4bd9-b824-b44f7b0e6eb6 · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Factorizable net: an efficient subgraph-based framework for scene graph generation
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 27cce3b6-fb70-4c0d-ba24-7b8a2dc85877 · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Scene graph generation from objects, phrases and region captions
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 195e6fd4-f7b0-4573-967a-ec81e199c2de · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations StableLLaVA: Enhanced Visual Instruction Tuning with Synthesized Image-Dialogue Data
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ac2b93e-21a5-4eef-9700-f62c2c0c8090 · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Binsformer: Revisiting adaptive bins for monocular depth estimation
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 184d2689-0012-46ea-bbdf-ba2f9adeaa3c · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Microsoft coco: Common objects in context
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation ef52ed51-0af0-4a9d-b0e9-6271e2fa447e · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Gps-net: Graph property sensing network for scene graph generation
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 2c555bb6-1d02-41c6-9066-5bbf5502c84f · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Visual spa- tial reasoning
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation d0be3b8e-dca5-4a8b-8c68-24d4a02c009c · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Improved baselines with visual instruction tuning
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 5caa1d29-3e0b-48b1-af09-dd80b0c8628f · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Visual instruction tuning
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 76d7b4f5-95ba-4e8b-9667-7db0c0ec5e21 · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Visual relationship detection with language priors
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 770f9547-6fa9-4961-98e3-16b7d98339fb · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Ocr-vqa: Visual question answering by reading text in images
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 388776f7-ad93-4878-9ab0-d4563d735b80 · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Atlas: End- to-end 3d scene reconstruction from posed images
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation b8ea3c9d-09e5-405a-803a-fddc0297c8d6 · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Gpt-4o system card, August 2024
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation fcbfdd98-0f94-4fdf-bfbb-869e7826715a · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Kosmos-2: Grounding Multimodal Large Language Models to the World
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 288bab9d-ed7f-4ea4-9257-110d2d00eda7 · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Spatial-temporal knowledge-embedded transformer for video scene graph generation
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation a643bf50-030d-4567-9baa-e9d6b10d7062 · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Qwen2.5: A party of foundation models, September 2024
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 169f72b4-033c-4f9e-8414-54abfebeb68a · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Learning transferable visual models from natural language supervi- sion
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0efdc85c-801b-4652-b9e9-409e725b6a86 · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Pixelwise view selection for unstructured multi-view stereo
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation b106f35c-b582-4689-ab6a-726bde78c5ca · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Structured query- based image retrieval using scene graphs
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 15eed070-521a-42e8-a319-925a3bdd0ea9 · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Nddepth: Normal-distance as- sisted monocular depth estimation
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 79c5c60c-2864-4287-8fd6-a76207b26152 · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Conceptual captions: A cleaned, hypernymed, im- age alt-text dataset for automatic image captioning
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 3e4fcaee-4e80-4299-a842-5c8d743c588c · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Towards vqa models that can read
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 7cf09185-a045-46bb-9aba-c76e1462d5e0 · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Neuralrecon: Real-time coherent 3d re- construction from monocular video
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 86281eb0-a7f4-4e34-aa7a-313b5666903b · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Learning to compose dynamic tree structures for visual contexts
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 33f13e65-90fc-43ab-9aad-e1ae3025a9a5 · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a669df6c-77bd-4c71-b69b-ea830c6f785f · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations The All-Seeing Project V2: Towards General Relation Comprehension of the Open World
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66a83fdb-39d3-41e2-a7e4-6ea556099b3d · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations The All-Seeing Project: Towards Panoptic Visual Recognition and Understanding of the Open World
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b975ad39-7d7f-4869-9957-1c16109e1168 · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Chain-of-thought prompting elicits reasoning in large lan- guage models
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1393d05e-ed00-4506-b704-174f2aeba5de · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Scene graph generation by iterative message passing
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 40815618-0aff-4db9-880d-3037723aaa58 · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Panoptic scene graph gen- eration
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 432d6546-9082-4975-9dc3-7d4a7950b967 · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Graph r-cnn for scene graph generation
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation e289b01d-f49f-4889-8f13-97769b19c86d · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Depth anything: Unleashing the power of large-scale unlabeled data
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 577ea4c0-40ea-49a8-958c-b66a7f600c46 · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations CoCa: Contrastive Captioners are Image-Text Foundation Models
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a34d4f19-8543-4087-8c2b-35f8ac3172b9 · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Neural motifs: Scene graph parsing with global con- text
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 579ed9c9-d3d0-4fd5-86d3-107fe9b2fdaa · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Conceptual and syntactical cross-modal alignment with cross-level consistency for image-text match- ing
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation ee6f49af-a391-40af-8c86-dab709790b46 · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Textpsg: Panoptic scene graph generation from textual descriptions
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 36274a89-045f-44c4-be4d-64fe38596e20 · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5bc93447-2041-4245-99cd-62e25ffd18a2 · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Unresolved cited work
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 012915af-4d29-4a0a-a1e5-4fab99d9bf9e · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations Unresolved cited work
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 576fa07c-ae81-4b70-9232-468f177291c2 · outbound
LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations house” as an example): object labels(“house
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 12af512c-a7af-4ad2-b8f3-0f9f707a467e · inbound
SceneGraphGrounder: Zero-Shot 3D Visual Grounding via Structured Scene Graph Matching LLaVA-SpaceSGG: Visual Instruct Tuning for Open-vocabulary Scene Graph Generation with Enhanced Spatial Relations
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.