Pith. sign in

Paper Citation Record · LEDGER

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation

As of 8 August 2026, this Paper Citation Record lists 72 of 72 outbound references and 2 inbound Pith citation observations for arXiv:2505.20106.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.20106 v1

Coverage vector

measured 72 of 72 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:06:09.067431Z

measured 74 of 74 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-27T17:30:48.848277Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T00:07:27.999553Z

Reference resolution

72 of 72 outbound references displayed

  • verified exact0
  • verified fuzzy58
  • unresolved14
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation cdee7b1c-498c-4ce6-9dd2-5ae7857f44d3 · outbound

This paper cites Scene graph generation by iterative message passing,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Scene graph generation by iterative message passing,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:19.512478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:01.386709Z digest=sha256:0c6ac491a7b1eaa7206fa2b9796a7080ca7a24e2b30c2b95c28fbd6c87bb1508

Observation bb9489d9-b38b-4f4e-8e75-e4d7fc3213c0 · outbound

This paper cites Neural motifs: Scene graph parsing with global context,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Neural motifs: Scene graph parsing with global context,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:19.353704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:01.463835Z digest=sha256:ddba58591cc74635926e41c0507b19967ddc1e81798e145463f3d8c9503399b8

Observation 2d989bd4-5f20-45b0-a41d-9f653430d3fa · outbound

This paper cites Learning to compose dynamic tree structures for visual contexts,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Learning to compose dynamic tree structures for visual contexts,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:19.190802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:01.611625Z digest=sha256:43c284b0c14cd899c4186ea7f7c36e2a7657e2828c9a1b6b9b123a8fe4853104

Observation 7d239a10-0c6d-433e-ad0c-e599fbc3e9ec · outbound

This paper cites Unbiased scene graph generation from biased training,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Unbiased scene graph generation from biased training,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:18.992266Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:01.746840Z digest=sha256:a095346f24ba41960609888dfc93c5f36c977308d015e6e1908cf7681b9d1eeb

Observation 698bbbe9-b9fb-4ded-8bb2-8a8601772183 · outbound

This paper cites Recovering the unbiased scene graphs from the biased ones,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Recovering the unbiased scene graphs from the biased ones,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:18.801403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:01.878744Z digest=sha256:3a6365d53a1a6ebc1efdcc638704f1cb225dcd5b1304782b533b6fea25955b94

Observation e858478d-b1ce-41e6-8749-f696b0d757bb · outbound

This paper cites Bipartite graph network with adaptive message passing for unbiased scene graph generation,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Bipartite graph network with adaptive message passing for unbiased scene graph generation,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:18.602538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:02.011294Z digest=sha256:a1019c7665a8d03e94681517332e5dcc7e8f0766d551826989434598801584a2

Observation cf3d4a51-d43c-49d7-8d61-f39be934a8fb · outbound

This paper cites Graphical contrastive losses for scene graph parsing,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Graphical contrastive losses for scene graph parsing,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:18.248992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:02.243722Z digest=sha256:1be0801b11bd320f00c30d8db93bf77a799f972795da001630ddbea4e7639827

Observation ba4e570b-ca0d-435a-960b-7f75fc2a087b · outbound

This paper cites Towards open-vocabulary scene graph generation with prompt-based finetuning,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Towards open-vocabulary scene graph generation with prompt-based finetuning,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:18.081171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:02.335243Z digest=sha256:fb4e8638f1ed80fd1ddd54fdc403fc659bc74fb25901f0c0e7526573e4a2ba3d

Observation f07808aa-fddb-448c-8d7f-0014dec725bb · outbound

This paper cites Learning to generate language-supervised and open-vocabulary scene graph using pre-trained visual-semantic space,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Learning to generate language-supervised and open-vocabulary scene graph using pre-trained visual-semantic space,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:17.945360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:02.447077Z digest=sha256:0b76b36b48114ed39eb677184f9eda182cbfd076dc297fb2a4c5b53759cebc76

Observation 102d9862-29e8-4f15-a2a4-8264c7c50420 · outbound

This paper cites Auto-encoding scene graphs for image captioning,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Auto-encoding scene graphs for image captioning,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:17.772322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:02.536362Z digest=sha256:b078501baf7f210e367bbd47792dc0b56b9507f0d17969d6727e79f564cccd1e

Observation c6b0d64b-7ea3-4113-a605-cfd7f34e8faf · outbound

This paper cites Say as you wish: Fine-grained control of image caption generation with abstract scene graphs,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Say as you wish: Fine-grained control of image caption generation with abstract scene graphs,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:17.595794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:02.622653Z digest=sha256:0c94a7fd1ea9d293523999886cd01379e66f24b766ee91bea194da1dfed38c19

Observation 6a7ba961-975e-4d05-8448-e7a2c2b27c9b · outbound

This paper cites Unpaired image captioning via scene graph alignments,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Unpaired image captioning via scene graph alignments,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:17.418253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:02.744274Z digest=sha256:82a0e91b720311c9ccc85dd5f5ae317be3c72f7061fd8ea01e1e138b216eb2c0

Observation 46c4d1ae-7e1b-40dd-b271-1061f3b29c61 · outbound

This paper cites On the role of scene graphs in image captioning,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation On the role of scene graphs in image captioning,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:17.242454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:02.813653Z digest=sha256:98e141fc00a43e7e63106b53ec59320a449bb104fa26e20f86c71c7f9e890a05

Observation d6126e6f-d065-4feb-b5c5-1032d596110b · outbound

This paper cites In defense of scene graphs for image captioning,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation In defense of scene graphs for image captioning,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:17.067575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:02.891163Z digest=sha256:8c6331136f1aac28477bd68f0c519b0dc3f1dc3cf528f15c100f93b5d9dceb11

Observation 712ef79c-afe3-406b-8a35-bbf77aa6428a · outbound

This paper cites Graph-structured representa- tions for visual question answering,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Graph-structured representa- tions for visual question answering,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:16.913736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:03.013076Z digest=sha256:1d3e75d6af2bee26b541b8ed9c5174da77eead8f12b8696da552f0c1606b8baf

Observation f8bd5f37-424c-4cda-bc51-95710d36ebea · outbound

This paper cites Lightweight visual question answering using scene graphs,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Lightweight visual question answering using scene graphs,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:16.784108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:03.149625Z digest=sha256:53fe2bf2d3e378fd15951d73f97736f7fc27846c0ff77c9dbe97b500be3c2023

Observation 0c0e1d6e-1ca4-412e-a57f-85b5bfa87696 · outbound

This paper cites Robotvqa - A scene-graph- and deep-learning-based visual question answering system for robot manipulation,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Robotvqa - A scene-graph- and deep-learning-based visual question answering system for robot manipulation,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:16.640986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:03.247905Z digest=sha256:d819dff78d38ac8117adfc64bd812790386c57c82924381d8099e0d66764e66a

Observation 6861e06a-199b-471a-b2b2-f093bb79c389 · outbound

This paper cites Visual question answering over scene graph,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Visual question answering over scene graph,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:16.492477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:03.343070Z digest=sha256:81b70934a478562bf8a91aadc9e57f329f297994332f95b3b8f9cc618155ad50

Observation 910f75e1-89e1-4fb6-a5b8-9f64030b3953 · outbound

This paper cites Image generation from scene graphs,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Image generation from scene graphs,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:16.178687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:03.474269Z digest=sha256:e39f36530860222708bc073b3c44aed4efd771cf0369d3af15f95b414886a6ac

Observation 44f50c3c-ef0f-4bad-a5ed-2cbb13c45c3b · outbound

This paper cites Diffusion-Based Scene Graph to Image Generation with Masked Contrastive Pre-Training.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Diffusion-Based Scene Graph to Image Generation with Masked Contrastive Pre-Training

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T14:06:03.652422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:06:03.652422Z digest=sha256:3454ebea4073b13baa5fe09ea067861e2d2d136cae2404738f5938fae15e38d6

Observation a7cb4a77-d9b7-4fb0-9725-bae17d463b84 · outbound

This paper cites Hierarchical Open-Vocabulary 3D Scene Graphs for Language-Grounded Robot Navigation.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Hierarchical Open-Vocabulary 3D Scene Graphs for Language-Grounded Robot Navigation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:06:03.814090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:06:03.814090Z digest=sha256:c574fa53bbc144e24c4229a367c1c01598df13133051f2e1f26d2295e9c602dd

Observation 7d4e668a-00db-4e27-869e-1692add75a64 · outbound

This paper cites SG-Nav: Online 3D Scene Graph Prompting for LLM-based Zero-shot Object Navigation.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation SG-Nav: Online 3D Scene Graph Prompting for LLM-based Zero-shot Object Navigation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T14:06:03.978487Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:06:03.978487Z digest=sha256:fee76400f1506ba1c0de73a3353932087ec441048f21c776698c610768891cf4

Observation bcceb78d-83be-4c1e-b2f9-841f4a6d43fc · outbound

This paper cites Scenegraphloc: Cross-modal coarse visual localization on 3d scene graphs,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Scenegraphloc: Cross-modal coarse visual localization on 3d scene graphs,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:15.875772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:04.132246Z digest=sha256:b706cb2068425571ea118c7ce10ad0f7e116aa9620fd1b6abe0f95f7bdb86206

Observation 719b043d-76b2-40c5-9fca-a47039705d96 · outbound

This paper cites Learning to generate scene graph from natural language supervision,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Learning to generate scene graph from natural language supervision,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:15.568533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:04.264776Z digest=sha256:15e7f2d0405077a0ff19111539dde98d787260402a996aff81a5d1ab96c049cf

Observation 13574a47-022e-49f5-9acd-662eaf112166 · outbound

This paper cites Integrating object-aware and interaction-aware knowledge for weakly supervised scene graph generation,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Integrating object-aware and interaction-aware knowledge for weakly supervised scene graph generation,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:15.328480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:04.360178Z digest=sha256:f50cf9d0326ebf746ecb3f7d5a14d6ca9fb8092361b45d9aa129fc305e685703

Observation 8f3bcb98-2e3b-4011-a13b-2d42917502c0 · outbound

This paper cites GPT4SGG: Synthesizing Scene Graphs from Holistic and Region-specific Narratives.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation GPT4SGG: Synthesizing Scene Graphs from Holistic and Region-specific Narratives

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:06:04.444777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:06:04.444777Z digest=sha256:52125d5827af9afd21b7779893ba00db6e15de75c8f8662fcd2f35a608d01c93

Observation 75f732c1-65c1-4422-93cc-875d30e8a8cb · outbound

This paper cites Open-vocabulary object detection using captions,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Open-vocabulary object detection using captions,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:15.118860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:04.521096Z digest=sha256:d1dd1d51aabfc3fb1c3842ee86889661db82dd2c45eba6f9df7fec4ddb900010

Observation 6daa01e2-ff05-4180-a8c1-a59046716e65 · outbound

This paper cites Aligning bag of regions for open-vocabulary object detection,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Aligning bag of regions for open-vocabulary object detection,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:14.978269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:04.611111Z digest=sha256:714d9b177ac133969a425f9bd0b51eefb4e2b39d2f7306d88d9a01fa690f724e

Observation f4b526ed-a1fb-4ca8-8971-ff943c67c27d · outbound

This paper cites Grounded language-image pre-training,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Grounded language-image pre-training,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:14.842015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:04.713894Z digest=sha256:5cc6fc8047800ff04d6af0228c22006c450ab38010028ec28a59464c875b165f

Observation 7c26f606-969e-48a4-9a84-85ae3be47776 · outbound

This paper cites Regionclip: Region-based language- image pretraining,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Regionclip: Region-based language- image pretraining,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:14.708747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:04.778671Z digest=sha256:3aceaf297dc0c517576c6cecc891a86e5122953c3d3d7f3a7e26088cac7ab4e2

Observation e5a12c88-11f3-40be-bd70-089f3c99d99d · outbound

This paper cites Learning to prompt for open-vocabulary object detection with vision-language model,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Learning to prompt for open-vocabulary object detection with vision-language model,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:14.563025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:04.865742Z digest=sha256:d7453be631918bb59b44956be2007497d1eea89880810196d723e365897322cc

Observation 5f143c4c-2a2f-447e-8384-30f7816d52a5 · outbound

This paper cites Scene graph parser,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Scene graph parser,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:14.419755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:04.975100Z digest=sha256:b708b5e5a960d01685691d27cd809196870ec8a5fcdb08c7c53915ee84652ac6

Observation 68bca05c-7128-4e54-a4e4-cf1a1aeb4dc5 · outbound

This paper cites Scene graph generation from objects, phrases and region captions,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Scene graph generation from objects, phrases and region captions,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:14.276576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:05.098519Z digest=sha256:ed973ea1794c90f770a3fc885efdcd6c9fb11570bb77f59f0d41d7f238f05590

Observation 7c904e98-17d2-45c8-9fa0-158f3aac65a6 · outbound

This paper cites Knowledge-embedded routing network for scene graph generation,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Knowledge-embedded routing network for scene graph generation,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:18.419833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:05.176216Z digest=sha256:aeecb1c79793fef0c64c6c2cec4459ec8ee4a65e75b5af023a1e768ccbfb414f

Observation d04a9a2d-7d63-405e-8f4c-2f942e33f315 · outbound

This paper cites Sgtr: End-to-end scene graph generation with transformer,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Sgtr: End-to-end scene graph generation with transformer,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:14.156470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:05.302846Z digest=sha256:7fa6db5918f4048772c0e0bde87973c815e1cd7e14b51cf42b4e53528b48208a

Observation fe14bb51-9adb-4ce5-bf79-a1f560062aab · outbound

This paper cites Iterative scene graph generation,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Iterative scene graph generation,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:14.009172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:05.397991Z digest=sha256:05fc06e16b96ade4fe9d461b58bd83c864b038aaff306f1da51f346bd66adebf

Observation d437da9e-d945-45b0-af75-5a85e280fe39 · outbound

This paper cites Reltr: Relation transformer for scene graph generation,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Reltr: Relation transformer for scene graph generation,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:13.910622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:05.483722Z digest=sha256:dcb4712b50454ca8f97d37eb422fd8e18382728bdf56aa91cad9899eb66e4ae6

Observation e935d187-0f34-43c2-aec2-846b981d37be · outbound

This paper cites Unbiased scene graph generation via two-stage causal modeling,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Unbiased scene graph generation via two-stage causal modeling,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:13.780426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:05.570808Z digest=sha256:414e6673fa23468df1820cc9f538dd9a8695eb51a872ff9a006fd7e508985b89

Observation 7e92f33b-87f3-45a3-9090-344ff2f68c8e · outbound

This paper cites Fast contextual scene graph generation with unbiased context augmentation,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Fast contextual scene graph generation with unbiased context augmentation,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:13.625301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:05.732333Z digest=sha256:4a98d8c28b5049fe73d966696256eef4eff8327756ef487013cbbf10ac597adc

Observation c1a8d885-c195-465f-8adf-40ba9bc32073 · outbound

This paper cites Semantic diversity-aware prototype-based learning for unbiased scene graph generation,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Semantic diversity-aware prototype-based learning for unbiased scene graph generation,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:13.494396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:05.861402Z digest=sha256:4c5128d150f77641c073e2a7efe2b1bf83aa98e52c169adc081eb145108775be

Observation 4dc38357-9ffa-4291-8bab-ddb47d7f6e35 · outbound

This paper cites Faster r-cnn: Towards real-time object detection with region proposal networks,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Faster r-cnn: Towards real-time object detection with region proposal networks,

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T14:06:05.948903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:06:05.948903Z digest=sha256:3a52ff3f080454a33dc98e795d1ee19d7149ed6b14a8ff15a5c4a2044d140194

Observation 9212e703-b1af-4a6b-bfc1-e00ae40ceed4 · outbound

This paper cites What Makes a Scene ? Scene Graph-based Evaluation and Feedback for Controllable Generation.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation What Makes a Scene ? Scene Graph-based Evaluation and Feedback for Controllable Generation

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T14:06:06.096668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:06:06.096668Z digest=sha256:5216dba9a5146327f3e6d114b0bd9da184defa53ec1ed81b905ff7f0887f6485

Observation daccf754-f33f-47b7-ba14-4939ac8e0264 · outbound

This paper cites Learning transferable visual models from natural language supervi- sion,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Learning transferable visual models from natural language supervi- sion,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:13.376646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:06.205169Z digest=sha256:98941fcff83793a48fa6af97e8ba46ecc3ba561897c3209ce3c0c030fbee3bfe

Observation ee357b2c-8fa4-48b8-a7f6-5ea001e9b4ae · outbound

This paper cites Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T14:06:06.308339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:06:06.308339Z digest=sha256:444d0e243df83bb146cc8edfd89ddfa1be8f6885f2e511a053b64f171951cc76

Observation 1ef1ab04-0b6d-42cb-a269-e4804826c615 · outbound

This paper cites Microsoft COCO Captions: Data Collection and Evaluation Server.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Microsoft COCO Captions: Data Collection and Evaluation Server

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T14:06:06.382309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:06:06.382309Z digest=sha256:c0d3c14e9c1c61779a5c56c167140bb7105704ee71538944989aee3b0aaeacd4

Observation df0a26ef-d565-4d4c-96c1-9d6182578379 · outbound

This paper cites Open-vocabulary object detection via vision and language knowledge distillation,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Open-vocabulary object detection via vision and language knowledge distillation,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:13.251945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:06.495719Z digest=sha256:63b5227fd5b553e2c64a2c18e0cd2a1f26cc679e14fa0b65dd6a60aacdbd9810

Observation ca382652-816d-4088-acf0-ae1b707b2be6 · outbound

This paper cites Scaling open-vocabulary image segmentation with image-level labels,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Scaling open-vocabulary image segmentation with image-level labels,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:13.088490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:06.578336Z digest=sha256:81f79f3126732afb2dec3d053133c2889f935d60b0edf00a0df0381dec1fd223

Observation ca03ad4d-a30e-4554-a856-6b3bd56dea65 · outbound

This paper cites ActionCLIP: A New Paradigm for Video Action Recognition.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation ActionCLIP: A New Paradigm for Video Action Recognition

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T14:06:06.655155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:06:06.655155Z digest=sha256:835e0067a7ea34b64788689bb4c8ecc7ce0df7dc10809ffe2a4732f2dd929e48

Observation b654f10c-9127-4f0b-b57d-e80deea9cac9 · outbound

This paper cites From pixels to graphs: Open-vocabulary scene graph generation with vision-language models,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation From pixels to graphs: Open-vocabulary scene graph generation with vision-language models,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:12.911201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:06.769749Z digest=sha256:2f66cadcdc855b91a05bd0b744f97f9d070bbc2a57691665c4f6f9f1c61b5231

Observation 9e5bd9b7-faed-407c-8ef9-4640feb433a9 · outbound

This paper cites Towards open vocabulary learning: A survey,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Towards open vocabulary learning: A survey,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:12.750669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:06.867087Z digest=sha256:af602f600b3d73af257e89efac6f6661db278c445c79a0bf869ea6f068c64e44

Observation b10b6357-aa9f-44c0-b104-27dd91218382 · outbound

This paper cites A Survey on Open-Vocabulary Detection and Segmentation: Past, Present, and Future.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation A Survey on Open-Vocabulary Detection and Segmentation: Past, Present, and Future

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T14:06:06.939775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:06:06.939775Z digest=sha256:980d929f15257859a0968c28581dc13f263b5a784e44556e804fd2d75e5e6ce8

Observation 706aa460-dc7a-4c11-b99d-3eef5e3ff097 · outbound

This paper cites LLM4SGG: Large language models for weakly supervised scene graph generation,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation LLM4SGG: Large language models for weakly supervised scene graph generation,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:12.565694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:07.025428Z digest=sha256:4843f2ebff7c349f0f268df6c33c3d53688d3c157057abe35b3c7a860c92ca7b

Observation cb899a64-2c3e-4981-a4a7-d4f20eb76682 · outbound

This paper cites GPT-4v(ision) System Card,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation GPT-4v(ision) System Card,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:12.359205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:07.157365Z digest=sha256:ef1abe1633e1e829c29b3da9cdedffdc21a60284f10a997248bea271662205d8

Observation 001b0fa4-b907-4ff9-9875-38fe4e234bd9 · outbound

This paper cites Expanding scene graph boundaries: fully open-vocabulary scene graph generation via visual-concept alignment and retention,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Expanding scene graph boundaries: fully open-vocabulary scene graph generation via visual-concept alignment and retention,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:12.102556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:07.233373Z digest=sha256:cda1bae6b51e30106f5d5c4d645796eebc17c4382c1fb23c2c6e0f65279528a9

Observation d40878fe-ea6a-4972-806f-235636f6ba4f · outbound

This paper cites Swin transformer: Hierarchical vision transformer using shifted windows,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Swin transformer: Hierarchical vision transformer using shifted windows,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:11.916722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:07.349389Z digest=sha256:6359655c8868bd162127d412fb438752a9a5553278e389f6ad9a3a7fcd3137c4

Observation 48a191a7-e43a-4b4b-8c9a-fc837c1bc116 · outbound

This paper cites BERT: pre-training of deep bidirectional transformers for language understanding,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation BERT: pre-training of deep bidirectional transformers for language understanding,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:11.731059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:07.436553Z digest=sha256:6c6da2b4073f5239d8231ed5ef16fe6153f318e6d94bc5da0f7b48ee07916f55

Observation 542926b9-d50a-43a2-859d-3864ecc202d8 · outbound

This paper cites Deformable DETR: deformable transformers for end-to-end object detection,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Deformable DETR: deformable transformers for end-to-end object detection,

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:11.532375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:07.557490Z digest=sha256:098c298ee9b5a8dbf0da5ec2103add7e12f7940221a29ecf0afcc2ed714e3812

Observation dc89014e-d49d-4cdb-8723-9b15616533ac · outbound

This paper cites Generalized intersection over union: A metric and a loss for bounding box regression,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Generalized intersection over union: A metric and a loss for bounding box regression,

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:11.333322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:07.672161Z digest=sha256:52badbecf19d669ec3232e4d06b2d5f5930aa1710f101850b5138de903564020

Observation 705cdc96-3f30-4e69-becb-f9b99d6bc8ef · outbound

This paper cites Focal loss for dense object detection,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Focal loss for dense object detection,

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:11.190363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:07.775201Z digest=sha256:1975aa68c47f4031c9c0381700c40915aa63f21957b45f71aada67b82d377427

Observation 069e383a-f387-4379-8edc-15c9d4d56a77 · outbound

This paper cites GPT-4 Technical Report.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation GPT-4 Technical Report

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T14:06:07.946971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:06:07.946971Z digest=sha256:68d26af1341bda5fc3f639acb86b01367b51300d4c240f41d1b35099e5ecd2bf

Observation 5d455cc7-4ad7-4b5a-bb31-c054fd9ef4b3 · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T14:06:08.035323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:06:08.035323Z digest=sha256:dbefe504856f37ce9dfd85021e7686abe4bdd3b7d1ba8e671455f64551727ccc

Observation 01046d1e-206e-46af-8ca5-23519cf25bc6 · outbound

This paper cites Gqa: A new dataset for real-world visual reasoning and compositional question answering,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Gqa: A new dataset for real-world visual reasoning and compositional question answering,

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:10.982766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:08.143995Z digest=sha256:ff4786a5384b6b2b38f0ea19bc4fe123b8cba49d9a05404b5282401d36adc691

Observation 1abba453-9b9d-4aaa-81cb-75ed808a26e4 · outbound

This paper cites Visual genome: Connecting language and vision using crowdsourced dense image annotations,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Visual genome: Connecting language and vision using crowdsourced dense image annotations,

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:10.820631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:08.245435Z digest=sha256:2117cd91c4c58fafd87eb33630b2447889bb832c010b9661d72d4cff1bbac54f

Observation a559b35e-0c62-4143-8680-c82ffe75e257 · outbound

This paper cites Stacked hybrid-attention and group collaborative learning for unbiased scene graph generation,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Stacked hybrid-attention and group collaborative learning for unbiased scene graph generation,

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:10.625054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:08.332388Z digest=sha256:87d2a726b0e15722ab639ad5ade768921372893bf9f752f2f982398e984234c5

Observation aa9ed081-3f84-4f1c-adbd-75939cc1d995 · outbound

This paper cites Vision relation transformer for unbiased scene graph generation,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Vision relation transformer for unbiased scene graph generation,

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:10.473040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:08.435412Z digest=sha256:e8571613c9c2566b24ce99a38250bdbdbd6965ffb69d4e5053a76c420608310b

Observation 51c3d11c-4363-4cc2-b94a-0a18b9469057 · outbound

This paper cites Decoupled weight decay regularization,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Decoupled weight decay regularization,

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T14:06:08.532477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:06:08.532477Z digest=sha256:4dcade2e0881385e15fff65b83c2e5bbeae8a4b7cbd966d1b4a74ddecab13f2c

Observation 18ad3b4f-4909-4273-97cf-ff2589428070 · outbound

This paper cites Linguistic structures as weak supervision for visual scene graph generation,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Linguistic structures as weak supervision for visual scene graph generation,

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:10.232457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:08.611359Z digest=sha256:a385aa4a3e9c278a4c18e7a8d1c7b917e5b36b96dda20423a4b2950a92f0be20

Observation 09da9128-4e40-436e-ab37-ceeb3caffa33 · outbound

This paper cites UNITER: universal image-text representation learning,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation UNITER: universal image-text representation learning,

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:10.006059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:08.685373Z digest=sha256:35fa2b8e590cce03a8005bdd8836ab28008a249cc5d8a1e0aa3f44d359353db0

Observation 86803cdc-bbcb-42d1-a26d-7f036b724630 · outbound

This paper cites Hl-net: Heterophily learning network for scene graph generation,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Hl-net: Heterophily learning network for scene graph generation,

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:09.840428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:08.795013Z digest=sha256:6c965916423ab4a61dcf2256c094ce8162702b6724a142557b106fcee34cca10

Observation b077492e-ffa8-4623-bdff-4eda17086f8c · outbound

This paper cites Fully convolutional scene graph generation,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Fully convolutional scene graph generation,

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:09.615403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:08.882042Z digest=sha256:f00cf6b69cc8470eae4c50a07f12cc923bed15f00e4e5f572e8c41b96533f247

Observation c1546d6c-a2e4-4259-bcaa-9eadf4f946b5 · outbound

This paper cites Leveraging predicate and triplet learning for scene graph generation,.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Leveraging predicate and triplet learning for scene graph generation,

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:06:09.449308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:06:08.966290Z digest=sha256:3e364253de812ac2f42917426513145d4ab40a29621a113a033bca2fe4e449ce

Observation 69142701-1c91-4690-a311-5b20cac60090 · outbound

This paper cites Visualizing data using t-sne.

From Data to Modeling: Fully Open-vocabulary Scene Graph Generation Visualizing data using t-sne

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-07T14:06:09.067431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:06:09.067431Z digest=sha256:72052c675d3ca17be3d98caa30ae43e355f9f72b857b599ecfcac8c8dc63cd01

Pith citing papers

Observation 6c41ff57-1027-4708-b114-5749d44553c7 · inbound

MORN: Metacognitive Object-Goal Regulation for Resource-Rational Long-Horizon Navigation cites this paper.

MORN: Metacognitive Object-Goal Regulation for Resource-Rational Long-Horizon Navigation From Data to Modeling: Fully Open-vocabulary Scene Graph Generation

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-19T20:27:49.619298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T20:26:59.407002Z digest=sha256:900932844d9118e634ff4522bc5a581b239819a90165650ad55d88441bf94886

Observation a4efc673-b638-4916-9f49-21b1e35a1f6d · inbound

PhysScene: A Scene Graph Dataset for Scientific Visual Reasoning in Physics Experiments cites this paper.

PhysScene: A Scene Graph Dataset for Scientific Visual Reasoning in Physics Experiments From Data to Modeling: Fully Open-vocabulary Scene Graph Generation

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-03T00:07:28.000970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T17:30:48.848277Z digest=sha256:80d64e35985d3756ef30c7b9f76593d8340335cc461ab9273b170b99e924c3f2