Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-14T12:55:11.572179Z
Paper Citation Record · LEDGER
As of 16 August 2026, this Paper Citation Record lists 55 of 55 outbound references and 0 inbound Pith citation observations for arXiv:1908.06354.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-14T12:55:11.572179Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
55 of 55 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 1f26aae9-ba06-4f3c-9546-f6b31eecd4ba · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Bottom-up and top-down attention for image captioning and visual question answering
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation e313c44d-687f-46c9-87e2-5681344257cb · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Msrc: Multimodal spatial regression with semantic context for phrase grounding
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 7f5ec710-b2f9-4891-997d-88a34ffdd60c · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Query-guided regression network with context policy for phrase ground- ing
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 29c11b80-f0a4-4ab8-8d5d-ec15b3a29c07 · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f9e55d75-ee69-4952-954c-2ad4c4a57dd7 · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Neural sequential phrase grounding (seqground)
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 8ad291f8-fb2f-4613-aab5-f888f5314bd2 · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding The segmented and annotated iapr tc-12 benchmark
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation f6417af6-6172-4b19-a970-61fbe3ae60c8 · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding The pascal visual object classes (voc) challenge
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 187aed7b-f426-4b90-bd66-208d7662b5df · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Unsupervised image captioning
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation cb9af571-ebff-461d-9807-b6d3577a408d · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Vqs: Linking segmentations to questions and answers for supervised attention in vqa and question-focused semantic segmentation
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 9ecd650c-6ffa-4d03-826e-fe17c81a5ff3 · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Fast r-cnn
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd625d7d-c626-4e1d-b36e-316cf87e4f57 · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Mask r-cnn
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 28fcd494-d121-4e93-9b04-ee4349de6536 · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Deep residual learning for image recognition
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8fc88fee-3d8b-4eb8-ac10-3c0ab2f1a0dc · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Seg- mentation from natural language expressions
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 222ae0c5-03b3-49f2-8d1c-2f78c3652b4d · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Natural language object re- trieval
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 84446249-5966-449f-809f-0f19c6a9ab14 · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Referitgame: Referring to objects in pho- tographs of natural scenes
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 0b43fd43-e575-4cd5-bf66-4f24bb53abfe · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Deep attribute-preserving metric learning for natural language object retrieval
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 4c223974-84cf-482f-98ec-037c79ddb607 · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Tell-and-answer: Towards explainable visual question an- swering using attributes and captions
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation d18cad78-f01e-4560-a100-11f0c9032862 · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Feature pyramid networks for object detection
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 6fac5db6-cba0-4b1e-a57a-dd20d40dcd04 · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Microsoft coco: Common objects in context
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 4c25bc21-62a8-4614-b7b5-2201777f2f44 · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Recurrent multimodal interaction for refer- ring image segmentation
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 19c774fe-cff0-429a-8e74-44bc1e117a10 · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Ssd: Single shot multibox detector
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 64f5e608-c7fc-4b1d-913a-78539321dca1 · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Improving referring expression grounding with cross-modal attention-guided erasing
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation a4fe9a4b-7cb3-4612-adee-495f3f231ec7 · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Comprehension- guided referring expressions
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 1941cfb0-e714-4f8f-91d3-048573786179 · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Generation and comprehension of unambiguous object descriptions
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 9f9cffbb-732e-48c1-ad5c-7ae37f593cf3 · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Dynamic multimodal instance segmentation guided by natural language queries
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 16ac2ed0-d065-45a1-8545-846d0f8d0415 · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Distributed representations of words and phrases and their compositionality
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 0a3d39b6-e267-4653-ba94-7f2ff6b1c7c3 · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Mod- eling context between objects for referring expression un- derstanding
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 18a9c807-fa59-4e85-ac6f-192335e743ac · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Im- proving the fisher kernel for large-scale image classification
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 96cb83fe-95ff-4af1-9aff-9515bfc4be25 · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Plummer, Paige Kordas, M
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 66963325-f148-41e7-b2b3-8e02c5019c84 · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Flickr30k entities: Collecting region-to-phrase corre- spondences for richer image-to-sentence models
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation d3ffc6f8-1ef1-432a-bf14-8a900638afb2 · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding 2, 3, 4, 5, 6
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 9e487ade-84fb-47af-b04a-dee03f7e1bb5 · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding You only look once: Unified, real-time object de- tection
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 9c23c48a-8026-4f80-a544-7eed6fe897f4 · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Yolo9000: better, faster, stronger
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bfcea0a0-7568-47da-a7a9-f3d425d8c0fd · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding YOLOv3: An Incremental Improvement
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 80296cc4-afb9-4024-92b0-585812587c2f · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Faster r-cnn: Towards real-time object detection with region proposal networks
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 8e3070fc-ac99-4c2b-90cf-5141e0268048 · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Grounding of textual phrases in images by reconstruction
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 67655938-7e62-4cb3-a054-765068d4ff21 · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Berg, and Li Fei-Fei
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0adf11b-18ea-4ee6-9647-9d95e1a1673f · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Faster r-cnn features for instance search
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 44ba2670-6622-49e1-9e43-d2a69daaab53 · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Very Deep Convolutional Networks for Large-Scale Image Recognition
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 40ea66b0-507d-4b0a-8047-4fb1352390eb · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Parsing with compositional vector grammars
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 0dfce907-d635-4953-a44a-9217be5e4aeb · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Lecture 6.5-rmsprop: Divide the gradient by a running average of its recent mag- nitude
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 31881fdf-a29e-4f6d-b080-ece6dc0f7da6 · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Selective search for ob- ject recognition
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 0e7ca21b-75fb-4315-8bc4-dd85f4d95cc7 · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Learning two-branch neural networks for image-text match- ing tasks
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation c037a184-0240-48ed-a8de-68f1429fe6bb · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Learning deep structure-preserving image-text embeddings
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation c8ab59d5-af30-4df5-a47a-acb6405928f1 · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Interpretable and globally optimal pre- diction for textual grounding using image concepts
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 695bafff-5894-4ec9-9066-06c88695cfda · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Image captioning with semantic attention
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 3431b645-d117-4ca9-85ab-5a477d600006 · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding From image descriptions to visual denotations: New similarity metrics for semantic inference over event descrip- tions
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 73927b3c-a610-4203-b97c-486d57ce6b54 · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Mattnet: Modular atten- tion network for referring expression comprehension
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation db96d21b-8c94-4b6d-9a64-8392ebff8f11 · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Modeling context in referring expres- sions
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation e6ca58b2-1d9f-4ede-93bc-5e407200082d · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding A joint speaker-listener-reinforcer model for referring expres- sions
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 03ca5762-0428-479b-bc44-6e35477e2fa6 · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Ground- ing referring expressions in images by variational context
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation f3832973-384d-4eb1-af96-7f5c0923950d · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Discriminative bimodal networks for visual localization and detection with natural language queries
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 079f9950-04a7-4442-a54c-4f7731c5c365 · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Weakly supervised phrase localization with multi-scale anchored transformer network
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 4f2fbce2-108d-4305-a01b-b37e19756b22 · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding Visual7w: Grounded question answering in images
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 5b78738a-f77f-4c7d-913d-ab60b6af7688 · outbound
A Fast and Accurate One-Stage Approach to Visual Grounding testA” contains images with multiple people and “testB
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
No inbound Pith citation observations are available.