Pith. sign in

Paper Citation Record · LEDGER

A Fast and Accurate One-Stage Approach to Visual Grounding

As of 16 August 2026, this Paper Citation Record lists 55 of 55 outbound references and 0 inbound Pith citation observations for arXiv:1908.06354.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1908.06354 v1

Coverage vector

measured 55 of 55 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-14T12:55:11.572179Z

measured 55 of 55 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

55 of 55 outbound references displayed

  • verified exact0
  • verified fuzzy46
  • unresolved8
  • parse uncertain1
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1f26aae9-ba06-4f3c-9546-f6b31eecd4ba · outbound

This paper cites Bottom-up and top-down attention for image captioning and visual question answering.

A Fast and Accurate One-Stage Approach to Visual Grounding Bottom-up and top-down attention for image captioning and visual question answering

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:12.440663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.335642Z digest=sha256:4f4b59cfe5701d40ab854db539934eb56748c0e8b83cef513d0b40217cee2c19

Observation e313c44d-687f-46c9-87e2-5681344257cb · outbound

This paper cites Msrc: Multimodal spatial regression with semantic context for phrase grounding.

A Fast and Accurate One-Stage Approach to Visual Grounding Msrc: Multimodal spatial regression with semantic context for phrase grounding

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:12.426027Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.340299Z digest=sha256:335a33dfde6a38495ae5e490365f556d7f0bdb6e95eae8103154b0870a6e61b6

Observation 7f5ec710-b2f9-4891-997d-88a34ffdd60c · outbound

This paper cites Query-guided regression network with context policy for phrase ground- ing.

A Fast and Accurate One-Stage Approach to Visual Grounding Query-guided regression network with context policy for phrase ground- ing

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:12.396695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.348979Z digest=sha256:1704b25cc8f05d7f0d784705f32a0799194cad3d7f688b776402e364c43a0841

Observation 29c11b80-f0a4-4ab8-8d5d-ec15b3a29c07 · outbound

This paper cites BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding.

A Fast and Accurate One-Stage Approach to Visual Grounding BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-14T12:55:11.353344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T12:55:11.353344Z digest=sha256:5b75344099d33acb050be1572ecaf499082cfbd89aad1b8e1abbd861d9eb9478

Observation f9e55d75-ee69-4952-954c-2ad4c4a57dd7 · outbound

This paper cites Neural sequential phrase grounding (seqground).

A Fast and Accurate One-Stage Approach to Visual Grounding Neural sequential phrase grounding (seqground)

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:12.381932Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.357435Z digest=sha256:8cab33f0d369429926997971b9e1b078a85edea7fdd10ddbc62368866b15b640

Observation 8ad291f8-fb2f-4613-aab5-f888f5314bd2 · outbound

This paper cites The segmented and annotated iapr tc-12 benchmark.

A Fast and Accurate One-Stage Approach to Visual Grounding The segmented and annotated iapr tc-12 benchmark

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:12.366049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.361607Z digest=sha256:a85b2bb46369ae7a7572e8571e2369ebfee381a8ffe764924b0e2d613234fc87

Observation f6417af6-6172-4b19-a970-61fbe3ae60c8 · outbound

This paper cites The pascal visual object classes (voc) challenge.

A Fast and Accurate One-Stage Approach to Visual Grounding The pascal visual object classes (voc) challenge

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:12.352034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.365667Z digest=sha256:ed6b67f26e0e6ed03d045bf2609dc446ac3b8d3edb03cf7afefd05b33889e5ee

Observation 187aed7b-f426-4b90-bd66-208d7662b5df · outbound

This paper cites Unsupervised image captioning.

A Fast and Accurate One-Stage Approach to Visual Grounding Unsupervised image captioning

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:12.336192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.369425Z digest=sha256:a74add250b2e2385129783b57838a61976c8da51e25e556e2ab901c73eb6c687

Observation cb9af571-ebff-461d-9807-b6d3577a408d · outbound

This paper cites Vqs: Linking segmentations to questions and answers for supervised attention in vqa and question-focused semantic segmentation.

A Fast and Accurate One-Stage Approach to Visual Grounding Vqs: Linking segmentations to questions and answers for supervised attention in vqa and question-focused semantic segmentation

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:12.319128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.373289Z digest=sha256:c6cf553cb5cfe84a8f263e90ec9a4731a61c454ae552cd5b5a6dcf9bcae7d7be

Observation 9ecd650c-6ffa-4d03-826e-fe17c81a5ff3 · outbound

This paper cites Fast r-cnn.

A Fast and Accurate One-Stage Approach to Visual Grounding Fast r-cnn

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-14T12:55:11.377370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T12:55:11.377370Z digest=sha256:4b18ed96e225aa4a0728f53083d27733113e0faea34a7bdd36990614bbc93cae

Observation fd625d7d-c626-4e1d-b36e-316cf87e4f57 · outbound

This paper cites Mask r-cnn.

A Fast and Accurate One-Stage Approach to Visual Grounding Mask r-cnn

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-14T12:55:11.381665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T12:55:11.381665Z digest=sha256:111ab9afcb4a24a16d6b13782d2e34496e3b41662b2280ba145062e10b144d56

Observation 28fcd494-d121-4e93-9b04-ee4349de6536 · outbound

This paper cites Deep residual learning for image recognition.

A Fast and Accurate One-Stage Approach to Visual Grounding Deep residual learning for image recognition

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-14T12:55:11.385359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T12:55:11.385359Z digest=sha256:631c0029b650be0dee22c065f92c75c664c7cad96b9e9b4bf8811376421362a7

Observation 8fc88fee-3d8b-4eb8-ac10-3c0ab2f1a0dc · outbound

This paper cites Seg- mentation from natural language expressions.

A Fast and Accurate One-Stage Approach to Visual Grounding Seg- mentation from natural language expressions

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:12.260855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.388957Z digest=sha256:270263106713392066782a56cc61e6679de67bfda8cbd4b50267ce09fda9fe51

Observation 222ae0c5-03b3-49f2-8d1c-2f78c3652b4d · outbound

This paper cites Natural language object re- trieval.

A Fast and Accurate One-Stage Approach to Visual Grounding Natural language object re- trieval

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:12.239661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.392648Z digest=sha256:cf1bb1d263059d5c29fab21b9a8add371100c930d64d4880ee89eb55c4993bf8

Observation 84446249-5966-449f-809f-0f19c6a9ab14 · outbound

This paper cites Referitgame: Referring to objects in pho- tographs of natural scenes.

A Fast and Accurate One-Stage Approach to Visual Grounding Referitgame: Referring to objects in pho- tographs of natural scenes

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:12.226138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.396644Z digest=sha256:beef788b823eacb4137945043d23d9f3500493fc1c7b9caa9af68e03ef43061c

Observation 0b43fd43-e575-4cd5-bf66-4f24bb53abfe · outbound

This paper cites Deep attribute-preserving metric learning for natural language object retrieval.

A Fast and Accurate One-Stage Approach to Visual Grounding Deep attribute-preserving metric learning for natural language object retrieval

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:12.206003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.401304Z digest=sha256:a8536458d699d6fafc7790e9577e2520f803d927fcc5c972a1a0db251c120e0f

Observation 4c223974-84cf-482f-98ec-037c79ddb607 · outbound

This paper cites Tell-and-answer: Towards explainable visual question an- swering using attributes and captions.

A Fast and Accurate One-Stage Approach to Visual Grounding Tell-and-answer: Towards explainable visual question an- swering using attributes and captions

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:12.191684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.405066Z digest=sha256:779e5e021d5fcd1354b8d06be0abf2c0c827f82935d43cb0bb6b11c0bca5412b

Observation d18cad78-f01e-4560-a100-11f0c9032862 · outbound

This paper cites Feature pyramid networks for object detection.

A Fast and Accurate One-Stage Approach to Visual Grounding Feature pyramid networks for object detection

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:12.175903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.409428Z digest=sha256:c08f3e45a59866c8dd1e3c47ae6fb52fad9797f669a907ca4b373188f68e985a

Observation 6fac5db6-cba0-4b1e-a57a-dd20d40dcd04 · outbound

This paper cites Microsoft coco: Common objects in context.

A Fast and Accurate One-Stage Approach to Visual Grounding Microsoft coco: Common objects in context

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:12.160594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.413337Z digest=sha256:8e7e59b6c2f779db5fd5eb6165dadc30888f590ad9506f5f723623db11f98288

Observation 4c25bc21-62a8-4614-b7b5-2201777f2f44 · outbound

This paper cites Recurrent multimodal interaction for refer- ring image segmentation.

A Fast and Accurate One-Stage Approach to Visual Grounding Recurrent multimodal interaction for refer- ring image segmentation

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:12.147129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.417176Z digest=sha256:7b68baab9019c146b3cad9753f2810a7e1c4b08385ea89c7c5abfa3c41af5fec

Observation 19c774fe-cff0-429a-8e74-44bc1e117a10 · outbound

This paper cites Ssd: Single shot multibox detector.

A Fast and Accurate One-Stage Approach to Visual Grounding Ssd: Single shot multibox detector

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:12.132601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.422367Z digest=sha256:1b29916ee70396967b2ceb71460af63a80d14131c355520f9388ef4679247da2

Observation 64f5e608-c7fc-4b1d-913a-78539321dca1 · outbound

This paper cites Improving referring expression grounding with cross-modal attention-guided erasing.

A Fast and Accurate One-Stage Approach to Visual Grounding Improving referring expression grounding with cross-modal attention-guided erasing

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:12.119086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.426931Z digest=sha256:113c69888051a2153d6a4468f3df1a65c099dab5c1959a455006e2216b6fb533

Observation a4fe9a4b-7cb3-4612-adee-495f3f231ec7 · outbound

This paper cites Comprehension- guided referring expressions.

A Fast and Accurate One-Stage Approach to Visual Grounding Comprehension- guided referring expressions

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:12.105696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.430923Z digest=sha256:4548fd12acc9200558f36cc2481b319ca982cfb23cf041a417208db971ad79ad

Observation 1941cfb0-e714-4f8f-91d3-048573786179 · outbound

This paper cites Generation and comprehension of unambiguous object descriptions.

A Fast and Accurate One-Stage Approach to Visual Grounding Generation and comprehension of unambiguous object descriptions

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:12.090608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.435330Z digest=sha256:7176e61025f2cbb22832b1ed60a8dc6f6e75f7dcbe273e74f1f97b62e0e9fd84

Observation 9f9cffbb-732e-48c1-ad5c-7ae37f593cf3 · outbound

This paper cites Dynamic multimodal instance segmentation guided by natural language queries.

A Fast and Accurate One-Stage Approach to Visual Grounding Dynamic multimodal instance segmentation guided by natural language queries

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:12.074576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.439580Z digest=sha256:d59ec19fdb105d3a521d479fd8a3aade4cc508d150cbd89abd1624606505a298

Observation 16ac2ed0-d065-45a1-8545-846d0f8d0415 · outbound

This paper cites Distributed representations of words and phrases and their compositionality.

A Fast and Accurate One-Stage Approach to Visual Grounding Distributed representations of words and phrases and their compositionality

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:12.056958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.443547Z digest=sha256:f688655b4ceb444788ee48055d058d0e0db9f31d1aca591dd3de66232adff28d

Observation 0a3d39b6-e267-4653-ba94-7f2ff6b1c7c3 · outbound

This paper cites Mod- eling context between objects for referring expression un- derstanding.

A Fast and Accurate One-Stage Approach to Visual Grounding Mod- eling context between objects for referring expression un- derstanding

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:12.009148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.447317Z digest=sha256:007f3aded7306f70497f7672c6aebd1a7a32b09fbd071beb0a984b326173fdc4

Observation 18a9c807-fa59-4e85-ac6f-192335e743ac · outbound

This paper cites Im- proving the fisher kernel for large-scale image classification.

A Fast and Accurate One-Stage Approach to Visual Grounding Im- proving the fisher kernel for large-scale image classification

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:11.992848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.452326Z digest=sha256:1ea6716ebdae57095355181ecbd08f294c3a3a35134924e54e528fb0cb17680c

Observation 96cb83fe-95ff-4af1-9aff-9515bfc4be25 · outbound

This paper cites Plummer, Paige Kordas, M.

A Fast and Accurate One-Stage Approach to Visual Grounding Plummer, Paige Kordas, M

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:11.978275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.455860Z digest=sha256:7d631d9df77b90b8d03d9a6a7f2db7afab86d63fddc39b7ce99d5bd277a30d37

Observation 66963325-f148-41e7-b2b3-8e02c5019c84 · outbound

This paper cites Flickr30k entities: Collecting region-to-phrase corre- spondences for richer image-to-sentence models.

A Fast and Accurate One-Stage Approach to Visual Grounding Flickr30k entities: Collecting region-to-phrase corre- spondences for richer image-to-sentence models

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:11.959536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.459668Z digest=sha256:e69f3c154f102f9169430a8f0ed36fd7a696758dc5e3caa9d24cf7dac5a996c1

Observation d3ffc6f8-1ef1-432a-bf14-8a900638afb2 · outbound

This paper cites 2, 3, 4, 5, 6.

A Fast and Accurate One-Stage Approach to Visual Grounding 2, 3, 4, 5, 6

Reference 31

Resolution
parse uncertain
raw_fallback, observed 2026-08-14T12:55:12.411425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.344484Z digest=sha256:484d7d94af53165b92fc4cc28f94fbc2c2efe5c68159ef94590579002e0375a3

Observation 9e487ade-84fb-47af-b04a-dee03f7e1bb5 · outbound

This paper cites You only look once: Unified, real-time object de- tection.

A Fast and Accurate One-Stage Approach to Visual Grounding You only look once: Unified, real-time object de- tection

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:11.944856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.464015Z digest=sha256:2a377e02f2a1900acf1e67d9acdad649bdbe2606eb0878baa4d9231033c16c65

Observation 9c23c48a-8026-4f80-a544-7eed6fe897f4 · outbound

This paper cites Yolo9000: better, faster, stronger.

A Fast and Accurate One-Stage Approach to Visual Grounding Yolo9000: better, faster, stronger

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-14T12:55:11.468168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T12:55:11.468168Z digest=sha256:2726de7a2ff0dba45400b17a6135057495e377d2797fc731f3327172f329e14d

Observation bfcea0a0-7568-47da-a7a9-f3d425d8c0fd · outbound

This paper cites YOLOv3: An Incremental Improvement.

A Fast and Accurate One-Stage Approach to Visual Grounding YOLOv3: An Incremental Improvement

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-14T12:55:11.472786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T12:55:11.472786Z digest=sha256:b3744568723900c9e3ce94bc09e2d6b6a27ebaf50f8898f0e5874a1d2c9f1074

Observation 80296cc4-afb9-4024-92b0-585812587c2f · outbound

This paper cites Faster r-cnn: Towards real-time object detection with region proposal networks.

A Fast and Accurate One-Stage Approach to Visual Grounding Faster r-cnn: Towards real-time object detection with region proposal networks

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:11.922562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.477553Z digest=sha256:70d95fc2018f438315b41925873460cb4bd9c3a3e0ac3e932e58a68889d36c17

Observation 8e3070fc-ac99-4c2b-90cf-5141e0268048 · outbound

This paper cites Grounding of textual phrases in images by reconstruction.

A Fast and Accurate One-Stage Approach to Visual Grounding Grounding of textual phrases in images by reconstruction

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:11.906492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.482659Z digest=sha256:e3d487ae4544c7a8a015f18eb853bbe280796516b096d69fdc42fd0df58ab230

Observation 67655938-7e62-4cb3-a054-765068d4ff21 · outbound

This paper cites Berg, and Li Fei-Fei.

A Fast and Accurate One-Stage Approach to Visual Grounding Berg, and Li Fei-Fei

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-14T12:55:11.486610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T12:55:11.486610Z digest=sha256:d215da361d5619b6fc24423a5eb36a71fbfe788678d176b658cd74f7852aaf57

Observation c0adf11b-18ea-4ee6-9647-9d95e1a1673f · outbound

This paper cites Faster r-cnn features for instance search.

A Fast and Accurate One-Stage Approach to Visual Grounding Faster r-cnn features for instance search

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:11.880617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.491009Z digest=sha256:c8d511ea1687627b35ee350267f67f34c89e1f80d3c805d54f4a42dee169f09a

Observation 44ba2670-6622-49e1-9e43-d2a69daaab53 · outbound

This paper cites Very Deep Convolutional Networks for Large-Scale Image Recognition.

A Fast and Accurate One-Stage Approach to Visual Grounding Very Deep Convolutional Networks for Large-Scale Image Recognition

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-14T12:55:11.495992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T12:55:11.495992Z digest=sha256:2786f6f21b0fa3c2adbd8a8dfbc4753b9f4e9b9af4dbbad5b14b9bbe33b30563

Observation 40ea66b0-507d-4b0a-8047-4fb1352390eb · outbound

This paper cites Parsing with compositional vector grammars.

A Fast and Accurate One-Stage Approach to Visual Grounding Parsing with compositional vector grammars

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:11.867146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.500807Z digest=sha256:c98c0844921108846b62d672cfcce44c274d71467831440d238bf3ea2c26a238

Observation 0dfce907-d635-4953-a44a-9217be5e4aeb · outbound

This paper cites Lecture 6.5-rmsprop: Divide the gradient by a running average of its recent mag- nitude.

A Fast and Accurate One-Stage Approach to Visual Grounding Lecture 6.5-rmsprop: Divide the gradient by a running average of its recent mag- nitude

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:11.847627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.505236Z digest=sha256:0e70e49547e3927478e245fa6503ee9f19dfb5501e91000609a5cddb62632151

Observation 31881fdf-a29e-4f6d-b080-ece6dc0f7da6 · outbound

This paper cites Selective search for ob- ject recognition.

A Fast and Accurate One-Stage Approach to Visual Grounding Selective search for ob- ject recognition

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:11.831025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.510325Z digest=sha256:14581b789843ac44a516c86365e3d8ac167d9d8e49f1759254acdf08859cec1f

Observation 0e7ca21b-75fb-4315-8bc4-dd85f4d95cc7 · outbound

This paper cites Learning two-branch neural networks for image-text match- ing tasks.

A Fast and Accurate One-Stage Approach to Visual Grounding Learning two-branch neural networks for image-text match- ing tasks

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:11.817024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.514598Z digest=sha256:fecbba5a13a09e0009599ee26242d22e3f8926a78051114ba1f70956b38084a9

Observation c037a184-0240-48ed-a8de-68f1429fe6bb · outbound

This paper cites Learning deep structure-preserving image-text embeddings.

A Fast and Accurate One-Stage Approach to Visual Grounding Learning deep structure-preserving image-text embeddings

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:11.804316Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.519533Z digest=sha256:8ee8ae9a3d4171b3e44852889d797ffc38a8f6ac5285261c1a304e90f524df3e

Observation c8ab59d5-af30-4df5-a47a-acb6405928f1 · outbound

This paper cites Interpretable and globally optimal pre- diction for textual grounding using image concepts.

A Fast and Accurate One-Stage Approach to Visual Grounding Interpretable and globally optimal pre- diction for textual grounding using image concepts

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:11.791462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.524365Z digest=sha256:c02cd9dbc18239f1c797406f1e6b63419a657404337f9fdb7253b1521a3cf0ea

Observation 695bafff-5894-4ec9-9066-06c88695cfda · outbound

This paper cites Image captioning with semantic attention.

A Fast and Accurate One-Stage Approach to Visual Grounding Image captioning with semantic attention

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:11.777497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.529345Z digest=sha256:52f0e6e4383977df35417e645320faf3d9aab839e84a987c0619413698e1679f

Observation 3431b645-d117-4ca9-85ab-5a477d600006 · outbound

This paper cites From image descriptions to visual denotations: New similarity metrics for semantic inference over event descrip- tions.

A Fast and Accurate One-Stage Approach to Visual Grounding From image descriptions to visual denotations: New similarity metrics for semantic inference over event descrip- tions

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:11.763882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.533648Z digest=sha256:c7a4964f6c876afaa6123d71489cb0b3507366fb5b823ea8c10664986fbf09f7

Observation 73927b3c-a610-4203-b97c-486d57ce6b54 · outbound

This paper cites Mattnet: Modular atten- tion network for referring expression comprehension.

A Fast and Accurate One-Stage Approach to Visual Grounding Mattnet: Modular atten- tion network for referring expression comprehension

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:11.750577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.539085Z digest=sha256:b48cda3040345fc230e0545fd61626437a6dd9fa5a9c79e377645604bf6ed0ea

Observation db96d21b-8c94-4b6d-9a64-8392ebff8f11 · outbound

This paper cites Modeling context in referring expres- sions.

A Fast and Accurate One-Stage Approach to Visual Grounding Modeling context in referring expres- sions

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:11.737131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.543205Z digest=sha256:dd7f5da9fc549846af7497c392b27c881234e1aa9d516d37b390a4e4c8c884d0

Observation e6ca58b2-1d9f-4ede-93bc-5e407200082d · outbound

This paper cites A joint speaker-listener-reinforcer model for referring expres- sions.

A Fast and Accurate One-Stage Approach to Visual Grounding A joint speaker-listener-reinforcer model for referring expres- sions

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:11.723865Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.550258Z digest=sha256:bc1aca765cbcb949fc486084d4bb66a99251572cbb88ca998b36036093570eea

Observation 03ca5762-0428-479b-bc44-6e35477e2fa6 · outbound

This paper cites Ground- ing referring expressions in images by variational context.

A Fast and Accurate One-Stage Approach to Visual Grounding Ground- ing referring expressions in images by variational context

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:11.708463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.554323Z digest=sha256:f70143d914af923452d2de3f7c69171c1d0d32ac04a7bb2ecc824ad349bab2bb

Observation f3832973-384d-4eb1-af96-7f5c0923950d · outbound

This paper cites Discriminative bimodal networks for visual localization and detection with natural language queries.

A Fast and Accurate One-Stage Approach to Visual Grounding Discriminative bimodal networks for visual localization and detection with natural language queries

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:11.692025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.559020Z digest=sha256:8da2695e4c5fd7eab2f313713cd9396aee5ae58452562dece7777fa7aede45e7

Observation 079f9950-04a7-4442-a54c-4f7731c5c365 · outbound

This paper cites Weakly supervised phrase localization with multi-scale anchored transformer network.

A Fast and Accurate One-Stage Approach to Visual Grounding Weakly supervised phrase localization with multi-scale anchored transformer network

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:11.677645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.563125Z digest=sha256:f58c18dc9d8891b45c3d82b95c6e7ebcbd12f5916fde5b1a1921de4893a344eb

Observation 4f2fbce2-108d-4305-a01b-b37e19756b22 · outbound

This paper cites Visual7w: Grounded question answering in images.

A Fast and Accurate One-Stage Approach to Visual Grounding Visual7w: Grounded question answering in images

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:11.663054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.567371Z digest=sha256:139e569d8b4e1abfc0a446bee3cb9088f06ade9d4ae250f7bd54f846008e4099

Observation 5b78738a-f77f-4c7d-913d-ab60b6af7688 · outbound

This paper cites testA” contains images with multiple people and “testB.

A Fast and Accurate One-Stage Approach to Visual Grounding testA” contains images with multiple people and “testB

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T12:55:11.648830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T12:55:11.572179Z digest=sha256:72139ae3284e3683a33dd3b0499515362644c0532181e2f52bd4835d309db380

Pith citing papers

No inbound Pith citation observations are available.