Pith. sign in

Paper Citation Record · LEDGER

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model

As of 9 August 2026, this Paper Citation Record lists 67 of 67 outbound references and 0 inbound Pith citation observations for arXiv:2505.23358.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.23358 v1

Coverage vector

measured 67 of 67 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:52:46.278075Z

measured 67 of 67 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

67 of 67 outbound references displayed

  • verified exact8
  • verified fuzzy21
  • unresolved37
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e70c888e-fa71-4752-9764-9110e71fe033 · outbound

This paper cites Anderson, X.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Anderson, X

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:38.225153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:38.225153Z digest=sha256:4bbfe305e041e14d939fde4b5c8a171fd3f1e0896749761825e518c7b9aa119a

Observation 4888885c-d95a-475b-9753-6f939a93b7e2 · outbound

This paper cites Cornia, M.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Cornia, M

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:58.633849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:38.329252Z digest=sha256:b80c85f7ef3cdb841ae9087685069227c77ba99fe442e1544da00e33b1de70fc

Observation 1e9e79fd-4936-415a-bfc2-6d29ce0e0484 · outbound

This paper cites Vinyals, A.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Vinyals, A

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:58.403872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:38.400669Z digest=sha256:713f621adae578b25ddd6409ea4af4ff4289db0b91cc34c5ba2d2871ee54d8ec

Observation e7a65481-e849-446c-9b33-c414de9c311c · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:58.054236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:38.489602Z digest=sha256:298081e7c95ad4f11900d986ec8461ce42a1def57dab9e2d2db8e22ca1709125

Observation f9435839-2957-4c63-9835-47d5448e576c · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:57.704900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:38.569375Z digest=sha256:20b04842c6b0ccabc31c975cdd2e9d5085635bb08c9e906814aee9257b2a48d1

Observation f6af9ec9-d65a-4cc5-a68e-deaa09cf7908 · outbound

This paper cites Stefanini, M.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Stefanini, M

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:57.397582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:38.647327Z digest=sha256:2a5302d6fd82a68953202391340f58e4806948cdb85cc856fab06ba24d6218b0

Observation fef9e153-5601-4eaf-968a-29db34242249 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 7

Resolution
verified exact
doi, observed 2026-08-07T12:52:46.868814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:38.727922Z digest=sha256:8c97ab7fd7b171e7bd2c3054fb92a85b68f981101dfc130e623e06a85b83a12a

Observation 785ad6f3-02cf-40b6-ae2b-bfaed1889868 · outbound

This paper cites Multi-Modal Image Captioning for the Visually Impaired.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Multi-Modal Image Captioning for the Visually Impaired

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:52:48.959440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:38.796139Z digest=sha256:fef4b413eb2f01554eec5aa93d04a76722303e4a13974509f0ea8d3f3252a508

Observation 029b2ba3-8538-4874-a84c-66bbcea5383d · outbound

This paper cites Image Captioning as an Assistive Technology: Lessons Learned from VizWiz 2020 Challenge.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Image Captioning as an Assistive Technology: Lessons Learned from VizWiz 2020 Challenge

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:52:48.712342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:38.871685Z digest=sha256:93b2efeba210a6faeb65b26e3d9fd557f29d5f6b1472d7477632e6d467cc9b6b

Observation 75d3e471-50d6-41aa-b64f-e410ff65e797 · outbound

This paper cites Captioning Images Taken by People Who Are Blind.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Captioning Images Taken by People Who Are Blind

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:38.976502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:38.976502Z digest=sha256:daf2da144f8b270781af090b3784c9e5129116a9a56754589fd11d155b74d4cb

Observation f159e2fb-340f-4bf6-957e-dfbea85de435 · outbound

This paper cites Faurina, A.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Faurina, A

Reference 11

Resolution
verified exact
doi, observed 2026-08-07T12:52:46.561064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:39.078368Z digest=sha256:974ea64bf3af96e78d26095d72d1d3159730a3f5b4c36d1f0d5862dc2b4cf164

Observation ac40d6e8-8e45-4d00-b667-8e002468660b · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:57.078394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:39.151464Z digest=sha256:7bdce641f6801fd4f2bfc645970910301c660b016d8a0d3f74a8d6e53ef7169f

Observation be2580ea-5407-4093-a27d-0affbbdd62ab · outbound

This paper cites Nikiforova, T.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Nikiforova, T

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:56.773631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:39.232914Z digest=sha256:fe7d72d4460bd0e1bdb0ddf6357bda9e23fbb708435d3d775c3e0080b347037f

Observation dd4c3f21-c962-4f7f-b8a0-c02c1272136c · outbound

This paper cites Informative Image Captioning with External Sources of Information.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Informative Image Captioning with External Sources of Information

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:39.321919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:39.321919Z digest=sha256:1ed022e50dc417966c33bad1879eeffd69ab4c29fbaf70274ca321b2e168d38a

Observation 5d799102-8f3a-42d4-bca7-8b91f9706c54 · outbound

This paper cites NOC-REK: Novel Object Captioning with Retrieved Vocabulary from External Knowledge.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model NOC-REK: Novel Object Captioning with Retrieved Vocabulary from External Knowledge

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:39.390400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:39.390400Z digest=sha256:2f57d33d716b746905018660ffcf2b90eb2f40bd311338f18e3a69d63c58f364

Observation 5fd0f2c5-6556-4c46-9c11-51f0167ff03a · outbound

This paper cites Boosting Entity-aware Image Captioning with Multi-modal Knowledge Graph.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Boosting Entity-aware Image Captioning with Multi-modal Knowledge Graph

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:52:48.458488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:39.476170Z digest=sha256:278ee3426b814d87315694a6d6b09bda6bc1c866b8fa1f6f30a8565420a123cd

Observation 60d63913-eb1e-4af9-abd2-d5caaf9684ea · outbound

This paper cites Entity-aware Image Caption Generation.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Entity-aware Image Caption Generation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:39.553652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:39.553652Z digest=sha256:266d3224dc9f1a2b84fe5e6f112d2a5e2cb8c6d64ecd8262accb41781493825a

Observation bbb191e8-7261-4676-8baf-c2bd74b465ac · outbound

This paper cites EVCap: Retrieval-Augmented Image Captioning with External Visual-Name Memory for Open-World Comprehension.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model EVCap: Retrieval-Augmented Image Captioning with External Visual-Name Memory for Open-World Comprehension

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:52:48.254957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:39.641589Z digest=sha256:5ef2cd15abc7c5c74c0d5a2a5685c3c402fe7d7f4bac85b0ebac5540bc5702a2

Observation 87ec3b0d-016b-4d07-a39a-87c165e74893 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:56.473669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:39.752067Z digest=sha256:95be6a924b39c895499fd2c830bfb032e3a455ce75df017612c430028aab4e66

Observation 0b82bc37-5fdc-46fd-8889-aa1c7f230e41 · outbound

This paper cites Zhang, X.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Zhang, X

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:56.216551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:39.866012Z digest=sha256:fb590ae31b65255040624892f5b279c86b4fb3a86703513839a61032ab937d05

Observation 903d26c9-c6ec-40ce-8b66-e5489d9accc2 · outbound

This paper cites Ayesha, J.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Ayesha, J

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:55.875785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:40.022083Z digest=sha256:8dccef6db77374c97140ab4eb345729d23e9aae1b873b75861bbc8f078a70b91

Observation f931b5b8-6cc6-44b8-9613-0f4090991634 · outbound

This paper cites Joint Image Captioning and Question Answering.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Joint Image Captioning and Question Answering

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:52:48.040856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:40.177749Z digest=sha256:0e7facceddde0cd503380e7ed26066a522472660b2eab522f323b04434cdcd04

Observation 554effaf-be40-4e95-9c68-f3535c39c000 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:55.602175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:40.237589Z digest=sha256:75ad8a43ca3b15b7aee8bc3f8805ac34bb988c49e77066a31601a6a17ed3ff14

Observation 54812251-928e-4475-90ea-ce0238a75fe7 · outbound

This paper cites Salaberria, G.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Salaberria, G

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:55.352987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:40.385666Z digest=sha256:8a9d4bacc682717bd7a6148186e6e61e36796169bcf41d3479aba3b98b87d15a

Observation 7c0678fa-8fff-4c8a-b8d4-380ad916ee57 · outbound

This paper cites Enhancing Visual Question Answering through Question-Driven Image Captions as Prompts.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Enhancing Visual Question Answering through Question-Driven Image Captions as Prompts

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:52:47.869346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:40.470416Z digest=sha256:e451cc0527ccb6a091dfca47de71bf2571d474302ad9f79f3417820a5a25057f

Observation 62d4644b-8acb-453a-ba90-f045a5abbfc8 · outbound

This paper cites Image Captioning and Visual Question Answering Based on Attributes and External Knowledge.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Image Captioning and Visual Question Answering Based on Attributes and External Knowledge

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:40.597487Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:40.597487Z digest=sha256:c0c91af30b192ba6f6ff0a8675df257c2444d1aba08097414b9052a9f944d0e3

Observation 5c33a43f-692b-48c9-a3a5-187a92206d95 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:55.015213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:40.746781Z digest=sha256:8cde971883e5f762ef408fee6b2473237658436b86ba99ee4523ee0e4636a1da

Observation bdc05ffb-3341-4d4f-a6b6-b089287f8993 · outbound

This paper cites Whitehead, H.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Whitehead, H

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:54.736213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:40.912671Z digest=sha256:edf2881735ff10bb664563da84027c6dc0f5f5782f017b56cb553f97ad9b4eb3

Observation f3755080-71f4-45c2-9779-3bd814c6b63d · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:54.551892Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:41.044280Z digest=sha256:ec4be5bb03c987a3371d5837b61ad7af911ea5e83371eada323d351c252bd653

Observation 785807f6-025b-422a-9fd9-ac259b927aaa · outbound

This paper cites Radford, J.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Radford, J

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:54.357001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:41.156367Z digest=sha256:8ca860236100a41f2b52a9fa6394f8bc057f9f56b2a3a15ab4636b311d145272

Observation 9ca17ee5-e151-4c6a-83e3-e9bf56c682a4 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:54.188963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:41.242756Z digest=sha256:130e1c2c05513c54b811d358281c861df4a973137b53fb813220c144cd834d97

Observation 9bbf83b5-4d06-42c1-a951-73890c52ff38 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:53.939757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:41.347860Z digest=sha256:7777a19ab53cd1de1184b092a3213e81179cc7aa7efe635a0153f4489ddaa43f

Observation 77bd1637-6379-4ef3-8d25-928d3b36ff97 · outbound

This paper cites SimVLM: Simple Visual Language Model Pretraining with Weak Supervision.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model SimVLM: Simple Visual Language Model Pretraining with Weak Supervision

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:41.477507Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:41.477507Z digest=sha256:6143ba0d4b19658487cfc8f9832d95baf334b150964b6750fff6ab8db57d470b

Observation c5dbae6e-3d28-445d-a8a7-7d60e9ac6351 · outbound

This paper cites Zhang, X.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Zhang, X

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:53.718800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:41.620139Z digest=sha256:d3769d159e9fba497d4efbfe79009c7a446a4597f7f0b1cef175816572d8072b

Observation e21a4c6f-2011-4bb3-83f2-fca4edc42986 · outbound

This paper cites Cheng, W.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Cheng, W

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:41.737519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:41.737519Z digest=sha256:3a20481b56c385e3d35612510d793324486a301e6bf065301de757d5d3d1a201

Observation 63038f23-68e0-4cf7-b811-4a6d07d8714b · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:41.863756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:41.863756Z digest=sha256:f36bc36d248e69df6d21a161d557c89ff29b703cecb6a8d5c6a45f6d2a9d5545

Observation 9bf901f3-42d4-4796-b3f9-57dca0f93082 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 37

Resolution
metadata mismatch
raw_fallback, observed 2026-08-07T12:52:47.355468Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:41.935467Z digest=sha256:95c21bf4d37a942c9d48fabc20c06a5cf15699f1224e3eca818510d42a664971

Observation aa067ddc-360f-49ec-ac90-0e63bab9460a · outbound

This paper cites Cheng, W.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Cheng, W

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:42.077219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:42.077219Z digest=sha256:48c89f1aa57c004b84e07ba1c44e55beaf7c5bb796e6f496a4c1fd4c80084e82

Observation e94ec38d-69ec-4738-a967-b5faadbeec74 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:53.500630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:42.260304Z digest=sha256:b3bdeed23b6149fa0c9dd3188cce411aa4d8d2bd737f89e8222e0df0107eb74f

Observation bb68489d-5d11-41d8-9982-ea5f287d7e21 · outbound

This paper cites Zhang, Z.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Zhang, Z

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:53.308178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:42.414916Z digest=sha256:98973fc11f45ddd078bcd53e79b61a4d4b05cbdcd71ff1f33ca94aac700210bb

Observation 317e716f-d6f1-4b0d-b7e0-5c577690f75f · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:53.118429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:42.564442Z digest=sha256:b043a786919bc6c336665dbf3b5c4a6473cfd2446aa3498c9512cb497a267938

Observation ead8fbd5-99c9-4163-9a60-66e4898bc5d6 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:52.872374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:42.746129Z digest=sha256:3dffc903123eda9e0eb5d3e94a62a788d0fda91fbf224ab935d6e40ad4326323

Observation 2ad7f8fa-a2e5-4653-9fc5-a25fb201a9c9 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:52.726239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:42.902878Z digest=sha256:65960a715f0d9cb455a7fe94d42760fa149a1caa2f581910b5b6fc340b1ecb19

Observation a1497428-a809-49e1-ba46-d42f066e3c26 · outbound

This paper cites Huang, W.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Huang, W

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:52.516230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:43.080542Z digest=sha256:f10fd22ee6157e2fdc30050b16279320f6d77418c2a6e99c89240bf0a560a824

Observation 7eb8d239-1849-4376-b212-2f30817b08e1 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:43.291761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:43.291761Z digest=sha256:db326f72cf8dc6ea3a574b8dcaef6ba24394584d92e1b0fadf5e94ba99c30323

Observation 939a91d4-e703-4647-8e57-dbec012f5be4 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:52.328504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:43.376452Z digest=sha256:1cb9742e19015806ce0ba38bd9e1f63a95fb1d18c60e78f483ad9848f3dd2508

Observation 83882436-fe97-4f28-b95c-1947616b957c · outbound

This paper cites Devlin, M.-W.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Devlin, M.-W

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:52.113059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:43.502994Z digest=sha256:c009bf7aa132fb02e6be68c22bff6bff792dd951b4a74db32654cfc96459d98c

Observation 0c5efa97-c4d9-4a2a-a5ba-41674afba9bc · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:51.880237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:43.665941Z digest=sha256:1815e5be1276dbacc85077da74f04a761b99b7fb2049922a3065c1ed4fb615be

Observation 68447c61-88fc-49cb-8413-4300f61ddb0a · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:51.717387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:43.799953Z digest=sha256:34d4444ab46c0711cc2766740cc4791eed632876d03ed09e58d88b6107d7437c

Observation e97e983c-750b-499f-b0fe-c605f3f115da · outbound

This paper cites Radford, J.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Radford, J

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:51.485927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:43.943310Z digest=sha256:f4f64cd8c1ead5c6cebfc572b4e71cf5b68a79fd08580ead7ca62b15ea03a927

Observation ab95742f-af91-4837-a8be-418c7492f131 · outbound

This paper cites ClipCap: CLIP Prefix for Image Captioning.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model ClipCap: CLIP Prefix for Image Captioning

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:44.090056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:44.090056Z digest=sha256:5f23caaba1af924e0e16e1b3742a389556e36919aacda0b6e945fbac112a3261

Observation 583f908b-8132-4b1d-a22e-58bdab4df765 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:51.296324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:44.217497Z digest=sha256:a54de90f754fa9d378a2687ae7534f3ec2ca997c8f7b06323cc5071dd10fcdcb

Observation d6987114-ce4b-4866-b9c4-b3af430e460a · outbound

This paper cites GIT: A Generative Image-to-text Transformer for Vision and Language.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model GIT: A Generative Image-to-text Transformer for Vision and Language

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:44.360480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:44.360480Z digest=sha256:9d7bf6fe3e94d537477297aa452f5fd4588ec9a38f65644c07fc04f061e53db5

Observation f33e1d90-fa99-4a87-ba0d-ebb038698a5e · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:51.143213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:44.524819Z digest=sha256:6a930135e440b687a953f44640a368dec09af8cd89857594fac3b22d1f53f771

Observation 619ce196-557d-4234-a37b-8be37228eae5 · outbound

This paper cites 13035–13045.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model 13035–13045

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:50.963362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:44.675623Z digest=sha256:e2959e429cdd944d0d338cbb1c2bea9301e5bc4ea915094cc4432e1cbf88baf2

Observation 771492b5-9dd1-4e5a-908c-30499c358417 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:50.768170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:44.829181Z digest=sha256:7522508ea15653bdf857ea3fbe980abe4ab06099d42bad31118924f1f256fe66

Observation eaa26ee6-5c18-4998-ad25-3464b4440290 · outbound

This paper cites Changpinyo, P.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Changpinyo, P

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:44.950358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:44.950358Z digest=sha256:043883cfea1a6b298865f06140c2ec84c192db406a93ed09ae4e7a3c67820bb1

Observation 1faceee6-58c1-42c1-a747-6b0968ee5f17 · outbound

This paper cites Kullback, R.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Kullback, R

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:50.556760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:45.091862Z digest=sha256:378baece4b72002426d06495c941ce9596524f9b174fa183be24c32297320186

Observation a9da7bcf-96d6-4e97-aa27-d303f3179f4c · outbound

This paper cites Microsoft COCO Captions: Data Collection and Evaluation Server.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Microsoft COCO Captions: Data Collection and Evaluation Server

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:45.252289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:45.252289Z digest=sha256:fb5074756fdf24032fd298a59a782f7f3e4beff55fc18766afa6d61837d7418d

Observation 7632d89f-de66-4896-af43-b06f96587e43 · outbound

This paper cites Papineni, S.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Papineni, S

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:50.405269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:45.322179Z digest=sha256:955abdfbadb93ecfac638c0637b7504e0c220947175254e8c88ced3680778f5a

Observation 3504e205-b35a-4e11-a653-83c1f2407f11 · outbound

This paper cites Banerjee, A.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Banerjee, A

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:50.178964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:45.478971Z digest=sha256:8515639aac39b2f424117a1d9dd02538e3a6ffdea00b136888486bdf3d16eb90

Observation 1aea2639-4e4d-4a37-ba4e-f76a1944a962 · outbound

This paper cites Lin, Rouge: A package for automatic evaluation of summaries, in: Text summarization branches out: Proceedings of the ACL-04 workshop, 2004, pp.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Lin, Rouge: A package for automatic evaluation of summaries, in: Text summarization branches out: Proceedings of the ACL-04 workshop, 2004, pp

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:50.009788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:45.626619Z digest=sha256:a74e23c7888571abc8c02e8b86b814615edacf630a767ac969fcfb3171226a4d

Observation 9225c830-2bce-45e3-a697-43c60fdde696 · outbound

This paper cites Vedantam, C.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Vedantam, C

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:49.793273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:45.757986Z digest=sha256:27ae744b957a2370f24188de41056da932275274995da73f2f7b1fa0c6d99933

Observation 301c5f23-0b3f-4bc1-8635-bf8b1b7805f3 · outbound

This paper cites Kirkpatrick, R.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Kirkpatrick, R

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:49.605423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:45.859012Z digest=sha256:3c2bc54078856e81efe4d109211ab61f264643190eda77cab22ef9f44585602e

Observation 210021e2-202f-45f3-968e-251efdbd32d1 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 65

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:49.376017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:45.945542Z digest=sha256:e8f2218e5f700db7369420be35e5776cef259f95e7ba03cb56bb07e8d16cd0d2

Observation 57f827d0-533d-40e0-9a1d-80b3884f9dbc · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 66

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:49.146830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:52:46.100120Z digest=sha256:35b1f69bc9163f2e1a9ca8e99bb5ef66a688dcf39993cb7d145498498071b258

Observation e0bfc597-ddae-4ac7-9bf9-2bf7e9f83ac8 · outbound

This paper cites CLIP-Adapter: Better Vision-Language Models with Feature Adapters.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model CLIP-Adapter: Better Vision-Language Models with Feature Adapters

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:46.278075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:46.278075Z digest=sha256:76934ca470376c6b6618fe2678914fb1c06a42e72f16461bfab65bc27472fad4

Pith citing papers

No inbound Pith citation observations are available.