Pith. sign in

Paper Citation Record · LEDGER

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model

As of 8 August 2026, this Paper Citation Record lists 67 of 67 outbound references and 0 inbound Pith citation observations for arXiv:2505.23358.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.23358 v1

Coverage vector

measured 67 of 67 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:52:46.278075Z

measured 67 of 67 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

67 of 67 outbound references displayed

  • verified exact8
  • verified fuzzy21
  • unresolved37
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e70c888e-fa71-4752-9764-9110e71fe033 · outbound

This paper cites Anderson, X.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Anderson, X

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:38.225153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:38.225153Z digest=sha256:4bbfe305e041e14d939fde4b5c8a171fd3f1e0896749761825e518c7b9aa119a

Observation 4888885c-d95a-475b-9753-6f939a93b7e2 · outbound

This paper cites Cornia, M.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Cornia, M

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:58.633849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:38.329252Z digest=sha256:37605d08b045fc3164bce78b3c1452b501f95383753cfe706f56a5435d69291c

Observation 1e9e79fd-4936-415a-bfc2-6d29ce0e0484 · outbound

This paper cites Vinyals, A.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Vinyals, A

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:58.403872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:38.400669Z digest=sha256:19d81bd80adc5851cea04b26916f2206d1844ef04874d4458615f4229e77cc16

Observation e7a65481-e849-446c-9b33-c414de9c311c · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:58.054236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:38.489602Z digest=sha256:71e80a6407f3abe8d99ecbe3ec88c0b02c9c30415ad2e57c127fff4791cd0e5b

Observation f9435839-2957-4c63-9835-47d5448e576c · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:57.704900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:38.569375Z digest=sha256:41d17b63379b967eda912673542f6889cbc6a8caf38789ac10b2128dbe1346e1

Observation f6af9ec9-d65a-4cc5-a68e-deaa09cf7908 · outbound

This paper cites Stefanini, M.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Stefanini, M

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:57.397582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:38.647327Z digest=sha256:36ee5576b1b176af5233966ee55370dd35cb054ce255f2bc25a05c84c640730b

Observation fef9e153-5601-4eaf-968a-29db34242249 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 7

Resolution
verified exact
doi, observed 2026-08-07T12:52:46.868814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:38.727922Z digest=sha256:d12eba8924c16acec4917b7562361e6832f19133ccaaca52107446ee85b2e513

Observation 785ad6f3-02cf-40b6-ae2b-bfaed1889868 · outbound

This paper cites Multi-Modal Image Captioning for the Visually Impaired.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Multi-Modal Image Captioning for the Visually Impaired

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:52:48.959440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:38.796139Z digest=sha256:0ef554ebfd3308b15e2de11235322bd0d4ccb1e336e2cd310637f28564817926

Observation 029b2ba3-8538-4874-a84c-66bbcea5383d · outbound

This paper cites Image Captioning as an Assistive Technology: Lessons Learned from VizWiz 2020 Challenge.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Image Captioning as an Assistive Technology: Lessons Learned from VizWiz 2020 Challenge

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:52:48.712342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:38.871685Z digest=sha256:1c02be3dda077b76ef5526893075967979da01fb848f94efecd5772d8bdb9078

Observation 75d3e471-50d6-41aa-b64f-e410ff65e797 · outbound

This paper cites Captioning Images Taken by People Who Are Blind.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Captioning Images Taken by People Who Are Blind

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:38.976502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:38.976502Z digest=sha256:1136afc08ae662e5474198f7a79205d36f665d7e75c5b1c74c9bf6c0f5d50f99

Observation f159e2fb-340f-4bf6-957e-dfbea85de435 · outbound

This paper cites Faurina, A.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Faurina, A

Reference 11

Resolution
verified exact
doi, observed 2026-08-07T12:52:46.561064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:39.078368Z digest=sha256:98f16fce0cec1ae8d2370020ca6bd8d9024b3ed321b799ebbbb1850b5df8615d

Observation ac40d6e8-8e45-4d00-b667-8e002468660b · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:57.078394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:39.151464Z digest=sha256:c647a5174764e4de0c4828891266db4017915c612438602d376ede0917500b0e

Observation be2580ea-5407-4093-a27d-0affbbdd62ab · outbound

This paper cites Nikiforova, T.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Nikiforova, T

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:56.773631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:39.232914Z digest=sha256:ff891f28f293cda6add49b1efa47c852f42fb3b86c6bae64d7057e79d762fb76

Observation dd4c3f21-c962-4f7f-b8a0-c02c1272136c · outbound

This paper cites Informative Image Captioning with External Sources of Information.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Informative Image Captioning with External Sources of Information

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:39.321919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:39.321919Z digest=sha256:1ed022e50dc417966c33bad1879eeffd69ab4c29fbaf70274ca321b2e168d38a

Observation 5d799102-8f3a-42d4-bca7-8b91f9706c54 · outbound

This paper cites NOC-REK: Novel Object Captioning with Retrieved Vocabulary from External Knowledge.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model NOC-REK: Novel Object Captioning with Retrieved Vocabulary from External Knowledge

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:39.390400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:39.390400Z digest=sha256:2f57d33d716b746905018660ffcf2b90eb2f40bd311338f18e3a69d63c58f364

Observation 5fd0f2c5-6556-4c46-9c11-51f0167ff03a · outbound

This paper cites Boosting Entity-aware Image Captioning with Multi-modal Knowledge Graph.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Boosting Entity-aware Image Captioning with Multi-modal Knowledge Graph

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:52:48.458488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:39.476170Z digest=sha256:bff4123ad170afac4df78dbae374f3cf7a7d83ca0b53b31d72f0ef33f511120f

Observation 60d63913-eb1e-4af9-abd2-d5caaf9684ea · outbound

This paper cites Entity-aware Image Caption Generation.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Entity-aware Image Caption Generation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:39.553652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:39.553652Z digest=sha256:611875da367b77090d165c5aaa71ca1bea068e07a344dfdae2e8af3cb470b12f

Observation bbb191e8-7261-4676-8baf-c2bd74b465ac · outbound

This paper cites EVCap: Retrieval-Augmented Image Captioning with External Visual-Name Memory for Open-World Comprehension.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model EVCap: Retrieval-Augmented Image Captioning with External Visual-Name Memory for Open-World Comprehension

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:52:48.254957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:39.641589Z digest=sha256:1192d834c481ce58f126fbcc099195359490d7e200408ac6b13febafdedb4af4

Observation 87ec3b0d-016b-4d07-a39a-87c165e74893 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:56.473669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:39.752067Z digest=sha256:62c14db805ba82f7ca34590c7777f19717f2273e08f7cd00b5584ba4ab469c49

Observation 0b82bc37-5fdc-46fd-8889-aa1c7f230e41 · outbound

This paper cites Zhang, X.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Zhang, X

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:56.216551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:39.866012Z digest=sha256:b248c5247183a55647f51939a485f57e1111b1245cdca69c06fb5fdfb61a5dc4

Observation 903d26c9-c6ec-40ce-8b66-e5489d9accc2 · outbound

This paper cites Ayesha, J.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Ayesha, J

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:55.875785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:40.022083Z digest=sha256:151630c4c8caa84912c12c17f7f459a72e29596ae0a3f0828268e97a6c6e860c

Observation f931b5b8-6cc6-44b8-9613-0f4090991634 · outbound

This paper cites Joint Image Captioning and Question Answering.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Joint Image Captioning and Question Answering

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:52:48.040856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:40.177749Z digest=sha256:0c62807bb1d3684d9aa2128be5aa009c64bf0316a85fe91c266a1734a0651444

Observation 554effaf-be40-4e95-9c68-f3535c39c000 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:55.602175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:40.237589Z digest=sha256:3f02a42412556841d5808a2f9da110f5338e0ea08c2b30cbed3946c1f7834dcf

Observation 54812251-928e-4475-90ea-ce0238a75fe7 · outbound

This paper cites Salaberria, G.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Salaberria, G

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:55.352987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:40.385666Z digest=sha256:0239e6b93536bb5f62deabdb2180a472cb724f8076386b8694ddce327f82a941

Observation 7c0678fa-8fff-4c8a-b8d4-380ad916ee57 · outbound

This paper cites Enhancing Visual Question Answering through Question-Driven Image Captions as Prompts.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Enhancing Visual Question Answering through Question-Driven Image Captions as Prompts

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:52:47.869346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:40.470416Z digest=sha256:bf01228f9887b4ec7eb0668ca09fde3f77eef5131f4be8cc8a50d28655d54434

Observation 62d4644b-8acb-453a-ba90-f045a5abbfc8 · outbound

This paper cites Image Captioning and Visual Question Answering Based on Attributes and External Knowledge.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Image Captioning and Visual Question Answering Based on Attributes and External Knowledge

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:40.597487Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:40.597487Z digest=sha256:c0c91af30b192ba6f6ff0a8675df257c2444d1aba08097414b9052a9f944d0e3

Observation 5c33a43f-692b-48c9-a3a5-187a92206d95 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:55.015213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:40.746781Z digest=sha256:a405851ddc00e3627252653139115ded65bcfae0dd7fd232ebed52bb3d330577

Observation bdc05ffb-3341-4d4f-a6b6-b089287f8993 · outbound

This paper cites Whitehead, H.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Whitehead, H

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:54.736213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:40.912671Z digest=sha256:53871d1d9958a04c0dc02b65bc2b0f40965bf423020b7bf366fe3948f6a373cf

Observation f3755080-71f4-45c2-9779-3bd814c6b63d · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:54.551892Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:41.044280Z digest=sha256:965f0490a3d726cb21344597f849bf6de04f29651c9e226bec0d635deb43bac3

Observation 785807f6-025b-422a-9fd9-ac259b927aaa · outbound

This paper cites Radford, J.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Radford, J

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:54.357001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:41.156367Z digest=sha256:3493c7cd7598a85bd01ad2e33ce7c067258dc02af22c063f10c78db658223f37

Observation 9ca17ee5-e151-4c6a-83e3-e9bf56c682a4 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:54.188963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:41.242756Z digest=sha256:497479aff3214bed86cf95f059cc60f585cdd4118f652b1b56eef041cfbc4d1a

Observation 9bbf83b5-4d06-42c1-a951-73890c52ff38 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:53.939757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:41.347860Z digest=sha256:ba34f8cb1419595964028c5af44092fa68a8777aa4a97cf100798948f9461e8f

Observation 77bd1637-6379-4ef3-8d25-928d3b36ff97 · outbound

This paper cites SimVLM: Simple Visual Language Model Pretraining with Weak Supervision.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model SimVLM: Simple Visual Language Model Pretraining with Weak Supervision

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:41.477507Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:41.477507Z digest=sha256:6143ba0d4b19658487cfc8f9832d95baf334b150964b6750fff6ab8db57d470b

Observation c5dbae6e-3d28-445d-a8a7-7d60e9ac6351 · outbound

This paper cites Zhang, X.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Zhang, X

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:53.718800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:41.620139Z digest=sha256:be1f4b8fc79093b81471f09687d756202960e901579b04e164d0c15f97270f50

Observation e21a4c6f-2011-4bb3-83f2-fca4edc42986 · outbound

This paper cites Cheng, W.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Cheng, W

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:41.737519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:41.737519Z digest=sha256:3a20481b56c385e3d35612510d793324486a301e6bf065301de757d5d3d1a201

Observation 63038f23-68e0-4cf7-b811-4a6d07d8714b · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:41.863756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:41.863756Z digest=sha256:f36bc36d248e69df6d21a161d557c89ff29b703cecb6a8d5c6a45f6d2a9d5545

Observation 9bf901f3-42d4-4796-b3f9-57dca0f93082 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 37

Resolution
metadata mismatch
raw_fallback, observed 2026-08-07T12:52:47.355468Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:41.935467Z digest=sha256:2a9eba88255f29a55f797d1adb86b52bdfe601ac34a909c6329e4c8ae36aa828

Observation aa067ddc-360f-49ec-ac90-0e63bab9460a · outbound

This paper cites Cheng, W.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Cheng, W

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:42.077219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:42.077219Z digest=sha256:48c89f1aa57c004b84e07ba1c44e55beaf7c5bb796e6f496a4c1fd4c80084e82

Observation e94ec38d-69ec-4738-a967-b5faadbeec74 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:53.500630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:42.260304Z digest=sha256:86b3b96a41a8f59b03ba8d8d31953cd5b79d6c6cdc01d698f8f67396517f4a33

Observation bb68489d-5d11-41d8-9982-ea5f287d7e21 · outbound

This paper cites Zhang, Z.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Zhang, Z

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:53.308178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:42.414916Z digest=sha256:fcd55a57159c0f7b3686b67895e0eaf668b0ad315cbd2f28448ff44c202311da

Observation 317e716f-d6f1-4b0d-b7e0-5c577690f75f · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:53.118429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:42.564442Z digest=sha256:061011ad67f55db6e9e3e7bd047456370040bb62f8c4be1f95ab1f3faebd2d5f

Observation ead8fbd5-99c9-4163-9a60-66e4898bc5d6 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:52.872374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:42.746129Z digest=sha256:0d72cb3799dcd529b1594cf14f91d0f0cc2ebf97996f0b7106ca3c0ba5d22803

Observation 2ad7f8fa-a2e5-4653-9fc5-a25fb201a9c9 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:52.726239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:42.902878Z digest=sha256:53a2de2e8fc16f8426fcd9c74c0463757d3e593e643d3e2e8d57b4a61c846527

Observation a1497428-a809-49e1-ba46-d42f066e3c26 · outbound

This paper cites Huang, W.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Huang, W

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:52.516230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:43.080542Z digest=sha256:d65958e72e911ab7fa7db9b9b3f356304aaeec7e7dfc7d842ece4dba0e812b0b

Observation 7eb8d239-1849-4376-b212-2f30817b08e1 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:43.291761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:43.291761Z digest=sha256:db326f72cf8dc6ea3a574b8dcaef6ba24394584d92e1b0fadf5e94ba99c30323

Observation 939a91d4-e703-4647-8e57-dbec012f5be4 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:52.328504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:43.376452Z digest=sha256:3c397da7f181162cac633ef09d97aa5a1b7ec7fd79faa3adfdbdc4044fb63ad3

Observation 83882436-fe97-4f28-b95c-1947616b957c · outbound

This paper cites Devlin, M.-W.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Devlin, M.-W

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:52.113059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:43.502994Z digest=sha256:65efc1328dbf14a1b2a8506c00867589ef706e4718e2d6445d872cedc87947b2

Observation 0c5efa97-c4d9-4a2a-a5ba-41674afba9bc · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:51.880237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:43.665941Z digest=sha256:64ae319d14d0d1d417aa9fe005f5a7c3a4ff84281c5afb9bbe8f26fbee2db67c

Observation 68447c61-88fc-49cb-8413-4300f61ddb0a · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:51.717387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:43.799953Z digest=sha256:d161774f2215749e4ad97ecb982999cc6d59bd805a7a6b6b0d551f0c9269030a

Observation e97e983c-750b-499f-b0fe-c605f3f115da · outbound

This paper cites Radford, J.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Radford, J

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:51.485927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:43.943310Z digest=sha256:a3b60fb0f9d465247551c8d80ac617675150b110671160163923a327db749a02

Observation ab95742f-af91-4837-a8be-418c7492f131 · outbound

This paper cites ClipCap: CLIP Prefix for Image Captioning.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model ClipCap: CLIP Prefix for Image Captioning

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:44.090056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:44.090056Z digest=sha256:5f23caaba1af924e0e16e1b3742a389556e36919aacda0b6e945fbac112a3261

Observation 583f908b-8132-4b1d-a22e-58bdab4df765 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:51.296324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:44.217497Z digest=sha256:ba0e5633d5fbdde789cc7dffbe2d2e81c0adce87ad78a4bb64a74ffba6252609

Observation d6987114-ce4b-4866-b9c4-b3af430e460a · outbound

This paper cites GIT: A Generative Image-to-text Transformer for Vision and Language.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model GIT: A Generative Image-to-text Transformer for Vision and Language

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:44.360480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:44.360480Z digest=sha256:9d7bf6fe3e94d537477297aa452f5fd4588ec9a38f65644c07fc04f061e53db5

Observation f33e1d90-fa99-4a87-ba0d-ebb038698a5e · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:51.143213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:44.524819Z digest=sha256:50fc5d03e689594b81521944e3b39024935b6312ff581de921820395a89ea43e

Observation 619ce196-557d-4234-a37b-8be37228eae5 · outbound

This paper cites 13035–13045.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model 13035–13045

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:50.963362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:44.675623Z digest=sha256:d6fc669d9f44bbfb5732c67ba73499603a338d0fd0dcb007dce4c4495f20acb8

Observation 771492b5-9dd1-4e5a-908c-30499c358417 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:50.768170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:44.829181Z digest=sha256:2473929a62ee28049b3dae9a3927cab4a74b0c47c445e5a9921a2aec960ea955

Observation eaa26ee6-5c18-4998-ad25-3464b4440290 · outbound

This paper cites Changpinyo, P.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Changpinyo, P

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:44.950358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:44.950358Z digest=sha256:043883cfea1a6b298865f06140c2ec84c192db406a93ed09ae4e7a3c67820bb1

Observation 1faceee6-58c1-42c1-a747-6b0968ee5f17 · outbound

This paper cites Kullback, R.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Kullback, R

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:50.556760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:45.091862Z digest=sha256:d097476563caf138fb6fd52c0ca19523c63f37f1bf5b803a73593bf20e4114d1

Observation a9da7bcf-96d6-4e97-aa27-d303f3179f4c · outbound

This paper cites Microsoft COCO Captions: Data Collection and Evaluation Server.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Microsoft COCO Captions: Data Collection and Evaluation Server

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:45.252289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:45.252289Z digest=sha256:fb5074756fdf24032fd298a59a782f7f3e4beff55fc18766afa6d61837d7418d

Observation 7632d89f-de66-4896-af43-b06f96587e43 · outbound

This paper cites Papineni, S.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Papineni, S

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:50.405269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:45.322179Z digest=sha256:8212d5343def3f55c581be7313aac5fb11a29927ab696cd13f836fef3b40f7e7

Observation 3504e205-b35a-4e11-a653-83c1f2407f11 · outbound

This paper cites Banerjee, A.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Banerjee, A

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:50.178964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:45.478971Z digest=sha256:c56e23a10fa7459fe43d91f165f59aa1d00cdf13c92f78b038f4624433d455c9

Observation 1aea2639-4e4d-4a37-ba4e-f76a1944a962 · outbound

This paper cites Lin, Rouge: A package for automatic evaluation of summaries, in: Text summarization branches out: Proceedings of the ACL-04 workshop, 2004, pp.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Lin, Rouge: A package for automatic evaluation of summaries, in: Text summarization branches out: Proceedings of the ACL-04 workshop, 2004, pp

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:50.009788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:45.626619Z digest=sha256:4ae2eb8c4890cbc6df41d451e28312121df51c27ba0c52393897af029b1f3fa6

Observation 9225c830-2bce-45e3-a697-43c60fdde696 · outbound

This paper cites Vedantam, C.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Vedantam, C

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:49.793273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:45.757986Z digest=sha256:66a787ff124ca129514d2de390bb7f63d0292c006beb712875627de8468956a2

Observation 301c5f23-0b3f-4bc1-8635-bf8b1b7805f3 · outbound

This paper cites Kirkpatrick, R.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Kirkpatrick, R

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:49.605423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:45.859012Z digest=sha256:fc8f37d0f31d12e2298473bf23958cb16c1a3c02aeb3200be05e11480a84e9b2

Observation 210021e2-202f-45f3-968e-251efdbd32d1 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 65

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:49.376017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:45.945542Z digest=sha256:b906bdef61d1a6de0e7514bae4694432526e7528b4693fbaf6086457ed6d027d

Observation 57f827d0-533d-40e0-9a1d-80b3884f9dbc · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 66

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:49.146830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:46.100120Z digest=sha256:a8d133c2d7c3b990010fbeb981e43d47c0c90cde8e9f8dba5661858227de4b1c

Observation e0bfc597-ddae-4ac7-9bf9-2bf7e9f83ac8 · outbound

This paper cites CLIP-Adapter: Better Vision-Language Models with Feature Adapters.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model CLIP-Adapter: Better Vision-Language Models with Feature Adapters

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:46.278075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:46.278075Z digest=sha256:76934ca470376c6b6618fe2678914fb1c06a42e72f16461bfab65bc27472fad4

Pith citing papers

No inbound Pith citation observations are available.