Pith. sign in

Paper Citation Record · LEDGER

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall

As of 21 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 0 inbound Pith citation observations for arXiv:2501.09155.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.09155 v2

Coverage vector

measured 44 of 44 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T20:15:57.460723Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

44 of 44 outbound references displayed

  • verified exact0
  • verified fuzzy24
  • unresolved20
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 65ebb3eb-4f73-4310-9b5b-8d357fb65d47 · outbound

This paper cites Springer New York, New York, NY, 2008.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall Springer New York, New York, NY, 2008

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:15:58.247236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:15:57.247270Z digest=sha256:74857c12fae2b9c6b85a2fed5411855a10f6a1182e11c47599a14b647a6847ff

Observation 5b4598df-9a43-4ffe-ab77-63c431e97c4a · outbound

This paper cites an unresolved cited work.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-10T20:15:58.231343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:15:57.254697Z digest=sha256:a8cb8e30c8508f91792e77cd70a55486432c8e3c2fc05e1248e86aa98bed7cfc

Observation b3fa7de7-5b49-4dff-8a8b-3d8b58925a7f · outbound

This paper cites Aditya, Y.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall Aditya, Y

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:15:58.216537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:15:57.260749Z digest=sha256:4150f2237f24f456d008f019865116f0507d4cbdb5fe79c3c86156e71a4b140a

Observation 557d2f3c-7760-4969-879f-2403280692a4 · outbound

This paper cites Anderson, B.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall Anderson, B

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:15:58.194350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:15:57.266727Z digest=sha256:383d61e8294fc26e156d962f6e67751e7613280196604fcf93fa20ecdf35ba6e

Observation f1beebc2-8f8e-43ee-9d06-28f57e4153f8 · outbound

This paper cites Aneja, A.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall Aneja, A

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:15:58.174182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:15:57.271352Z digest=sha256:5b5dded5e1d8ffd8cc9f9cf41583eb23ffc7edb4cf198e4b9c1c5a56245c148b

Observation 9e8a0d37-c3d0-45ec-a70f-7333868b19dc · outbound

This paper cites Carion, F.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall Carion, F

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:15:58.156389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:15:57.276188Z digest=sha256:61c92406b71d6519b277b8b14a4d81d1a2160bd30aa0b6948579bd94c08a3ae0

Observation cc9a1d59-7d73-4d88-b85f-2ae53b22e3d1 · outbound

This paper cites Cornia, M.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall Cornia, M

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:15:58.138784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:15:57.281106Z digest=sha256:61d503484ddb87fe23f764dbd75eb6433f6830e6353407732b086870d6b43733

Observation 2117796b-1b92-4efe-912f-c23573234fe0 · outbound

This paper cites an unresolved cited work.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-10T20:15:58.114738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:15:57.286019Z digest=sha256:e8a631d9990c8222df2ebe4ef3e18dbfbc5d58bf2cbba6eea6ddf6982a680590

Observation 4381c911-4e5c-4b39-9cb3-66db67b41973 · outbound

This paper cites Devlin, M.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall Devlin, M

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:15:58.096576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:15:57.290272Z digest=sha256:8288cfcb3f84c5af574b94af3e0e28b2f426447dc7cfe25cc61fde213652661b

Observation 5af8ffbc-1dd6-4f45-91a5-52879b9b5a47 · outbound

This paper cites an unresolved cited work.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-10T20:15:58.079781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:15:57.294793Z digest=sha256:6bc6f98c0b3e0ddc5271adc2f74f5d189895a51ee0866fea9edb57f8eaec146d

Observation a09aec31-3044-44f8-9856-7a63ad2b0507 · outbound

This paper cites an unresolved cited work.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-10T20:15:58.064380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:15:57.299636Z digest=sha256:602e284bac6195d508bf2f757f207e807c91aa3f1f34e8657da102ecdcdbe944

Observation 47d7930b-3290-4e14-b065-b89416f9759b · outbound

This paper cites Gonz´ alez-Ch´ avez, G.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall Gonz´ alez-Ch´ avez, G

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:15:58.043912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:15:57.304071Z digest=sha256:dc9bfd6ddef6347482676849faa1134c1b954c90386361f8c8ff103e68809733

Observation dee7aec3-8330-41f7-bd85-cc5b9bde2d5f · outbound

This paper cites an unresolved cited work.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-10T20:15:58.026461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:15:57.309515Z digest=sha256:21598410e73730e9c3f9406329419019c207166f42848ca72337ad05ed34add7

Observation 3c9d3916-7670-4853-8863-b649ab24e444 · outbound

This paper cites Hodosh, P.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall Hodosh, P

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:15:58.005410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:15:57.317543Z digest=sha256:d2df6efdf4ce0d81654c8663f7422dcfd4df3ce33c22a6d019212f99bb250a80

Observation 9e50a53c-1121-4a36-a052-a3bbe84f2b33 · outbound

This paper cites Kasai, K.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall Kasai, K

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:15:57.986157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:15:57.321804Z digest=sha256:6e587240ac5f0eaa66bf009ac14e2d96d27b6c3b57f52b5a1d88d6cf4a5dc564

Observation 4d6ba391-c6c6-42d3-8cf6-9f49928df913 · outbound

This paper cites an unresolved cited work.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-10T20:15:57.963864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:15:57.326591Z digest=sha256:a1a3eff3ba729adfffc3fea7d3fbeb4817f4dce18f37f6493a6fa53b3b2af550

Observation 02f0908e-911d-4a3a-b821-dce02eb069c5 · outbound

This paper cites an unresolved cited work.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-10T20:15:57.945190Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:15:57.331813Z digest=sha256:2a5f4e2cf1eb084709ee33b34df2c1953f22cbea46bb28df92c95289658c97cd

Observation b454c05e-481c-436c-bfe6-4df6c811887c · outbound

This paper cites Krippendorff.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall Krippendorff

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:15:57.925982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:15:57.336628Z digest=sha256:3b2aa19032d8c8156c0cff72d6bfbacdb283b143be490ca83a6a567daf4b1444

Observation 1649a973-3f22-4334-bb7d-a0ce0c26ca16 · outbound

This paper cites Krippendorff.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall Krippendorff

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:15:57.907003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:15:57.341156Z digest=sha256:b40b8bc42a0cc886a770e8f8de3f79cb08c705043f02e710e3d0e905cc9fe158

Observation 9088461f-61ed-44d1-aeca-929d35000d3c · outbound

This paper cites Lavie and A.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall Lavie and A

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:15:57.888941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:15:57.345662Z digest=sha256:3a728961dd031eaaba132d54372024ee8c9be5fb39d2fa018c9ad0f42012db6a

Observation fb159076-4cea-4149-a0ba-e4b28b5ee87f · outbound

This paper cites Levinboim, A.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall Levinboim, A

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:15:57.867204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:15:57.350160Z digest=sha256:47b50c008f2b124ddbd026fde41296f7052a51af454174eec832fb539f8a2d94

Observation d1c3cae0-2169-4ec2-b920-b7a98d027277 · outbound

This paper cites an unresolved cited work.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-10T20:15:57.355863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:15:57.355863Z digest=sha256:421cde8df41307c873b927f827e39d14cb332bedb5815251534cdedb0be9050c

Observation e9f5bffc-e62a-4565-9af0-516dd8fc398e · outbound

This paper cites an unresolved cited work.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-10T20:15:57.838405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:15:57.360682Z digest=sha256:214c660ddf44eeef3d3fa70c099e84b8627cbc1e38030e0975db4f584fe7c1f7

Observation 772b51ef-5b30-4dba-b984-64907d6908b0 · outbound

This paper cites an unresolved cited work.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-10T20:15:57.366327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:15:57.366327Z digest=sha256:8c71edb5c0dff1ec6b2174d39361acc920def4a1680f4d12b5ec369bbfaea079

Observation 6d3915eb-1b89-48bc-831a-440902a8c71f · outbound

This paper cites Lin and F.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall Lin and F

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:15:57.807380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:15:57.370175Z digest=sha256:8233d79a0454297936939e67edee786e8771f3e046f00616ed2d586d96266f7d

Observation 1ba33e88-375d-47a0-99aa-fef0a018492f · outbound

This paper cites an unresolved cited work.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-10T20:15:57.790960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:15:57.374598Z digest=sha256:cd882b0b0c31ff1a83381c89c73d4b5f90de313b87926a890d9f8b9d1bf80fac

Observation 4cde52b5-4ce0-4038-8096-d4bc2ac5d7eb · outbound

This paper cites an unresolved cited work.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-10T20:15:57.767531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:15:57.378590Z digest=sha256:86080ff50f5f9d8f826dd93bab28f825186db845d41f2f19ac24a650f0fb25be

Observation 9ddb9518-4557-4bc6-be62-adc6643d47c0 · outbound

This paper cites Moctezuma, T.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall Moctezuma, T

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:15:57.749495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:15:57.383350Z digest=sha256:a321a59e0b0977c973b330ceb005615c40c1d975fadbf0b2c6cf5e6023f2226f

Observation 0d93616d-5668-47ef-b8fc-9cc047206a94 · outbound

This paper cites an unresolved cited work.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-10T20:15:57.733545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:15:57.387863Z digest=sha256:1705781b0e86e4601c5e92b5fc6e91cf7cfa8dba483bc0364ce05dc28116383b

Observation d7c165a0-03a1-4fe5-8b98-329c53af196c · outbound

This paper cites an unresolved cited work.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-10T20:15:57.712620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:15:57.391993Z digest=sha256:32d1b5172df8415979d0c88fcb2b49a86995634510ef3cec33e0e3e650f42961

Observation 3289be5b-6e5b-4b2d-b8db-e1d0a10bbb35 · outbound

This paper cites Papineni, S.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall Papineni, S

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:15:57.696719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:15:57.396570Z digest=sha256:f8f58e650a7c41465e7d3260a374ff7943211413d73fe8310753c74e945afb89

Observation 0ed91ce9-5d1b-4e2c-89a0-6be33c8a410b · outbound

This paper cites an unresolved cited work.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-10T20:15:57.681779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:15:57.402073Z digest=sha256:d2e63e3df70dc1d0c926e971ce3ff2304edcfcaa96c20d20334ac4b6c6d08d0c

Observation 21ff81ff-e9e1-4fa6-808a-edf03c311a97 · outbound

This paper cites Radford, J.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall Radford, J

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T20:15:57.406166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:15:57.406166Z digest=sha256:e0545bb38f7dcb37f474344d2ccdbabac21370e01c7826091f33218b15bbf66f

Observation 9b705bf6-9c45-45ed-984d-87d7a3e151fb · outbound

This paper cites Learning Transferable Visual Models From Natural Language Supervision.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall Learning Transferable Visual Models From Natural Language Supervision

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-10T20:15:57.410119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:15:57.410119Z digest=sha256:3ed45a314c0c3451b9d9f1d580abcd3d5e20d23ec8aa146188d9d96d8f51668b

Observation d65728e6-742e-4dcb-a59e-1ccab7b95f2a · outbound

This paper cites Rahman, N.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall Rahman, N

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:15:57.656746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:15:57.414356Z digest=sha256:4fd84c4ac9eb149b6008eb989fd62b350a1a78e55f8057dc711070f03c3297fd

Observation 36ccfc71-6eb9-4787-b84c-524f62ca6e89 · outbound

This paper cites Schall, K.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall Schall, K

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:15:57.640510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:15:57.419586Z digest=sha256:be787690f3d067b44f01cba54f2183c54ea2b99f5d8dc746f088602b240296a6

Observation fb540e5a-364b-48e9-bc70-483b70f8c72c · outbound

This paper cites Schuster, R.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall Schuster, R

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:15:57.625359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:15:57.423542Z digest=sha256:02e4520fe8df47273654009ef258f97daa91617cb23a1c3616dd3087541b8efa

Observation 781df4f6-5772-4246-a79f-bb353f05f86a · outbound

This paper cites Stanojevi´ c and K.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall Stanojevi´ c and K

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:15:57.610281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:15:57.427689Z digest=sha256:02405c4be5afbad3f1e5b5a67a344a11839257ee68d13aebd0588cc210b833e8

Observation 7bb619fd-3ae7-4b92-be39-030ef025cdd1 · outbound

This paper cites Vedantam, C.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall Vedantam, C

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:15:57.594375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:15:57.431952Z digest=sha256:8313b4155b7eab0df5789d161bd464832729be7f64be359c39825e47c6ec472f

Observation cfd59bb3-1d50-4eda-b3ef-a63716c5f551 · outbound

This paper cites OFA: Unifying Architectures, Tasks, and Modalities Through a Simple Sequence-to-Sequence Learning Framework.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall OFA: Unifying Architectures, Tasks, and Modalities Through a Simple Sequence-to-Sequence Learning Framework

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-10T20:15:57.436866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:15:57.436866Z digest=sha256:0c79b76af391f2857451349af097ae805e315a31511d7f6e33404444eb42fd61

Observation 6f297336-97ea-49cb-ad28-15f75c0a3faa · outbound

This paper cites an unresolved cited work.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-10T20:15:57.579246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:15:57.442081Z digest=sha256:24879f4f2eb75b556e67103aecc8e98482915c9f5d8821fd748a97ba29adbc3e

Observation 3b56e3a8-1d5b-43c9-b3c7-34f17722c229 · outbound

This paper cites Wiebe, R.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall Wiebe, R

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:15:57.564531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:15:57.447767Z digest=sha256:f065b55ee5fb598f2bdf3e23b3a6d0e0f4e6c8e2abc966aa3f611e81ae5c7c8f

Observation c52b34b8-83bb-4507-9afc-5489490037ff · outbound

This paper cites an unresolved cited work.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-10T20:15:57.550294Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:15:57.453140Z digest=sha256:ae75302f683abe4a88cea99377deb237a48de9b91bc9bfdacea5e3fa330e47c4

Observation 13dc7e3b-ed98-4c18-81b6-f4a132436a48 · outbound

This paper cites Zhang*, V.

VCRScore: Image captioning metric based on V\&L Transformers, CLIP, and precision-recall Zhang*, V

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:15:57.535594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:15:57.460723Z digest=sha256:8b02b6a46d96ec2be4f89ba29dacac3969c7b2f8b13535f0c3b73a9f433bf7bd

Pith citing papers

No inbound Pith citation observations are available.