Pith. sign in

Paper Citation Record · LEDGER

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models?

As of 9 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 0 inbound Pith citation observations for arXiv:2502.06600.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.06600 v2

Coverage vector

measured 43 of 43 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T14:58:48.251828Z

measured 43 of 43 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

43 of 43 outbound references displayed

  • verified exact2
  • verified fuzzy0
  • unresolved41
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 83f9ae22-6664-4e3e-9ae7-facd58bf7ce4 · outbound

This paper cites From Images to Sentences through Scene Description Graphs using Commonsense Reasoning and Knowledge.

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models? From Images to Sentences through Scene Description Graphs using Commonsense Reasoning and Knowledge

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-08T14:58:48.107848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:58:48.107848Z digest=sha256:23e89627770c63ac95928af8c5046a32c29a7f5f2cba182cf415e8d0a9759ff9

Observation 8d4fc923-288d-469e-8913-2e82bf926c33 · outbound

This paper cites an unresolved cited work.

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models? Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-08T14:58:48.703743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:58:48.112701Z digest=sha256:9a4570a89f6411563107bb84eb0b8fb690992c76f91a9f5177d17e95f01c3c46

Observation 6b138ef1-d3f6-4927-8fbb-ec2a0c8e06e4 · outbound

This paper cites Tower: An Open Multilingual Large Language Model for Translation-Related Tasks.

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models? Tower: An Open Multilingual Large Language Model for Translation-Related Tasks

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-08T14:58:48.116638Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:58:48.116638Z digest=sha256:0b77c6e09e4bf936a6711835094b58842c1385ef74d0dc503320d9b5e4ba2553

Observation 4e69acfb-352d-46d3-80bc-8845f2411477 · outbound

This paper cites an unresolved cited work.

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models? Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-08T14:58:48.693345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:58:48.120954Z digest=sha256:8bad91291e464b41944ab187f3ba0f0d088f2a641f359231aefbdd657681e33f

Observation ab46c83d-b001-4dd1-8c64-eeadb081bf10 · outbound

This paper cites an unresolved cited work.

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models? Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-08T14:58:48.683627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:58:48.124736Z digest=sha256:70a6925568bc84384bbee7bb4790461af181a456ff34dbf4c96a7c9c931f3683

Observation 58f9782a-809f-4fa1-933c-d12dff79cff4 · outbound

This paper cites Evaluating Image Caption via Cycle-consistent Text-to-Image Generation.

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models? Evaluating Image Caption via Cycle-consistent Text-to-Image Generation

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-08-08T14:58:48.351746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:58:48.128379Z digest=sha256:86255bd8d2564ae8a1b5df50e9ece511e52485063927c93dd559c7b281ada3bd

Observation 1c9b6edf-f166-431e-9e0e-68708a488229 · outbound

This paper cites an unresolved cited work.

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models? Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-08T14:58:48.673976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:58:48.132436Z digest=sha256:46de42c646977a11fa21a64b0d6e1d6be118ab4b5e943d45d6b696ef660b7d18

Observation 7d6d3166-20b1-43c9-ad06-3b9f14f4478f · outbound

This paper cites mBLIP: Efficient Bootstrapping of Multilingual Vision-LLMs.

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models? mBLIP: Efficient Bootstrapping of Multilingual Vision-LLMs

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-08T14:58:48.135855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:58:48.135855Z digest=sha256:219ad0b04707464a7a62906c4c4efbbe7694c5b0efdf35996d63eb84b6ce69fc

Observation 054bdc68-448b-4a6b-900d-c853408c8bd1 · outbound

This paper cites an unresolved cited work.

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models? Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-08T14:58:48.664055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:58:48.139766Z digest=sha256:de55761095b593895e7de149931d9cfb4196ec12964f6ccd71f8e2422fce4fbd

Observation dcd3ac4a-2f04-42d5-86ba-ed55c2335154 · outbound

This paper cites an unresolved cited work.

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models? Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-08T14:58:48.653900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:58:48.143162Z digest=sha256:e55cf9d0d48caf66ba5aac0e74e082ed5283407b36ab02f4e7933e238999c1fa

Observation e8b7e6cb-d43e-4f60-bd21-d1e3534c6313 · outbound

This paper cites an unresolved cited work.

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models? Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-08T14:58:48.643350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:58:48.146463Z digest=sha256:368625bf321540192b21a99391b53a8d79af14b09be7a5dd37e58863263a9a53

Observation f4cd076b-50e7-4c45-88e4-127b8d59cb26 · outbound

This paper cites an unresolved cited work.

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models? Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-08T14:58:48.632497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:58:48.149925Z digest=sha256:60d72ed92be150652c06c8943ea038739741903c1d8f56f0b84029f00486adc0

Observation f86b738c-e88b-462f-84b7-616c8fbf42da · outbound

This paper cites an unresolved cited work.

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models? Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-08T14:58:48.622305Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:58:48.152564Z digest=sha256:14e9673686f90f183168b9ba9b42171a5ea9d3eb867dbb9dd4865f66ecc91124

Observation 0509c59d-0ce0-4f90-b056-583aad488a4d · outbound

This paper cites an unresolved cited work.

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models? Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-08T14:58:48.611322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:58:48.155242Z digest=sha256:2ab69cb608c88bf4e430fa86f77b38db1664d8f6b89fe15acf10e9f0d0d964cb

Observation 262e7b40-e123-4ea2-a546-18028f092554 · outbound

This paper cites FLEUR: An Explainable Reference-Free Evaluation Metric for Image Captioning Using a Large Multimodal Model.

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models? FLEUR: An Explainable Reference-Free Evaluation Metric for Image Captioning Using a Large Multimodal Model

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-08T14:58:48.157795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:58:48.157795Z digest=sha256:40282f9c236bfdaa9f42be01ecd84053f95a58f044b2746afca8e33398a4d79c

Observation e94ee4ea-f377-4b0f-8732-a7ba9c3bb5ab · outbound

This paper cites an unresolved cited work.

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models? Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-08T14:58:48.600724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:58:48.160659Z digest=sha256:13960d80731d5c3879bacc585f1bbb7986b8ab4e3aa1e8738bbf6a34f07726e8

Observation 5dd98dfd-ec45-4b48-b16e-190a8b5ee264 · outbound

This paper cites an unresolved cited work.

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models? Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-08T14:58:48.589632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:58:48.163384Z digest=sha256:1b8b7cdbe7bbb09b9d6b814f3060d37a8f1601ce5c4f0bc2689bcf3724af07d8

Observation 2a499ef0-c798-4326-958d-05772ebe347c · outbound

This paper cites an unresolved cited work.

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models? Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-08T14:58:48.578526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:58:48.166210Z digest=sha256:155341170c592873c8113cf305bafee52175a53dacb462cd8d563cd6b17e2475

Observation dcabe155-8dab-4e76-91c0-868983d29dd0 · outbound

This paper cites an unresolved cited work.

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models? Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-08T14:58:48.567322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:58:48.168871Z digest=sha256:11e38b5597828267371047585d3f399abc19b5c08b054f324b0628c4cee11b17

Observation 52e30fe1-49a0-4c7a-a590-48eef4db58df · outbound

This paper cites an unresolved cited work.

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models? Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-08T14:58:48.556817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:58:48.171937Z digest=sha256:b956f7052c2fcdd9809cc97bec605229929695b9fdfac04966067b742dabc5ba

Observation e827993b-975c-49bd-b84a-0efffdec7d0e · outbound

This paper cites an unresolved cited work.

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models? Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-08T14:58:48.545931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:58:48.175223Z digest=sha256:be365d8ed5ed3cf57364d1b62706eb3acf7e73c4ceea968cf335bc4f2b84deba

Observation cd7d2f62-94a8-4d85-a4a6-1f32bfd8f481 · outbound

This paper cites an unresolved cited work.

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models? Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-08T14:58:48.178520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:58:48.178520Z digest=sha256:33ba3628270e68652a7566fb42dce52081c365b7eb902c0894614e5f2055c324

Observation 720e87c9-ba5b-4ca1-9252-aee9ddfa6290 · outbound

This paper cites an unresolved cited work.

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models? Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-08T14:58:48.529185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:58:48.181934Z digest=sha256:c0ad57c581b59769d2fbe50229d0e7c1d520359a743df373f4eaa9d9d902f7d4

Observation 2c560504-c6a2-4af7-a220-b923ba8bf3ad · outbound

This paper cites an unresolved cited work.

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models? Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-08T14:58:48.520419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:58:48.185359Z digest=sha256:cfe8f10ab12edc5bb0ee6581de5b2cb65451959ac2a71f580e5164226dc21d95

Observation 91ae5dad-3086-4095-902b-c279a8d85fa4 · outbound

This paper cites an unresolved cited work.

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models? Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-08T14:58:48.511567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:58:48.188730Z digest=sha256:08104e72126a643b79679ac9e9f377f3159a13f0c44895caa05f648b2a391965

Observation 93864a7b-81d6-4f6b-9c5d-fe781abf2d33 · outbound

This paper cites an unresolved cited work.

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models? Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-08T14:58:48.502102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:58:48.192388Z digest=sha256:719f447878d1c03759725ef0ab236073f78c8c3b1ce89d5d070c8af37898cef9

Observation 3ff3180b-db37-4e66-a805-e6ba78042af0 · outbound

This paper cites an unresolved cited work.

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models? Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-08T14:58:48.492448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:58:48.195667Z digest=sha256:8ea279db8e98a499ece5d68ff7e9c579188cd146bd133bf42c2cf377b2d4cdeb

Observation 0128b44b-877f-49ea-b537-4d2c28710619 · outbound

This paper cites an unresolved cited work.

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models? Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-08T14:58:48.482354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:58:48.199242Z digest=sha256:869e14866cf38ce9226f1dd426dfd73ef24904a50b58dc399de92ba6806c0fc6

Observation f4db2265-c362-4c07-bb56-8c486bd1e2b1 · outbound

This paper cites an unresolved cited work.

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models? Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-08T14:58:48.472530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:58:48.202653Z digest=sha256:4493857dc5ab4a1e1e434d2c1f8f2f801b056afca85ce25fdaeb857bfc776c2f

Observation 88661f33-1a84-4df6-8210-b8a6f49236f4 · outbound

This paper cites an unresolved cited work.

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models? Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-08T14:58:48.462397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:58:48.206031Z digest=sha256:091844be1713ba3bb128a77e2f38682d167bbda2702d8f25da159041c9a5f8c3

Observation 70bab504-fbe1-4dcd-94b0-de0d35278fbd · outbound

This paper cites Positive-Augmented Contrastive Learning for Vision-and-Language Evaluation and Training.

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models? Positive-Augmented Contrastive Learning for Vision-and-Language Evaluation and Training

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-08T14:58:48.209337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:58:48.209337Z digest=sha256:68f0c7c4d0893c92077328d110f3a0c96fe2b11599f540026003c88d100ccbd5

Observation 30ead303-1f8d-4752-8115-bc8cd0becdc4 · outbound

This paper cites an unresolved cited work.

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models? Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-08T14:58:48.452291Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:58:48.213093Z digest=sha256:6eba660cc95f1c9c61ad6ec47838ea52818e1e1bb77a576ed696797bfc23efad

Observation a4d74aca-a523-4c46-b549-d18939fdc182 · outbound

This paper cites an unresolved cited work.

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models? Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-08T14:58:48.442298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:58:48.216626Z digest=sha256:d148562aa73e20ae712144dbf5c5e418c895208d6479167f4a5845b6e5997f3b

Observation 91a1aeb2-7791-4bb0-806e-49a13e7c416f · outbound

This paper cites an unresolved cited work.

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models? Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-08T14:58:48.431991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:58:48.219896Z digest=sha256:446585131e6f2761a6ac2291a34210bd1b81b5037415710e7f4bcb8d90c34643

Observation 5444d6e4-1a51-4718-8f98-e699647f718d · outbound

This paper cites G-VEval: A Versatile Metric for Evaluating Image and Video Captions Using GPT-4o.

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models? G-VEval: A Versatile Metric for Evaluating Image and Video Captions Using GPT-4o

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-08-08T14:58:48.309144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:58:48.223313Z digest=sha256:21c2a67d04c64fd4fccabef46c4f65d2bae6d18a76bf2135e55cf8560c881c19

Observation 80b7e105-15a2-437c-83b2-7746e24bfdb2 · outbound

This paper cites an unresolved cited work.

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models? Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-08T14:58:48.421535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:58:48.227281Z digest=sha256:d04e793b39ea6feb298019f83d839e161076459357513b3ea5978d696c345c9f

Observation 8e4b798c-3d69-4815-acbf-c5c5db183cf4 · outbound

This paper cites an unresolved cited work.

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models? Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-08T14:58:48.411384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:58:48.230725Z digest=sha256:d7c8df74ccfa786111202c22c97bff7a2aea6babc81cadb2af8db6ad8cd93c15

Observation 9fede902-8fa7-4281-a23a-e75100d5d95e · outbound

This paper cites an unresolved cited work.

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models? Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-08T14:58:48.401688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:58:48.234045Z digest=sha256:0c25e20a7952369b45daa258f0c32687b22eb52c1974a5f95995488201de7e03

Observation e0a843d9-4994-4c1b-8a96-0b8dd8c8be83 · outbound

This paper cites Visual Entailment: A Novel Task for Fine-Grained Image Understanding.

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models? Visual Entailment: A Novel Task for Fine-Grained Image Understanding

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-08T14:58:48.237478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:58:48.237478Z digest=sha256:69995adcb8f5c3d3794d5550c2b15aa3d5da7d1d0d99705a819abc7e7e146c5c

Observation 3e996417-493e-4955-a628-82ecb0dba27f · outbound

This paper cites When are Lemons Purple? The Concept Association Bias of Vision-Language Models.

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models? When are Lemons Purple? The Concept Association Bias of Vision-Language Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-08T14:58:48.240959Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:58:48.240959Z digest=sha256:3aaedce081c34b01bb595fefd3ffcf862bc5c302f781977d4e7cf1b1ca6076e4

Observation 55d6fdf1-22a0-49e7-b462-f7684214929e · outbound

This paper cites an unresolved cited work.

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models? Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-08T14:58:48.392417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:58:48.244606Z digest=sha256:1f77c46b57ebd2aadf9e14df775c3c1d04c4ac541db67f10cae03b46e20b2aa1

Observation d455749d-bc74-4cf7-b9b3-4288580f2fab · outbound

This paper cites online" 'onlinestring :=.

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models? online" 'onlinestring :=

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-08T14:58:48.248032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:58:48.248032Z digest=sha256:ab66f12f86351c62a7e0aaf8cc164c3be920ac101c1a2e37c7394c5515c521b3

Observation ecbced9a-afae-4de3-ac64-1f0c38fbdb2e · outbound

This paper cites write newline.

Evaluation of Multilingual Image Captioning: How far can we get with CLIP models? write newline

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-08T14:58:48.251828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:58:48.251828Z digest=sha256:4822dfc2b96b519d2b3c11f8e5d3ea0a4f92866430a5b40177707e2442a07059

Pith citing papers

No inbound Pith citation observations are available.