Pith. sign in

Paper Citation Record · LEDGER

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models

As of 7 August 2026, this Paper Citation Record lists 86 of 86 outbound references and 0 inbound Pith citation observations for arXiv:2507.08410.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.08410 v1

Coverage vector

measured 86 of 86 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:27:47.245370Z

measured 86 of 86 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

86 of 86 outbound references displayed

  • verified exact3
  • verified fuzzy43
  • unresolved39
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9156d484-00be-4473-948a-c2cad2f6b3c9 · outbound

This paper cites Learning transferable visual models from natural language supervision,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Learning transferable visual models from natural language supervision,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T18:27:46.747091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:27:46.747091Z digest=sha256:1da55586f099b15ff8c23032efe274419e94da2da03780f006fd6e32f61e07ef

Observation 2f4deb52-f0f0-4b76-8427-1491df391ee8 · outbound

This paper cites Why are Visually-Grounded Language Models Bad at Image Classification?.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Why are Visually-Grounded Language Models Bad at Image Classification?

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T18:27:46.828449Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:27:46.828449Z digest=sha256:18fca6fdcde45c69f3739a8ddf72246f7879df39bd5815349194e6f8759af672

Observation 13029e43-a1b6-49e4-bafc-42fb665b7bfb · outbound

This paper cites Eva: Exploring the limits of masked visual representation learning at scale,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Eva: Exploring the limits of masked visual representation learning at scale,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T18:27:46.858025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:27:46.858025Z digest=sha256:f262990e7955b1b91fbfceac23feb12a76d0603a269e5f1c196a4d1dc8d6e4ef

Observation 39aa69af-70b4-4a4f-8183-f1198ec0dbeb · outbound

This paper cites Conditional prompt learning for vision-language models,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Conditional prompt learning for vision-language models,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T18:27:46.862663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:27:46.862663Z digest=sha256:7a924b3302d5723821e2bd955aecf123bc10187900fcdfa688a11611854b6a7a

Observation 46276c91-e114-4954-aa13-5a3a05cb3661 · outbound

This paper cites Learning to prompt for vision-language models,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Learning to prompt for vision-language models,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T18:27:46.867601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:27:46.867601Z digest=sha256:4620ad9418b4db87df68719df63dea70b87ca3972dffba279d656c5119d8548d

Observation e285e8a3-1e6c-46ea-beb8-b17c138e9631 · outbound

This paper cites Maple: Multi-modal prompt learning,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Maple: Multi-modal prompt learning,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T18:27:46.871905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:27:46.871905Z digest=sha256:0486873c81485ce9a84e9f8dab3c5497c1f7334cd668637d549958f3f58c5aab

Observation a6fc409e-cd32-491d-b6b6-0a58406c2ab0 · outbound

This paper cites Tcp: Textual-based class-aware prompt tuning for visual-language model,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Tcp: Textual-based class-aware prompt tuning for visual-language model,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T18:27:46.876620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:27:46.876620Z digest=sha256:f314e666acb75362893f4fda3bdf308399f132a5c4ff47e3086d8f221ed03abf

Observation d1b1b93a-719a-4ddf-8302-3d48af181791 · outbound

This paper cites Dept: Decoupled prompt tuning,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Dept: Decoupled prompt tuning,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:27:53.308142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:46.881369Z digest=sha256:a1e5eeb03d49a79be731eb58e3f197de172af3535a3face7e27052af1b7b5e31

Observation 1e11ad71-9bdb-44c6-96e9-9c28f9822ee4 · outbound

This paper cites PromptKD: Unsupervised Prompt Distillation for Vision-Language Models.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models PromptKD: Unsupervised Prompt Distillation for Vision-Language Models

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-06T18:27:47.628086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:46.885780Z digest=sha256:8fb9b4823c5c2b7a595d76b8ae351e35fcb47a933ed5d093904cca57214e3d45

Observation ff40152d-8d67-4828-ba13-342a9de8c0cb · outbound

This paper cites Self-regulating prompts: Foundational model adaptation without forgetting,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Self-regulating prompts: Foundational model adaptation without forgetting,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T18:27:46.890805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:27:46.890805Z digest=sha256:844a45e15d1e9b254740d773dad500d3b268ce281272229e80e0ef4b31397271

Observation 6bf813c6-152a-4bcf-9010-af13b40d0145 · outbound

This paper cites Consistency-guided Prompt Learning for Vision-Language Models.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Consistency-guided Prompt Learning for Vision-Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T18:27:46.895300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:27:46.895300Z digest=sha256:275d269e44588525afe2ad2359b4dc4cdabe3890e8c43fd8b25a5a38a7719e41

Observation beabbd0c-6a93-456d-884e-253b41e88a75 · outbound

This paper cites Large Language Models are Good Prompt Learners for Low-Shot Image Classification.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Large Language Models are Good Prompt Learners for Low-Shot Image Classification

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-08-06T18:27:47.588821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:46.900106Z digest=sha256:79129703e3eb49b62df54c85905859dc94af9a1aba81be0718ca59935f31dfe9

Observation b0972253-29f7-495d-b171-ee7e2656d84a · outbound

This paper cites Improved zero-shot classification by adapting vlms with text descriptions,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Improved zero-shot classification by adapting vlms with text descriptions,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:27:53.132833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:46.904781Z digest=sha256:1f7916b927b59a2f672d50ed3748c4844953cb8f78182dd6a25bca890510d184

Observation b51b309f-f8ee-4a08-aa8c-88b0b494d25e · outbound

This paper cites Bilateral adaptive cross-modal fusion prompt learning for clip,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Bilateral adaptive cross-modal fusion prompt learning for clip,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:27:53.023998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:46.909676Z digest=sha256:e86a5d2e6cf033076673f31e661e138a35b3a5cb72683a37a904ac0a5effa9bc

Observation 238e3490-284d-4bb2-961f-47d8d62720ee · outbound

This paper cites Unified Vision and Language Prompt Learning.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Unified Vision and Language Prompt Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T18:27:46.914058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:27:46.914058Z digest=sha256:859b1fce580920deaf65a9c6aadde86dabf7358be6c381c707b33a2af50b5591

Observation df39cd68-4808-4b93-84a3-376673743d68 · outbound

This paper cites Visual prompt tuning,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Visual prompt tuning,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T18:27:46.918620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:27:46.918620Z digest=sha256:90a14bfb0db02fee7dc4e5c55a1e33015d11262e060acc2890620030df62907d

Observation 8de8356e-06a5-4660-bfaa-dcb20f95c1c9 · outbound

This paper cites Enhancing clip with gpt-4: Harnessing visual descriptions as prompts,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Enhancing clip with gpt-4: Harnessing visual descriptions as prompts,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:27:52.914404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:46.922850Z digest=sha256:a45850d02eccbd741cf16a883f9423dddbedd8ad27d4ca586c88b85a2a3d4ca9

Observation 3bd84a38-a290-439d-98ea-de7097bf09bf · outbound

This paper cites Prompt distribution learning,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Prompt distribution learning,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:27:52.819718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:46.927549Z digest=sha256:6ab4a7fc03cd78178a813c66cd57f92f2896fb6aabade7e4dbba25f18bad6426

Observation 746ad6f6-bf5f-4710-a27b-837d46252273 · outbound

This paper cites Prompt-aligned gradient for prompt tuning,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Prompt-aligned gradient for prompt tuning,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:27:52.688397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:46.931748Z digest=sha256:55b1249dab79d9e2c0b4fe4b851e3f90c49b56044f4272bf8fdd5bfb23f0a62f

Observation ce721eb8-229b-4418-bd7f-ddf5a5307220 · outbound

This paper cites Visual-language prompt tuning with knowledge-guided context optimization,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Visual-language prompt tuning with knowledge-guided context optimization,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:27:52.589453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:46.936059Z digest=sha256:5b12e40a2142567a312874ed70056a007c81dc62fcdd649c809fa9f6be20980d

Observation bd6d3c8b-1362-4dc4-9d32-66c8364755c7 · outbound

This paper cites Gradient-regulated meta-prompt learning for generalizable vision-language models,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Gradient-regulated meta-prompt learning for generalizable vision-language models,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:27:52.463189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:46.940199Z digest=sha256:7d0ffe7ad05e2f76ef2147665ad9ccabfb697eaef9368a627005bba88f2fda0e

Observation 10ae3ddd-ef59-4170-a0e7-6ca883cffe58 · outbound

This paper cites Overcoming the Pitfalls of Vision-Language Model Finetuning for OOD Generalization.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Overcoming the Pitfalls of Vision-Language Model Finetuning for OOD Generalization

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T18:27:46.944447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:27:46.944447Z digest=sha256:0d64c3685848a88a3a10884a0cb73be741cbf7f774fa57cb0d0ce0a69276af5d

Observation 9a37ab55-322c-41fb-aa5d-67c6ecef6e1f · outbound

This paper cites Prompt Learning via Meta-Regularization.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Prompt Learning via Meta-Regularization

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-08-06T18:27:47.531123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:46.950350Z digest=sha256:4fa794db6a18836c61f41e6b2b48411e2d0675adf493be9c4f2057d91817037a

Observation 220f93bd-0fb4-4ca8-9531-dfae0fce96f5 · outbound

This paper cites Extract free dense labels from clip,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Extract free dense labels from clip,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T18:27:46.954920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:27:46.954920Z digest=sha256:1a817a27fc6947c576e297c691a833770883d60ddf13df28071c6e432dad6909

Observation 30df0cf1-407e-485a-b5cf-7dffa62d7dc8 · outbound

This paper cites Maskclip: Masked self-distillation advances contrastive language-image pretraining,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Maskclip: Masked self-distillation advances contrastive language-image pretraining,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:27:52.347774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:46.959853Z digest=sha256:d05ad4ee992f3d5e67da478fc6efd28870425b002d83160fd33a63df09cdeace

Observation 656ca6ad-35e3-44cb-8f23-7b855be7738e · outbound

This paper cites Scaling open-vocabulary image segmentation with image-level labels,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Scaling open-vocabulary image segmentation with image-level labels,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:27:52.187378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:46.964096Z digest=sha256:de5199eb7f44eea8e67774896333fe4ae5a6eafe1f0c688d9f0f33defb8e7e72

Observation 91fa11dc-3e8a-4aa1-a9ca-edfed4077d34 · outbound

This paper cites Clip-actor: Text-driven recom- mendation and stylization for animating human meshes,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Clip-actor: Text-driven recom- mendation and stylization for animating human meshes,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:27:52.049810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:46.968647Z digest=sha256:a80c474f44525c622280c087d6c6a1b8c077bc364877eec8d26b09414e22664a

Observation 0a060441-440b-47f6-a6c8-5f8ac15e06e3 · outbound

This paper cites Motionclip: Exposing human motion generation to clip space,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Motionclip: Exposing human motion generation to clip space,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:27:51.926665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:46.972909Z digest=sha256:09a13e29f261a9e33c0be8e1b3bdd688269e9a27f34dbc24a02d2f746c1bee84

Observation 4eecf22f-9870-4ade-b773-d53cbb9d45ec · outbound

This paper cites Lerf: Language embedded radiance fields,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Lerf: Language embedded radiance fields,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:27:51.794526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:46.977678Z digest=sha256:4197fa9ec0ca9b9d8b7ac87e3efe4344661f0b4448f7561b8e3a5e15b072d436

Observation 02d23dc6-b701-44f5-ada5-0e86715b5721 · outbound

This paper cites Zero-shot learning—a comprehensive evaluation of the good, the bad and the ugly,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Zero-shot learning—a comprehensive evaluation of the good, the bad and the ugly,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:27:51.672169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:46.982662Z digest=sha256:b061889c3d1abb4084a08c5b997c99e2e895250dfbd877ac932e5a1f0878a6d8

Observation 23b94d3d-e2e8-4244-8bbb-6e32f8929bca · outbound

This paper cites Zero-shot text-to-image generation,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Zero-shot text-to-image generation,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T18:27:46.986836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:27:46.986836Z digest=sha256:f24c667f3c147dc844eb9554f53ce72e5f67f056ded58328e5b38970e33bd18d

Observation 35d1dcec-ae34-4f33-bb68-d42d1619d1db · outbound

This paper cites High- resolution image synthesis with latent diffusion models,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models High- resolution image synthesis with latent diffusion models,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T18:27:46.991834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:27:46.991834Z digest=sha256:f87c16e340395d05cacd6503f1c5b31ff739ae0021b0f7c90f6c79728fda5bc3

Observation 6abaeceb-78db-4f94-a602-10c36459ebab · outbound

This paper cites Convolutions die hard: Open-vocabulary segmentation with single frozen convolutional clip,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Convolutions die hard: Open-vocabulary segmentation with single frozen convolutional clip,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:27:51.558422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:46.996424Z digest=sha256:8ebae9e1bc7c91b9510ee1c2264680207a2e6d4f0bd621aa9761761118bf0197

Observation dc50ac75-8e87-4e05-b27d-bdd526939afd · outbound

This paper cites Learning multiple visual do- mains with residual adapters,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Learning multiple visual do- mains with residual adapters,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:27:51.422176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:47.000463Z digest=sha256:2b4112ba5246f7e794133e21e1d5a909c64bceb188b3ad4d72c7299db1b066b8

Observation df3884a4-e868-4a13-b3f5-779d240ad034 · outbound

This paper cites Task residual for tuning vision-language models,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Task residual for tuning vision-language models,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T18:27:47.004756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:27:47.004756Z digest=sha256:42fa59e05cf24fab838191122d7d7f1edfb0ba5e87e576f609489987764c2585

Observation 1049f89e-7bdc-42f0-842a-40ef4abe2cfb · outbound

This paper cites Graphadapter: Tuning vision-language models with dual knowledge graph,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Graphadapter: Tuning vision-language models with dual knowledge graph,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:27:51.287187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:47.009320Z digest=sha256:f6bb1be1f2e5ca8bd426e1fae75c9b3aff656090bfd9bfa96e65efe911cc559c

Observation 1eda3c94-5ad5-4f4d-be25-959d22145c73 · outbound

This paper cites Tip-adapter: Training-free adaption of clip for few-shot classification,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Tip-adapter: Training-free adaption of clip for few-shot classification,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:27:51.184126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:47.014458Z digest=sha256:c7b71cc9b9163fc7e6d1320516b7d9096b4e33c0c9e355c3a7af9cfeab37d78c

Observation 1abc598d-c2f6-4616-ad26-a737a108eb74 · outbound

This paper cites Prompt, generate, then cache: Cascade of foundation models makes strong few-shot learners,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Prompt, generate, then cache: Cascade of foundation models makes strong few-shot learners,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:27:51.062183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:47.020022Z digest=sha256:0efca87375f5960e56e78443aa57fd6c89bd4bcbae47885c67c1ac9865cec10c

Observation 9f914995-6d80-4a1f-9ae2-225555a4edfd · outbound

This paper cites Scaling up visual and vision-language representation learning with noisy text supervision,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Scaling up visual and vision-language representation learning with noisy text supervision,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:27:50.934926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:47.024837Z digest=sha256:cb4efff7ea8050d75d2263fe00fb8e562caf5ea5ad5dbba8d200577d2eaf90f0

Observation a0da0837-f247-4fa4-961c-15e60e0e2d10 · outbound

This paper cites Clip-adapter: Better vision-language models with feature adapters,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Clip-adapter: Better vision-language models with feature adapters,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T18:27:47.029881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:27:47.029881Z digest=sha256:f59a66dba335132de12fe6bb4bdf8fc46cbfe512f5df0039e47635bc8556d264

Observation c4e96674-b6bb-4465-8a7b-65835d098491 · outbound

This paper cites Visual instruction tuning,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Visual instruction tuning,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:27:50.825910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:47.034432Z digest=sha256:25bcbf780381a836d81e5959fdb2b4516da494b5464d2336ccc2b289c5e7a85e

Observation ef0bc9f0-d96f-414c-8020-02d97cf13d3b · outbound

This paper cites Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:27:50.620422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:47.040924Z digest=sha256:d5c7a255603fbac55e5d17a2c9102bdca894b9ff02624bf269ce41b02b241c9f

Observation 28ae536f-a7b7-4040-a6a4-6e8d39f811f1 · outbound

This paper cites BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T18:27:47.045952Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:27:47.045952Z digest=sha256:8d4697d121ad042081f3f236147a3c9612d854daafc4c3d9f68eef6822ef4293

Observation fb9e63f8-2e1f-419d-83b6-5ea2e4e0e2bc · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T18:27:47.051880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:27:47.051880Z digest=sha256:16b4d8386db48c2d521a7d631c2f713c9f6c79d1211a70793f01d63563412ed6

Observation b6fbb7e0-9f7a-4505-a77c-8edc06cd4e2c · outbound

This paper cites MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T18:27:47.056475Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:27:47.056475Z digest=sha256:f084024eabf089715ea4b39bf94df054e2c7a447956630e92faba562ccca0433

Observation db69644f-4f1c-46dd-892f-a28717e3872e · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models LLaMA: Open and Efficient Foundation Language Models

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T18:27:47.061316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:27:47.061316Z digest=sha256:4586cc0551befb16934dce1fd69a99e6fb2b8788459c7fffa20cb4a1d0f61625

Observation 4ea8671e-27eb-4810-9f07-15b354a3fbe0 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T18:27:47.067589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:27:47.067589Z digest=sha256:adef62a0e4e30e6ea99331edcf8822eeacb615546a2a31f4b00e32c1291bca8f

Observation df2878d6-2386-40f8-a3be-cf412156b8f0 · outbound

This paper cites Language models are few-shot learners,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Language models are few-shot learners,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:27:50.483391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:47.072777Z digest=sha256:509c7878cc0a009dffdc25d7488b487a565cba08b98b4d0dc0ed721f5ffb61aa

Observation e21620ed-086e-4534-9205-d2ba8e2d569d · outbound

This paper cites GPT-4 Technical Report.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models GPT-4 Technical Report

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T18:27:47.077647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:27:47.077647Z digest=sha256:b0157458a2e6c837ade5cd67f757cd3a946ee67b25f9b1a9ac319211b8ef61c3

Observation da45060a-6976-45f4-af72-cf3165bc0341 · outbound

This paper cites Lit: Zero-shot transfer with locked-image text tuning,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Lit: Zero-shot transfer with locked-image text tuning,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:27:50.324687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:47.082212Z digest=sha256:797e12320d62ab78238302997f36cef3c53ff01afafd948f7ebce4b256fe5d53

Observation 35d68594-d399-4b80-9387-415d325d1bc9 · outbound

This paper cites CoCa: Contrastive Captioners are Image-Text Foundation Models.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models CoCa: Contrastive Captioners are Image-Text Foundation Models

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T18:27:47.087432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:27:47.087432Z digest=sha256:41c41f434360a8ac088989dad0710a7ea19894dcbf52d1397b901b884061c494

Observation 9f233f02-4d06-4e35-acbb-1f53b16ebe0c · outbound

This paper cites Compound text- guided prompt tuning via image-adaptive cues,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Compound text- guided prompt tuning via image-adaptive cues,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:27:50.217897Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:47.091790Z digest=sha256:c6b811c77504b09ad6aa1015a45a848a104c86498cc1645f380fa209d1e07a58

Observation e3af5be0-e7e6-4bbf-ad16-cf73bf0f6d94 · outbound

This paper cites What does a platypus look like? generating customized prompts for zero-shot image classification,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models What does a platypus look like? generating customized prompts for zero-shot image classification,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:27:50.121431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:47.095841Z digest=sha256:0c95c9ef9f8c813c451a6b304c52aa8652cca8b765922903ef730defbc1ad81d

Observation 058280ef-4874-4509-bcbf-2194a2a50869 · outbound

This paper cites Knowledge- aware prompt tuning for generalizable vision-language models,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Knowledge- aware prompt tuning for generalizable vision-language models,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:27:49.956199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:47.099862Z digest=sha256:0ef98216b092335cb38d63e530937ca7bcb9a21c73f911e58bc02c8e88c7480c

Observation 1d8720a8-38d4-4b3a-8478-00c46d016f16 · outbound

This paper cites Visual Classification via Description from Large Language Models.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Visual Classification via Description from Large Language Models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T18:27:47.103905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:27:47.103905Z digest=sha256:d0c3e1c84381ad74696acef987d21b4818d03094116d64d81637b5d240f8a209

Observation 75668555-2309-4fa6-bb4c-f2118b5d3851 · outbound

This paper cites Learning concise and descriptive attributes for visual recognition,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Learning concise and descriptive attributes for visual recognition,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:27:49.775912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:47.108072Z digest=sha256:a81c058ddaca61d681a529fd0f4fd8e65bce7b918c69a64a98ebd87113b4dcbe

Observation 34ada827-974f-4242-ba9b-9fd737e8d60b · outbound

This paper cites Language in a bottle: Language model guided concept bottlenecks for interpretable image classification,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Language in a bottle: Language model guided concept bottlenecks for interpretable image classification,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:27:49.621856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:47.112614Z digest=sha256:1f9e073e58d0aa35ec5e17c45d3017fbffab7384cbc554a59464634416dde182

Observation 50027c03-64e6-4f04-9744-23d2afa54312 · outbound

This paper cites CommonCanvas: An Open Diffusion Model Trained with Creative-Commons Images.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models CommonCanvas: An Open Diffusion Model Trained with Creative-Commons Images

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T18:27:47.118712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:27:47.118712Z digest=sha256:e1b82730eefece7a74124e2cfa2b83cf42fcbe618ae5c19fee801179888f3e09

Observation c6b5550b-2977-42f3-9ad0-d2c49cf999a3 · outbound

This paper cites Enhancing clip with a third modality,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Enhancing clip with a third modality,

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:27:49.445333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:47.123856Z digest=sha256:ce901c7c65fde438e5065ca289080486a06445a15de14de2dc168d6d7692aa30

Observation ec1c2b0c-570f-417f-8122-21e4f8930b80 · outbound

This paper cites Democratizing Fine-grained Visual Recognition with Large Language Models.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Democratizing Fine-grained Visual Recognition with Large Language Models

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T18:27:47.128776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:27:47.128776Z digest=sha256:3adae0f742d73080e010d884f79f2cf30e94a0d79ee4176d7e74397742574af9

Observation 47bfcda6-c303-4016-97c9-7eb8c868765a · outbound

This paper cites Fine- tuned clip models are efficient video learners,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Fine- tuned clip models are efficient video learners,

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:27:49.292751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:47.133211Z digest=sha256:3f15f7214e5964023b3830798b6b4ae5e58f04fdb2cfeea6aab124a952cdaa1d

Observation 14e7dd2a-a6e4-44ef-83af-2c44af5d7ca2 · outbound

This paper cites Imagenet: A large-scale hierarchical image database,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Imagenet: A large-scale hierarchical image database,

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-06T18:27:47.137318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:27:47.137318Z digest=sha256:a13b5798687e34bfe7c8b995b9cd7e351123038173728a86998925f26855ab11

Observation 475c5e0c-cfcc-4dfa-988d-dfcd5af801cf · outbound

This paper cites 3d object representations for fine-grained categorization,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models 3d object representations for fine-grained categorization,

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:27:49.180107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:47.141493Z digest=sha256:1a85abbf6488f58bfe330df4c2f33f82852ab6de4091613e9377f64db98b3783

Observation 9857915e-5ab4-4c8a-8e06-b24ec7478460 · outbound

This paper cites UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-06T18:27:47.145588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:27:47.145588Z digest=sha256:29010ea6c0e85067df290bf59c391ace3ee1e669d1c39c284352e473d2a8bdf8

Observation 15ca188b-e6c5-41de-ac53-07fcd31661af · outbound

This paper cites Learning generative visual models from few training examples: An incremental bayesian approach tested on 101 object categories,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Learning generative visual models from few training examples: An incremental bayesian approach tested on 101 object categories,

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T18:27:47.150373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:27:47.150373Z digest=sha256:9833d7f102841f4d2fb8492efadb3f5475ebfd1a0b7853f5a47f661ea1c8e993

Observation 0a4b130c-298f-4fa9-8c3b-a67156111844 · outbound

This paper cites Automated flower classification over a large number of classes,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Automated flower classification over a large number of classes,

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:27:48.995277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:47.155160Z digest=sha256:b40d3c54d6418c6faddeade24c980e03ee0747d561be7bb53f497d6a2608783e

Observation cfb0970c-d0ef-4a1e-9a4e-cdb51eac85ba · outbound

This paper cites Sun database: Large-scale scene recognition from abbey to zoo,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Sun database: Large-scale scene recognition from abbey to zoo,

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:27:48.855739Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:47.159461Z digest=sha256:46fff227129c150cba08d1bc05a6028374c46b0e331c1835e0989662d34b23b7

Observation cb1de294-5148-4188-80ce-d383d634eec2 · outbound

This paper cites Describing textures in the wild,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Describing textures in the wild,

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-06T18:27:47.164738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:27:47.164738Z digest=sha256:5cd92ee25a963389f8095dcf65a06512af041b12a8965f1f2f134c96a81e08a1

Observation a67ce6cd-e736-4acf-ab61-68a8ea592866 · outbound

This paper cites Eurosat: A novel dataset and deep learning benchmark for land use and land cover classification,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Eurosat: A novel dataset and deep learning benchmark for land use and land cover classification,

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-06T18:27:47.169601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:27:47.169601Z digest=sha256:d768f579ee8e605aa5b99425277b3b2da32449ac7177fddfb66a1c3f0b8493cf

Observation e59cc4cc-4243-4123-a5ec-3092aa3d8d3a · outbound

This paper cites Fine-Grained Visual Classification of Aircraft.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Fine-Grained Visual Classification of Aircraft

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-06T18:27:47.173763Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:27:47.173763Z digest=sha256:ff502bdecf1bd1aeb188355496826b1ed22608be18c1ceda5b57272e790c5405

Observation b3085d30-3809-40f2-8b98-d7c610e63ed1 · outbound

This paper cites Cats and dogs,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Cats and dogs,

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:27:48.630071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:47.178390Z digest=sha256:ca56eaf5dc1f7efde05aa804228dc4ca924b985c1e006b28966c9b0b49b70dac

Observation 909a32ad-7675-47a7-b735-0ae79adae21f · outbound

This paper cites Food-101–mining discriminative components with random forests,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Food-101–mining discriminative components with random forests,

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:27:48.360216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:47.182947Z digest=sha256:cb66a7f9633d565b29d8c4a8edbf74d9c91b7951afd6c2b9a696f4690d37fdf2

Observation 72fc57ec-122e-432a-9988-717872c231ea · outbound

This paper cites Natural adversarial examples,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Natural adversarial examples,

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-06T18:27:47.187502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:27:47.187502Z digest=sha256:4d08b7c3eff99e381abb453981f57e9c8a4b43eee25637b566dcfa304d948cc7

Observation cc9ce98a-bbfc-4d28-8fe6-d2c3c6a19bb2 · outbound

This paper cites Do imagenet classifiers generalize to imagenet?.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Do imagenet classifiers generalize to imagenet?

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:27:48.217435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:47.191888Z digest=sha256:907f4fc451b943205044ea6769e0d6793b26ac95a525dac125212cd92cfb70ed

Observation 65100d47-d282-481d-8bd0-386ace54c4d1 · outbound

This paper cites Learning robust global representations by penalizing local predictive power,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Learning robust global representations by penalizing local predictive power,

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:27:48.073216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:47.196049Z digest=sha256:2e352d0fdf187fa88fb7a7fb43327316b038766152332b3ce55d4ad4662692b0

Observation 3f9b258a-8bca-4d7d-a7a9-79e5628ee05f · outbound

This paper cites Decoupled Weight Decay Regularization.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Decoupled Weight Decay Regularization

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-06T18:27:47.200213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:27:47.200213Z digest=sha256:215f0f438ac6a50e1880a1c9181004be8800e4491f795b18c36d8d924805cef3

Observation ae1a09a8-1962-4d48-a8aa-019755f1e3df · outbound

This paper cites VL-Mamba: Exploring State Space Models for Multimodal Learning.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models VL-Mamba: Exploring State Space Models for Multimodal Learning

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-06T18:27:47.204582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:27:47.204582Z digest=sha256:98d86046a25d18efc6d212105a541e798d1026fd2615bb968da1b64d70454c83

Observation fbef7d83-68a3-4d39-aaeb-6e6f10e51d9c · outbound

This paper cites Grad-cam: Visual explanations from deep networks via gradient-based localization,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Grad-cam: Visual explanations from deep networks via gradient-based localization,

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-06T18:27:47.208713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:27:47.208713Z digest=sha256:884d9e0be31f1e5de588b75a2b61503ce4960db2f6e6e6743f6a5dd3765c4088

Observation c553cb0d-8a20-4659-9ea2-1ba1c4a0a9f7 · outbound

This paper cites Efficiently scaling transformer inference,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Efficiently scaling transformer inference,

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-06T18:27:47.212991Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:27:47.212991Z digest=sha256:f94d685bf4e99440ebf3b144dcb01a5ef081b73d1d067d1962ab045b41ee0ceb

Observation 5a4c514e-fda3-44e4-acaf-7f7137416859 · outbound

This paper cites Transformers: State-of-the-art natural language processing,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Transformers: State-of-the-art natural language processing,

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:27:47.920034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:47.217153Z digest=sha256:abc4093d8c7eafb1e67ca099c4e9092fd19129ce742b5dd096315280d88ac322

Observation 934244d5-50ea-42da-a768-6dbe291ce291 · outbound

This paper cites Modality- consistent prompt tuning with optimal transport,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Modality- consistent prompt tuning with optimal transport,

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:27:47.791480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:47.222518Z digest=sha256:0d76af68f07fb0680f65e1e335735ef2571ab5bf803d0256671d373610d83a04

Observation c1c080de-e9d9-4f7d-8580-e4b6fec02d7c · outbound

This paper cites Hierarchy- aware interactive prompt learning for few-shot classification,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Hierarchy- aware interactive prompt learning for few-shot classification,

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:27:47.761692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:47.226756Z digest=sha256:c91dc0697f00ab1aad5db9afbb22621c31f8dd92394f07ee00f1cb4001cd00fd

Observation b8cd94c3-82c5-4186-96e5-eb32b176a837 · outbound

This paper cites Language-driven visual consensus for zero-shot semantic segmentation,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Language-driven visual consensus for zero-shot semantic segmentation,

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:27:47.737604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:47.232099Z digest=sha256:6d9ed07dbeced8971988a35a217df00400726604ac7553fff16948f5e806bb93

Observation 0626ab68-3e72-4287-8bd3-3356f62aaa44 · outbound

This paper cites Pedestrian attribute recognition via clip based prompt vision-language fusion,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Pedestrian attribute recognition via clip based prompt vision-language fusion,

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:27:47.711586Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:47.236731Z digest=sha256:6cadaf8eed9fb0bced9dbe16d1b7c8b8542ef35f83ee28193f911d95d01fd160

Observation d0247ef0-75df-4c82-9424-44f458c37556 · outbound

This paper cites Understanding and mitigating overfitting in prompt tuning for vision-language models,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Understanding and mitigating overfitting in prompt tuning for vision-language models,

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-06T18:27:47.240681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:27:47.240681Z digest=sha256:27ee79d0f99a166eeb35c79cc40d03d58b940c8f32f65f78f2c49c1e56ee380d

Observation 536fa8b1-0817-4ccd-873a-84eb3ba6cabb · outbound

This paper cites Clipood: Generalizing clip to out-of-distributions,.

Multi-modal Mutual-Guidance Conditional Prompt Learning for Vision-Language Models Clipood: Generalizing clip to out-of-distributions,

Reference 86

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T18:27:47.672562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:27:47.245370Z digest=sha256:1630ae7ecf77c03e9458859b48b2de8d075f52d13d52925b21b6b3e341c9398f

Pith citing papers

No inbound Pith citation observations are available.