Pith. sign in

Paper Citation Record · LEDGER

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs

As of 5 August 2026, this Paper Citation Record lists 63 of 63 outbound references and 0 inbound Pith citation observations for arXiv:2511.16107.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2511.16107 v4

Coverage vector

measured 63 of 63 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-03T21:15:48.046423Z

measured 63 of 63 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

63 of 63 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved62
  • parse uncertain1
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 28cc01dc-8820-46e3-915b-38811a7d9e54 · outbound

This paper cites an unresolved cited work.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Unresolved cited work

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:43.561379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:43.561379Z digest=sha256:3f643ab578cf1369350f1ac9c9fd01f2a205dc32e7e169860a0e51ed8867a267

Observation e05665d2-dedb-4bcd-96a9-4d3dcf824644 · outbound

This paper cites an unresolved cited work.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:43.632692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:43.632692Z digest=sha256:3bad4dd01a5473014504010c052261ebbbc3a4ba89d564c6e863b491399d919d

Observation ad298996-07d7-4aad-a32f-45c0197d4585 · outbound

This paper cites Fowlkes, Ste- fano Soatto, and Pietro Perona.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Fowlkes, Ste- fano Soatto, and Pietro Perona

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:43.740549Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:43.740549Z digest=sha256:d454ba1d9d74d63355dd66304de31063879824354949e76a0507b775e8b1efda

Observation cbc73e69-be17-4c88-987d-235fbe38c222 · outbound

This paper cites NTIRE 2017 chal- lenge on single image super-resolution: Dataset and study.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs NTIRE 2017 chal- lenge on single image super-resolution: Dataset and study

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:43.789703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:43.789703Z digest=sha256:9534dd8944fde940697d07bfb67908212d6963b9b1a7cf2ae59c79181ebfd032

Observation 7bf6fd37-55b3-406c-9817-51df7aef5cd6 · outbound

This paper cites Ancuti, and Christophe De Vleeschouwer.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Ancuti, and Christophe De Vleeschouwer

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:43.845221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:43.845221Z digest=sha256:e90d24e4c3b414ff4d7f6270dfa282a7dd47075023834b7b05fae5d547292fcf

Observation eb5b3e28-8234-47b6-81fd-3c4a806dea1e · outbound

This paper cites an unresolved cited work.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:43.932541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:43.932541Z digest=sha256:b1c48b38bcd92a93c9d2d72f2d85e9f80746305d82434bdb15cd3114cd9d6f2f

Observation b8c0e265-fee8-4971-ad9f-fac7562bc3af · outbound

This paper cites an unresolved cited work.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:44.004358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:44.004358Z digest=sha256:92adf3eafe3be88b90f70fe6e10f1f4e6fba32f0fca0bf71ae2ff6875a77dc82

Observation 30f2fecd-ae61-46ca-b295-fe4bf9ed72d9 · outbound

This paper cites an unresolved cited work.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:44.056904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:44.056904Z digest=sha256:d255695f0f59f427c550944d8cf544514daefd80d7a1d2875e0bbe6fe40be064

Observation 02c556d1-acba-4d79-89dc-27849733a6c6 · outbound

This paper cites Fleet, and Geoffrey E.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Fleet, and Geoffrey E

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:44.151771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:44.151771Z digest=sha256:e9f478ce0926f7915ffe22db03c9c6cd65549a8c3f7f12549012c5e440d25a83

Observation f48c2da2-5215-4fb5-a9b8-216c79953cd9 · outbound

This paper cites an unresolved cited work.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:44.220036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:44.220036Z digest=sha256:3595a65875bf814809cc2ceebaad458cf78429cc1b3eb665270bb6db5539c529

Observation 1fda4edb-a382-4d55-90c9-9e7289bbc99c · outbound

This paper cites Conde, Gregor Geigle, and Radu Timofte.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Conde, Gregor Geigle, and Radu Timofte

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:44.267141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:44.267141Z digest=sha256:5792451bae85b49bbd350134a382fe7d01a971d5da8965a09c83a9c86e856cfc

Observation dedd65a9-3cb9-4b54-80ec-392d16977c99 · outbound

This paper cites Image Harmonization Dataset iHarmony4: HCOCO, HAdobe5k, HFlickr, and Hday2night.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Image Harmonization Dataset iHarmony4: HCOCO, HAdobe5k, HFlickr, and Hday2night

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:44.327675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:44.327675Z digest=sha256:af5e51967544b7b7baf2d495fb612dc3212eef1e770b0b69f613cd32ca15c424

Observation bafd3f54-ebf3-4ae8-9b83-14d5250f73f3 · outbound

This paper cites A survey on in-context learning.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs A survey on in-context learning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:44.379544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:44.379544Z digest=sha256:768632be0935ba95e5a530b993e25612de6e5153a7e4c5f1905bfa79190a4c9e

Observation 3397c0f9-b21f-4a9f-9c94-ef1a9ada6b65 · outbound

This paper cites Representation similar- ity analysis for efficient task taxonomy & transfer learning.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Representation similar- ity analysis for efficient task taxonomy & transfer learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:44.432405Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:44.432405Z digest=sha256:4d5255a6e7e4529ab6f1a17ba094c5bb4dd03fcfea0c96f700bdecd50498a3ff

Observation 72661d17-18a1-4816-8091-3798fdb59900 · outbound

This paper cites an unresolved cited work.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:44.508829Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:44.508829Z digest=sha256:ae8fe7a284850ec00064187e55535c640543c2be6437d44cea9210d9e2a4b8d9

Observation ac818eb8-3eea-4060-986f-c7476f267631 · outbound

This paper cites Instructdiffusion: A generalist modeling interface for vision tasks.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Instructdiffusion: A generalist modeling interface for vision tasks

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:44.596562Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:44.596562Z digest=sha256:3f6eed29611e20f80fbc6259f758bb9a80753c60da85b39f317e32cf3b8217b3

Observation aa9f8100-31d0-405a-9489-09ccdd5f34df · outbound

This paper cites In-context learning in large language mod- els: A comprehensive survey.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs In-context learning in large language mod- els: A comprehensive survey

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:44.645302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:44.645302Z digest=sha256:c13d913fb3a59b934ae526db1192428de93ccd2847c109c12a58635f9c60ffc3

Observation 1977744a-8b8e-4478-97e3-c93019014b04 · outbound

This paper cites Depth-attentional features for single-image rain removal.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Depth-attentional features for single-image rain removal

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:44.736003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:44.736003Z digest=sha256:4ac12ba2499a2f74b792fc4c12d27d3ca4393290f1c83e0726f0b40cbcd1fd44

Observation 5e4da351-d396-4a80-898c-94053d21f4cd · outbound

This paper cites Belongie, Bharath Hariharan, and Ser-Nam Lim.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Belongie, Bharath Hariharan, and Ser-Nam Lim

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:44.804254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:44.804254Z digest=sha256:037da81bd95b69fee641a094f5edb3b23acbc0190476850f12730e16eaa02403

Observation 47959791-ac2d-43ea-b410-8b12081eb1c3 · outbound

This paper cites an unresolved cited work.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:44.862806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:44.862806Z digest=sha256:36e39ec81f42b85169af2379ca803c964b6a407a8b75fd840c1555f37223b629

Observation 7197528e-8865-4c55-a962-cea153970c09 · outbound

This paper cites Uvim: A unified modeling approach for vision with learned guid- ing codes.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Uvim: A unified modeling approach for vision with learned guid- ing codes

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:44.959526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:44.959526Z digest=sha256:3409b8d5241973c89ffc72751c60ccc559ceae1545e34aa92b4026e930a86607

Observation c10bcbfa-7062-4328-8372-ae8fd975ed1e · outbound

This paper cites Viescore: Towards explainable metrics for condi- tional image synthesis evaluation.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Viescore: Towards explainable metrics for condi- tional image synthesis evaluation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:45.018624Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:45.018624Z digest=sha256:6c202cf24ccd228c0daa0de89b6d926404dbc7877d5acf0f42a255f82c0d259f

Observation 4297a91d-dec2-4dfb-ab27-109907b2221a · outbound

This paper cites Transient attributes for high-level under- standing and editing of outdoor scenes.ACM Transactions on Graphics (TOG), 33(4):149:1–149:11, 2014.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Transient attributes for high-level under- standing and editing of outdoor scenes.ACM Transactions on Graphics (TOG), 33(4):149:1–149:11, 2014

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:45.074204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:45.074204Z digest=sha256:1d00e13ea407e8f84c919327b6f39b14184ad06a328923570c200381123a1149

Observation 6654e740-bde6-4ed1-8771-c07b88371e69 · outbound

This paper cites an unresolved cited work.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:45.148597Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:45.148597Z digest=sha256:5d3c1186de0d6e388e2769727e7214f919a4c1ea6e8c3d76e780a319de9d51b8

Observation eb9aae68-33fd-4f0c-a762-5a757f05efcb · outbound

This paper cites Large lan- guage model-aware in-context learning for code generation.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Large lan- guage model-aware in-context learning for code generation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:45.245867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:45.245867Z digest=sha256:45f250ad8ac4381065973daf256bfbe045817b73b6a9961bfe554fd07d0cd7cd

Observation 0179b96d-1bae-4b4e-9b7f-0114cfeaa83f · outbound

This paper cites Unifying image processing as visual prompting question answering.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Unifying image processing as visual prompting question answering

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:45.319044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:45.319044Z digest=sha256:3713ee507ed4a69dadf8fb0a97658a9f3059c84c14e7c10637df86a8ce88497d

Observation e5da5647-e731-4f53-ac1b-4c3a95dcc400 · outbound

This paper cites UNIFIED-IO: A unified model for vision, language, and multi-modal tasks.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs UNIFIED-IO: A unified model for vision, language, and multi-modal tasks

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:45.387266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:45.387266Z digest=sha256:f0ba83c577776e15c4b40e8cd23b58771591fdff9d926554840c76c0373dbb32

Observation a0d16ea5-200b-45fb-acc8-003ba827f9b5 · outbound

This paper cites Vi- sion language models are in-context value learners.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Vi- sion language models are in-context value learners

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:45.461712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:45.461712Z digest=sha256:90ab0264b36dd7872bb2ff3f225668f9fb1c84a2512fa19787da1743eafabcc4

Observation cb23f0a2-d01f-46a5-a70e-71b911273863 · outbound

This paper cites an unresolved cited work.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:45.510210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:45.510210Z digest=sha256:73495e5e29b6cad96173fdcaccffa19a98cb5a34ca3c33fb358081860b8d6f06

Observation 3cfc4889-8bfc-4c39-9965-172efbe9c667 · outbound

This paper cites Deep multi-scale convolutional neural network for dynamic scene deblurring.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Deep multi-scale convolutional neural network for dynamic scene deblurring

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:45.583626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:45.583626Z digest=sha256:406e8379e821bfdbe7aa97ea0c72f1a5f022a64d7edc885b3a3330ca55278192

Observation 678ca0c8-c38f-4fe7-bc54-6f713fff697f · outbound

This paper cites GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:45.641226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:45.641226Z digest=sha256:8836202d0c0c9e322cda513d65c2535974964dc0c1d90c69840a6e3a8a85a5dd

Observation fa748d22-7bf6-42df-a122-31476a660873 · outbound

This paper cites Balasubramanian.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Balasubramanian

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:45.725184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:45.725184Z digest=sha256:e13f82dcedbec0198be5bfe4904b4bc80788ab9988d9aad584df263577ac4b02

Observation 8c3b158a-2492-401b-af83-b11e09f39a3e · outbound

This paper cites Khan, and Fahad Shahbaz Khan.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Khan, and Fahad Shahbaz Khan

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:45.785859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:45.785859Z digest=sha256:458ea59093777e046d7447f4c0f7d1f2371acbdf733d37af51afbfccf0be0e1e

Observation ba005813-f78a-48e0-ba8d-c53019c4f659 · outbound

This paper cites SPIRE: semantic prompt-driven image restoration.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs SPIRE: semantic prompt-driven image restoration

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:45.864189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:45.864189Z digest=sha256:703bfd3c460ef6ddfe04005d28367db2a0f39da12862cb14dcbf070491fe8f0d

Observation a07499ed-7e28-4423-9e9d-7d43d9c3e33a · outbound

This paper cites Visual Text Meets Low-level Vision: A Comprehensive Survey on Visual Text Processing.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Visual Text Meets Low-level Vision: A Comprehensive Survey on Visual Text Processing

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:45.940245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:45.940245Z digest=sha256:c2cde3507ff86d24d6fd0d569a7bf11a73c80a2bdc8de362bfec44b6568c37d7

Observation 207b2e47-9340-48b7-8767-11b854b4681a · outbound

This paper cites an unresolved cited work.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Unresolved cited work

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:46.011936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:46.011936Z digest=sha256:5d20aa0e8bb1b2c6164da7ff6a5bef09f99fde8fb734abb90969d554630ebf65

Observation cf765669-a146-415a-8ac7-bf6338b4576f · outbound

This paper cites Guibas, Jitendra Malik, and Silvio Savarese.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Guibas, Jitendra Malik, and Silvio Savarese

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:46.079384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:46.079384Z digest=sha256:33e7f031ce2f767baf0a14cf3e52cff9defe5382b9543a2b9c485a6405a73b2d

Observation 8ea03b61-530b-4012-bfa5-3d966b8f53f3 · outbound

This paper cites Emu: Generative pretraining in multimodality.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Emu: Generative pretraining in multimodality

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:46.155658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:46.155658Z digest=sha256:c15c310fc8e25428a82bd1984824fa64c7d07a2cc3b251ce55ddbf4b1bbe1cde

Observation bf4091d7-c1f8-4a74-b56b-357c74f87fa0 · outbound

This paper cites X-Prompt: Towards Universal In-Context Image Generation in Auto-Regressive Vision Language Foundation Models.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs X-Prompt: Towards Universal In-Context Image Generation in Auto-Regressive Vision Language Foundation Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:46.208071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:46.208071Z digest=sha256:47a5c3b7c01b7345d7f2640af398ae14fb9644e97055467a498e21af4af56588

Observation 0dc6d36e-b72d-405a-b2fc-d2f5b7163455 · outbound

This paper cites Axiomatic attribution for deep networks.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Axiomatic attribution for deep networks

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:46.256952Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:46.256952Z digest=sha256:2e9682ce02fa2d9276d368e786f0fae8b42ef797075340eaf10896b229adf75b

Observation 90eeabf9-f3db-4aa1-89bd-c17285ba073b · outbound

This paper cites Chameleon: Mixed-Modal Early-Fusion Foundation Models.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Chameleon: Mixed-Modal Early-Fusion Foundation Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:46.316714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:46.316714Z digest=sha256:01f14f72a525c45e07d43b80100d4e82e85d5b5572ee07eeb23f9c0424407b80

Observation 4b85c6e3-a3a1-46ad-bdeb-1df327130680 · outbound

This paper cites A comprehensive survey of deep learning approaches in image processing.Sensors, 25 (2):531, 2025.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs A comprehensive survey of deep learning approaches in image processing.Sensors, 25 (2):531, 2025

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:46.388924Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:46.388924Z digest=sha256:abf377ef099365d306b7d8ffd5f9c6a3ef654fbae84757bfdc112335b5208b1d

Observation 229b67db-c76b-46d6-a19a-a3cf267ee381 · outbound

This paper cites an unresolved cited work.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Unresolved cited work

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:46.453165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:46.453165Z digest=sha256:4fa160f202dfd4605fca998a17996756c32f9a7a8babb5e2fb9e4ff0e048c2d4

Observation ecc30141-159b-48d8-b24b-7de26501ea61 · outbound

This paper cites Stacked conditional generative adversarial networks for jointly learning shadow detection and shadow removal.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Stacked conditional generative adversarial networks for jointly learning shadow detection and shadow removal

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:46.529451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:46.529451Z digest=sha256:0c4effb7f0a845afd00a04428c9348627b12cbe79dde039b39510e46eeb844fc

Observation 27b1c682-db7e-4ad9-9d35-8824d86c81b8 · outbound

This paper cites OFA: unifying architectures, tasks, and modalities through a simple sequence-to-sequence learning framework.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs OFA: unifying architectures, tasks, and modalities through a simple sequence-to-sequence learning framework

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:46.579921Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:46.579921Z digest=sha256:cc4ae2360cecb91454fd90ef9a8fb7dcfc345cfd2330be766a410e908ba81cf6

Observation 34f364e2-02df-4a36-b3ad-8a792821a7d8 · outbound

This paper cites Images speak in images: A generalist painter for in-context visual learning.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Images speak in images: A generalist painter for in-context visual learning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:46.634753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:46.634753Z digest=sha256:834f0fb35d254a09f96cb9f449fc354d2df571d7859d8e7675addb429618bea0

Observation 820d4cc0-7a52-4178-91d8-4ae6e38852f2 · outbound

This paper cites Large language models are latent variable models: Explaining and finding good demonstra- tions for in-context learning.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Large language models are latent variable models: Explaining and finding good demonstra- tions for in-context learning

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:46.689960Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:46.689960Z digest=sha256:e3a9a030341ab2d853e0fb03ef34bb52ac85dfee8657e1f4a68e6561eba47955

Observation 9c86a16e-fa58-4310-847c-cef2c6081df0 · outbound

This paper cites Emu3: Next-Token Prediction is All You Need.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Emu3: Next-Token Prediction is All You Need

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:46.744566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:46.744566Z digest=sha256:3265d5eed65ea854abf895e3ec7a29974c071544eebda4f08673f351243e7104

Observation ce5e4c5a-dabd-4d72-aac5-5342361ba74a · outbound

This paper cites In-context learning unlocked for diffusion models.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs In-context learning unlocked for diffusion models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:46.793661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:46.793661Z digest=sha256:380d5d25301d0041c42757cf80c5fefdb6932d0683048f5d597876d4672a96b9

Observation 01eb31f2-3bd1-41f6-a5a8-b9a5db72465f · outbound

This paper cites Deep retinex decomposition for low-light enhancement.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Deep retinex decomposition for low-light enhancement

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:46.844948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:46.844948Z digest=sha256:b0c5c64d2cdaa0cd10a3c5acfde3b4baab2a0e79bc20ea0c95e825ab464885c9

Observation fe972cb7-c539-45b3-980a-ee39da3cef36 · outbound

This paper cites Florence-2: Advancing a unified representation for a vari- ety of vision tasks.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Florence-2: Advancing a unified representation for a vari- ety of vision tasks

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:46.957592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:46.957592Z digest=sha256:e46d2b3c3978ebc2ad29f74d4c895e33170cb154b28fd37f677a128259140286

Observation b4ab59c7-bf54-4621-a769-df361d11dcca · outbound

This paper cites Show-o: One single transformer to unify multimodal understanding and generation.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Show-o: One single transformer to unify multimodal understanding and generation

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:47.015077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:47.015077Z digest=sha256:fabd10fba694f0e41d68d19e9f1a4df32ae6296c5880219eed3d4fae9145d420

Observation b4b0403c-385d-4e5c-b232-bb8bb58551c5 · outbound

This paper cites Towards efficient and scale-robust ultra-high-definition image demoir ´eing.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Towards efficient and scale-robust ultra-high-definition image demoir ´eing

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:47.085316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:47.085316Z digest=sha256:b74a31d2e3c7665644c2e84e9ffd9e6b29b1fb21c2759f268d61d3c7e641fd0a

Observation 13c4481e-41f8-4bd2-b327-0c3cb4bf9bfc · outbound

This paper cites Promptfix: You prompt and we fix the photo.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Promptfix: You prompt and we fix the photo

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:47.178044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:47.178044Z digest=sha256:114b5e0d6f110a10e425a829ff35ecad1b2ba5ac7c84fddcf1b0ca4b665365e9

Observation f026a556-cb9a-431c-9e74-ac672192f7f0 · outbound

This paper cites Zamir, Alexander Sax, William B.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Zamir, Alexander Sax, William B

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:47.279237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:47.279237Z digest=sha256:45be41ede39e3cf14fa4e761251e44a7707edad8087954872d00cd43addd9da6

Observation 58cded30-3857-4cf6-8aef-1e96315c1b84 · outbound

This paper cites A comprehensive evaluation of full reference image quality as- sessment algorithms.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs A comprehensive evaluation of full reference image quality as- sessment algorithms

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:47.389513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:47.389513Z digest=sha256:ac1161596366767941d6399cd1cbe342ae2cedb6f440c3e362eca27a50432015

Observation 298a173f-4222-4a55-b2d3-3cf425b47cd2 · outbound

This paper cites Perceive-ir: Learning to perceive degrada- tion better for all-in-one image restoration.IEEE Transac- tions on Image Processing, 2025.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Perceive-ir: Learning to perceive degrada- tion better for all-in-one image restoration.IEEE Transac- tions on Image Processing, 2025

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:47.491890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:47.491890Z digest=sha256:64f342959eea34d49d501aebb723acf42f7d6911673954c66e6659502c5fbefe

Observation 21f9beef-f07c-4655-abed-0bd42d236740 · outbound

This paper cites Transfusion: Pre- dict the next token and diffuse images with one multi- modal model.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Transfusion: Pre- dict the next token and diffuse images with one multi- modal model

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:47.600267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:47.600267Z digest=sha256:e001200b98722440342d5731a7b9d6ba88f009e42d3130b25f3d67a203d3f5da

Observation 49631359-3237-468b-8fa6-b2d38e3fb04f · outbound

This paper cites Learning to prompt for vision-language models.In- ternational Journal of Computer Vision, 130(9):2337–2348,.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Learning to prompt for vision-language models.In- ternational Journal of Computer Vision, 130(9):2337–2348,

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:47.707533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:47.707533Z digest=sha256:c82068d67d79d44e16c2a20725669ae5c4d35471efd2e677f3ee49dacfdc3b07

Observation 0f1bdd2c-17de-42f7-9ea8-7f98346f5fb8 · outbound

This paper cites Seeing the unseen: A fre- quency prompt guided transformer for image restoration.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Seeing the unseen: A fre- quency prompt guided transformer for image restoration

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:47.778161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:47.778161Z digest=sha256:10b6433b74e6a7c45fed492109b266d4cabca52fcdb4a9305b3a534a88640936

Observation b3dfca24-33d9-4f43-a878-d7c309e98f60 · outbound

This paper cites Visual in-context learning for large vision-language models.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Visual in-context learning for large vision-language models

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:47.892191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:47.892191Z digest=sha256:b164d708df9dfa15589425656b40f33d134e28a4e8761a6a124a061aa736eec4

Observation 2d099f1f-78f2-4100-8599-1ec84cc60fc9 · outbound

This paper cites Uni-perceiver: Pre- training unified architecture for generic perception for zero- shot and few-shot tasks.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Uni-perceiver: Pre- training unified architecture for generic perception for zero- shot and few-shot tasks

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:48.046423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:48.046423Z digest=sha256:d20720d761de33f2e902c4803f270b165f6661269000a98fe0f8a4209965ad03

Observation 51033315-3767-4c82-bb7d-eed2d1292d07 · outbound

This paper cites an unresolved cited work.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs Unresolved cited work

Reference 155

Resolution
parse uncertain
no resolver link, observed 2026-08-03T21:15:46.877076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:46.877076Z digest=sha256:bae6521dff59ac5776baff25313d72675aa57494c17c9428ef66895c6c4f05b9

Pith citing papers

No inbound Pith citation observations are available.