Pith. sign in

Paper Citation Record · LEDGER

Zero-Shot Vision Encoder Grafting via LLM Surrogates

As of 8 August 2026, this Paper Citation Record lists 53 of 53 outbound references and 0 inbound Pith citation observations for arXiv:2505.22664.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.22664 v2

Coverage vector

measured 53 of 53 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:10:11.740399Z

measured 53 of 53 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

53 of 53 outbound references displayed

  • verified exact0
  • verified fuzzy37
  • unresolved16
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 260503fb-70c9-419e-9346-d45496269371 · outbound

This paper cites Understanding Inter- mediate Layers Using Linear Classifier Probes.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Understanding Inter- mediate Layers Using Linear Classifier Probes

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:24.280642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:10:03.570614Z digest=sha256:d77d8b9226a0877f6b5c3e805fdba9a9be023cfc29a91da0e318572fff4acd24

Observation a5d9bd1d-3cfd-476b-8415-2836dd75b673 · outbound

This paper cites Flamingo: a Visual Language Model for Few-Shot Learning.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Flamingo: a Visual Language Model for Few-Shot Learning

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:23.894969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:10:03.714649Z digest=sha256:3cc6b19d77df218e49a4267f0c82bfd3ca245f3de00dee2612c74082d4aef018

Observation 9544eb4d-59cd-40dc-a31a-a6342a5cd088 · outbound

This paper cites Eliciting Latent Predictions from Transformers with the Tuned Lens.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Eliciting Latent Predictions from Transformers with the Tuned Lens

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:03.928054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:03.928054Z digest=sha256:8f5b23424291c862eeed0c5b51c849a9abe85f580c559bc82d7bd16305eaf6ed

Observation 362a970d-0fc7-457c-9579-d940edc8e5a0 · outbound

This paper cites PIQA: Reasoning about Physical Commonsense in Natural Language.

Zero-Shot Vision Encoder Grafting via LLM Surrogates PIQA: Reasoning about Physical Commonsense in Natural Language

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:23.407238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:10:04.074749Z digest=sha256:23a4b6587423b6984984fd702f3aae98f81c3ed88ebc42e804f5624e9e5afb54

Observation 0038ccda-7977-4d22-928b-75fbc3ad31fe · outbound

This paper cites GenQA: Generating Millions of Instructions from a Handful of Prompts.

Zero-Shot Vision Encoder Grafting via LLM Surrogates GenQA: Generating Millions of Instructions from a Handful of Prompts

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:04.258432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:04.258432Z digest=sha256:712b1e14be8e112df6c28e6249f0be59caa4efb4ede26737b3ff4a0a16442201

Observation 23ecf2b9-8082-4abd-9ce3-97b81ab43e7a · outbound

This paper cites InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks.

Zero-Shot Vision Encoder Grafting via LLM Surrogates InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:23.087790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:10:04.453895Z digest=sha256:b121424c14e77f417084351bc4c507ef84a1c8f17865465c1eb17f1203280261

Observation 15b07d9f-64ed-41f7-9924-daad416ba5bc · outbound

This paper cites BoolQ: Exploring the Surprising Difficulty of Natural Yes/No Questions.

Zero-Shot Vision Encoder Grafting via LLM Surrogates BoolQ: Exploring the Surprising Difficulty of Natural Yes/No Questions

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:22.825369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:10:04.575468Z digest=sha256:8e6b91d35ca7ef4adf22d182b23fc232d6e7f2f7df3758f523da3bfa4cab7ac4

Observation d5b5b0a8-89d3-4666-8664-009a8d53eee2 · outbound

This paper cites Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:04.718039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:04.718039Z digest=sha256:88fa47f968e8988161b8633308bd5dd28f9b4f63608d4db611e753555f37a3ac

Observation 01fa7e38-1e5d-4ef9-93a5-901f89510876 · outbound

This paper cites The Llama 3 Herd of Models.

Zero-Shot Vision Encoder Grafting via LLM Surrogates The Llama 3 Herd of Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:04.898522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:04.898522Z digest=sha256:9fa52b05e6a1dc3bea04ee0d520e75f4114536f27a2344893eac06c01338e135

Observation 5b4a33fc-eb7a-4283-b220-365723fedb31 · outbound

This paper cites MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models.

Zero-Shot Vision Encoder Grafting via LLM Surrogates MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:04.974166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:04.974166Z digest=sha256:55eac41af4f9761251ff374323dac34c80b10b0065f6df13ed8c2c85ce325982

Observation 551b002b-ce61-45cd-b6e5-ac28a7b6aded · outbound

This paper cites A Framework for Few-Shot Language Model Evaluation, 2024.

Zero-Shot Vision Encoder Grafting via LLM Surrogates A Framework for Few-Shot Language Model Evaluation, 2024

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:22.629018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:10:05.094039Z digest=sha256:b39bfdaab85d4cde2bc4ebc0c176df2ac67050fc512eac5637f8ddc9dbac2852

Observation d1a63bc4-cf48-4255-bdb5-778f03175f2e · outbound

This paper cites The Unreasonable Ineffectiveness of the Deeper Layers.

Zero-Shot Vision Encoder Grafting via LLM Surrogates The Unreasonable Ineffectiveness of the Deeper Layers

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:22.422088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:10:05.218644Z digest=sha256:849e85cffc6b285d2e2597736f12b58c0513417b109ae41f96a41f491ae773f2

Observation b946c6a1-1a01-4cf3-a6f8-e9132388e955 · outbound

This paper cites VizWiz Grand Challenge: Answering Visual Questions from Blind People.

Zero-Shot Vision Encoder Grafting via LLM Surrogates VizWiz Grand Challenge: Answering Visual Questions from Blind People

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:22.118148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:10:05.387611Z digest=sha256:724c325ebe668b71e0d6dba40f83a9a66541aad0ec6b45c07af3b38aea122ac1

Observation b82948d8-89bc-4133-8fea-961338574cef · outbound

This paper cites Word Embed- dings Are Steers for Language Models.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Word Embed- dings Are Steers for Language Models

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:21.805678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:10:05.521073Z digest=sha256:8fb814fc78200aa99d6427bb00ade0a819829bf856f568bc8829ba861c8d9e2d

Observation 59cedae0-882b-4c12-8981-88ae6f07fd93 · outbound

This paper cites Measur- ing Massive Multitask Language Understanding.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Measur- ing Massive Multitask Language Understanding

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:21.475285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:10:05.648401Z digest=sha256:d1b31f288fb5d9762d48bc2d018bd8bb668b47a5f29986ad9db10e7f093218ab

Observation 23bde0c7-c291-4500-a284-599a88ae5017 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Zero-Shot Vision Encoder Grafting via LLM Surrogates LoRA: Low-Rank Adaptation of Large Language Models

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:21.140630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:10:05.846778Z digest=sha256:c8fdb09ff555ecb656045f9bc0324b9d6b179a14c423cc88cf41a049939546c7

Observation 649794ba-112a-4f81-bc14-897e75708ce8 · outbound

This paper cites GQA: A New Dataset for Real-World Visual Reasoning and Compositional Question Answering.

Zero-Shot Vision Encoder Grafting via LLM Surrogates GQA: A New Dataset for Real-World Visual Reasoning and Compositional Question Answering

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:20.845832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:10:06.057813Z digest=sha256:e13c2c1445c972ced1d52c157cef0eef2f1f0276828810f0cae064f904e06e16

Observation eb6c8812-05c5-41e1-95f8-1d5eee1132ef · outbound

This paper cites InternVL2: Better than the Best—Expanding Performance Boundaries of Open-Source Multimodal Mod- els with the Progressive Scaling Strategy.

Zero-Shot Vision Encoder Grafting via LLM Surrogates InternVL2: Better than the Best—Expanding Performance Boundaries of Open-Source Multimodal Mod- els with the Progressive Scaling Strategy

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:20.536709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:10:06.173724Z digest=sha256:40d7a8a336cf5a5955ae9803eb7a1404351c99f440c32688261bba31f425ada7

Observation c7be7e24-b793-42cf-a693-fd05281938c6 · outbound

This paper cites A Diagram Is Worth a Dozen Images.

Zero-Shot Vision Encoder Grafting via LLM Surrogates A Diagram Is Worth a Dozen Images

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:20.238302Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:10:06.293534Z digest=sha256:3aef6681b371587bef96c259326c83969d202e5f267408c49f9e93537d4cc9a5

Observation 51e4d9b8-f8e7-47e3-85ef-fb41ceeaaa29 · outbound

This paper cites Propulsion: Steering LLM with Tiny Fine-Tuning.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Propulsion: Steering LLM with Tiny Fine-Tuning

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:19.880653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:10:06.462436Z digest=sha256:c178d16cb427153640f2bf16f09eadcb06aff09cf413b06b18b20bd2ac8effe3

Observation dc51ade7-865f-494a-813a-e740ae092e60 · outbound

This paper cites SEED-Bench: Benchmarking Multi- modal LLMs with Generative Comprehension.

Zero-Shot Vision Encoder Grafting via LLM Surrogates SEED-Bench: Benchmarking Multi- modal LLMs with Generative Comprehension

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:19.640428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:10:06.605423Z digest=sha256:1bd4e88763b36e09d644372e87245512cd548a5e122a4a461838a3c35ee82c0d

Observation ef148f84-4a47-4d82-8ed6-8b250a7c0ea5 · outbound

This paper cites LLaV A-Next: Stronger Llms Supercharge Multimodal Capabilities in the Wild.

Zero-Shot Vision Encoder Grafting via LLM Surrogates LLaV A-Next: Stronger Llms Supercharge Multimodal Capabilities in the Wild

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:19.339366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:10:06.744498Z digest=sha256:e45f8f1d973e33d6f5c495456f744503a4840def1fce4156b3c648920a4a8f0b

Observation 7832c4fa-c5e8-4888-b070-8974240f4229 · outbound

This paper cites LMMs-Eval: Accelerating the Development of Large Multimodal Models.

Zero-Shot Vision Encoder Grafting via LLM Surrogates LMMs-Eval: Accelerating the Development of Large Multimodal Models

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:19.028093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:10:06.928443Z digest=sha256:57913e9ee69be6dbd02a76a0d09514cb4fb35b8ff5c7286a1433b7cee259b7f0

Observation c0410ddb-23dc-45dd-848c-d1419c58cf19 · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

Zero-Shot Vision Encoder Grafting via LLM Surrogates LLaVA-OneVision: Easy Visual Task Transfer

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:07.072452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:07.072452Z digest=sha256:1c8b0b1290361298303ed46cb1cb74c9ea628c2bcf93394acc1808bdd6f4513f

Observation 0755a51f-eaac-4b00-9b4c-3239e119c8f5 · outbound

This paper cites BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models.

Zero-Shot Vision Encoder Grafting via LLM Surrogates BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:18.689949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:10:07.202405Z digest=sha256:d5bd988e29af9a51307185e17d4f1ad0b5496e323e948e0f69ab6f5e40c0b0ed

Observation 3e3532c1-4f67-4bb3-a78f-3d463911b73a · outbound

This paper cites Evaluating Object Hallucination in Large Vision-Language Models.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Evaluating Object Hallucination in Large Vision-Language Models

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:18.346769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:10:07.335320Z digest=sha256:8bfb053af7f704cc053e94170e896d2989b0cfb57756f98237c28fdd1e7ed75b

Observation c9b6ce0f-78ed-4883-bbb3-dec8e561454a · outbound

This paper cites Improved Baselines with Visual Instruction Tuning.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Improved Baselines with Visual Instruction Tuning

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:18.009253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:10:07.486743Z digest=sha256:e9537f1804701ad7c0871083b7600c9223a7ea7597475886beedcae6155a8527

Observation 9dd22051-41b2-4fad-9ab1-324a36a2e708 · outbound

This paper cites LLaV A-Next: Im- proved Reasoning, Ocr, and World Knowledge.

Zero-Shot Vision Encoder Grafting via LLM Surrogates LLaV A-Next: Im- proved Reasoning, Ocr, and World Knowledge

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:17.719998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:10:07.663602Z digest=sha256:fb637ad121ef79bbe5e780d2b0443441abc457fdd5f968c3cee2de311f691e26

Observation b193148e-e3df-4f21-a0fd-15b2cdf9772e · outbound

This paper cites Visual Instruction Tuning.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Visual Instruction Tuning

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:17.431218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:10:07.830359Z digest=sha256:2abf37c961a0ceacbf5ff5117114bdbf30b7eb576424c722993cb85cc9c7ab5b

Observation c245efa1-3580-47e7-ba77-2bcc3340a6bd · outbound

This paper cites MMBench: Is Your Multi-modal Model an All-around Player? In ECCV, 2025.

Zero-Shot Vision Encoder Grafting via LLM Surrogates MMBench: Is Your Multi-modal Model an All-around Player? In ECCV, 2025

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:17.038114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:10:07.935523Z digest=sha256:c647824dab423ff206a5328809d555e35e9421260c0eb954520e7a735bdb2872

Observation 54401c49-46c3-4af6-b992-1fdc4d582771 · outbound

This paper cites Chartqa: A Benchmark for Question Answering About Charts With Visual and Logical Reason- ing.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Chartqa: A Benchmark for Question Answering About Charts With Visual and Logical Reason- ing

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:16.743903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:10:08.060825Z digest=sha256:d0710afdda5c229b505950fc678d5e21ac2cc4382e935337a47ca686d35982d8

Observation 4c64e431-059d-4041-88f4-0bf1cb177edd · outbound

This paper cites Docvqa: A Dataset for VQA on Document Images.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Docvqa: A Dataset for VQA on Document Images

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:16.373407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:10:08.282593Z digest=sha256:e0e139458e841887028cc988cf0e36eccbef8802195b40b0ac8c75267490152b

Observation a9d22fad-abc7-49d9-810f-a284bbfe4824 · outbound

This paper cites Infograph- icVQA.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Infograph- icVQA

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:16.063658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:10:08.400143Z digest=sha256:05fade42fb58fee96e996d8bd12e9d506b627dfac2382061f12b6e13a7912e7a

Observation b2bdd531-cbc6-4173-bb5f-290a138497d0 · outbound

This paper cites ShortGPT: Layers in Large Language Models are More Redundant Than You Expect.

Zero-Shot Vision Encoder Grafting via LLM Surrogates ShortGPT: Layers in Large Language Models are More Redundant Than You Expect

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:08.563542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:08.563542Z digest=sha256:3fc9c37724825058c4848cc2d0c6ccd52645cbca0f7a45c749e14b0b520e9406

Observation 6cdc855b-7bfd-4691-801c-425a5f9c46f5 · outbound

This paper cites Can a Suit of Armor Conduct Electricity? A New Dataset for Open Book Question Answering.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Can a Suit of Armor Conduct Electricity? A New Dataset for Open Book Question Answering

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:15.729807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:10:08.737225Z digest=sha256:60b208db62773c93f859be7dd5c551dabf9af7807dcee1cc5139c94d66196bf1

Observation 232146ff-0a88-4f50-855c-fb474f1916d3 · outbound

This paper cites interpreting GPT: the logit lens.

Zero-Shot Vision Encoder Grafting via LLM Surrogates interpreting GPT: the logit lens

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:15.404758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:10:08.896020Z digest=sha256:1399fd09b2852b33d50268a49d8da41b3dea0954635c7cd96f4522832998b0f0

Observation 70fe4b8e-5a8c-4c2e-ac06-9159051c07fa · outbound

This paper cites Learning Transferable Visual Models From Natural Language Super- vision.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Learning Transferable Visual Models From Natural Language Super- vision

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:15.037846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:10:09.084874Z digest=sha256:b6c77cfe1e2f525909aa4f834a0ad20649c1ce2b1e20cbd097191f47fa64ebf4

Observation a54574e4-0195-4a87-b106-2c4102115b3f · outbound

This paper cites WinoGrande: An Adversarial Winograd Schema Challenge at Scale.

Zero-Shot Vision Encoder Grafting via LLM Surrogates WinoGrande: An Adversarial Winograd Schema Challenge at Scale

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:14.688023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:10:09.261308Z digest=sha256:3324f62ddf1d2c8516dc52180d46d5eb6ffa5f2027f3e593ae3a459ff0fc0a26

Observation 919f7776-77bc-4d98-9e36-159147c403cf · outbound

This paper cites Open Problems in Mechanistic Interpretability.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Open Problems in Mechanistic Interpretability

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:09.455799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:09.455799Z digest=sha256:ee77f24348e726711bb5d75a24f5c9eb2aa6fad52bb6bb15a090460a0bcd5e81

Observation 78f61f07-327e-420c-a7aa-9e2c43699f45 · outbound

This paper cites Towards VQA Models That Can Read.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Towards VQA Models That Can Read

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:14.377820Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:10:09.639043Z digest=sha256:68125ee033112388feb714114ab871028400873a041e30aa24462efb00e55c85

Observation ae21e43a-ad87-459a-96a4-d61fe2ab610b · outbound

This paper cites Transformer Layers as Painters.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Transformer Layers as Painters

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:09.868806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:09.868806Z digest=sha256:1862c3bcb43b4450983829f513ab45da6f47361f28ee5dca79375d1ba9bb5f6b

Observation c2eeae5e-c64c-4045-9a0d-fe116ec56739 · outbound

This paper cites an unresolved cited work.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:10:14.074467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:10:10.023098Z digest=sha256:be052fc22d541cf2833ff3214e74c23c8842016d24c69cef30a5887533b51a62

Observation 8b80db9e-b15a-49f8-9293-6e6b9916aaad · outbound

This paper cites an unresolved cited work.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:10:13.774364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:10:10.154755Z digest=sha256:0ea661155a6c3c42ef0550d2057e4949cf0f8962b1db150022e47ba61e866050

Observation a5233e1d-bf72-455f-9bde-cbcffb59e415 · outbound

This paper cites Qwen2.5 Technical Report.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Qwen2.5 Technical Report

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:10.316669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:10.316669Z digest=sha256:2d460b2e7f80da8fa42a438f9c748d4fc5ea9244104856bbf60a5e178fdc35be

Observation 8cac2992-8ccd-4fe4-b59d-3433f20693fb · outbound

This paper cites Qwen3 Technical Report.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Qwen3 Technical Report

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:10.411364Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:10.411364Z digest=sha256:08173e71e416a9d933ce60034fa0c45e3129424888f5df39d6fde9a8399b4e9a

Observation 9a9ab4b5-bf59-4450-99bf-fbc4ddfbb156 · outbound

This paper cites Cambrian- 1: A Fully Open, Vision-Centric Exploration of Multimodal LLMs.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Cambrian- 1: A Fully Open, Vision-Centric Exploration of Multimodal LLMs

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:13.491774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:10:10.548879Z digest=sha256:78517554d086f2780031cd384b6d98844df896a404019e58b7d7510d1c4f5187

Observation bbd238ab-8814-446f-beb0-b6b14efeb168 · outbound

This paper cites SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features.

Zero-Shot Vision Encoder Grafting via LLM Surrogates SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:10.711035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:10.711035Z digest=sha256:31f1a5d9a217ef51c5aa4c3b372160a26e559b20076588ff728ca02bb7836c2d

Observation 02bcb860-639e-41f6-b365-d8f131348935 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:10.917149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:10.917149Z digest=sha256:6c99d89d0bf887fb1ecce5007aa5e6ae9435ba72754296e1fe39cffef2c92ed5

Observation 90da24bf-cbc8-4556-ad96-05125bcc443d · outbound

This paper cites CogVLM: Visual Expert for Pretrained Lan- guage Models.

Zero-Shot Vision Encoder Grafting via LLM Surrogates CogVLM: Visual Expert for Pretrained Lan- guage Models

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:13.221500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:10:11.082351Z digest=sha256:1fe7f379ced102598446dca35f2c7330dacc6c533c2838ee807babb6cade7f34

Observation 5a1b519c-ef56-4100-9df7-56f6a70460de · outbound

This paper cites MM-Vet: Evaluating Large Multimodal Models for Inte- grated Capabilities.

Zero-Shot Vision Encoder Grafting via LLM Surrogates MM-Vet: Evaluating Large Multimodal Models for Inte- grated Capabilities

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:12.806130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:10:11.239273Z digest=sha256:808f085f890e3adceba952abc3e4df324e44419258cccfa0d45b05d93bb45373

Observation 9f4c83da-0827-4fb2-b03c-2602aeeaa9b4 · outbound

This paper cites HellaSwag: Can a Machine Really Finish Your Sentence? In ACL Anthology, 2019.

Zero-Shot Vision Encoder Grafting via LLM Surrogates HellaSwag: Can a Machine Really Finish Your Sentence? In ACL Anthology, 2019

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:12.491448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:10:11.437478Z digest=sha256:8b30de692a5642e250409d206c280ae346a01e90d8596de2d92f213606e0fd2b

Observation 652f85e5-fdfb-400c-be5a-0ee75c053fec · outbound

This paper cites Sigmoid Loss for Language Image Pre- Training.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Sigmoid Loss for Language Image Pre- Training

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:12.138345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:10:11.570195Z digest=sha256:97baf9926715761cd6f81cbcd3857441bd594d75d2d93bd12932aac982900fb7

Observation f1dca64a-039e-41c8-b879-848cd520fce3 · outbound

This paper cites PyTorch FSDP: Experiences on Scaling Fully Sharded Data Parallel.

Zero-Shot Vision Encoder Grafting via LLM Surrogates PyTorch FSDP: Experiences on Scaling Fully Sharded Data Parallel

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:11.740399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:11.740399Z digest=sha256:ab2e6a06293ca50487c913a816979c57749654f018f8f2b761119ff09396b444

Pith citing papers

No inbound Pith citation observations are available.