Pith. sign in

InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

citation-role summary

background 1

citation-polarity summary

fields

cs.CV 1

years

2025 1

verdicts

CONDITIONAL 1

roles

background 1

polarities

unclear 1

representative citing papers

Zero-Shot Vision Encoder Grafting via LLM Surrogates

cs.CV · 2025-05-28 · conditional · novelty 7.0

Training a vision encoder against a small surrogate made from a target LLM's early layers lets the encoder be grafted into the full LLM with no fine-tuning, matching some full-training results at roughly half the cost.

citing papers explorer

Showing 1 of 1 citing paper.

  • Zero-Shot Vision Encoder Grafting via LLM Surrogates cs.CV · 2025-05-28 · conditional · none · ref 6

    Training a vision encoder against a small surrogate made from a target LLM's early layers lets the encoder be grafted into the full LLM with no fine-tuning, matching some full-training results at roughly half the cost.