Pith. sign in

Paper Citation Record · LEDGER

Treat Visual Tokens as Text? But Your MLLM Only Needs Fewer Efforts to See

As of 4 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 4 inbound Pith citation observations for arXiv:2410.06169.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.06169 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 4 of 4 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-22T23:58:57.819555Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-23T00:02:17.737999Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 15d8b9b0-f2f0-40a0-a0ee-27ca5b52ac8a · inbound

Growing a Multi-head Twig via Distillation and Reinforcement Learning to Accelerate Large Vision-Language Models cites this paper.

Growing a Multi-head Twig via Distillation and Reinforcement Learning to Accelerate Large Vision-Language Models Treat Visual Tokens as Text? But Your MLLM Only Needs Fewer Efforts to See

Reference 67

Resolution
verified exact
arxiv_id, observed 2026-05-23T00:02:17.741875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T23:58:57.819555Z digest=sha256:d677478eb06d9bdeff4c4fdc588d68c63a9b5e2bbf8f0a1c67d1affbe597fb96

Observation c88f045b-b0bf-4adc-8807-714f44e7571d · inbound

Can VLMs Truly Forget? Benchmarking Training-Free Visual Concept Unlearning cites this paper.

Can VLMs Truly Forget? Benchmarking Training-Free Visual Concept Unlearning Treat Visual Tokens as Text? But Your MLLM Only Needs Fewer Efforts to See

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T19:48:11.181219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T19:48:03.278822Z digest=sha256:c5cd042732e14cc84fac9f90fa0ddbd14b42b471051367f7b51ffa464ddc1687

Observation 0455d4c3-b8ee-4ffa-aa19-99096315cec7 · inbound

Counting to Four is still a Chore for VLMs cites this paper.

Counting to Four is still a Chore for VLMs Treat Visual Tokens as Text? But Your MLLM Only Needs Fewer Efforts to See

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:06:03.307146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:39:06.327888Z digest=sha256:68dd6166ffbbc94b349f8c666003ae8b2af1b0f25d254cc453584c2b485d6e12

Observation d331525c-62db-428e-af46-17f1a622b9c0 · inbound

EvoComp: Learning Visual Token Compression for Multimodal Large Language Models via Semantic-Guided Evolutionary Labeling cites this paper.

EvoComp: Learning Visual Token Compression for Multimodal Large Language Models via Semantic-Guided Evolutionary Labeling Treat Visual Tokens as Text? But Your MLLM Only Needs Fewer Efforts to See

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-10T07:01:49.126917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T07:00:36.870817Z digest=sha256:78f42513d500598d3b9000fa5ea8b85b57a05eed35201a9ce4cef734b6326511