Pith. sign in

Paper Citation Record · LEDGER

WildVision: Evaluating Vision-Language Models in the Wild with Human Preferences

As of 5 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2406.11069.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.11069 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T12:27:27.424386Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T22:36:16.753738Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation f535af8a-a6cb-43cc-927f-a89c9d3e7798 · inbound

MME-RealWorld: Could Your Multimodal LLM Challenge High-Resolution Real-World Scenarios that are Difficult for Humans? cites this paper.

MME-RealWorld: Could Your Multimodal LLM Challenge High-Resolution Real-World Scenarios that are Difficult for Humans? WildVision: Evaluating Vision-Language Models in the Wild with Human Preferences

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-16T07:59:32.834674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T07:59:32.638758Z digest=sha256:affa91811992b81cc945f3ec24517d30aa3f8d630c6798ef192b32d53b492ea7

Observation 7624be59-6f43-4a48-b087-3d8d57ad9fc4 · inbound

Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling cites this paper.

Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling WildVision: Evaluating Vision-Language Models in the Wild with Human Preferences

Reference 171

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:23:57.801367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:23:57.588851Z digest=sha256:213ba11bc33a848b5bd6aa866cb6adcbbc4cbdbe20c1a35642308cf44ba2e460

Observation d17ec5f6-459a-455b-b4b0-3829971dd4ef · inbound

Unified Reward Model for Multimodal Understanding and Generation cites this paper.

Unified Reward Model for Multimodal Understanding and Generation WildVision: Evaluating Vision-Language Models in the Wild with Human Preferences

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-14T00:44:30.617150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-14T00:44:30.558048Z digest=sha256:e444da8657078193aa203b529b42ffb36a1be788220b447e12b9d18acbb05cde

Observation 2b254b4a-fbad-4a33-8c1f-dc89cd8c24f3 · inbound

InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models cites this paper.

InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models WildVision: Evaluating Vision-Language Models in the Wild with Human Preferences

Reference 86

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:41:08.056502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T13:41:07.991012Z digest=sha256:3782049458428d860ab82a804bc038dc1f5f989d103a12b76037e155ef8cf76a

Observation 81477c90-a19e-45d7-9e93-13cbe146a047 · inbound

InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency cites this paper.

InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency WildVision: Evaluating Vision-Language Models in the Wild with Human Preferences

Reference 78

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:58:59.198792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:58:58.660564Z digest=sha256:8be59bac035009af15d9007277a5a466ca45013366e456b474a5ae696bd21151

Observation 13178546-7ca2-40d5-bc30-fc30b9c5f40e · inbound

Improving Large Vision and Language Models by Learning from a Panel of Peers cites this paper.

Improving Large Vision and Language Models by Learning from a Panel of Peers WildVision: Evaluating Vision-Language Models in the Wild with Human Preferences

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-05T12:27:27.424386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:27:27.424386Z digest=sha256:b683d76444af7f6b37382116c59866deb80dda7b8701224b354188b34b693025

Observation 22678309-66f4-49ea-b1f6-3cb9ed156487 · inbound

Understanding Space Is Rocket Science -- Only Top Reasoning Models Can Solve Spatial Understanding Tasks cites this paper.

Understanding Space Is Rocket Science -- Only Top Reasoning Models Can Solve Spatial Understanding Tasks WildVision: Evaluating Vision-Language Models in the Wild with Human Preferences

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-05T11:56:10.304534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:56:10.304534Z digest=sha256:ddc13a2d17abb499dd32e95bf7483c453726c919ff05e99c5a5cd02f305ad48e

Observation f0a16358-ba02-47c4-8dd6-b1a4b9d12506 · inbound

Assessing Privacy Preservation and Utility in Online Vision-Language Models cites this paper.

Assessing Privacy Preservation and Utility in Online Vision-Language Models WildVision: Evaluating Vision-Language Models in the Wild with Human Preferences

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:30:50.217647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T19:49:04.957341Z digest=sha256:debe2090ec709e46b96ea024a7ff59d73f83c4b1b541646fb0e51ac34713e56a

Observation f621129b-a43f-41fe-8df7-91606e63403c · inbound

When Meaning Travels: A Granular Lens on Hybrid-MoE's Role in Idiomatic Understanding for Language Models cites this paper.

When Meaning Travels: A Granular Lens on Hybrid-MoE's Role in Idiomatic Understanding for Language Models WildVision: Evaluating Vision-Language Models in the Wild with Human Preferences

Reference 71

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T22:36:16.756023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-28T15:19:26.983760Z digest=sha256:2811f5ec6361e05764c7ea2d01a0619d685c0ae284f0b82be2d8e5edd675a0d8