Pith. sign in

Paper Citation Record · LEDGER

DyFo: A Training-Free Dynamic Focus Visual Search for Enhancing LMMs in Fine-Grained Visual Understanding

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2504.14920.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.14920 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:40:11.502823Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T19:20:07.242082Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation ca30a7ad-6ad2-4fd2-92f0-9978ef0cbcd1 · inbound

Zoom-Refine: Boosting High-Resolution Multimodal Understanding via Localized Zoom and Self-Refinement cites this paper.

Zoom-Refine: Boosting High-Resolution Multimodal Understanding via Localized Zoom and Self-Refinement DyFo: A Training-Free Dynamic Focus Visual Search for Enhancing LMMs in Fine-Grained Visual Understanding

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T11:40:11.502823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:40:11.502823Z digest=sha256:f2e3100e0e31adac329557936d0d7f11ff8c08d4dc52533c643ad24fedae55c1

Observation fbc78e49-d107-4875-9227-0be3f22a816d · inbound

Reinforcing VLMs to Use Tools for Detailed Visual Reasoning Under Resource Constraints cites this paper.

Reinforcing VLMs to Use Tools for Detailed Visual Reasoning Under Resource Constraints DyFo: A Training-Free Dynamic Focus Visual Search for Enhancing LMMs in Fine-Grained Visual Understanding

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T04:57:50.419860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:57:50.419860Z digest=sha256:e68e9f058d5733911a01b8ad017213cd351fecb1123861b4ecc648144dc7d58f

Observation ce7eca06-4527-4969-bc2d-cd8ca7de113a · inbound

Mini-o3: Scaling Up Reasoning Patterns and Interaction Turns for Visual Search cites this paper.

Mini-o3: Scaling Up Reasoning Patterns and Interaction Turns for Visual Search DyFo: A Training-Free Dynamic Focus Visual Search for Enhancing LMMs in Fine-Grained Visual Understanding

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-18T01:17:55.682640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T01:17:55.500268Z digest=sha256:ea5077ccc92101ea7cbc306145ffdf5cf2e3cf30e7146fd65d3e8d660d7b97da

Observation 838ca301-c881-476a-a8e6-f781ed760830 · inbound

MAG-3D: Multi-Agent Grounded Reasoning for 3D Understanding cites this paper.

MAG-3D: Multi-Agent Grounded Reasoning for 3D Understanding DyFo: A Training-Free Dynamic Focus Visual Search for Enhancing LMMs in Fine-Grained Visual Understanding

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:51:23.766620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T17:25:31.097385Z digest=sha256:fb17b830c6e32c8bd8247b031aa10e4e89064077ec49f71d88274a7964cda6b7

Observation 84ff8670-a189-4058-b858-55be341df460 · inbound

Self-Prophetic Decoding to Unlock Visual Search in LVLMs cites this paper.

Self-Prophetic Decoding to Unlock Visual Search in LVLMs DyFo: A Training-Free Dynamic Focus Visual Search for Enhancing LMMs in Fine-Grained Visual Understanding

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T13:03:26.911173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T12:53:40.783281Z digest=sha256:094fd34d3a4e35503c9393fcbdb046df720f7f93a1c31181bb8b07bc9125e40a

Observation 06625fc2-de45-4963-9f90-d050c97a79f7 · inbound

V-Zero: Answer-Label-Free On-Policy Distillation with Contrastive Evidence Gating for Fine-Grained Visual Reasoning cites this paper.

V-Zero: Answer-Label-Free On-Policy Distillation with Contrastive Evidence Gating for Fine-Grained Visual Reasoning DyFo: A Training-Free Dynamic Focus Visual Search for Enhancing LMMs in Fine-Grained Visual Understanding

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-07-04T19:20:07.243798Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-25T21:23:10.051805Z digest=sha256:a3f3e05f83f754a9eabe860973ca1c7996aa0b52ca71f460a292d14941fc9b49