Pith. sign in

Paper Citation Record · LEDGER

DyFo: A Training-Free Dynamic Focus Visual Search for Enhancing LMMs in Fine-Grained Visual Understanding

As of 12 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2504.14920.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.14920 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:40:11.502823Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T19:20:07.242082Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation ca30a7ad-6ad2-4fd2-92f0-9978ef0cbcd1 · inbound

Zoom-Refine: Boosting High-Resolution Multimodal Understanding via Localized Zoom and Self-Refinement cites this paper.

Zoom-Refine: Boosting High-Resolution Multimodal Understanding via Localized Zoom and Self-Refinement DyFo: A Training-Free Dynamic Focus Visual Search for Enhancing LMMs in Fine-Grained Visual Understanding

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T11:40:11.502823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:40:11.502823Z digest=sha256:6a96d4bed31059d2f9acf5e04e5383c36ff7d4fdd3286d58b1a0f06d7f2c4aaf

Observation fbc78e49-d107-4875-9227-0be3f22a816d · inbound

Reinforcing VLMs to Use Tools for Detailed Visual Reasoning Under Resource Constraints cites this paper.

Reinforcing VLMs to Use Tools for Detailed Visual Reasoning Under Resource Constraints DyFo: A Training-Free Dynamic Focus Visual Search for Enhancing LMMs in Fine-Grained Visual Understanding

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T04:57:50.419860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:57:50.419860Z digest=sha256:f8291ed38717520b5a640354bdd190afa0835e356d1c67646c02c80fb52b18e7

Observation ce7eca06-4527-4969-bc2d-cd8ca7de113a · inbound

Mini-o3: Scaling Up Reasoning Patterns and Interaction Turns for Visual Search cites this paper.

Mini-o3: Scaling Up Reasoning Patterns and Interaction Turns for Visual Search DyFo: A Training-Free Dynamic Focus Visual Search for Enhancing LMMs in Fine-Grained Visual Understanding

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-18T01:17:55.682640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-18T01:17:55.500268Z digest=sha256:d777e3cb9031eb39a981593f7f6011f709db443f909820c5392f0f64a4b7c405

Observation 838ca301-c881-476a-a8e6-f781ed760830 · inbound

MAG-3D: Multi-Agent Grounded Reasoning for 3D Understanding cites this paper.

MAG-3D: Multi-Agent Grounded Reasoning for 3D Understanding DyFo: A Training-Free Dynamic Focus Visual Search for Enhancing LMMs in Fine-Grained Visual Understanding

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:51:23.766620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-10T17:25:31.097385Z digest=sha256:4d58126a45727a4604ea697edd5f8640cd632f2ad096f8d21f2e7024d0ca7c74

Observation 84ff8670-a189-4058-b858-55be341df460 · inbound

Self-Prophetic Decoding to Unlock Visual Search in LVLMs cites this paper.

Self-Prophetic Decoding to Unlock Visual Search in LVLMs DyFo: A Training-Free Dynamic Focus Visual Search for Enhancing LMMs in Fine-Grained Visual Understanding

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T13:03:26.911173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-29T12:53:40.783281Z digest=sha256:8eeb9b87b15fa21063513139f9d2f23e5bc60e37f8a0edd5f6fbbbdf066ed5ae

Observation 06625fc2-de45-4963-9f90-d050c97a79f7 · inbound

V-Zero: Answer-Label-Free On-Policy Distillation with Contrastive Evidence Gating for Fine-Grained Visual Reasoning cites this paper.

V-Zero: Answer-Label-Free On-Policy Distillation with Contrastive Evidence Gating for Fine-Grained Visual Reasoning DyFo: A Training-Free Dynamic Focus Visual Search for Enhancing LMMs in Fine-Grained Visual Understanding

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-07-04T19:20:07.243798Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-06-25T21:23:10.051805Z digest=sha256:b8879cb3e1b2d7555caa159b7609e5603c9c5708bc355eb874990fa226c8fe87