Pith. sign in

Paper Citation Record · LEDGER

DualFocus: Integrating Macro and Micro Perspectives in Multi-modal Large Language Models

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2402.14767.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.14767 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:39:47.677291Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T15:09:55.154167Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 325da1dc-64dd-4d87-8500-d5990eb198e4 · inbound

InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output cites this paper.

InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output DualFocus: Integrating Macro and Micro Perspectives in Multi-modal Large Language Models

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-17T10:46:28.820987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T10:46:28.447347Z digest=sha256:7c0007f3117602ff875790113da3e25476e1d964352a38d84d3f74763e0b65f2

Observation 3692d56d-5932-4216-8d21-bc4b56c42684 · inbound

Zoom-Refine: Boosting High-Resolution Multimodal Understanding via Localized Zoom and Self-Refinement cites this paper.

Zoom-Refine: Boosting High-Resolution Multimodal Understanding via Localized Zoom and Self-Refinement DualFocus: Integrating Macro and Micro Perspectives in Multi-modal Large Language Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T11:39:47.677291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:39:47.677291Z digest=sha256:7d176ee92e7dfbb7705f0598c29c1c78ef6c9899e39002206e2916bb75c1c169

Observation 491dde10-fbc3-42fd-8000-9f0e7860b452 · inbound

VisCodex: Unified Multimodal Code Generation via Merging Vision and Coding Models cites this paper.

VisCodex: Unified Multimodal Code Generation via Merging Vision and Coding Models DualFocus: Integrating Macro and Micro Perspectives in Multi-modal Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T20:44:35.572631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:44:35.572631Z digest=sha256:af31166719487db938d4e4ab5800ca71bfa8ddde12458c973110073baee8b7bf

Observation 3faa70b6-057a-4ec8-9d96-6bfb364208ec · inbound

VisReason: A Large-Scale Dataset for Visual Chain-of-Thought Reasoning cites this paper.

VisReason: A Large-Scale Dataset for Visual Chain-of-Thought Reasoning DualFocus: Integrating Macro and Micro Perspectives in Multi-modal Large Language Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-03T20:57:28.598985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T20:57:28.598985Z digest=sha256:31355b4f076ad17c56d0e5d47af0301edf0462f64ba640e94169c83179fca250

Observation d2d131a7-b406-46aa-a39b-8d4e4564250f · inbound

From Structure to Synergy: A Survey of Vision-Language Perception Paradigm Evolution in Multimodal Large Language Models cites this paper.

From Structure to Synergy: A Survey of Vision-Language Perception Paradigm Evolution in Multimodal Large Language Models DualFocus: Integrating Macro and Micro Perspectives in Multi-modal Large Language Models

Reference 129

Resolution
verified exact
arxiv_id, observed 2026-07-04T15:09:55.156272Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T01:50:54.242508Z digest=sha256:e2237eb038e3a1c8e838420f6ee383bde3e66ee34d48df20fc857e441fc54ec4

Observation 29b544ec-42d6-43bb-9f82-75ec45e7f9b8 · inbound

Event-Aware Instructed Assistant for Referring Video Segmentation cites this paper.

Event-Aware Instructed Assistant for Referring Video Segmentation DualFocus: Integrating Macro and Micro Perspectives in Multi-modal Large Language Models

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-07-04T13:29:51.270731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T05:13:26.233621Z digest=sha256:4bc8795399a88213b6ed50aa59cacb78d46a776f5f0d961223c28b9f0be6548f

Observation f0b83103-a2b6-42af-b36e-ab776fcb22f6 · inbound

BVS: Bayesian Visual Search with Multimodal Large Language Model for Fine-grained Perception cites this paper.

BVS: Bayesian Visual Search with Multimodal Large Language Model for Fine-grained Perception DualFocus: Integrating Macro and Micro Perspectives in Multi-modal Large Language Models

Reference 213

Resolution
unresolved
no resolver link, observed 2026-07-12T04:17:40.198357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-12T04:17:40.198357Z digest=sha256:17904cb698b081611c03f07e56dec5015e576b92ca0b642b692d85718b69172c

Observation 13c7c5eb-e327-4dd2-b04d-41223cc03bbc · inbound

MentalThink: Shaping Thoughts in Mental SVG World cites this paper.

MentalThink: Shaping Thoughts in Mental SVG World DualFocus: Integrating Macro and Micro Perspectives in Multi-modal Large Language Models

Reference 224

Resolution
unresolved
no resolver link, observed 2026-07-12T01:50:59.184754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-12T01:50:59.184754Z digest=sha256:4986533f6c047ffb5d398875e1bb6d592ce430f42fb4a7054f65e13b83bebf37