Pith. sign in

Paper Citation Record · LEDGER

BRAVE: Broadening the visual encoding of vision-language models

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2404.07204.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2404.07204 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T13:35:05.087875Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T02:16:26.476658Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 793f627f-7938-45aa-9020-81c0af7cad9f · inbound

PaliGemma 2: A Family of Versatile VLMs for Transfer cites this paper.

PaliGemma 2: A Family of Versatile VLMs for Transfer BRAVE: Broadening the visual encoding of vision-language models

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-15T09:15:07.598735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T09:15:07.523565Z digest=sha256:5f6a95b0370fd9fb22e64dfac1e4b518ee2e0cf50b51851438edffabc8758f48

Observation 59305fb5-9d11-45e6-aa55-6f55bdb4bcfd · inbound

NanoVLMs: How small can we go and still make coherent Vision Language Models? cites this paper.

NanoVLMs: How small can we go and still make coherent Vision Language Models? BRAVE: Broadening the visual encoding of vision-language models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-08T13:35:05.087875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T13:35:05.087875Z digest=sha256:e34a2dd7c4fbeb4c16beaaa8b1d4994030736b4f7371e84ac67254f6f8e71312

Observation bf309773-992a-4ce3-a6c3-4332c53bc974 · inbound

Towards Multimodal Understanding via Stable Diffusion as a Task-Aware Feature Extractor cites this paper.

Towards Multimodal Understanding via Stable Diffusion as a Task-Aware Feature Extractor BRAVE: Broadening the visual encoding of vision-language models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T18:52:31.767820Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:52:31.767820Z digest=sha256:a77e9cd30a0c446827d4a25c524c4fe71541be8d738b99b8633bf77f54232e4c

Observation efa875fc-b398-4e67-abed-731a7b7b199a · inbound

Multi-Agent Interactive Question Generation Framework for Long Document Understanding cites this paper.

Multi-Agent Interactive Question Generation Framework for Long Document Understanding BRAVE: Broadening the visual encoding of vision-language models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T13:45:18.512376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:45:18.512376Z digest=sha256:801311bed757e4fd612844f766a4bd35425c5314412ee2d670028750fcf59fae

Observation bd6221a3-afeb-45c4-8f76-2b12957543ce · inbound

Beyond Encoder Accumulation: Measuring Encoder Roles in Multi-Encoder VLMs cites this paper.

Beyond Encoder Accumulation: Measuring Encoder Roles in Multi-Encoder VLMs BRAVE: Broadening the visual encoding of vision-language models

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-02T02:16:26.478330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T11:08:36.451147Z digest=sha256:270f92e57118e1d114ea6f1cd0cfe4075d489c48eb2cdd81ab786b30a7f94e03

Observation e4c19394-d85b-40ee-be29-e51c05c7d730 · inbound

An Exam for Active Observers cites this paper.

An Exam for Active Observers BRAVE: Broadening the visual encoding of vision-language models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-01T21:12:04.844747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T21:12:04.844747Z digest=sha256:34c92357c2390f1de65e7c247fe9c65ef1670675c93e08e96ae122f31feab5ae