Pith. sign in

Paper Citation Record · LEDGER

POINTS: Improving Your Vision-language Model with Affordable Strategies

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2409.04828.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2409.04828 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:40:22.349276Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T15:09:55.125524Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 4c5a3925-a823-4d99-babe-30ce143d52db · inbound

Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling cites this paper.

Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling POINTS: Improving Your Vision-language Model with Affordable Strategies

Reference 157

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:23:57.730768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T13:23:57.588851Z digest=sha256:49fa541a6b3b2bd75fcf377d7d97b4f78067e80e4e33275f0a9d7e6bef52a6ef

Observation 0c7b6a80-6563-476d-bf84-ec4f198ca983 · inbound

AutoJudger: An Agent-Driven Framework for Efficient Benchmarking of MLLMs cites this paper.

AutoJudger: An Agent-Driven Framework for Efficient Benchmarking of MLLMs POINTS: Improving Your Vision-language Model with Affordable Strategies

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T13:40:22.349276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:40:22.349276Z digest=sha256:4f3b7935348e6d2029c43315f4b26db6c03039bfbad65e59cbd7e04a12b72146

Observation 4b05dd3c-7398-4b9e-85fc-a52325fd1f34 · inbound

Affordance Benchmark for MLLMs cites this paper.

Affordance Benchmark for MLLMs POINTS: Improving Your Vision-language Model with Affordable Strategies

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T11:59:54.147203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:59:54.147203Z digest=sha256:ebb849cbc0dbf5fcf15a3b0562e15104cf55cd4ca968715b2f41aa5679e9c077

Observation e4f59be0-c7f5-4511-a5ac-9530c98c863c · inbound

POINTS-Long: Adaptive Dual-Mode Visual Reasoning in MLLMs cites this paper.

POINTS-Long: Adaptive Dual-Mode Visual Reasoning in MLLMs POINTS: Improving Your Vision-language Model with Affordable Strategies

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:41:03.983349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T15:23:08.671342Z digest=sha256:cf333e612ec4a6b96d08f8d2edfd7a38252990614eced9b93fa108720d6bb850

Observation 422cf898-3ab1-48a3-9d34-8a4b832dc3d6 · inbound

From Structure to Synergy: A Survey of Vision-Language Perception Paradigm Evolution in Multimodal Large Language Models cites this paper.

From Structure to Synergy: A Survey of Vision-Language Perception Paradigm Evolution in Multimodal Large Language Models POINTS: Improving Your Vision-language Model with Affordable Strategies

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-07-04T15:09:55.127564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T01:50:54.242508Z digest=sha256:fb6c5363b4bf5c8837edc316f995be20d2b777a68cd60a6c1511b0a65fa5fd25