Pith. sign in

Paper Citation Record · LEDGER

Fine-grained Image Captioning with CLIP Reward

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 4 inbound Pith citation observations for arXiv:2205.13115.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2205.13115 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 4 of 4 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:03:20.260110Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-18T14:41:30.171123Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 983bf124-daaa-4ea4-b362-a13ce4860e92 · inbound

Seeing Far and Clearly: Mitigating Hallucinations in MLLMs with Attention Causal Decoding cites this paper.

Seeing Far and Clearly: Mitigating Hallucinations in MLLMs with Attention Causal Decoding Fine-grained Image Captioning with CLIP Reward

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T15:03:20.260110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:03:20.260110Z digest=sha256:161ad73cc4feb736742c73aa563661a3f1a49f734ec9f4bbfa96ac891aa92e20

Observation 276f4202-75b9-43d5-b962-9ca78f2917a6 · inbound

InverTune: Removing Backdoors from Multimodal Contrastive Learning Models via Trigger Inversion and Activation Tuning cites this paper.

InverTune: Removing Backdoors from Multimodal Contrastive Learning Models via Trigger Inversion and Activation Tuning Fine-grained Image Captioning with CLIP Reward

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T00:58:06.678094Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:58:06.678094Z digest=sha256:2ddc1c007a2ad2b00149e9cdfbadb7a645fbb752fac09c66bfdf9b5e374cb542

Observation fae5e868-8fd5-43e6-92a9-ed7f29afb72c · inbound

Modality Bias in LVLMs: Analyzing and Mitigating Object Hallucination via Attention Lens cites this paper.

Modality Bias in LVLMs: Analyzing and Mitigating Object Hallucination via Attention Lens Fine-grained Image Captioning with CLIP Reward

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T05:04:22.275679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T05:04:22.275679Z digest=sha256:1ffb411549faeca14fa3c1d1b1e42aa23943382d13b568607d20977907f268d9

Observation 9fd07743-f4d3-4744-b35b-0ba5248616c4 · inbound

Long Story Short: Disentangling Compositionality and Long-Caption Understanding in Contrastive VLMs cites this paper.

Long Story Short: Disentangling Compositionality and Long-Caption Understanding in Contrastive VLMs Fine-grained Image Captioning with CLIP Reward

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-18T14:41:30.173527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T14:41:20.403259Z digest=sha256:2d76596bd9edd94a98f2e0768d8b988443a7648d88120841678f19c80dd32144