Pith. sign in

Paper Citation Record · LEDGER

SLIP: Self-supervision meets Language-Image Pre-training

As of 17 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2112.12750.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2112.12750 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T01:06:20.852001Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-24T11:09:22.375297Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a08922a2-2835-489c-83c5-e3a7ccc6d006 · inbound

Hierarchical Text-Conditional Image Generation with CLIP Latents cites this paper.

Hierarchical Text-Conditional Image Generation with CLIP Latents SLIP: Self-supervision meets Language-Image Pre-training

Reference 31

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T16:55:57.837312Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-10T16:55:57.612364Z digest=sha256:b798b6a02f8202e93014d9ddde391c888223a7c432ae72e0a7e0799ddad77374

Observation 2af4e0bb-9e70-4c8c-ba09-1940b83fb103 · inbound

DetailCLIP: Injecting Image Details into CLIP's Feature Space cites this paper.

DetailCLIP: Injecting Image Details into CLIP's Feature Space SLIP: Self-supervision meets Language-Image Pre-training

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-24T11:09:22.379272Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-24T11:08:20.298043Z digest=sha256:d25e37d5e36db22d560f3b6a53f04af593064570833b25594db00355775a7b82

Observation aea50e06-111b-4c59-92a7-a0ed59beed2e · inbound

Demystifying CLIP Data cites this paper.

Demystifying CLIP Data SLIP: Self-supervision meets Language-Image Pre-training

Reference 79

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T09:20:20.320666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-05-16T09:20:20.143143Z digest=sha256:73a2e02608af9a087179d665d3b172c972d17168e3ce838e491d05bba0e1c24d

Observation c9fd9868-860c-4876-ba31-ae48f8af8efd · inbound

Visual Pre-Training on Unlabeled Images using Reinforcement Learning cites this paper.

Visual Pre-Training on Unlabeled Images using Reinforcement Learning SLIP: Self-supervision meets Language-Image Pre-training

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T01:06:20.852001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T01:06:20.852001Z digest=sha256:034e5fe628443b462611246ef1c39a12c6d31775d91de4cbcde3e32959413e6e

Observation 3d51cdbc-cacf-482e-9d25-a439e1f36727 · inbound

CXR-CML: Improved zero-shot classification of long-tailed multi-label diseases in Chest X-Rays cites this paper.

CXR-CML: Improved zero-shot classification of long-tailed multi-label diseases in Chest X-Rays SLIP: Self-supervision meets Language-Image Pre-training

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T14:24:51.259176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:24:51.259176Z digest=sha256:62ad5476f6f02ecbe0cc6618c4acbd119894468f457f15c5255aa791a02b525c

Observation 026b57e2-414c-42b3-ac8e-6a91ed6ce1f7 · inbound

Meta CLIP 2: A Worldwide Scaling Recipe cites this paper.

Meta CLIP 2: A Worldwide Scaling Recipe SLIP: Self-supervision meets Language-Image Pre-training

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T12:08:23.020444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:08:23.020444Z digest=sha256:6c4097262cf8f6727bf9e4089fb9fb61143ff644091b4c21a218e485891cc5b3

Observation 7b19e53f-3bc4-47e8-b13b-5492d1d5e755 · inbound

Bottleneck Tokens for Unified Multimodal Retrieval cites this paper.

Bottleneck Tokens for Unified Multimodal Retrieval SLIP: Self-supervision meets Language-Image Pre-training

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T11:01:03.336386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-10T15:14:00.615638Z digest=sha256:defbbd395b67c67e1f25eb67e8e578dfaeb6e73f0ac936746bd9df81b6bc36f2