Pith. sign in

Paper Citation Record · LEDGER

T-MARS: Improving Visual Representations by Circumventing Text Feature Learning

As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2307.03132.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2307.03132 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T12:12:30.745548Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T09:49:44.890982Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 7ab61e17-d0ee-4519-8ef5-cf6420d81bda · inbound

Scaling Pre-training to One Hundred Billion Data for Vision Language Models cites this paper.

Scaling Pre-training to One Hundred Billion Data for Vision Language Models T-MARS: Improving Visual Representations by Circumventing Text Feature Learning

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-08T12:12:30.745548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T12:12:30.745548Z digest=sha256:0f7ebb90942fb8a68d5e6bd7633bd1654b4bbe99aa0b91d1929bb5c8fc1e25e2

Observation 77f46a19-d285-4d68-a19c-1866b54a68c0 · inbound

Quality over Quantity: Boosting Data Efficiency Through Ensembled Multimodal Data Curation cites this paper.

Quality over Quantity: Boosting Data Efficiency Through Ensembled Multimodal Data Curation T-MARS: Improving Visual Representations by Circumventing Text Feature Learning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-08T06:06:17.557918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T06:06:17.557918Z digest=sha256:64a4010ab2079e5271018c0976510b609e1316af69a88919a58e8a1440ea1d22

Observation d9da09a4-3fe4-4be0-b30a-ed005f0366b4 · inbound

What Does the Caption Really Say? Counterfactual Phrase Intervention for Compositional Data Selection in Vision-Language Pretraining cites this paper.

What Does the Caption Really Say? Counterfactual Phrase Intervention for Compositional Data Selection in Vision-Language Pretraining T-MARS: Improving Visual Representations by Circumventing Text Feature Learning

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-22T06:21:10.270477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T06:19:54.322211Z digest=sha256:5949ee9f50b5a35dc9f1a78e820eff034b001c628ba3ecf4d6c1d7c90d0dc317

Observation f069d1dc-4508-474e-b932-b754e49d125b · inbound

Data Selection Through Iterative Self-Filtering for Vision-Language Settings cites this paper.

Data Selection Through Iterative Self-Filtering for Vision-Language Settings T-MARS: Improving Visual Representations by Circumventing Text Feature Learning

Reference 77

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T09:49:44.892311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-06-26T09:22:47.537137Z digest=sha256:20a3fb28d1a75b723d2cf529b534c0fb22a23d73eab857994bab2fbc68e4e271

Observation 82be6d52-1da5-4d4c-ba4e-2ed518577dd6 · inbound

DataComp-VLM: Improved Open Datasets for Vision-Language Models cites this paper.

DataComp-VLM: Improved Open Datasets for Vision-Language Models T-MARS: Improving Visual Representations by Circumventing Text Feature Learning

Reference 200

Resolution
verified exact
arxiv_id, observed 2026-07-01T15:45:47.689741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-30T01:16:16.834861Z digest=sha256:797a9d675c05cae4e24a07b36662d5e1a0c3f6b45a631588a92a7d2a4de70ea0

Observation f75d8715-046b-4d2c-9685-fdcde13bd79b · inbound

DataComp-VLM: Improved Open Datasets for Vision-Language Models cites this paper.

DataComp-VLM: Improved Open Datasets for Vision-Language Models T-MARS: Improving Visual Representations by Circumventing Text Feature Learning

Reference 200

Resolution
verified exact
arxiv_id, observed 2026-07-02T21:17:23.898703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-07-02T21:10:10.548489Z digest=sha256:ed1cf0d0b5f5984863f39afabed3309628a62b68431b81d71b236b246099d5ec