Pith. sign in

Paper Citation Record · LEDGER

DataComp: In search of the next generation of multimodal datasets

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 36 inbound Pith citation observations for arXiv:2304.14108.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2304.14108 v5

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 36 of 36 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:42:32.345850Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

74
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 53e4f70b-9858-4a6d-889c-4aa69b2da6f1 · inbound

Objaverse-XL: A Universe of 10M+ 3D Objects cites this paper.

Objaverse-XL: A Universe of 10M+ 3D Objects DataComp: In search of the next generation of multimodal datasets

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-17T13:02:11.604349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T13:02:11.512409Z digest=sha256:1fdeff3418e8a1ae893617b7d02dbb66c8443ff63a7a0026c7b0dd2d80f093f7

Observation 5b0763c5-fe42-473d-ae4b-031e297c053a · inbound

InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation cites this paper.

InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation DataComp: In search of the next generation of multimodal datasets

Reference 71

Resolution
verified exact
arxiv_id, observed 2026-05-15T06:30:22.712246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T06:30:22.431538Z digest=sha256:11e395e0343951d60d394514e6908191e498945df74c280f0db5c47dc8f08260

Observation b7853a04-19fe-49b5-bcb6-5fa3b35955f9 · inbound

OpenFlamingo: An Open-Source Framework for Training Large Autoregressive Vision-Language Models cites this paper.

OpenFlamingo: An Open-Source Framework for Training Large Autoregressive Vision-Language Models DataComp: In search of the next generation of multimodal datasets

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-14T01:52:01.415728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-14T01:52:01.163900Z digest=sha256:345528f1900db94cb8970f0e37145640dcdfe3936f8b04ec4ac0b2fc35503dd0

Observation 50537838-c88e-44dd-add6-32810a4947d5 · inbound

mPLUG-Owl2: Revolutionizing Multi-modal Large Language Model with Modality Collaboration cites this paper.

mPLUG-Owl2: Revolutionizing Multi-modal Large Language Model with Modality Collaboration DataComp: In search of the next generation of multimodal datasets

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-18T03:18:51.844203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T03:18:51.582340Z digest=sha256:7718817ca86beb1e893a5070425491c5a8cd299d2b3dd3beed0a6b68c52bffa5

Observation dbd967c6-4515-481d-8b9a-4f0905432019 · inbound

ShareGPT4V: Improving Large Multi-Modal Models with Better Captions cites this paper.

ShareGPT4V: Improving Large Multi-Modal Models with Better Captions DataComp: In search of the next generation of multimodal datasets

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-13T17:08:12.842273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T17:08:12.727773Z digest=sha256:387637a2c0eaec37621324a9eea1034b8240cb681cf6b67348a2385a49770562

Observation 16a27061-3fdc-4a50-b658-94fc2332560b · inbound

DataComp-LM: In search of the next generation of training sets for language models cites this paper.

DataComp-LM: In search of the next generation of training sets for language models DataComp: In search of the next generation of multimodal datasets

Reference 66

Resolution
verified exact
arxiv_id, observed 2026-05-17T22:58:16.961906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T22:58:16.523267Z digest=sha256:cda2e376d07e9fac15602727c39469f9a450f2068d466f230a553f565a400d0b

Observation 26e6ffb9-7845-4953-af59-dcce8109799e · inbound

mPLUG-Owl3: Towards Long Image-Sequence Understanding in Multi-Modal Large Language Models cites this paper.

mPLUG-Owl3: Towards Long Image-Sequence Understanding in Multi-Modal Large Language Models DataComp: In search of the next generation of multimodal datasets

Reference 211

Resolution
verified exact
arxiv_id, observed 2026-05-20T06:20:36.377313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-20T06:20:36.235304Z digest=sha256:58830e8d69b4a0345150829c9563941601aa60343b81c1f9e1b60f9ef50eee2b

Observation 30ac765f-d39d-43fa-9d45-b57bc71a3f6f · inbound

FLARE: Fully Integration of Vision-Language Representations for Deep Cross-Modal Understanding cites this paper.

FLARE: Fully Integration of Vision-Language Representations for Deep Cross-Modal Understanding DataComp: In search of the next generation of multimodal datasets

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-22T19:52:01.977529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T19:49:00.961388Z digest=sha256:c2105d9c3a150c62929fce8259063509c54564d2e5d7892ce694a0d6bfda931b

Observation aadfe9c0-a28f-4092-b195-fc2eae3e4cd4 · inbound

Beginning with You: Perceptual-Initialization Improves Vision-Language Representation and Alignment cites this paper.

Beginning with You: Perceptual-Initialization Improves Vision-Language Representation and Alignment DataComp: In search of the next generation of multimodal datasets

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:32.345850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:42:32.345850Z digest=sha256:5a3bcbd5419c2ee8dc8ca714d878af647b234b9f434bd533689b2cbf21df0754

Observation e6df2df8-30ed-41f0-b903-f650583f6d3b · inbound

Spatial-MLLM: Boosting MLLM Capabilities in Visual-based Spatial Intelligence cites this paper.

Spatial-MLLM: Boosting MLLM Capabilities in Visual-based Spatial Intelligence DataComp: In search of the next generation of multimodal datasets

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-16T08:34:36.875949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T08:34:36.824053Z digest=sha256:4f66e2e919756017fe64dec6be74d81817d7260f1673c20352cdde8292e1e801

Observation 38170498-c1f5-4e46-8004-1602482eff3c · inbound

Spatial-MLLM: Boosting MLLM Capabilities in Visual-based Spatial Intelligence cites this paper.

Spatial-MLLM: Boosting MLLM Capabilities in Visual-based Spatial Intelligence DataComp: In search of the next generation of multimodal datasets

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-22T01:00:51.274211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T00:59:13.826054Z digest=sha256:1c02051205bbb7824570a21b471508b1e88479b7b0e3c5bc627013551ad9bdf3

Observation fceae8d0-9b28-4509-a9a7-255461e1da22 · inbound

Smoothed Preference Optimization via ReNoise Inversion for Aligning Diffusion Models with Varied Human Preferences cites this paper.

Smoothed Preference Optimization via ReNoise Inversion for Aligning Diffusion Models with Varied Human Preferences DataComp: In search of the next generation of multimodal datasets

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T11:26:56.445045Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:26:56.445045Z digest=sha256:c0c46ac8fd66fca28dbdc0292b4df6fca1044426a488ce1fdcc49c26f925a9b8

Observation 1a413808-640f-485c-a0d7-4112960d33e7 · inbound

CuRe: Cultural Gaps in the Long Tail of Text-to-Image Systems cites this paper.

CuRe: Cultural Gaps in the Long Tail of Text-to-Image Systems DataComp: In search of the next generation of multimodal datasets

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:16.025685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:16.025685Z digest=sha256:86bb3aa8a7b338dc91264976f480a39047ab243850caaa8ac6a960c7a2f3c042

Observation 28da12cc-c86a-4b77-a6b4-f5ab6e92c7cb · inbound

An Open-Source Software Toolkit & Benchmark Suite for the Evaluation and Adaptation of Multimodal Action Models cites this paper.

An Open-Source Software Toolkit & Benchmark Suite for the Evaluation and Adaptation of Multimodal Action Models DataComp: In search of the next generation of multimodal datasets

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:40.265102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:59:40.265102Z digest=sha256:42c87dc43c8a7dde3171865bf7528a88304e33c8bbd5850c54ec9ed004ca474d

Observation 12352945-800f-405c-90d6-e72bb1f3bac9 · inbound

Ambient Diffusion Omni: Training Good Models with Bad Data cites this paper.

Ambient Diffusion Omni: Training Good Models with Bad Data DataComp: In search of the next generation of multimodal datasets

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T05:01:13.736890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:01:13.736890Z digest=sha256:1be07da3458e46f0d2fbe206afd86cd01fae394bec7c199b6e969e2a92301730

Observation 8ca22854-6d9e-43e6-a931-eeb949699ba6 · inbound

CLIP-like Model as a Foundational Density Ratio Estimator cites this paper.

CLIP-like Model as a Foundational Density Ratio Estimator DataComp: In search of the next generation of multimodal datasets

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T22:03:46.175499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:03:46.175499Z digest=sha256:8db9d7318abdd3bb688516c1b60d80342c1b23e92e7ff1bf158c919fcb7e0464

Observation 6823ab5d-aa99-4459-8da0-55367a9087ae · inbound

MobileCLIP2: Improving Multi-Modal Reinforced Training cites this paper.

MobileCLIP2: Improving Multi-Modal Reinforced Training DataComp: In search of the next generation of multimodal datasets

Reference 2013

Resolution
unresolved
no resolver link, observed 2026-08-05T14:59:19.528028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:59:19.528028Z digest=sha256:043a6205dafc59cca6bb825e476b4a022e98ae4f52e0249191d6a206fde3d4c1

Observation 98faa756-30e0-4b35-a4e4-807594854e0b · inbound

OpenVision 2: A Family of Generative Pretrained Visual Encoders for Multimodal Learning cites this paper.

OpenVision 2: A Family of Generative Pretrained Visual Encoders for Multimodal Learning DataComp: In search of the next generation of multimodal datasets

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T12:22:33.659440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:22:33.659440Z digest=sha256:13d04f1899bf57773807e93eb6d4e578b43bab04bf2777d974452565ad613f1f

Observation ba5b2e99-95be-4cbb-8b2c-e65ec2e8cb55 · inbound

QuiLL: An LLM-Based Vulnerability Assessment Framework for the Wild cites this paper.

QuiLL: An LLM-Based Vulnerability Assessment Framework for the Wild DataComp: In search of the next generation of multimodal datasets

Reference 50

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T11:01:17.173725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T10:56:28.973065Z digest=sha256:779e3abf633405655236fee2fcf87a65117515c692ca6f89994d700a5527c0fb

Observation 36a23a66-588a-4566-86fd-137999944328 · inbound

MixAtlas: Uncertainty-aware Data Mixture Optimization for Multimodal LLM Midtraining cites this paper.

MixAtlas: Uncertainty-aware Data Mixture Optimization for Multimodal LLM Midtraining DataComp: In search of the next generation of multimodal datasets

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T20:08:12.623943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T20:07:29.544613Z digest=sha256:4f42a1f4ca014f0ac074f1fa54500508628500b9b1ea1faf1bd9995931f93e22

Observation 354b2574-26db-433b-b021-4ab0150da4c1 · inbound

Prior-Aligned Data Cleaning for Tabular Foundation Models cites this paper.

Prior-Aligned Data Cleaning for Tabular Foundation Models DataComp: In search of the next generation of multimodal datasets

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:31:17.041197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-07T16:47:31.905149Z digest=sha256:4cc96cc5413fcd4916749ea6e43ef12735a17015ba352cde82029c76f4200524

Observation 44239207-3e69-40cd-b826-4e2a54fcd34f · inbound

From Cradle to Cloud: A Life Cycle Review of AI's Environmental Footprint cites this paper.

From Cradle to Cloud: A Life Cycle Review of AI's Environmental Footprint DataComp: In search of the next generation of multimodal datasets

Reference 38

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T18:31:12.927636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-08T15:43:50.422887Z digest=sha256:ec0c33bda83108f0dc005f9b19d623e0b1cbb6a2fd97de0993d9305fa8aff648

Observation dd4dc2db-6d56-4ac3-be94-13ed77e75d64 · inbound

Offline Preference Optimization for Rectified Flow with Noise-Tracked Pairs cites this paper.

Offline Preference Optimization for Rectified Flow with Noise-Tracked Pairs DataComp: In search of the next generation of multimodal datasets

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-12T02:11:15.568450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T02:10:27.595446Z digest=sha256:df44f2fd43ed80d2a82f8d5ddba6fe2de144ad9be1e5d1c78912d68ce6be7eff

Observation fb96e050-6f1b-4a26-8252-11400aca9b15 · inbound

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison cites this paper.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison DataComp: In search of the next generation of multimodal datasets

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-21T08:39:53.732465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-21T08:36:27.676888Z digest=sha256:5cef9080627c44b8c73f523e2fcf4d17622f564f27d15d5f0f9a37391d671650

Observation cc80062b-c6f6-46c5-be8d-15b48b4c6989 · inbound

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison cites this paper.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison DataComp: In search of the next generation of multimodal datasets

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-06-30T18:04:58.323685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:b51a7cf3df419c94fe4aa38f00fcc124b4aa7c11e5244bf54bc62de02a4e84b2

Observation bffc5f5e-6bfb-4ee6-a4e5-d39898cd29a4 · inbound

GPIC: A Giant Permissive Image Corpus for Visual Generation cites this paper.

GPIC: A Giant Permissive Image Corpus for Visual Generation DataComp: In search of the next generation of multimodal datasets

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:43:14.139724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T07:36:21.262064Z digest=sha256:4bb514a198b86cf768600a362dfb200b40966b282f4c9848e2b025c49534f7fe

Observation 7fab5c0d-92cc-4308-b2b2-1ece77cb957d · inbound

Principles and Practice of Deep Representation Learning: or a Mathematical Theory of Memory cites this paper.

Principles and Practice of Deep Representation Learning: or a Mathematical Theory of Memory DataComp: In search of the next generation of multimodal datasets

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T11:46:55.266557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T03:07:52.730713Z digest=sha256:dc9684d589b6cb0dee858375777dd4254bc437e51f6add40030def681dc60aa9

Observation 8becee99-1461-4517-9165-db60f18cc0b9 · inbound

Instrumented data for causal scientific machine learning cites this paper.

Instrumented data for causal scientific machine learning DataComp: In search of the next generation of multimodal datasets

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-06-27T22:31:21.232799Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T22:24:28.643956Z digest=sha256:cc9183ee18081aaeacc3b57b641d165bd8db02393b5d59b66c2f49e921bb16e5

Observation 02ffbe7a-5687-4cdd-a96d-6257e73f5924 · inbound

Qwen-RobotWorld Technical Report: Unifying Embodied World Modeling through Language-Conditioned Video Generation cites this paper.

Qwen-RobotWorld Technical Report: Unifying Embodied World Modeling through Language-Conditioned Video Generation DataComp: In search of the next generation of multimodal datasets

Reference 59

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T17:18:43.855348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T04:19:26.332718Z digest=sha256:71bc797a824dfbaeeccc548f344e7de8627b176e15fcc6bf64d755fc8af81013

Observation 1ed58a5c-454a-49fe-9e34-bd224feed00c · inbound

TuringViT: Making SOTA Vision Transformers Accessible to All cites this paper.

TuringViT: Making SOTA Vision Transformers Accessible to All DataComp: In search of the next generation of multimodal datasets

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-07-04T16:29:57.904493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T00:29:41.291832Z digest=sha256:b045163be97c8229e8028b5e17a07dd16cba9894d5e73e575c480ac8e0003d16

Observation 60f42309-0369-402a-a0f7-a289360b21fd · inbound

TuringViT: Making SOTA Vision Transformers Accessible to All cites this paper.

TuringViT: Making SOTA Vision Transformers Accessible to All DataComp: In search of the next generation of multimodal datasets

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-06-29T15:03:32.239700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T05:32:26.746776Z digest=sha256:09b2f726331c54c5508dcdbb64bd5e35e5b88001dc65a7e0579087c55320ad6b

Observation 06c80a0a-6a21-4901-9a71-0ac0cf997d2b · inbound

MIRAGE: Protecting against Malicious Image Editing via False Moderation cites this paper.

MIRAGE: Protecting against Malicious Image Editing via False Moderation DataComp: In search of the next generation of multimodal datasets

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-07-04T15:09:54.596740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T01:52:12.291420Z digest=sha256:f05eac0629769ffa1c34af74dbfee266155a1a6ef6b4d72530b43da9967c700c

Observation ad1a47ec-70b8-44bb-9cfa-7ce7b22c1616 · inbound

MIRAGE: Protecting against Malicious Image Editing via False Moderation cites this paper.

MIRAGE: Protecting against Malicious Image Editing via False Moderation DataComp: In search of the next generation of multimodal datasets

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-06-29T19:13:53.521084Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T04:46:56.552601Z digest=sha256:5460a4993483cbb0d665a402b947bbe78a0bbed938a7a5f9022c5c9388defc39

Observation 64f9987d-9b2e-495e-97c9-e375d3f97e63 · inbound

RADIO1D: Elastic Representations for Condensed Vision Modeling cites this paper.

RADIO1D: Elastic Representations for Condensed Vision Modeling DataComp: In search of the next generation of multimodal datasets

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-12T01:07:20.766474Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:07:20.766474Z digest=sha256:5ae25fe4e0fb9cd62efc366457e3ebc0b1f044f34621fc95b8d4659e30efcac1

Observation 139999be-369b-4d71-8716-c21c3f3d0e95 · inbound

Qwen-Audio-VAE Technical Report cites this paper.

Qwen-Audio-VAE Technical Report DataComp: In search of the next generation of multimodal datasets

Reference 74

Resolution
unresolved
no resolver link, observed 2026-07-14T03:31:19.309532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T03:31:19.309532Z digest=sha256:8f467f558af2fac4e940a677fd5d9955506e14e3fac72a2f64873afdddc3cd65

Observation cf286f86-7e59-4e71-9d37-597deed39d0f · inbound

RSIBench-Data: Benchmarking Data-Centric Research for Recursive Self-Improvement cites this paper.

RSIBench-Data: Benchmarking Data-Centric Research for Recursive Self-Improvement DataComp: In search of the next generation of multimodal datasets

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.149546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T01:15:07.149546Z digest=sha256:e16b90aa7292199b833b44fb0db37f08ea2899639a62a2b46f0a03fe691ef15b