Pith. sign in

Paper Citation Record · LEDGER

DataComp: In search of the next generation of multimodal datasets

As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 40 inbound Pith citation observations for arXiv:2304.14108.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2304.14108 v5

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 40 of 40 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 40 of 40 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T14:59:01.673140Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

74
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 53e4f70b-9858-4a6d-889c-4aa69b2da6f1 · inbound

Objaverse-XL: A Universe of 10M+ 3D Objects cites this paper.

Objaverse-XL: A Universe of 10M+ 3D Objects DataComp: In search of the next generation of multimodal datasets

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-17T13:02:11.604349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-17T13:02:11.512409Z digest=sha256:102bdaedde20d7531df1b897834c7bf0c980eff1a3019157a74ede7da76ca36c

Observation 5b0763c5-fe42-473d-ae4b-031e297c053a · inbound

InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation cites this paper.

InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation DataComp: In search of the next generation of multimodal datasets

Reference 71

Resolution
verified exact
arxiv_id, observed 2026-05-15T06:30:22.712246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-15T06:30:22.431538Z digest=sha256:f803c85f784e93875cc97ef365686bf60de5b4fc6dfefe7c00ac9fc434fdeb12

Observation b7853a04-19fe-49b5-bcb6-5fa3b35955f9 · inbound

OpenFlamingo: An Open-Source Framework for Training Large Autoregressive Vision-Language Models cites this paper.

OpenFlamingo: An Open-Source Framework for Training Large Autoregressive Vision-Language Models DataComp: In search of the next generation of multimodal datasets

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-14T01:52:01.415728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-14T01:52:01.163900Z digest=sha256:386acd79e781330db560dcafd61255e60a557926bac8ee0a048827b99527a40e

Observation 50537838-c88e-44dd-add6-32810a4947d5 · inbound

mPLUG-Owl2: Revolutionizing Multi-modal Large Language Model with Modality Collaboration cites this paper.

mPLUG-Owl2: Revolutionizing Multi-modal Large Language Model with Modality Collaboration DataComp: In search of the next generation of multimodal datasets

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-18T03:18:51.844203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-18T03:18:51.582340Z digest=sha256:092179ae67d1ff9c31029bdedc025e3f138b7c53fc52a58b9c4a8a635261b182

Observation dbd967c6-4515-481d-8b9a-4f0905432019 · inbound

ShareGPT4V: Improving Large Multi-Modal Models with Better Captions cites this paper.

ShareGPT4V: Improving Large Multi-Modal Models with Better Captions DataComp: In search of the next generation of multimodal datasets

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-13T17:08:12.842273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-13T17:08:12.727773Z digest=sha256:45d61ad585aca7e7a91fa9102a28c7a1c3279b513a6fbc21e2f77b9882847864

Observation 16a27061-3fdc-4a50-b658-94fc2332560b · inbound

DataComp-LM: In search of the next generation of training sets for language models cites this paper.

DataComp-LM: In search of the next generation of training sets for language models DataComp: In search of the next generation of multimodal datasets

Reference 66

Resolution
verified exact
arxiv_id, observed 2026-05-17T22:58:16.961906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-17T22:58:16.523267Z digest=sha256:911fd0ae33e56e25c55b6988d245bc358a4b3c48afca8c72b375c8fbe0ed5c57

Observation 26e6ffb9-7845-4953-af59-dcce8109799e · inbound

mPLUG-Owl3: Towards Long Image-Sequence Understanding in Multi-Modal Large Language Models cites this paper.

mPLUG-Owl3: Towards Long Image-Sequence Understanding in Multi-Modal Large Language Models DataComp: In search of the next generation of multimodal datasets

Reference 211

Resolution
verified exact
arxiv_id, observed 2026-05-20T06:20:36.377313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-20T06:20:36.235304Z digest=sha256:9315a00b30b093868c3c9c8950f209d9b053261fa1941c1b12766742735866aa

Observation 4e3bda14-adab-4da1-99bc-5aef939eb0d6 · inbound

FastVLM: Efficient Vision Encoding for Vision Language Models cites this paper.

FastVLM: Efficient Vision Encoding for Vision Language Models DataComp: In search of the next generation of multimodal datasets

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T13:19:23.176565Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:19:23.176565Z digest=sha256:c7000bd6b38466b0c28fdd4999765876a7ef219ca946174baa8176e48a783a03

Observation fa64068f-00f8-4ec1-990d-c6e4c2d63e7e · inbound

Next Token Prediction Towards Multimodal Intelligence: A Comprehensive Survey cites this paper.

Next Token Prediction Towards Multimodal Intelligence: A Comprehensive Survey DataComp: In search of the next generation of multimodal datasets

Reference 127

Resolution
unresolved
no resolver link, observed 2026-08-11T14:59:01.673140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:59:01.673140Z digest=sha256:8ab3dd6a7e55e912628a6a710c02bd31d72b55237c9caffb716de7e250da232c

Observation 7dc61a48-3420-4b62-82db-fccc3d0d19ae · inbound

LR0.FM: Low-Res Benchmark and Improving Robustness for Zero-Shot Classification in Foundation Models cites this paper.

LR0.FM: Low-Res Benchmark and Improving Robustness for Zero-Shot Classification in Foundation Models DataComp: In search of the next generation of multimodal datasets

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-09T00:12:36.209585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T00:12:36.209585Z digest=sha256:70a6927cbe973ba494b6b26b5f59282601a8c984fc49758504cd0a063cfa3f67

Observation 30ac765f-d39d-43fa-9d45-b57bc71a3f6f · inbound

FLARE: Fully Integration of Vision-Language Representations for Deep Cross-Modal Understanding cites this paper.

FLARE: Fully Integration of Vision-Language Representations for Deep Cross-Modal Understanding DataComp: In search of the next generation of multimodal datasets

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-22T19:52:01.977529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T19:49:00.961388Z digest=sha256:531d43a9a7ac94888d5f3d662483ba747b8812a07cafeb9af1a54cfe8af7c06f

Observation aadfe9c0-a28f-4092-b195-fc2eae3e4cd4 · inbound

Beginning with You: Perceptual-Initialization Improves Vision-Language Representation and Alignment cites this paper.

Beginning with You: Perceptual-Initialization Improves Vision-Language Representation and Alignment DataComp: In search of the next generation of multimodal datasets

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:32.345850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:42:32.345850Z digest=sha256:069fdae0433be518c5363d2291912f13c0f23aad7cbe665abd0ab59999a12222

Observation e6df2df8-30ed-41f0-b903-f650583f6d3b · inbound

Spatial-MLLM: Boosting MLLM Capabilities in Visual-based Spatial Intelligence cites this paper.

Spatial-MLLM: Boosting MLLM Capabilities in Visual-based Spatial Intelligence DataComp: In search of the next generation of multimodal datasets

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-16T08:34:36.875949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-16T08:34:36.824053Z digest=sha256:ac77a8e10ca17a39e0997474934a3335ca2df1ef77bdac35603c4151deeb4508

Observation 38170498-c1f5-4e46-8004-1602482eff3c · inbound

Spatial-MLLM: Boosting MLLM Capabilities in Visual-based Spatial Intelligence cites this paper.

Spatial-MLLM: Boosting MLLM Capabilities in Visual-based Spatial Intelligence DataComp: In search of the next generation of multimodal datasets

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-22T01:00:51.274211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T00:59:13.826054Z digest=sha256:c8f4b6f16dae4b8cec7255fa323784bbdb62b6ec821862d7bdc40495e604d2dd

Observation fceae8d0-9b28-4509-a9a7-255461e1da22 · inbound

Smoothed Preference Optimization via ReNoise Inversion for Aligning Diffusion Models with Varied Human Preferences cites this paper.

Smoothed Preference Optimization via ReNoise Inversion for Aligning Diffusion Models with Varied Human Preferences DataComp: In search of the next generation of multimodal datasets

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T11:26:56.445045Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:26:56.445045Z digest=sha256:4237ef5afed16eb688df8fedd156059150871ed7577dffb044cc8efa59ac6aca

Observation 1a413808-640f-485c-a0d7-4112960d33e7 · inbound

CuRe: Cultural Gaps in the Long Tail of Text-to-Image Systems cites this paper.

CuRe: Cultural Gaps in the Long Tail of Text-to-Image Systems DataComp: In search of the next generation of multimodal datasets

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:16.025685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:16.025685Z digest=sha256:1152e6a07213be97f6898acfd184a931061fc300c4d35f83c7ab5f9b86fd7926

Observation 28da12cc-c86a-4b77-a6b4-f5ab6e92c7cb · inbound

An Open-Source Software Toolkit & Benchmark Suite for the Evaluation and Adaptation of Multimodal Action Models cites this paper.

An Open-Source Software Toolkit & Benchmark Suite for the Evaluation and Adaptation of Multimodal Action Models DataComp: In search of the next generation of multimodal datasets

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:40.265102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:59:40.265102Z digest=sha256:119da92b33e55d416c7054f6be15b10ee26bd628e2db592d70cd6b60a1e7b912

Observation 12352945-800f-405c-90d6-e72bb1f3bac9 · inbound

Ambient Diffusion Omni: Training Good Models with Bad Data cites this paper.

Ambient Diffusion Omni: Training Good Models with Bad Data DataComp: In search of the next generation of multimodal datasets

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T05:01:13.736890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:01:13.736890Z digest=sha256:fe06ebbac191d870d5b60e6526a3566b0832528bb5713d038b989f447f439986

Observation 8ca22854-6d9e-43e6-a931-eeb949699ba6 · inbound

CLIP-like Model as a Foundational Density Ratio Estimator cites this paper.

CLIP-like Model as a Foundational Density Ratio Estimator DataComp: In search of the next generation of multimodal datasets

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T22:03:46.175499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:03:46.175499Z digest=sha256:8db9d7318abdd3bb688516c1b60d80342c1b23e92e7ff1bf158c919fcb7e0464

Observation 6823ab5d-aa99-4459-8da0-55367a9087ae · inbound

MobileCLIP2: Improving Multi-Modal Reinforced Training cites this paper.

MobileCLIP2: Improving Multi-Modal Reinforced Training DataComp: In search of the next generation of multimodal datasets

Reference 2013

Resolution
unresolved
no resolver link, observed 2026-08-05T14:59:19.528028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:59:19.528028Z digest=sha256:043a6205dafc59cca6bb825e476b4a022e98ae4f52e0249191d6a206fde3d4c1

Observation 98faa756-30e0-4b35-a4e4-807594854e0b · inbound

OpenVision 2: A Family of Generative Pretrained Visual Encoders for Multimodal Learning cites this paper.

OpenVision 2: A Family of Generative Pretrained Visual Encoders for Multimodal Learning DataComp: In search of the next generation of multimodal datasets

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T12:22:33.659440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:22:33.659440Z digest=sha256:13d04f1899bf57773807e93eb6d4e578b43bab04bf2777d974452565ad613f1f

Observation ba5b2e99-95be-4cbb-8b2c-e65ec2e8cb55 · inbound

QuiLL: An LLM-Based Vulnerability Assessment Framework for the Wild cites this paper.

QuiLL: An LLM-Based Vulnerability Assessment Framework for the Wild DataComp: In search of the next generation of multimodal datasets

Reference 50

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T11:01:17.173725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-18T10:56:28.973065Z digest=sha256:ed2e677592d38e8e81ec8d77cb373c9620489e8a83e2436476313b48465fd4c0

Observation 36a23a66-588a-4566-86fd-137999944328 · inbound

MixAtlas: Uncertainty-aware Data Mixture Optimization for Multimodal LLM Midtraining cites this paper.

MixAtlas: Uncertainty-aware Data Mixture Optimization for Multimodal LLM Midtraining DataComp: In search of the next generation of multimodal datasets

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T20:08:12.623943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-13T20:07:29.544613Z digest=sha256:a77ace0db64c5c72fbd3d3e8a9984ed0fb8cc96974c8597d73dde7ffaa769a13

Observation 354b2574-26db-433b-b021-4ab0150da4c1 · inbound

Prior-Aligned Data Cleaning for Tabular Foundation Models cites this paper.

Prior-Aligned Data Cleaning for Tabular Foundation Models DataComp: In search of the next generation of multimodal datasets

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:31:17.041197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-07T16:47:31.905149Z digest=sha256:cce5696f3ed69c751ebdc653ea022c81c9e2369950fa82708b6e18f0583a3537

Observation 44239207-3e69-40cd-b826-4e2a54fcd34f · inbound

From Cradle to Cloud: A Life Cycle Review of AI's Environmental Footprint cites this paper.

From Cradle to Cloud: A Life Cycle Review of AI's Environmental Footprint DataComp: In search of the next generation of multimodal datasets

Reference 38

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T18:31:12.927636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T15:43:50.422887Z digest=sha256:59bda4260bf75249919b3c30e7a97bda71d17225a61841f184b4994e9b60ab28

Observation dd4dc2db-6d56-4ac3-be94-13ed77e75d64 · inbound

Offline Preference Optimization for Rectified Flow with Noise-Tracked Pairs cites this paper.

Offline Preference Optimization for Rectified Flow with Noise-Tracked Pairs DataComp: In search of the next generation of multimodal datasets

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-12T02:11:15.568450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-12T02:10:27.595446Z digest=sha256:09ebcbc8e25e9a0b02ddb4ab07b93dc61cf9124e2fa4cc293273942ae2c63836

Observation fb96e050-6f1b-4a26-8252-11400aca9b15 · inbound

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison cites this paper.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison DataComp: In search of the next generation of multimodal datasets

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-21T08:39:53.732465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-21T08:36:27.676888Z digest=sha256:b649efd151eac68355e0c783f10862bccc1fe7484ff05bd0508fc515c9db19fe

Observation cc80062b-c6f6-46c5-be8d-15b48b4c6989 · inbound

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison cites this paper.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison DataComp: In search of the next generation of multimodal datasets

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-06-30T18:04:58.323685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:4f4f617d65b4d3987c4a29e4fadba1e4877648c7cc4044a9830cbce50dc63abc

Observation bffc5f5e-6bfb-4ee6-a4e5-d39898cd29a4 · inbound

GPIC: A Giant Permissive Image Corpus for Visual Generation cites this paper.

GPIC: A Giant Permissive Image Corpus for Visual Generation DataComp: In search of the next generation of multimodal datasets

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:43:14.139724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-29T07:36:21.262064Z digest=sha256:70b9af7a35854b41abd9ec9011070a8635eed85b0ac59bc6f29455cb9ba07806

Observation 7fab5c0d-92cc-4308-b2b2-1ece77cb957d · inbound

Principles and Practice of Deep Representation Learning: or a Mathematical Theory of Memory cites this paper.

Principles and Practice of Deep Representation Learning: or a Mathematical Theory of Memory DataComp: In search of the next generation of multimodal datasets

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T11:46:55.266557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-28T03:07:52.730713Z digest=sha256:f88514676ca0f66a197902b48aa75e3bbcf4596dd36dae2e742d1dfdba54d391

Observation 8becee99-1461-4517-9165-db60f18cc0b9 · inbound

Instrumented data for causal scientific machine learning cites this paper.

Instrumented data for causal scientific machine learning DataComp: In search of the next generation of multimodal datasets

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-06-27T22:31:21.232799Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-27T22:24:28.643956Z digest=sha256:60336768b3dba72c4fa7c7fa956b650de41e64f114477662201a40f067f66760

Observation 02ffbe7a-5687-4cdd-a96d-6257e73f5924 · inbound

Qwen-RobotWorld Technical Report: Unifying Embodied World Modeling through Language-Conditioned Video Generation cites this paper.

Qwen-RobotWorld Technical Report: Unifying Embodied World Modeling through Language-Conditioned Video Generation DataComp: In search of the next generation of multimodal datasets

Reference 59

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T17:18:43.855348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-06-27T04:19:26.332718Z digest=sha256:c9da3bc1d6975d64142d42a8b4bfc82520b8d4d545f899528a50be72d0ba04d0

Observation 1ed58a5c-454a-49fe-9e34-bd224feed00c · inbound

TuringViT: Making SOTA Vision Transformers Accessible to All cites this paper.

TuringViT: Making SOTA Vision Transformers Accessible to All DataComp: In search of the next generation of multimodal datasets

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-07-04T16:29:57.904493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-26T00:29:41.291832Z digest=sha256:cd4344c54b243fd3a237b73a0bdbbb83531673c02cb7620790bda7837b73f981

Observation 60f42309-0369-402a-a0f7-a289360b21fd · inbound

TuringViT: Making SOTA Vision Transformers Accessible to All cites this paper.

TuringViT: Making SOTA Vision Transformers Accessible to All DataComp: In search of the next generation of multimodal datasets

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-06-29T15:03:32.239700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-29T05:32:26.746776Z digest=sha256:a4d20e391a569c2f9b3787cd34ba6ea475eb7279d4cf13f5b9069c68313075e5

Observation 06c80a0a-6a21-4901-9a71-0ac0cf997d2b · inbound

MIRAGE: Protecting against Malicious Image Editing via False Moderation cites this paper.

MIRAGE: Protecting against Malicious Image Editing via False Moderation DataComp: In search of the next generation of multimodal datasets

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-07-04T15:09:54.596740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-26T01:52:12.291420Z digest=sha256:c5cc3d67d9533c4cd14451f26d2fec8c95133db856d48118d8e2646931d726bf

Observation ad1a47ec-70b8-44bb-9cfa-7ce7b22c1616 · inbound

MIRAGE: Protecting against Malicious Image Editing via False Moderation cites this paper.

MIRAGE: Protecting against Malicious Image Editing via False Moderation DataComp: In search of the next generation of multimodal datasets

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-06-29T19:13:53.521084Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-29T04:46:56.552601Z digest=sha256:7fe9d6d0b0d955ffa5f58ed40555a17264fc40a3d5e42e324f8366921c554256

Observation 64f9987d-9b2e-495e-97c9-e375d3f97e63 · inbound

RADIO1D: Elastic Representations for Condensed Vision Modeling cites this paper.

RADIO1D: Elastic Representations for Condensed Vision Modeling DataComp: In search of the next generation of multimodal datasets

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-12T01:07:20.766474Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:07:20.766474Z digest=sha256:5ae25fe4e0fb9cd62efc366457e3ebc0b1f044f34621fc95b8d4659e30efcac1

Observation 139999be-369b-4d71-8716-c21c3f3d0e95 · inbound

Qwen-Audio-VAE Technical Report cites this paper.

Qwen-Audio-VAE Technical Report DataComp: In search of the next generation of multimodal datasets

Reference 74

Resolution
unresolved
no resolver link, observed 2026-07-14T03:31:19.309532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T03:31:19.309532Z digest=sha256:8f467f558af2fac4e940a677fd5d9955506e14e3fac72a2f64873afdddc3cd65

Observation cf286f86-7e59-4e71-9d37-597deed39d0f · inbound

RSIBench-Data: Benchmarking Data-Centric Research for Recursive Self-Improvement cites this paper.

RSIBench-Data: Benchmarking Data-Centric Research for Recursive Self-Improvement DataComp: In search of the next generation of multimodal datasets

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-01T01:15:07.149546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T01:15:07.149546Z digest=sha256:e16b90aa7292199b833b44fb0db37f08ea2899639a62a2b46f0a03fe691ef15b

Observation 96d3de4c-1eb1-4273-ba32-125730c3fefc · inbound

Gaze Behavior in Visual World Experiments Can be Modeled With Off-the-shelf Language-Vision Encoders cites this paper.

Gaze Behavior in Visual World Experiments Can be Modeled With Off-the-shelf Language-Vision Encoders DataComp: In search of the next generation of multimodal datasets

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-10T11:09:10.532529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T11:09:10.532529Z digest=sha256:1c39dddb51838eb793c3d5dba287ef23601a5bacc0da4cd7337c09c74e31e6de