Pith. sign in

Paper Citation Record · LEDGER

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation

As of 13 August 2026, this Paper Citation Record lists 71 of 71 outbound references and 3 inbound Pith citation observations for arXiv:2501.04155.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.04155 v1

Coverage vector

measured 71 of 71 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T21:45:46.148987Z

measured 74 of 74 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-14T21:28:37.680681Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

71 of 71 outbound references displayed

  • verified exact0
  • verified fuzzy41
  • unresolved29
  • parse uncertain1
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation e0def824-6a1a-4cca-a2cd-8c401bd92ee2 · outbound

This paper cites an unresolved cited work.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:45:47.698235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:45.674255Z digest=sha256:7d8935d0e4ef7bbeb9d980e4f667d89159a61b9d9a321a849715c2ef5542e44e

Observation 16a9f92a-8eaf-46fc-a474-aa8f9bf4390d · outbound

This paper cites an unresolved cited work.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:45:47.679392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:45.679645Z digest=sha256:04c80da1099ef65ca61b780a511f9de7c49a1a8fdd560586bbd1420fedd3dde4

Observation fa811d3f-279e-405b-87ec-5b725f5caf2c · outbound

This paper cites Eureka: Evaluating and Understanding Large Foundation Models.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Eureka: Evaluating and Understanding Large Foundation Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:45.685534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:45.685534Z digest=sha256:98289312868f776b1f7620b9deea7cc7ed6ec0eab2939f27667fb61cadda25bc

Observation 6b99cb32-86f6-4adc-b1f8-35f0b0144e5b · outbound

This paper cites Introduction to Natural Language Pro- cessing.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Introduction to Natural Language Pro- cessing

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.653822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:45.691150Z digest=sha256:138816b193d65b877bc0b7c0d4913fb71d72641eb4b5189e3f4c2c57a5c8831b

Observation f90a3c62-72a6-41df-bf1e-1a24d80661fe · outbound

This paper cites Conceptual 12m: Pushing web-scale image-text pre- training to recognize long-tail visual concepts.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Conceptual 12m: Pushing web-scale image-text pre- training to recognize long-tail visual concepts

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.636329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:45.703431Z digest=sha256:036fe232723beb6562ac723a600aa9debed2ed3ae1bdf0084cfdf96e2641d01d

Observation d597aca8-aa56-4094-9aa8-b3306e2aed41 · outbound

This paper cites Sharegpt4v: Improving large multi-modal models with better captions,.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Sharegpt4v: Improving large multi-modal models with better captions,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.621454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:45.708945Z digest=sha256:49d6566727f81f8d35b85a4fedcd30132b90be10a365fba66b3265ce3120327c

Observation 2f826d38-c6e4-4022-a905-0f2c1dde5257 · outbound

This paper cites AlpaGasus: Training A Better Alpaca with Fewer Data.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation AlpaGasus: Training A Better Alpaca with Fewer Data

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:45.714805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:45.714805Z digest=sha256:6349e6f7edf49b6a419539e9017118267febd34871392f949605d744b7cd4b26

Observation f57deb5f-8094-49ef-aa8e-8c43bd611be7 · outbound

This paper cites Selection via Proxy: Efficient Data Selection for Deep Learning.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Selection via Proxy: Efficient Data Selection for Deep Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:45.721082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:45.721082Z digest=sha256:7c9cd7328a50fe982b5ef68185a7b0840ccf1ae2a744c6bbf17b6231952d204e

Observation 3f0ef15c-d78f-4b4b-8729-b705e4a8e8c2 · outbound

This paper cites A survey on in-context learning, 2024.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation A survey on in-context learning, 2024

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.597592Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:45.725412Z digest=sha256:e2adc808f25a73ed8b7b8a0025a910e857bee3d4164017aad67ce2c67c8e259f

Observation b88e0666-15ca-46fa-9cea-85783dd9ac06 · outbound

This paper cites The Llama 3 Herd of Models.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation The Llama 3 Herd of Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:45.729541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:45.729541Z digest=sha256:f194e72620047f58dc25fc797c3a7604d72f5468562fca5e6093bce9809e545f

Observation a7f45de2-aae3-463c-b563-90662e2ae71b · outbound

This paper cites Tinystories: How small can language models be and still speak coherent english?, 2023.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Tinystories: How small can language models be and still speak coherent english?, 2023

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.579297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:45.733543Z digest=sha256:cd5ead7a7921a57efe8ae4fa96dcf6a2f7deed7c44f820cc5f188290325b7bbf

Observation 91ce6dd0-283b-4568-9cc0-10540dec98d6 · outbound

This paper cites Data curation via joint example selection further accelerates multimodal learning.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Data curation via joint example selection further accelerates multimodal learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:45.737944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:45.737944Z digest=sha256:9ae227505b3f7be887d0013760f0e8ef33a6f6d1c103e6b098b1853e68c7dc95

Observation a6175318-33b7-43c8-89f7-70956a29f416 · outbound

This paper cites Improving clip training with language rewrites.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Improving clip training with language rewrites

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.538874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:45.743101Z digest=sha256:d03ccd33e09a606db7c19b837f952b0cb57d6c50ee4c4df83d7d0ee370257624

Observation 91bceeee-3b9b-485b-ada7-45b350f679a6 · outbound

This paper cites Data fil- tering networks, 2023.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Data fil- tering networks, 2023

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.520855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:45.747389Z digest=sha256:beec72e543ba632d8ef767adaabe1d752f3fad05a9e825128d8f67804c8c7f49

Observation 480ff99d-ec12-41ec-8a86-7267c4fe9fef · outbound

This paper cites Blink: Multimodal large language models can see but not perceive.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Blink: Multimodal large language models can see but not perceive

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.492428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:45.752067Z digest=sha256:f34ccae760d3e9be8790e20215734a0395e97d9c5a339f7b07295acf489d990a

Observation d06e20d5-c06c-417c-9eb6-8ad535de0467 · outbound

This paper cites Dat- acomp: In search of the next generation of multimodal datasets.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Dat- acomp: In search of the next generation of multimodal datasets

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.475310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:45.757274Z digest=sha256:7df762fa754bf4df87839c411e48e3e2c78403439c5ff272a5d8f0d77cb90f2e

Observation f3951993-f15d-49da-8fe6-8f24fb16161f · outbound

This paper cites Textbooks are all you need, 2023.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Textbooks are all you need, 2023

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.453986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:45.762665Z digest=sha256:29661447588a2e0bd85c91351c934325f134c081146558ce4e98a6c8ac0a521e

Observation 9a90dad8-22f3-4995-9be3-23a89620a890 · outbound

This paper cites Statistical Methods for Speech Recogni- tion.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Statistical Methods for Speech Recogni- tion

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.433242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:45.766460Z digest=sha256:b9c0bf1401f13e3ca8c295480bd57ad0eeb5ad29578d125712070c661b1e29ae

Observation 2d18b204-ead3-48a5-987f-7d9df402fbe2 · outbound

This paper cites Data-efficient contrastive self-supervised learning: Most beneficial exam- ples for supervised learning contribute the least.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Data-efficient contrastive self-supervised learning: Most beneficial exam- ples for supervised learning contribute the least

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.418100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:45.771697Z digest=sha256:4b45ac37fec832400b5261fe81ba86076b25f545fe3d2304c8fa2dc19eb606e6

Observation 26c40066-ca40-4111-ba37-05090a8ab588 · outbound

This paper cites Data-efficient contrastive language-image pre- training: Prioritizing data quality over quantity, 2024.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Data-efficient contrastive language-image pre- training: Prioritizing data quality over quantity, 2024

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.399558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:45.906802Z digest=sha256:5e8466b8dba097af9cfcf50d598014846baa911ceb64eae1e8642c665aeccdb5

Observation 9fc42298-ffae-4ff3-b920-8f8db743d20e · outbound

This paper cites What’s ”up” with vision-language models? investigating their strug- gle with spatial reasoning, 2023.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation What’s ”up” with vision-language models? investigating their strug- gle with spatial reasoning, 2023

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.383268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:45.911771Z digest=sha256:3ddb1ad32b9cf1279dac755195da10226819e52f4d4b1bf01fd9c2c0f2b62056

Observation d9d34882-3d13-45ae-91ff-e3fc602c97ed · outbound

This paper cites Not all sam- ples are created equal: Deep learning with importance sam- pling.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Not all sam- ples are created equal: Deep learning with importance sam- pling

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.360864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:45.915930Z digest=sha256:bd47905681817a581635a9e315c61188566aa1ee0388a8d52f811a5231a0c382

Observation 80fbdd5b-f870-4798-9d68-55de9976f92d · outbound

This paper cites A diagram is worth a dozen images.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation A diagram is worth a dozen images

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:45.920097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:45.920097Z digest=sha256:eb0ae1824a08f4aa75ebace8db3cb5b9d0bcaf89d742434fb4edde4da3a16449

Observation 94ca5eeb-8653-4765-8e0f-abcd00236f8e · outbound

This paper cites Grad-match: Gradient matching based data subset selection for efficient deep model training.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Grad-match: Gradient matching based data subset selection for efficient deep model training

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.309657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:45.928292Z digest=sha256:2919ae296b0be179d1a2bfbd4f764c70a9ac1e30794c3d31f214dda0953eaf5f

Observation cd7275d1-fe30-4ccc-adeb-5b68524de13b · outbound

This paper cites Revisit large-scale image-caption data in pre-training multi- modal foundation models, 2024.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Revisit large-scale image-caption data in pre-training multi- modal foundation models, 2024

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.290284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:45.932593Z digest=sha256:be2e0e27356e2a26cdea66ac77ed314ce55f89ee04e29a0dc6fe197e4f0695a1

Observation c6bc482d-75a7-4721-8036-7e202af722ea · outbound

This paper cites Veclip: Improving clip training via visual-enriched captions,.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Veclip: Improving clip training via visual-enriched captions,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:45.936558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:45.936558Z digest=sha256:91a65ec11f78c7b9609cc9c33a19eae2f069fcd2801bf396c0b8626553d72674

Observation 7dd9a7a5-e192-4923-9288-c98bfaceeb94 · outbound

This paper cites M$^3$IT: A Large-Scale Dataset towards Multi-Modal Multilingual Instruction Tuning.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation M$^3$IT: A Large-Scale Dataset towards Multi-Modal Multilingual Instruction Tuning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:45.940806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:45.940806Z digest=sha256:efbf1ca65cfe4123b3b71bc5080614645d5939dc1b0953db158fbdee15c6e603

Observation f6ccc2c0-cf5c-4081-b7b4-db4d30d78697 · outbound

This paper cites Textbooks are all you need ii: phi-1.5 technical report, 2023.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Textbooks are all you need ii: phi-1.5 technical report, 2023

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:45.945149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:45.945149Z digest=sha256:41c30554a945f94265d872ff82143a9a025c749510338190ef01b166f94d1ce3

Observation a9b8a560-b0a3-4fe4-bfe3-246c2a309289 · outbound

This paper cites Visual instruction tuning, 2023.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Visual instruction tuning, 2023

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.246927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:45.948900Z digest=sha256:51fb3247a87788e2777cd676531767d22017c2f6e642fbf3b89c11b47e94e7a8

Observation 6d8ee901-ecfc-4de0-818e-394d550b92a5 · outbound

This paper cites T-MARS: Improving visual representations by circumventing text feature learning.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation T-MARS: Improving visual representations by circumventing text feature learning

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.232888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:45.952814Z digest=sha256:1e1d8cd1a6a784ee051feda95f96c253aad932b389144ecff564ef3347c8b545

Observation 81b1b235-bad7-4ea9-9c48-ef461e92a2e8 · outbound

This paper cites When Less is More: Investigating Data Pruning for Pretraining LLMs at Scale.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation When Less is More: Investigating Data Pruning for Pretraining LLMs at Scale

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:45.956267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:45.956267Z digest=sha256:0e25335c161c82b5a02004754f15fe2a247ab85f40bee470b76a35e3534b74d7

Observation 97b13408-1408-497c-bbcd-e2ec1173d76f · outbound

This paper cites Joty, and Enamul Hoque.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Joty, and Enamul Hoque

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.217798Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:45.960156Z digest=sha256:fbea503627267be823c26ccc65f1edef2b535a2305250041f9c29a32a87c9245

Observation 585c7a07-f294-4e00-ba51-6e49d2fc65f0 · outbound

This paper cites Chartinstruct: Instruction tuning for chart comprehension and reasoning,.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Chartinstruct: Instruction tuning for chart comprehension and reasoning,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.184369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:45.969503Z digest=sha256:9234bd678464cc6848959bb65b36093b75e076c1b73d611032be94402a0f43f8

Observation 5f923886-9885-455c-a8c6-9fe06caed7df · outbound

This paper cites Mmiu: Multimodal multi- image understanding for evaluating large vision-language models, 2024.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Mmiu: Multimodal multi- image understanding for evaluating large vision-language models, 2024

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.168981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:45.974012Z digest=sha256:cdb47eea5b32efe1c274ebab7e3858bc4330e57abf132532e7638a925e7941b4

Observation 375f069d-e6e4-4116-adcf-8aabd8afa165 · outbound

This paper cites Coresets for data-efficient training of machine learning mod- els.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Coresets for data-efficient training of machine learning mod- els

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.147986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:45.978496Z digest=sha256:ccd10797045bd8b57b79e2e06c36e0d729af35ed1196126f2e0ad0d9b8bd872e

Observation d6d30720-6c6f-4300-a80a-8500b3f2ce19 · outbound

This paper cites Orca 2: Teaching small language models how to reason, 2023.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Orca 2: Teaching small language models how to reason, 2023

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.128426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:45.982471Z digest=sha256:1ab2167c2b5d33cd25304fc38a782539977e2aaec1bc40fe7b267edc72a73fda

Observation 6772d19d-d32f-4bf5-a53f-9cf2e2f4baf9 · outbound

This paper cites Agentinstruct: Toward generative teaching with agentic flows, 2024.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Agentinstruct: Toward generative teaching with agentic flows, 2024

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.111690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:45.987055Z digest=sha256:75127972ee32ee4943243bde1eb31fce70c3fb8c42c492712445ce86d5d64cca

Observation 7d9775f9-a91c-48ae-94b0-4db88dd951a6 · outbound

This paper cites Orca: Progressive learning from complex explanation traces of gpt- 4, 2023.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Orca: Progressive learning from complex explanation traces of gpt- 4, 2023

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.086437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:45.991377Z digest=sha256:69d622eee5782223d693435f40e2346223e4d5600806376bd27c075621b23ca3

Observation 966db625-5b0a-4672-ba1c-ccbe55accc66 · outbound

This paper cites Improving multimodal datasets with image captioning.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Improving multimodal datasets with image captioning

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.060679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:45.995650Z digest=sha256:4553afd0d239ea7521fd4ec3e20711fd87c466437149e5af324e3b6f8eeb35fd

Observation 0d1b369f-d3c2-455c-9aa6-90b60dcc1cdb · outbound

This paper cites GPT-4 Technical Report.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation GPT-4 Technical Report

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:46.000150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:46.000150Z digest=sha256:a091e0c7afee48dacdb259076a4eae12e7de200efc9fed1044e7c22feb06fd17

Observation b37b0661-1dbe-4c0b-a23b-623d3d5e449d · outbound

This paper cites Deep learning on a data diet: Finding important ex- amples early in training.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Deep learning on a data diet: Finding important ex- amples early in training

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:46.004308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:46.004308Z digest=sha256:11e439110d4c167833ea272b6079a2979946da74025212e7a32b8f8f90dcf30b

Observation 7441493e-9901-42cd-8f78-b9741ec5c742 · outbound

This paper cites Adaptive second order coresets for data-efficient machine learning.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Adaptive second order coresets for data-efficient machine learning

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.028134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:46.008764Z digest=sha256:5e8b91362b77d860cd40fd70e5da438b2ab725dcd4d8aa4edbc471e36516d416

Observation 0c9f476e-2b9f-404c-bc01-aeda92c41599 · outbound

This paper cites Learning transferable visual models from natural language supervision, 2021.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Learning transferable visual models from natural language supervision, 2021

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:46.012853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:46.012853Z digest=sha256:579ce152a57ecfe6316c30ef15db06e0bd0e458126d74c0c501f0a241c7bae92

Observation e09306b9-615d-45fb-96bf-a95e914fc6e9 · outbound

This paper cites FuseCap: Leveraging Large Language Models for Enriched Fused Image Captions.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation FuseCap: Leveraging Large Language Models for Enriched Fused Image Captions

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:46.017427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:46.017427Z digest=sha256:1cc92064d3b7002be9677065cc3da91d6073c9fb176b6ffc1771a323db3feb70

Observation 1fe657d6-43b0-4aa6-98ff-41a1eadeca3e · outbound

This paper cites Is a caption worth a thousand im- ages? a study on representation learning.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Is a caption worth a thousand im- ages? a study on representation learning

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:46.992347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:46.022417Z digest=sha256:2de97cb3c01e8c80b9a7c124ab125b6f6d7186f48a3a4649d8009fa497c100ba

Observation 39e4111d-ce49-4428-b350-7cedebdd518d · outbound

This paper cites Laion-5b: An open large-scale dataset for training next generation image-text models.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Laion-5b: An open large-scale dataset for training next generation image-text models

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:46.965511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:46.027100Z digest=sha256:1e98e39e4dea0c4b031e33ae150b523b0dd506446e1835df18bbf3487db14aa0

Observation 36f9edc2-e3cb-4aa2-a14b-42e62cb9f5b1 · outbound

This paper cites Conceptual captions: A cleaned, hypernymed, im- age alt-text dataset for automatic image captioning.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Conceptual captions: A cleaned, hypernymed, im- age alt-text dataset for automatic image captioning

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:46.943332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:46.031188Z digest=sha256:736bc7e2967b5262dbb8378f6f9e67d9ff7b5e47396cebd7a843088d345dc777

Observation d94aea82-d231-4228-bf57-7325fe8e7948 · outbound

This paper cites Math- llava: Bootstrapping mathematical reasoning for multimodal large language models, 2024.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Math- llava: Bootstrapping mathematical reasoning for multimodal large language models, 2024

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:46.915219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:46.035329Z digest=sha256:924feb0e149936867cd65a018016e66ae60afb0b06ceedbdc505d86a804e4b64

Observation f96b1b11-bae4-4952-b324-f9d0c1fd5603 · outbound

This paper cites Chatgpt-4 vision struggles with radiologic image interpretation.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Chatgpt-4 vision struggles with radiologic image interpretation

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:46.896538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:46.039698Z digest=sha256:f4a41e58f2c1c5d56c789200378ce969e0aee43df04496619597d3130e7589a8

Observation 8d4114e7-e92e-4315-919c-15eae6eea1b7 · outbound

This paper cites Dataset Cartography: Mapping and Diagnosing Datasets with Training Dynamics.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Dataset Cartography: Mapping and Diagnosing Datasets with Training Dynamics

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:46.044678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:46.044678Z digest=sha256:a74c430c07b967a7b3b104d4ef3485cbb6348fdde9c52b137ba02376fc735e77

Observation f7827bac-5b2f-49c4-9ae3-25caf3d7bdb2 · outbound

This paper cites an unresolved cited work.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:45:46.872048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:46.049655Z digest=sha256:9cc87d48828f0c528e21d37b2dacc8985b91de12ed2c668efa1939b36ea78196

Observation 780bbae5-20fd-4acc-98d1-5ea42e275b6a · outbound

This paper cites An Empirical Study of Example Forgetting during Deep Neural Network Learning.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation An Empirical Study of Example Forgetting during Deep Neural Network Learning

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:46.054263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:46.054263Z digest=sha256:0d27d4c3f9429f276cc6a609f680066d1c2534556eb51e8ba594ca0b39018f53

Observation a1565f15-2808-4542-a8ff-29620c750f04 · outbound

This paper cites Dynamic data selection for efficient ssl via coarse-to-fine re- finement.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Dynamic data selection for efficient ssl via coarse-to-fine re- finement

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:46.841057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:46.058675Z digest=sha256:5c3496d7bf6c31b03532a65b89403953bf943c03771fb07f1b904117cb520fe0

Observation 352f8531-7729-42b3-ac28-45ac41f85937 · outbound

This paper cites Show and tell: Lessons learned from the 2015 mscoco image captioning challenge.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Show and tell: Lessons learned from the 2015 mscoco image captioning challenge

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:46.063180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:46.063180Z digest=sha256:f40dc9c8472c4059ef610772c01e7afccf58de2e8e4cf077d534795bb3589947

Observation 1715ed42-84bd-4452-97d9-33cd6d7d81a4 · outbound

This paper cites Is a picture worth a thou- sand words? delving into spatial reasoning for vision lan- guage models.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Is a picture worth a thou- sand words? delving into spatial reasoning for vision lan- guage models

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:46.818213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:46.068355Z digest=sha256:a77435792d60253921451a40def8ad2695a71310c2da3ec99efc04e569bf2556

Observation 082548be-0f79-4cb3-b979-4b187a08643c · outbound

This paper cites Decoding Data Quality via Synthetic Corruptions: Embedding-guided Pruning of Code Data.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Decoding Data Quality via Synthetic Corruptions: Embedding-guided Pruning of Code Data

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:46.075086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:46.075086Z digest=sha256:1006b0c0c70f809ee9f750ef35c57dc479bc0db690848f62889faff8e154e241

Observation 0728e6dd-9bae-4dee-8845-0769751b2e0b · outbound

This paper cites Capsfu- sion: Rethinking image-text data at scale, 2024.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Capsfu- sion: Rethinking image-text data at scale, 2024

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:46.803460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:46.081145Z digest=sha256:1b667a7360401e064341dee343f0762779f0579b5f519bf437ea324b5155e676

Observation 4c2bfe0e-b3a2-465f-98a9-51acff9f3453 · outbound

This paper cites Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for ex- pert agi.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for ex- pert agi

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:46.788371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:46.086087Z digest=sha256:149add174ad674dc9d3f6841fb409a33977bb8d49136eeda910fa12a18dcf141

Observation 42f49a25-c3fb-495b-876c-e260f90e1e37 · outbound

This paper cites Multimodal self-instruct: Synthetic abstract image and visual reasoning instruction using language model, 2024.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Multimodal self-instruct: Synthetic abstract image and visual reasoning instruction using language model, 2024

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:46.771221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:46.091233Z digest=sha256:1dc39837940406b9de2df52e5fc29b522240aee1d61ce8fd865c39dd6e87a557

Observation 338dfdf7-4491-485b-b8a4-aa512523bbea · outbound

This paper cites Lima: Less is more for alignment.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Lima: Less is more for alignment

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:46.755567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:46.096926Z digest=sha256:fc4b55e1ba95e051a239f314d54938e33ab0105e25908646a42c9753e28fed46

Observation 0384b6f5-82c5-4c52-82d0-a9ce40c7e24e · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:46.102791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:46.102791Z digest=sha256:cc18b9bb67c56eed4d04ccdb768e29fbf5d478f0b81cc31793d7532129302602

Observation fca5074e-2878-4323-b264-106069c62193 · outbound

This paper cites an unresolved cited work.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Unresolved cited work

Reference 64

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:45:46.737796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:46.109989Z digest=sha256:f5dc6952c3b63b13fec6e2f681eb0d5b1ad64a5f8fdf56d45c600a06f8377ad1

Observation 7d031cdb-4b33-4c4b-9006-b195c8611540 · outbound

This paper cites an unresolved cited work.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Unresolved cited work

Reference 65

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:45:46.718986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:46.115378Z digest=sha256:ecb06af040a40f8cf93e9099ab145245dcb9d213939b899f9bf373649521f52f

Observation 3622dcb6-4e75-4d0a-a05b-a258c69a8ee3 · outbound

This paper cites an unresolved cited work.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Unresolved cited work

Reference 66

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:45:46.701823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:46.120785Z digest=sha256:411748a5023179efa2a2a499e4ed282c80a699a54997e4380de855fafd6100de

Observation 2a2cd15e-85ae-4343-99ba-77929c6c5cd2 · outbound

This paper cites an unresolved cited work.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Unresolved cited work

Reference 67

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:45:46.683180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:46.125900Z digest=sha256:c0a05970aa4a71723020fbc19ff8ea26a0e2ff883028a43c4e8d0595226e09c0

Observation 59f8e91d-9f74-4057-80de-d888344960c2 · outbound

This paper cites Q": The generated question (include options if it’s multiple-choice). -.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Q": The generated question (include options if it’s multiple-choice). -

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:46.446901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:46.131769Z digest=sha256:f37743985fde4653ba51ed90c9bc0f3364ea1de0e11d2336fba2e04f3ef881d3

Observation 842fba42-ce9e-4cd6-a8e7-53cf0c9ec6a4 · outbound

This paper cites an unresolved cited work.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Unresolved cited work

Reference 69

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:45:46.428751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:46.137605Z digest=sha256:4ab69232ad0c4ae6cca7bd14ffc3093a0e2977961b83f0c0e767dea9b616a7bb

Observation cbbbe3ac-33bd-45f5-9780-c2d150fa0176 · outbound

This paper cites an unresolved cited work.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Unresolved cited work

Reference 70

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:45:46.409958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:46.142195Z digest=sha256:c2367d993b2fb9561878e2e4e768da6e51e1840666a4bf30d92d12ddc92476cd

Observation 363f0474-8c9a-4ed0-a547-9ffa06d47600 · outbound

This paper cites an unresolved cited work.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Unresolved cited work

Reference 71

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:45:46.393032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:46.148987Z digest=sha256:83080c8a3db691f3215ce407ec4876ea1b0e87a25b70220e80884a399f116d9c

Observation 26f1a52b-f90d-4bdc-b54b-ae663d6e4b1a · outbound

This paper cites an unresolved cited work.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Unresolved cited work

Reference 251

Resolution
parse uncertain
raw_fallback, observed 2026-08-10T21:45:47.331313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:45.923926Z digest=sha256:815328686f392348a7313acfb5052b4ada26e3736f79dff242dd848eb7b55841

Observation bc4665e9-da03-44f8-9961-72f23475c1bb · outbound

This paper cites 3, 4, 5, 6.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation 3, 4, 5, 6

Reference 2279

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.201313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:45:45.964422Z digest=sha256:320e6895d620d04c0472638b331cf71b836b179bb1270539c21a2822bc89acca

Pith citing papers

Observation ac973bde-508f-42a9-8b61-26faac0367b8 · inbound

20/20 Vision Language Models: A Prescription for Better VLMs through Data Curation Alone cites this paper.

20/20 Vision Language Models: A Prescription for Better VLMs through Data Curation Alone MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-13T02:57:09.442144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-13T02:52:43.674969Z digest=sha256:70ea69ffb3bbd894f59aedc06485a901c8a758a8e1016367b641eb30fa11bac3

Observation e8383d39-89ac-468f-83f8-6624a3821a07 · inbound

20/20 Vision Language Models: A Prescription for Better VLMs through Data Curation Alone cites this paper.

20/20 Vision Language Models: A Prescription for Better VLMs through Data Curation Alone MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:29:28.681673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-14T21:28:37.680681Z digest=sha256:b19f6c79b2f82aab4b7307fbcbf5635e6e510d35a12f044959bb08c933eec7c6

Observation f93a6f5e-3930-4dc7-ad66-56309254a29c · inbound

Unlocking UML Class Diagram Understanding in Vision Language Models cites this paper.

Unlocking UML Class Diagram Understanding in Vision Language Models MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:42:02.618815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-13T01:39:50.027264Z digest=sha256:fca55ec33a57c1186664654c2d61ef19b4fbdcaa9b29e79781affbfd80954090