Pith. sign in

Paper Citation Record · LEDGER

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey

As of 11 August 2026, this Paper Citation Record lists 100 of 300 outbound references and 3 inbound Pith citation observations for arXiv:2501.18648.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.18648 v2

Coverage vector

measured 100 of 300 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T04:36:37.393483Z

measured 103 of 103 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:28:56.015260Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-22T14:14:55.612445Z

Reference resolution

100 of 300 outbound references displayed

  • verified exact2
  • verified fuzzy0
  • unresolved98
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 94d0501c-c7b6-4496-bedb-c82058bfdcd7 · outbound

This paper cites Data aug- mentation techniques in time series domain: a survey and taxonomy.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Data aug- mentation techniques in time series domain: a survey and taxonomy

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.882692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.882692Z digest=sha256:4e3576e3313350766f6dca4621c2e6a339b142f148a97a84606487a0339bb928

Observation 20bd4bed-e38f-4823-824d-d3e311f414ef · outbound

This paper cites A survey on image data augmentation for deep learning.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey A survey on image data augmentation for deep learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.889081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.889081Z digest=sha256:7f1cc9d3bace7e183b7c28702e7214ad03c8916b82b6113bd310590db7534b4b

Observation b01d5ada-6bf2-4c3e-ac15-12d20e9640d5 · outbound

This paper cites Learning to compose domain-specific transformations for data augmentation.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Learning to compose domain-specific transformations for data augmentation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.894427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.894427Z digest=sha256:2204f9d0bf2f34420d61756a20b23c205edf5f718f09125247503e7477dd7e77

Observation b26f48a1-fb4c-4c98-9440-e59df99fe6b4 · outbound

This paper cites Data augmentation: A comprehensive survey of modern approaches.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Data augmentation: A comprehensive survey of modern approaches

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.899931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.899931Z digest=sha256:aab9bd5d5fe7c65b190de96817b210800b6512f80af0a331325c482b38627a1e

Observation 65bc3357-0870-4050-a371-68321bc4a7d5 · outbound

This paper cites An empirical survey of data augmentation for time series classification with neural networks.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey An empirical survey of data augmentation for time series classification with neural networks

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.905473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.905473Z digest=sha256:e90f4ab64ddbb6ce828875a21537692d60321f4add47bf9920e8476351b18d11

Observation 69544c12-a2d4-4e00-9be4-f5b29c07e2ca · outbound

This paper cites A review: Data pre-processing and data augmentation techniques.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey A review: Data pre-processing and data augmentation techniques

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.910502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.910502Z digest=sha256:d96c2dceb3e59730b01041f55b97051f551395a9b822856d83336619615bd3a1

Observation 23dfb845-987a-42cb-ae4e-9133f8a8eedc · outbound

This paper cites Data augmentation in natural language processing: a novel text generation approach for long and short text classifiers.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Data augmentation in natural language processing: a novel text generation approach for long and short text classifiers

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.916917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.916917Z digest=sha256:2da5f5d384235fa63cc3f3de160743e1cb54ec99e605478d1c7224890b4b36ca

Observation ac373589-e0e0-4b5a-a2da-78972ce89953 · outbound

This paper cites Data augmentation for object detection: A review.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Data augmentation for object detection: A review

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.922131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.922131Z digest=sha256:fe3b097cb5af165b295741c4e4675942a37f9e031642bb94b1ed6de505fda966

Observation 1656c9c8-34cb-4acc-b7bb-3aaee5d37680 · outbound

This paper cites Toward text data augmentation for sentiment analysis.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Toward text data augmentation for sentiment analysis

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.927850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.927850Z digest=sha256:fd21567435f9c395e9cfaebabd357786b12490e179501655a13ca20f7e054ccb

Observation d1b250b3-82a4-481d-845d-12d74d746c04 · outbound

This paper cites Incorporating noise robustness in speech command recognition by noise augmentation of training data.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Incorporating noise robustness in speech command recognition by noise augmentation of training data

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.932916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.932916Z digest=sha256:dc4b7ebb1fc26d02018b4180e9bceb3ff05cb8262a45dd6301ba0d5eefdfe5ec

Observation 6c546a1b-7839-4b22-bcd0-701a925fd8ee · outbound

This paper cites Look once to hear: Target speech hearing with noisy examples.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Look once to hear: Target speech hearing with noisy examples

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.937860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.937860Z digest=sha256:f628307db8f7ee78ca19162df4ea6f54d502db385940b73e1c91c4b029a681b2

Observation 3ec26437-5cca-4fa3-a71c-013d6283463c · outbound

This paper cites Perception and sensing for autonomous vehicles under adverse weather conditions: A survey.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Perception and sensing for autonomous vehicles under adverse weather conditions: A survey

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.942752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.942752Z digest=sha256:d82d37658ec015dddf3c154774e1e39a38387fed0252bc1ecb878923c3050c89

Observation 44d4164b-6f92-4d7a-9b79-409a3bf233fd · outbound

This paper cites Safe traffic sign recognition through data augmentation for autonomous vehicles software.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Safe traffic sign recognition through data augmentation for autonomous vehicles software

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.947984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.947984Z digest=sha256:3cec574bee223f579e774a92fdfdacf7d31943ef8b680bc2dfa8020b498092b6

Observation aeeb538b-4661-475d-92d7-3e5b9b28a6c9 · outbound

This paper cites Medical image synthesis for data augmentation and anonymization using generative adversarial networks.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Medical image synthesis for data augmentation and anonymization using generative adversarial networks

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.952931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.952931Z digest=sha256:e973d671b9fc838e70d6914c2728ce368702525a45a0a770472b01cebdc8d75b

Observation 0b48cd95-74fd-4f39-94ad-1ebdb353cf5a · outbound

This paper cites A review of medical image data augmentation techniques for deep learning applications.Journal of Medical Imaging and Radiation Oncology, 65(5):545–563, 2021.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey A review of medical image data augmentation techniques for deep learning applications.Journal of Medical Imaging and Radiation Oncology, 65(5):545–563, 2021

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.957673Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.957673Z digest=sha256:61200fabe6545332c05525251626e5dbc5832c4d5b642327d9d2131ed6df9cfc

Observation ffbb7c91-e59f-417c-a1d2-cb3347b4b00f · outbound

This paper cites To augment or not to augment? a comparative study on text augmentation techniques for low-resource nlp.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey To augment or not to augment? a comparative study on text augmentation techniques for low-resource nlp

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.962899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.962899Z digest=sha256:52b4faddaa0493083afdaced60b4de32924ccd1d22a8ad85f0c23580af982168

Observation f6eb9db6-a210-4066-83fc-3733966f8f9b · outbound

This paper cites An empirical survey of data augmentation for limited data learning in nlp.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey An empirical survey of data augmentation for limited data learning in nlp

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.967816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.967816Z digest=sha256:c1bc3eeb5c431b2213f477fa363c060a316312bb9685be87d4fc6016a555b924

Observation e70e85cb-7f39-4938-a2ca-adf29bf514b5 · outbound

This paper cites A survey on face data augmentation for the training of deep neural networks.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey A survey on face data augmentation for the training of deep neural networks

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.972680Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.972680Z digest=sha256:e5dae2ccb014ef422f519231e3345ef72c0ed53de0eba602e99c1d4e128a9c6f

Observation 3781bc13-09fe-446d-a57b-804a6e91993d · outbound

This paper cites Data augmentation using llms: Data perspectives, learning paradigms and challenges.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Data augmentation using llms: Data perspectives, learning paradigms and challenges

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.977773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.977773Z digest=sha256:87626a2428b15e3f5b1b8d46ec27f40be33482f9e52d9b529bacb36afab44698

Observation f981ae2a-9272-418b-bcd4-aabf4b8a2005 · outbound

This paper cites Generative pre-trained transformer (gpt) in research: A systematic review on data augmentation.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Generative pre-trained transformer (gpt) in research: A systematic review on data augmentation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.982770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.982770Z digest=sha256:84f3626db7bad2041f94f5ae376ae1aef19f81e92e3ce424429752eceef554a3

Observation d4c3769f-f12c-4f90-a753-1620f27d0569 · outbound

This paper cites A survey of knowledge enhanced pre-trained language models.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey A survey of knowledge enhanced pre-trained language models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.988561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.988561Z digest=sha256:458b2b89dc44764e7b79ed41ccb0f4acae6e703a0262700bbed899d145cc4fb0

Observation 7069f3ff-85e4-41dd-938e-12130db40877 · outbound

This paper cites Data aug- mentation techniques for machine learning applied to optical spectroscopy datasets in agrifood applications: A comprehensive review.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Data aug- mentation techniques for machine learning applied to optical spectroscopy datasets in agrifood applications: A comprehensive review

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.994495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.994495Z digest=sha256:4c8d190002abe557872e00b805b463347111d3f5be4f2d84f42859a2a7d889b7

Observation 94144259-c452-47de-ade0-308dff5e95d1 · outbound

This paper cites Speech recognition utilizing deep learning: A systematic review of the latest developments.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Speech recognition utilizing deep learning: A systematic review of the latest developments

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.000143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.000143Z digest=sha256:22b9a197e8f8d8afbea0c7477e6f51061f913a0f7982439a9539dd5116a3929b

Observation 047f96a9-2391-4926-b513-b54fab76ed98 · outbound

This paper cites Data augmentation and deep learning methods in sound classification: A systematic review.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Data augmentation and deep learning methods in sound classification: A systematic review

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.005219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.005219Z digest=sha256:aeba60c027100575445a578882b14141c2bfde8d751ee3242eea64ddaf788c23

Observation 2e211e99-4f1e-48de-bc6c-6631d6603018 · outbound

This paper cites A survey of text data augmentation.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey A survey of text data augmentation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.010677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.010677Z digest=sha256:9044a4cff1ba67f8dc20835741b5c203af5cc12959203ceede2ec97ca821cab2

Observation a98588ea-9db6-47de-9db7-3847882f0d71 · outbound

This paper cites Survey on videos data augmentation for deep learning models.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Survey on videos data augmentation for deep learning models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.015809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.015809Z digest=sha256:f9a4565d4336dd6074c13062309cf6649bbeb7478d6fbd099e421d2d6fb40f27

Observation 22afe267-0dcc-4ff7-8be7-07b0fbbba6f9 · outbound

This paper cites A survey on data augmentation for text classification.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey A survey on data augmentation for text classification

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.020680Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.020680Z digest=sha256:d7c31d4cb342cca131f1e04c3ef596bbc03539f3791b74167f2bc808f14d7da9

Observation e4ba067a-8f49-467d-b3c6-70d28121070e · outbound

This paper cites Image data augmentation approaches: A comprehensive survey and future directions.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Image data augmentation approaches: A comprehensive survey and future directions

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.025611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.025611Z digest=sha256:7c4ce4380e965ede12f4a310adabd6651b926354f0be10166447c81fcf211cd8

Observation b279c55f-c18b-411e-96fa-36716654ebe6 · outbound

This paper cites Advancements in data augmentation and transfer learning: A comprehensive survey to address data scarcity challenges.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Advancements in data augmentation and transfer learning: A comprehensive survey to address data scarcity challenges

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.030778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.030778Z digest=sha256:3d12cbf78ccdb75f569fe96c3489e7b44d8959d06d23fdce9481c0615256bbad

Observation 85ef3018-772c-4e36-be50-720de1cf302b · outbound

This paper cites A survey of synthetic data augmentation methods in machine vision.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey A survey of synthetic data augmentation methods in machine vision

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.035750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.035750Z digest=sha256:2c0d00d666fe2acfa9af83423cb9cca1b117fa039176656978ee4158a42652aa

Observation 64c7f8ed-f1b6-4290-ba28-baee0b8c38bd · outbound

This paper cites Data augmentation using conditional generative adversarial networks for robust speech recognition.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Data augmentation using conditional generative adversarial networks for robust speech recognition

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.041084Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.041084Z digest=sha256:e50ebb309d34c49b51a4479cf16d30a91d608421fab6a4bc7c2376423b7291f3

Observation 812a16ff-185c-428a-96de-98b1c32daeb4 · outbound

This paper cites Data augmentation using generative adversarial networks for robust speech recognition.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Data augmentation using generative adversarial networks for robust speech recognition

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.046168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.046168Z digest=sha256:717232233f710584ad5db5911d1dbbbaf7dcccc595247f79766ded11b5909e45

Observation 4f9ad1ec-e938-4979-ab92-8d86744c1885 · outbound

This paper cites Generative adversarial networks for speech processing: A review.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Generative adversarial networks for speech processing: A review

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.051070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.051070Z digest=sha256:0de5b3e1f403bab95789cd2fda14b9f735d0b69b83750354049632a67b800895

Observation 42ed4e35-1a8a-418d-b404-a16de0ce5541 · outbound

This paper cites How-to conduct a systematic literature review: A quick guide for computer science research.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey How-to conduct a systematic literature review: A quick guide for computer science research

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.055982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.055982Z digest=sha256:2b97ac76f0bd0336ab6decead2905893eb4cad8266a8466848208b2ac836417e

Observation 812304f3-c063-445b-87c9-f5ed85d9604a · outbound

This paper cites Guiding principles for ethical research, n.d.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Guiding principles for ethical research, n.d

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.061218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.061218Z digest=sha256:889e74ed66dc09c7c2b2ab5b43ba998f09cdf71222f96a109942aac7f2b19a89

Observation fa3b37fd-cdbc-4130-9030-b38d857da8b2 · outbound

This paper cites Colour retinal image enhancement based on domain knowledge.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Colour retinal image enhancement based on domain knowledge

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.066104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.066104Z digest=sha256:4fe2e94e16ca747e50af2cfae16e4bdcd4a879c2e916f893828ab26b77d1facf

Observation fe7313cd-b819-4aca-af93-c941c1c12e9b · outbound

This paper cites Gray and color image contrast enhancement by the curvelet transform.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Gray and color image contrast enhancement by the curvelet transform

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.071219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.071219Z digest=sha256:f35c5895306c64045a14d189a2feddf6c91099dbaa8be133f987bdb11617b2c2

Observation b34d0f07-e213-4148-9a67-c9d361340df7 · outbound

This paper cites Image enhancement by histogram hyperbolization.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Image enhancement by histogram hyperbolization

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.076068Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.076068Z digest=sha256:7475a47d8e3a57c03213f987d9a3eb81472cddce95ec742c9931a04b09d14bbd

Observation 528e136e-bb63-4323-ab48-cabce6eef727 · outbound

This paper cites Digital image enhancement and noise filtering by use of local statistics.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Digital image enhancement and noise filtering by use of local statistics

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.081172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.081172Z digest=sha256:adf8278f76586ac42796333e58b77da4970e082047a0156a163ccb291437afe1

Observation 96a247ae-f5a8-4d27-94ab-b2de65b06d5a · outbound

This paper cites Real-time image enhancement techniques.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Real-time image enhancement techniques

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.085957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.085957Z digest=sha256:aea949281f5da23180bd14f9722f61c70820333275e40a7b77a8877b34c729a9

Observation 7faa3157-fb83-4560-96fd-7b399af47c1c · outbound

This paper cites Feature-oriented image enhancement using shock filters.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Feature-oriented image enhancement using shock filters

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.090697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.090697Z digest=sha256:7c0b00c2231ad38f2f5db46ff6d04db38899518619f729c818891f62641b6dff

Observation 83004d1f-c87c-483f-b52d-ab6a1d90613a · outbound

This paper cites Image enhancement using fuzzy set.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Image enhancement using fuzzy set

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.095664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.095664Z digest=sha256:b7dd8f4934107da0d13aa9ce884de4f7c0eaa70a59c0247ef7552e413b7f3f27

Observation 175827d8-b896-42ac-ba4c-d582dcb0519f · outbound

This paper cites Super-resolution image reconstruction: a technical overview.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Super-resolution image reconstruction: a technical overview

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.100884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.100884Z digest=sha256:afe1de0428cb94c9a82163b3249b4396084c48d4ea7fa7b3313e3addd0e910c0

Observation 990bbeb4-8203-4ee4-9592-d2ef5784367d · outbound

This paper cites Accurate camera calibration for off-line, video-based augmented reality.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Accurate camera calibration for off-line, video-based augmented reality

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.107949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.107949Z digest=sha256:f2622f4e33142540cbed38a092007487199530d20c8acb5b4f749668bfc94789

Observation 7da54937-9cf4-4220-9d59-6b4b00b0a8e2 · outbound

This paper cites A statistical approach to material classification using image patch exemplars.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey A statistical approach to material classification using image patch exemplars

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.113472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.113472Z digest=sha256:c401ed1909624055127d921f2744167ce51eb07b77ca61727036b6e86e19d236

Observation f6a059e3-fde3-47d8-a5a3-4524f00033c2 · outbound

This paper cites Jittering reduction in marker-based augmented reality systems.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Jittering reduction in marker-based augmented reality systems

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.118287Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.118287Z digest=sha256:4c126d42c551d8f804b90b14fa238724401c19c135e93e55ab0750b8ca49570d

Observation f1fc3810-cf00-4a21-8c32-a4b9896f447b · outbound

This paper cites Transform image enhancement.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Transform image enhancement

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.123096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.123096Z digest=sha256:590d279ee6ecf18f5c66edf06abd54abda67ca87719cb7201a28d48a5796b3da

Observation 2e56f12f-af01-43bc-96ae-73e94f361680 · outbound

This paper cites Camera identification from cropped and scaled images.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Camera identification from cropped and scaled images

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.128344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.128344Z digest=sha256:97080ffeb3c7046874e394cadc4e38f02b5f3505b36f0a4d0a14169dbd48181c

Observation b9eed5c1-c58d-4b34-8d6a-ce3d7a217034 · outbound

This paper cites Image enhancement by nonlinear extrapolation in frequency space.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Image enhancement by nonlinear extrapolation in frequency space

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.133661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.133661Z digest=sha256:d233c8550f4f90bb71946b390bcecfc551ec262e2a738575e254abc3aff3f048

Observation edc5880b-f68e-44e1-9a5c-0e4d3c784343 · outbound

This paper cites A text-to-picture synthesis system for augmenting communication.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey A text-to-picture synthesis system for augmenting communication

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.138683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.138683Z digest=sha256:370629c8b494c47a8046f66b1290bc5bd181777fc3cb979fbaabf7fd43e8a3fd

Observation 79d4184e-1a34-4194-a742-04b9b0229ffa · outbound

This paper cites Semantic representations of near-synonyms for automatic lexical choice.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Semantic representations of near-synonyms for automatic lexical choice

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.143570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.143570Z digest=sha256:98c6c7ab4349f30587a11c4ebf247e6b5a7a4ed1d20f4de1af1f49bc35dab946

Observation cbf2b730-d0cc-4065-94a7-b3923a72f6ee · outbound

This paper cites Word sense disambiguation with pictures.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Word sense disambiguation with pictures

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.148422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.148422Z digest=sha256:378b04ef8607980adc4ff58289222fb97914748e04b72d9a4bacaccbee3c4649

Observation e82c4bbf-ffc6-4448-a768-c78a6d2236c7 · outbound

This paper cites Text classification by augmenting the bag-of-words representation with redundancy-compensated bigrams.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Text classification by augmenting the bag-of-words representation with redundancy-compensated bigrams

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.153908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.153908Z digest=sha256:40d48408359ce4ae0d603f2dd8f6dd8c6120b148d912d033dc0b251f6c133845

Observation 6c2e3b0b-3789-4881-a5e8-4e2aef11c6c5 · outbound

This paper cites Addition–deletion networks.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Addition–deletion networks

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.158661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.158661Z digest=sha256:c5be44c564b3644f2c3eb43e4bb6fbf48226ea27bef8621eca0fa5dae473102d

Observation efb7ac88-e8ea-4fee-835d-16114cba5a36 · outbound

This paper cites Are good texts always better? interactions of text coherence, background knowledge, and levels of understanding in learning from text.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Are good texts always better? interactions of text coherence, background knowledge, and levels of understanding in learning from text

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.163568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.163568Z digest=sha256:cf3f19253f03ea4e408e7cfd09cc485e6b35c191e1685ed42728de3517d20a49

Observation 47aa5e3a-a54b-4d19-8362-55c6779ca1db · outbound

This paper cites Missing inaction: the dangers of ignoring missing data.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Missing inaction: the dangers of ignoring missing data

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.168609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.168609Z digest=sha256:5d13e95ceab2fa889756ecbfb65fb8abdac640de49283e68aeeb2e5e8ee7925e

Observation 0fe8770a-c501-4df3-90a4-29f580b48d33 · outbound

This paper cites Techniques for automatically correcting words in text.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Techniques for automatically correcting words in text

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.173447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.173447Z digest=sha256:be3c86ddf5b2a6d6f9064f786374d00ec7d3d14409f0b13378256ada76e31b86

Observation 51312975-2988-4c94-b3aa-4588c66283ea · outbound

This paper cites An augmented template-based approach to text realization.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey An augmented template-based approach to text realization

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.178231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.178231Z digest=sha256:c00e6330681798a74d2e3ac13be6b8f3a0d28d3d87a1d597ef397e3013bb4236

Observation 6a8205ec-b85d-44a5-9935-364ac6061e59 · outbound

This paper cites Scene text extraction and translation for handheld devices.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Scene text extraction and translation for handheld devices

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.183331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.183331Z digest=sha256:bf72a0b1aadcbfb27e5f32700313ef786937e4bdca24e71951f89a623c5e88b2

Observation f51130ce-83e6-459e-b0fc-e0630a9dd286 · outbound

This paper cites Augmenting the power of lsi in text retrieval: Singular value rescaling.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Augmenting the power of lsi in text retrieval: Singular value rescaling

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.188607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.188607Z digest=sha256:0a07b6b864faa9336d939535af0628f591e2c139612d4894ca74704ff2e5e0ac

Observation 7f78fc6c-2303-47e7-8799-8b34c9d84090 · outbound

This paper cites The proper place of men and machines in language translation.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey The proper place of men and machines in language translation

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.194122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.194122Z digest=sha256:acfdd64437bd2241dbbcfe080d4839983033b5d7f609f8d8701ab5b2b92ce192

Observation b29a1f95-5189-4d32-bc61-34e6b7ac6288 · outbound

This paper cites Embedding web-based statistical translation models in cross-language information retrieval.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Embedding web-based statistical translation models in cross-language information retrieval

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.199312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.199312Z digest=sha256:aa9aaa452149327d5f98febaadcbf1601bf534342bf5f08e54bf1c5663be09ff

Observation f764078f-f7fd-4fb1-ba20-082ce830bde4 · outbound

This paper cites A technical word-and term-translation aid using noisy parallel corpora across language groups.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey A technical word-and term-translation aid using noisy parallel corpora across language groups

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.204671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.204671Z digest=sha256:ace48228ff7bd83ff376e4fef9d9a262df05786c349603bbeb3b15b313d013cc

Observation caf2b31b-d28b-446a-b3f2-405c6ce696dc · outbound

This paper cites Iterative clustering of high dimensional text data augmented by local search.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Iterative clustering of high dimensional text data augmented by local search

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.210201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.210201Z digest=sha256:ddb0ac667eee0a8f46381e71bdc2071b4dcfe42f6f972e18e6e35a199f831ed9

Observation 1df9d0f5-65cb-41d9-9e85-cefb82e68153 · outbound

This paper cites Augmented audio reality: Telepresence/vr hybrid acoustic environments.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Augmented audio reality: Telepresence/vr hybrid acoustic environments

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.215357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.215357Z digest=sha256:8d38254d3fc052244dd7b94beb5aed1fab3daa3232f7ff55d80081779a7ccbe2

Observation 42897cc0-5516-46e4-94ed-d98d133e5635 · outbound

This paper cites Vocal cord augmentation with autogenous fat.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Vocal cord augmentation with autogenous fat

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.220186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.220186Z digest=sha256:feb5c71dc6dbce7231a456da014cb04105f105020bf0f70cff0c791fb3d2415a

Observation 6b01ec0d-1268-4c79-8737-fb9ee07b3ad1 · outbound

This paper cites Audio and visually augmented teleconferencing.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Audio and visually augmented teleconferencing

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.225307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.225307Z digest=sha256:ed6c2764e14407407403ae1cd424bf46dfbfe206f33d0e7bda87bf10cffdd3e1

Observation ef72fd8c-b6d2-4a92-9412-4f156063c2f0 · outbound

This paper cites Ackerman, and Debby Hindus.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Ackerman, and Debby Hindus

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.230099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.230099Z digest=sha256:0790d371267fbe10ca12a89c60cbe38c5ad7e7fc7a42fba20167eb876d572c89

Observation d30e73a8-294e-48d7-a27e-2f431bf18d8e · outbound

This paper cites Can the lombard effect be used to improve low voice intensity in parkinson’s disease? European Journal of Disorders of Communication , 27(2):121–127, 1992.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Can the lombard effect be used to improve low voice intensity in parkinson’s disease? European Journal of Disorders of Communication , 27(2):121–127, 1992

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.235099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.235099Z digest=sha256:9035a221e23cdad21d9d159bce2f3fadde302686dcf76eef67d708b8b3e93691

Observation 32902113-038e-4a72-b45f-6e7cbd72e8a6 · outbound

This paper cites Speech perception, localization, and lateralization with bilateral cochlear implants.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Speech perception, localization, and lateralization with bilateral cochlear implants

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.239690Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.239690Z digest=sha256:39452a2c9cda696631a6387f6f8474da502a603ed1be1b822d796e315b5dcbc9

Observation 7e7c8182-2a5d-41a7-baab-7d5730cb3611 · outbound

This paper cites Improving performance in noise for hearing aids and cochlear implants using coherent modulation filtering.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Improving performance in noise for hearing aids and cochlear implants using coherent modulation filtering

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.244655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.244655Z digest=sha256:0e258e76ef959faeb987ae153000eb1df0d163b332831e22527fa3446160fb2a

Observation 4d824749-df9e-4229-9065-ec4133a6bfbc · outbound

This paper cites Speech recognition for a humanoid with motor noise utilizing missing feature theory.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Speech recognition for a humanoid with motor noise utilizing missing feature theory

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.249717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.249717Z digest=sha256:5af8bf4d276dfe77e8dd41a50b8334c45483f14c4efdb5225f62244e41fce409

Observation 80b8c125-affd-40cb-b290-e20e686ed734 · outbound

This paper cites Augmented reality audio for mobile and wearable appliances.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Augmented reality audio for mobile and wearable appliances

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.254610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.254610Z digest=sha256:4c3027c4dfd4def7e051044207a8df356b28036b102902202470ab562faa4063

Observation 61b3996c-e4c7-4e28-bf19-f9ae3f4af0e4 · outbound

This paper cites Power supply noise in analog audio class d amplifiers.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Power supply noise in analog audio class d amplifiers

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.259505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.259505Z digest=sha256:c73636abf45187ff5d5eeec73a4662a6aca7fd224aae9c806d750da82a147cab

Observation b7b396fc-ee80-4798-a315-eb60b9c05f85 · outbound

This paper cites Deep convolutional neural network based medical image classification for disease diagnosis.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Deep convolutional neural network based medical image classification for disease diagnosis

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.264379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.264379Z digest=sha256:1b7d3a511e0ffb2fc585bbaf0e9afc7f316077c85db3a94636348b6701ce32ff

Observation 25e9b7e2-f5c9-4c1f-be9a-6c3398ab068f · outbound

This paper cites Going deep in medical image analysis: concepts, methods, challenges, and future directions.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Going deep in medical image analysis: concepts, methods, challenges, and future directions

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.269219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.269219Z digest=sha256:808d3c4c81b93c270caf3529e09fec7a0509b6241d34bd9c00248645ded99d46

Observation ea5177f1-14c0-444e-98ed-48c19828a46d · outbound

This paper cites Adversarial differentiable data augmentation for autonomous systems.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Adversarial differentiable data augmentation for autonomous systems

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.274303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.274303Z digest=sha256:1c08e41d1de61b856e8c6875fe159f8b128e5fa31af85ee7e043ce7919e635f3

Observation 11a311c2-d345-4bf3-a521-b44d9567cfa4 · outbound

This paper cites Data augmentation for deep learning based semantic segmentation and crop-weed classification in agricultural robotics.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Data augmentation for deep learning based semantic segmentation and crop-weed classification in agricultural robotics

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.279774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.279774Z digest=sha256:c85a3a458ab6bd68862178a3470328424bc3e7f59b69cfdbb9e79ff4812a80db

Observation f338ce1e-4f3a-4453-9a37-0b07c3278872 · outbound

This paper cites DAGAM: Data Augmentation with Generation And Modification.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey DAGAM: Data Augmentation with Generation And Modification

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.284791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.284791Z digest=sha256:cf96dfedf4356bf5f5fb9a68071f896769ef5d8bcdd88ab8c907ada15ec5309b

Observation 7ae1d42d-27a9-4f11-a1f6-dd6f5df6df43 · outbound

This paper cites Generative AI in Agriculture: Creating Image Datasets Using DALL.E's Advanced Large Language Model Capabilities.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Generative AI in Agriculture: Creating Image Datasets Using DALL.E's Advanced Large Language Model Capabilities

Reference 80

Resolution
verified exact
local_arxiv, observed 2026-08-10T04:36:40.551209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T04:36:37.290650Z digest=sha256:a7259e270323a958975e3b3550574663095dcf23492f90c9e31cafaa50826dbb

Observation 66cda6a2-0cf3-4af4-8763-b9000de2769f · outbound

This paper cites Synthetic data generation for tabular health records: A systematic review.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Synthetic data generation for tabular health records: A systematic review

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.296356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.296356Z digest=sha256:edb3c6ed194dd217de830d1835721750b44800e7550a651e4e39e1a408cb6233

Observation ef86f27d-84d5-4c4d-bfe1-1b3f042ebc67 · outbound

This paper cites Data augmentation for medical imaging: A systematic literature review.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Data augmentation for medical imaging: A systematic literature review

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.301684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.301684Z digest=sha256:fef87890a72e7f4ab125e719b0b5f8453c30cd32d42da240464887266500b6c2

Observation c848dc83-e0e6-4821-8644-be026470ab82 · outbound

This paper cites Multi-modal llms in agriculture: A comprehensive review.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Multi-modal llms in agriculture: A comprehensive review

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.306903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.306903Z digest=sha256:76278714a1568441da98daf60a59fb8ab683d862694d8d31732ea85a27392736

Observation 1859460f-82db-41c6-bd72-383f27b9a4bf · outbound

This paper cites Research on data augmentation for image classification based on convolution neural networks.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Research on data augmentation for image classification based on convolution neural networks

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.312065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.312065Z digest=sha256:e920745216e683ea16b413e5234d5e87c972441cb5959575e523a4d9b17be3bd

Observation 7cf4a101-8fd4-4295-8906-671ef30ae74d · outbound

This paper cites Data augmentation and generative machine learning on the cloud platform.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Data augmentation and generative machine learning on the cloud platform

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.317193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.317193Z digest=sha256:167183d8c38bde637ed4f7ce9552e43d49ac5bb46df8899f314ac2b609c57f5c

Observation f99bdc6d-2f1f-4985-aa55-124ffe2e1be6 · outbound

This paper cites Data augmentation using deep generative models for embedding based speaker recognition.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Data augmentation using deep generative models for embedding based speaker recognition

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.322050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.322050Z digest=sha256:6aaecef14b3a4251138a900f2cb075163b34b904845ab719c6ca91646406ba15

Observation 93ce9540-afb3-4093-bf4a-d7df71200239 · outbound

This paper cites Statistical Augmentation of a Chinese Machine-Readable Dictionary.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Statistical Augmentation of a Chinese Machine-Readable Dictionary

Reference 87

Resolution
verified exact
local_arxiv, observed 2026-08-10T04:36:40.524795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T04:36:37.326820Z digest=sha256:9ef7de0b6d3d6ef12046b9ef22b60dd252267193dfaedf6d8f0db31cfd1268b1

Observation 25744ac4-66a5-4abf-bf79-8ed70b2a9c03 · outbound

This paper cites A comparison of id3 and backpropagation for english text-to-speech mapping.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey A comparison of id3 and backpropagation for english text-to-speech mapping

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.332059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.332059Z digest=sha256:8051fc89feb3c63bf893e2f537b77553158dd758f5a67f44b3bb8d7d2f8840d3

Observation 5b0ba7af-0f83-4446-9faf-6b1d4923bf61 · outbound

This paper cites an unresolved cited work.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Unresolved cited work

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.336892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.336892Z digest=sha256:fe1e8383e55b965505a4038ceb155d8ce0f309fe5a3110f0db969f145be4f106

Observation 5ef8e37c-f0ff-4371-8ae0-70c261709fb3 · outbound

This paper cites Srl-aco: A text augmentation framework based on semantic role labeling and ant colony optimization.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Srl-aco: A text augmentation framework based on semantic role labeling and ant colony optimization

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.341637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.341637Z digest=sha256:153d9d9c9352d817e13ee436356d3836d016e037d5bba376cf6e4634b019c396

Observation 24e63926-3e4d-468b-a93e-1c35a2c781f9 · outbound

This paper cites From theories on styles to their transfer in text: Bridging the gap with a hierarchical survey.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey From theories on styles to their transfer in text: Bridging the gap with a hierarchical survey

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.346759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.346759Z digest=sha256:16ce700ffd57dc4ce282ece30c48dc40ad3dc71c719d496aec8b0d673e68ee2e

Observation 8662a7b1-9908-44b4-b1d8-d66254b59665 · outbound

This paper cites Summary of chatgpt-related research and perspective towards the future of large language models.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Summary of chatgpt-related research and perspective towards the future of large language models

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.351795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.351795Z digest=sha256:5407dad2b37de8cbcbc478123594eba38b2e1ad86eacc2f77fdf7d36a7d6edd7

Observation 25ef9681-2fbe-4811-8d35-af8604763111 · outbound

This paper cites Survey on deep neural networks in speech and vision systems.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Survey on deep neural networks in speech and vision systems

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.357623Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.357623Z digest=sha256:0e3c8c88342a2b7b4a2fc60bead96fa77b33e8459137185aae39efdafdedee7c

Observation 7fa32c4f-8dbe-4dab-a294-3ee124a44298 · outbound

This paper cites an unresolved cited work.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Unresolved cited work

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.363135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.363135Z digest=sha256:e165d6481a7b4d182cfeb45b57b41467d2e42ff83ee63c413b50c696fa085cfc

Observation abd27d9c-19ff-481e-8596-2c3c65e3ebca · outbound

This paper cites Automated building damage assessment and large-scale mapping by integrating satellite imagery, gis, and deep learning.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Automated building damage assessment and large-scale mapping by integrating satellite imagery, gis, and deep learning

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.368046Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.368046Z digest=sha256:db4bf2ad12b467d36458693123eb9eae87c3669ecbc5c7d7729814f697978cd2

Observation e117fe4c-cbf5-48d1-9a64-f181e685b8ac · outbound

This paper cites Machine learning for advanced emission monitoring and reduction strategies in fossil fuel power plants.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Machine learning for advanced emission monitoring and reduction strategies in fossil fuel power plants

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.373656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.373656Z digest=sha256:daa9aacac7f77709bd975cdacbdc1c0fb2e58e979f7e97f011680c5e5d67ff32

Observation 8a93cebc-465f-4130-8bcd-1fd05bad4371 · outbound

This paper cites A review on large language models: Architectures, applications, taxonomies, open issues and challenges.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey A review on large language models: Architectures, applications, taxonomies, open issues and challenges

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.378456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.378456Z digest=sha256:d97c81a76f4b209277379ced0b1ceeaf484565a3b90ee531d3c0ca86e771ed81

Observation c7d05b4c-179d-4d19-aedf-b3ae5ddff980 · outbound

This paper cites A comparison on data augmentation methods based on deep learning for audio classification.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey A comparison on data augmentation methods based on deep learning for audio classification

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.383521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.383521Z digest=sha256:4efac05872b829a63d01d3834381a395ea003fd21117e712155aac6e736befd7

Observation 6fcc3640-d25c-4f29-83d7-16ef18effa2c · outbound

This paper cites A study on data augmentation in voice anti-spoofing.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey A study on data augmentation in voice anti-spoofing

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.388689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.388689Z digest=sha256:fc2aef53b241afcd9c679dfb61c36d2dfe4938e9d275f592740eefae8b8010d5

Observation 052846a0-90c6-4626-aa5d-9913de8dbf8d · outbound

This paper cites On the analysis of data augmentation methods for spectral imaged based heart sound classification using convolutional neural networks.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey On the analysis of data augmentation methods for spectral imaged based heart sound classification using convolutional neural networks

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.393483Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.393483Z digest=sha256:e72e9a8bcadb28c8508ca081e30620a8d1f1873a98f88504027ea42e9d5742a5

Pith citing papers

Observation 763b5aaf-b75e-4ada-8887-4dfad38ffd6c · inbound

SafeTrans: LLM-assisted Transpilation from C to Rust cites this paper.

SafeTrans: LLM-assisted Transpilation from C to Rust Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-22T14:14:55.616435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T14:12:55.504993Z digest=sha256:d5c72b195c59bf83c08e76a33d7675c26764f8e6b3d0744ece4f3695052e67d0

Observation 80ddc295-914b-43f5-864b-888679f4acd2 · inbound

Breaking the Barriers of Text-Hungry and Audio-Deficient AI cites this paper.

Breaking the Barriers of Text-Hungry and Audio-Deficient AI Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey

Reference 112

Resolution
unresolved
no resolver link, observed 2026-08-07T11:28:56.015260Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:28:56.015260Z digest=sha256:be637069efa57bdf6f0c38442db3fe8c21c036b7e6b8a9ae7051052b152da2cc

Observation 73b99071-00fa-46f8-a614-09d889d6ee33 · inbound

ORBIT: Guided Agentic Orchestration for Autonomous C-to-Rust Transpilation cites this paper.

ORBIT: Guided Agentic Orchestration for Autonomous C-to-Rust Transpilation Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:46:05.073472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T15:19:09.020005Z digest=sha256:3961c2317577f46f24618f7025a5722893641cb53304092bb2dee95f66fb858f