Pith. sign in

Paper Citation Record · LEDGER

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey

As of 10 August 2026, this Paper Citation Record lists 100 of 300 outbound references and 3 inbound Pith citation observations for arXiv:2501.18648.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.18648 v2

Coverage vector

measured 100 of 300 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T04:36:37.393483Z

measured 103 of 103 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:28:56.015260Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-22T14:14:55.612445Z

Reference resolution

100 of 300 outbound references displayed

  • verified exact2
  • verified fuzzy0
  • unresolved98
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 94d0501c-c7b6-4496-bedb-c82058bfdcd7 · outbound

This paper cites Data aug- mentation techniques in time series domain: a survey and taxonomy.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Data aug- mentation techniques in time series domain: a survey and taxonomy

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.882692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.882692Z digest=sha256:1a377b033820adc4b557a2d1021ace8d40cc73e4956b92b1983a29356f7d2a49

Observation 20bd4bed-e38f-4823-824d-d3e311f414ef · outbound

This paper cites A survey on image data augmentation for deep learning.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey A survey on image data augmentation for deep learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.889081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.889081Z digest=sha256:bb98b49605f2daed246ca81b0536eef7c369ac490936df122e5b85586834afe6

Observation b01d5ada-6bf2-4c3e-ac15-12d20e9640d5 · outbound

This paper cites Learning to compose domain-specific transformations for data augmentation.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Learning to compose domain-specific transformations for data augmentation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.894427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.894427Z digest=sha256:f2d72e9064ddadf0bbfa1adc63319fbf80fbaf088fb5d39de4fb9e9675fdb633

Observation b26f48a1-fb4c-4c98-9440-e59df99fe6b4 · outbound

This paper cites Data augmentation: A comprehensive survey of modern approaches.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Data augmentation: A comprehensive survey of modern approaches

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.899931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.899931Z digest=sha256:3cdbaef1e689f0d2e04681a1317850f166cf8c9c051df6769cabeff787c738dd

Observation 65bc3357-0870-4050-a371-68321bc4a7d5 · outbound

This paper cites An empirical survey of data augmentation for time series classification with neural networks.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey An empirical survey of data augmentation for time series classification with neural networks

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.905473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.905473Z digest=sha256:d04133fb40241b6c792343f4835ce51fec0e27c77a9a656f978496efac95f38d

Observation 69544c12-a2d4-4e00-9be4-f5b29c07e2ca · outbound

This paper cites A review: Data pre-processing and data augmentation techniques.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey A review: Data pre-processing and data augmentation techniques

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.910502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.910502Z digest=sha256:f5d353495aecaa7a5003f4f7097c2b62466f70d96b14ec9d54827da499a6c7d5

Observation 23dfb845-987a-42cb-ae4e-9133f8a8eedc · outbound

This paper cites Data augmentation in natural language processing: a novel text generation approach for long and short text classifiers.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Data augmentation in natural language processing: a novel text generation approach for long and short text classifiers

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.916917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.916917Z digest=sha256:f117565c837af86ac34d6586978caa7bb27541433c55ef9c9f84f5235c6ed834

Observation ac373589-e0e0-4b5a-a2da-78972ce89953 · outbound

This paper cites Data augmentation for object detection: A review.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Data augmentation for object detection: A review

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.922131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.922131Z digest=sha256:095ce2ee38edbb469a8cc2c0a1dd5f6c7d6c0cf7ea8cff392790dcaaa56600f2

Observation 1656c9c8-34cb-4acc-b7bb-3aaee5d37680 · outbound

This paper cites Toward text data augmentation for sentiment analysis.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Toward text data augmentation for sentiment analysis

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.927850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.927850Z digest=sha256:8d86a6108c23003828d2c56878727b2295bf95e1e9596a591f20d6ddddd2eee6

Observation d1b250b3-82a4-481d-845d-12d74d746c04 · outbound

This paper cites Incorporating noise robustness in speech command recognition by noise augmentation of training data.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Incorporating noise robustness in speech command recognition by noise augmentation of training data

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.932916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.932916Z digest=sha256:b93395f97247b88e4d4c9f242f86e8a064b2c6411a045c9d65b4404a2bc91675

Observation 6c546a1b-7839-4b22-bcd0-701a925fd8ee · outbound

This paper cites Look once to hear: Target speech hearing with noisy examples.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Look once to hear: Target speech hearing with noisy examples

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.937860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.937860Z digest=sha256:cbcf486fbc8f6de9a36d6495f15e2724c8359514f54a1a108e6cf6ccd6b1841a

Observation 3ec26437-5cca-4fa3-a71c-013d6283463c · outbound

This paper cites Perception and sensing for autonomous vehicles under adverse weather conditions: A survey.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Perception and sensing for autonomous vehicles under adverse weather conditions: A survey

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.942752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.942752Z digest=sha256:985c6f232797061f1c9520640794e60df51b65935b96f49ece8398120b73ba01

Observation 44d4164b-6f92-4d7a-9b79-409a3bf233fd · outbound

This paper cites Safe traffic sign recognition through data augmentation for autonomous vehicles software.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Safe traffic sign recognition through data augmentation for autonomous vehicles software

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.947984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.947984Z digest=sha256:35fba3903a70698c2d6719c5bb4a04715c066d48e2c1511bd5559a11d36e653c

Observation aeeb538b-4661-475d-92d7-3e5b9b28a6c9 · outbound

This paper cites Medical image synthesis for data augmentation and anonymization using generative adversarial networks.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Medical image synthesis for data augmentation and anonymization using generative adversarial networks

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.952931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.952931Z digest=sha256:468ae102b9ccb001ff09bb808ff926ca920b581fb6c8363660ed4d5f95fe8a95

Observation 0b48cd95-74fd-4f39-94ad-1ebdb353cf5a · outbound

This paper cites A review of medical image data augmentation techniques for deep learning applications.Journal of Medical Imaging and Radiation Oncology, 65(5):545–563, 2021.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey A review of medical image data augmentation techniques for deep learning applications.Journal of Medical Imaging and Radiation Oncology, 65(5):545–563, 2021

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.957673Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.957673Z digest=sha256:a378fffc550f8bc6a3db7f2def67bdb1dce18a81f768e63fc941cc851c347e30

Observation ffbb7c91-e59f-417c-a1d2-cb3347b4b00f · outbound

This paper cites To augment or not to augment? a comparative study on text augmentation techniques for low-resource nlp.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey To augment or not to augment? a comparative study on text augmentation techniques for low-resource nlp

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.962899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.962899Z digest=sha256:fc00b4a09ed53f69922db65a81aaa21683a8ce5ffb9adc544dac59c779c32ed3

Observation f6eb9db6-a210-4066-83fc-3733966f8f9b · outbound

This paper cites An empirical survey of data augmentation for limited data learning in nlp.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey An empirical survey of data augmentation for limited data learning in nlp

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.967816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.967816Z digest=sha256:7636254da8888ead92a1e309c6843ad04914d1a77f944ffde82644b3ce6aedbc

Observation e70e85cb-7f39-4938-a2ca-adf29bf514b5 · outbound

This paper cites A survey on face data augmentation for the training of deep neural networks.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey A survey on face data augmentation for the training of deep neural networks

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.972680Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.972680Z digest=sha256:df94c0074cbf15eaef845fd88dd507f82439f523dbeaa61ed0d1201816ff8db7

Observation 3781bc13-09fe-446d-a57b-804a6e91993d · outbound

This paper cites Data augmentation using llms: Data perspectives, learning paradigms and challenges.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Data augmentation using llms: Data perspectives, learning paradigms and challenges

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.977773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.977773Z digest=sha256:48b46c79f98221107aec1aac39b73c9b4bb65b75d96b7c6d41a31d6d34834d37

Observation f981ae2a-9272-418b-bcd4-aabf4b8a2005 · outbound

This paper cites Generative pre-trained transformer (gpt) in research: A systematic review on data augmentation.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Generative pre-trained transformer (gpt) in research: A systematic review on data augmentation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.982770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.982770Z digest=sha256:eaf1834128efd1c163a1e2447193ec439471e0d0a76be2521fa7c2020cce5fdb

Observation d4c3769f-f12c-4f90-a753-1620f27d0569 · outbound

This paper cites A survey of knowledge enhanced pre-trained language models.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey A survey of knowledge enhanced pre-trained language models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.988561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.988561Z digest=sha256:a1f509dca96da4a95445053f9c692b6462925ccc5a80fdbb7b2d3865bd54fa7c

Observation 7069f3ff-85e4-41dd-938e-12130db40877 · outbound

This paper cites Data aug- mentation techniques for machine learning applied to optical spectroscopy datasets in agrifood applications: A comprehensive review.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Data aug- mentation techniques for machine learning applied to optical spectroscopy datasets in agrifood applications: A comprehensive review

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:36.994495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:36.994495Z digest=sha256:a93c0be187c45dee9a53ac574ded8d8cf401c69576099bd3d2a907e30aad7226

Observation 94144259-c452-47de-ade0-308dff5e95d1 · outbound

This paper cites Speech recognition utilizing deep learning: A systematic review of the latest developments.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Speech recognition utilizing deep learning: A systematic review of the latest developments

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.000143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.000143Z digest=sha256:967877b36dd15b8ca2489458849958e67f1c887a78cf1e3d2312a9ea39f5db85

Observation 047f96a9-2391-4926-b513-b54fab76ed98 · outbound

This paper cites Data augmentation and deep learning methods in sound classification: A systematic review.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Data augmentation and deep learning methods in sound classification: A systematic review

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.005219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.005219Z digest=sha256:94c61fea59d25a788f844bdb3185657ac56883cd9d30caaa0b917740aaf1ee66

Observation 2e211e99-4f1e-48de-bc6c-6631d6603018 · outbound

This paper cites A survey of text data augmentation.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey A survey of text data augmentation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.010677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.010677Z digest=sha256:fe46e75a211c36b2554fb0f79abb678b413b725a354dedeb8187df4abaa6549d

Observation a98588ea-9db6-47de-9db7-3847882f0d71 · outbound

This paper cites Survey on videos data augmentation for deep learning models.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Survey on videos data augmentation for deep learning models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.015809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.015809Z digest=sha256:6b1953d442d772a5a6d28677f010a6ba0e5cf0357801b345c44d9d57ad39b32c

Observation 22afe267-0dcc-4ff7-8be7-07b0fbbba6f9 · outbound

This paper cites A survey on data augmentation for text classification.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey A survey on data augmentation for text classification

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.020680Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.020680Z digest=sha256:65225c14a642de7bb00d773e7a220c960b0d4f5fed3cf058e519af476ec5728b

Observation e4ba067a-8f49-467d-b3c6-70d28121070e · outbound

This paper cites Image data augmentation approaches: A comprehensive survey and future directions.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Image data augmentation approaches: A comprehensive survey and future directions

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.025611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.025611Z digest=sha256:c2859078af4cea7d45c8277dd53bbfda134d7cab3cbcca7d6f93f9a1f551328c

Observation b279c55f-c18b-411e-96fa-36716654ebe6 · outbound

This paper cites Advancements in data augmentation and transfer learning: A comprehensive survey to address data scarcity challenges.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Advancements in data augmentation and transfer learning: A comprehensive survey to address data scarcity challenges

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.030778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.030778Z digest=sha256:0ed89c367f592a44280ab6128a2015768f744435fed5947b178840f978689089

Observation 85ef3018-772c-4e36-be50-720de1cf302b · outbound

This paper cites A survey of synthetic data augmentation methods in machine vision.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey A survey of synthetic data augmentation methods in machine vision

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.035750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.035750Z digest=sha256:eb67f60e00f45100172bd931070f0e264c1c101d57f35c264055e093c3a87f55

Observation 64c7f8ed-f1b6-4290-ba28-baee0b8c38bd · outbound

This paper cites Data augmentation using conditional generative adversarial networks for robust speech recognition.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Data augmentation using conditional generative adversarial networks for robust speech recognition

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.041084Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.041084Z digest=sha256:346760f56bda8c125adbe3ab0a33da4e5e61e8d54d80b5935cec7939691a60d3

Observation 812a16ff-185c-428a-96de-98b1c32daeb4 · outbound

This paper cites Data augmentation using generative adversarial networks for robust speech recognition.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Data augmentation using generative adversarial networks for robust speech recognition

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.046168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.046168Z digest=sha256:173812c8b662cc964bfab30c4bc4196b957760d8a9a993498455cc6d26799e00

Observation 4f9ad1ec-e938-4979-ab92-8d86744c1885 · outbound

This paper cites Generative adversarial networks for speech processing: A review.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Generative adversarial networks for speech processing: A review

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.051070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.051070Z digest=sha256:9de809744f858cefdd4ee4a47f4ad80737cc9c48572b355bdb36000f4a5937f2

Observation 42ed4e35-1a8a-418d-b404-a16de0ce5541 · outbound

This paper cites How-to conduct a systematic literature review: A quick guide for computer science research.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey How-to conduct a systematic literature review: A quick guide for computer science research

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.055982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.055982Z digest=sha256:2906f28680bf900cf8a28ab1e97e24bdfef2727ad7cdd00a9339311220eff15f

Observation 812304f3-c063-445b-87c9-f5ed85d9604a · outbound

This paper cites Guiding principles for ethical research, n.d.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Guiding principles for ethical research, n.d

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.061218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.061218Z digest=sha256:607ed7beed85ebabedce667541a9a642938be01de37c9ff630f1118029cb0dd1

Observation fa3b37fd-cdbc-4130-9030-b38d857da8b2 · outbound

This paper cites Colour retinal image enhancement based on domain knowledge.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Colour retinal image enhancement based on domain knowledge

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.066104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.066104Z digest=sha256:b14571299424067a2eb187e0b0ab4efd6e6cd056fbd914a3a224016c2ca1a83b

Observation fe7313cd-b819-4aca-af93-c941c1c12e9b · outbound

This paper cites Gray and color image contrast enhancement by the curvelet transform.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Gray and color image contrast enhancement by the curvelet transform

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.071219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.071219Z digest=sha256:05dd31e45d779b8e921d585a8b26ac1a60f03d3ad751402943b1fcc5c07585d5

Observation b34d0f07-e213-4148-9a67-c9d361340df7 · outbound

This paper cites Image enhancement by histogram hyperbolization.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Image enhancement by histogram hyperbolization

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.076068Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.076068Z digest=sha256:1b44219f0b22d1fa0dfd99baf34cff9d15f9267b13ac1e86a7aeabfa42050a99

Observation 528e136e-bb63-4323-ab48-cabce6eef727 · outbound

This paper cites Digital image enhancement and noise filtering by use of local statistics.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Digital image enhancement and noise filtering by use of local statistics

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.081172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.081172Z digest=sha256:67f7be48b88188bb520fc7d74ce9d0082be0782019e6176f1e92579aafded044

Observation 96a247ae-f5a8-4d27-94ab-b2de65b06d5a · outbound

This paper cites Real-time image enhancement techniques.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Real-time image enhancement techniques

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.085957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.085957Z digest=sha256:0801903b55e2e7824f11d4085367537c1cde0d6e88bc9c1ef992f68c5d63ff2e

Observation 7faa3157-fb83-4560-96fd-7b399af47c1c · outbound

This paper cites Feature-oriented image enhancement using shock filters.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Feature-oriented image enhancement using shock filters

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.090697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.090697Z digest=sha256:d27ece54bc26fba82ab51c9c757fe9ca28b0eb52c8271f1ffc494e552bbdd7a0

Observation 83004d1f-c87c-483f-b52d-ab6a1d90613a · outbound

This paper cites Image enhancement using fuzzy set.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Image enhancement using fuzzy set

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.095664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.095664Z digest=sha256:329a59400f363104eb0f71b61f33334bdc56e356f30a84ee845b1e533315a009

Observation 175827d8-b896-42ac-ba4c-d582dcb0519f · outbound

This paper cites Super-resolution image reconstruction: a technical overview.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Super-resolution image reconstruction: a technical overview

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.100884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.100884Z digest=sha256:ff8c29799b33c5803a272b5eddee1377aaefa73048e17528fefd5ff3a0abfb15

Observation 990bbeb4-8203-4ee4-9592-d2ef5784367d · outbound

This paper cites Accurate camera calibration for off-line, video-based augmented reality.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Accurate camera calibration for off-line, video-based augmented reality

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.107949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.107949Z digest=sha256:96a6175e782222aa22e8dc01ff40926d9bf2d0bea8999819e9b80f81aef32324

Observation 7da54937-9cf4-4220-9d59-6b4b00b0a8e2 · outbound

This paper cites A statistical approach to material classification using image patch exemplars.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey A statistical approach to material classification using image patch exemplars

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.113472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.113472Z digest=sha256:77fce2555aafc225f3d549e7dccd0e0811a56b1761d2a15822d1719dd096c4cc

Observation f6a059e3-fde3-47d8-a5a3-4524f00033c2 · outbound

This paper cites Jittering reduction in marker-based augmented reality systems.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Jittering reduction in marker-based augmented reality systems

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.118287Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.118287Z digest=sha256:eebd78d941b2bc671fd54a2bc33b7aec122ce0a147fd2362274943f14f74f739

Observation f1fc3810-cf00-4a21-8c32-a4b9896f447b · outbound

This paper cites Transform image enhancement.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Transform image enhancement

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.123096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.123096Z digest=sha256:833ffe3ede29f3cb61b8f166e2fac3615fa0f21c6f79d163c258833c8ab1b447

Observation 2e56f12f-af01-43bc-96ae-73e94f361680 · outbound

This paper cites Camera identification from cropped and scaled images.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Camera identification from cropped and scaled images

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.128344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.128344Z digest=sha256:0f9b5cb81c32f20f6e7ff5a282648ec1602a219b52b8fddff3bae546ec233f15

Observation b9eed5c1-c58d-4b34-8d6a-ce3d7a217034 · outbound

This paper cites Image enhancement by nonlinear extrapolation in frequency space.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Image enhancement by nonlinear extrapolation in frequency space

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.133661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.133661Z digest=sha256:1625c1ab32282ee17c5440cd647ef49f1436d8442c9465c4faddc7eb6d55e9da

Observation edc5880b-f68e-44e1-9a5c-0e4d3c784343 · outbound

This paper cites A text-to-picture synthesis system for augmenting communication.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey A text-to-picture synthesis system for augmenting communication

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.138683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.138683Z digest=sha256:ded6e3d1a683a8eb187a64b47564360343920be207d0b9307cecd9f5a248d7b6

Observation 79d4184e-1a34-4194-a742-04b9b0229ffa · outbound

This paper cites Semantic representations of near-synonyms for automatic lexical choice.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Semantic representations of near-synonyms for automatic lexical choice

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.143570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.143570Z digest=sha256:4ed4ba2884eda386c9e314619366f81cff19c312e2e3982048e1d6806c3f4c1f

Observation cbf2b730-d0cc-4065-94a7-b3923a72f6ee · outbound

This paper cites Word sense disambiguation with pictures.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Word sense disambiguation with pictures

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.148422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.148422Z digest=sha256:8fe8a62c5f369fa64924b3c60ebc3a2103bf9fe73a9451ba4664998d3a1bc87d

Observation e82c4bbf-ffc6-4448-a768-c78a6d2236c7 · outbound

This paper cites Text classification by augmenting the bag-of-words representation with redundancy-compensated bigrams.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Text classification by augmenting the bag-of-words representation with redundancy-compensated bigrams

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.153908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.153908Z digest=sha256:5eefd8207d8afbb499c1a588436356f517f1617285b7c179c430fc9c7519e7e3

Observation 6c2e3b0b-3789-4881-a5e8-4e2aef11c6c5 · outbound

This paper cites Addition–deletion networks.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Addition–deletion networks

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.158661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.158661Z digest=sha256:f033e7916d7c98a279822ffec35b71da4ea528896003716b911a1446913a86ad

Observation efb7ac88-e8ea-4fee-835d-16114cba5a36 · outbound

This paper cites Are good texts always better? interactions of text coherence, background knowledge, and levels of understanding in learning from text.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Are good texts always better? interactions of text coherence, background knowledge, and levels of understanding in learning from text

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.163568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.163568Z digest=sha256:b246cb5ca48ad733b320dc714bb92a0ae11a43ac9727385e399676e29b90d8ee

Observation 47aa5e3a-a54b-4d19-8362-55c6779ca1db · outbound

This paper cites Missing inaction: the dangers of ignoring missing data.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Missing inaction: the dangers of ignoring missing data

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.168609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.168609Z digest=sha256:3bb771578eebeea0f0dadc5fff09796664488d00ed9fbcbc8f8a741c82216ee0

Observation 0fe8770a-c501-4df3-90a4-29f580b48d33 · outbound

This paper cites Techniques for automatically correcting words in text.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Techniques for automatically correcting words in text

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.173447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.173447Z digest=sha256:9f42035f8f99b0e40c584ae0d5660adf35c56126b0d2a7f6f340b031b03b0a9d

Observation 51312975-2988-4c94-b3aa-4588c66283ea · outbound

This paper cites An augmented template-based approach to text realization.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey An augmented template-based approach to text realization

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.178231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.178231Z digest=sha256:c68944dc7887eadb1e1526a395da048b8e11b980a6fede6f78cc1535eabc83af

Observation 6a8205ec-b85d-44a5-9935-364ac6061e59 · outbound

This paper cites Scene text extraction and translation for handheld devices.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Scene text extraction and translation for handheld devices

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.183331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.183331Z digest=sha256:f1f8b4e2701dbeff62b72f395641e956b767b29ee09c36d8927d20ded9b4b7f9

Observation f51130ce-83e6-459e-b0fc-e0630a9dd286 · outbound

This paper cites Augmenting the power of lsi in text retrieval: Singular value rescaling.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Augmenting the power of lsi in text retrieval: Singular value rescaling

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.188607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.188607Z digest=sha256:39ec654f0f4ab16b659c23652b722cfac624d7e1bc68e14ad0895931ad010f2e

Observation 7f78fc6c-2303-47e7-8799-8b34c9d84090 · outbound

This paper cites The proper place of men and machines in language translation.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey The proper place of men and machines in language translation

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.194122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.194122Z digest=sha256:100da170c42b45640f077e1a25dcb6a3d1784f92772916eaa5b2adb8e80e5d97

Observation b29a1f95-5189-4d32-bc61-34e6b7ac6288 · outbound

This paper cites Embedding web-based statistical translation models in cross-language information retrieval.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Embedding web-based statistical translation models in cross-language information retrieval

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.199312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.199312Z digest=sha256:34d4c9e6fa2180116bd1a65e8b2ad471ce2be76054e1818698cdbea6643b4430

Observation f764078f-f7fd-4fb1-ba20-082ce830bde4 · outbound

This paper cites A technical word-and term-translation aid using noisy parallel corpora across language groups.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey A technical word-and term-translation aid using noisy parallel corpora across language groups

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.204671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.204671Z digest=sha256:01bb92c9f17f3a7d5d0a95f3f56e39abb216dddc85ca08007fd601bb2b2e6bdf

Observation caf2b31b-d28b-446a-b3f2-405c6ce696dc · outbound

This paper cites Iterative clustering of high dimensional text data augmented by local search.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Iterative clustering of high dimensional text data augmented by local search

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.210201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.210201Z digest=sha256:1a50fe5aceaaba18d8ac6b58e09d50f8b842affce24044df658a7f548f2b1e7e

Observation 1df9d0f5-65cb-41d9-9e85-cefb82e68153 · outbound

This paper cites Augmented audio reality: Telepresence/vr hybrid acoustic environments.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Augmented audio reality: Telepresence/vr hybrid acoustic environments

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.215357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.215357Z digest=sha256:6d460b15c91ab92ed880e6b6a2ba4c348f7b794cad7db651201227aeeecf42c4

Observation 42897cc0-5516-46e4-94ed-d98d133e5635 · outbound

This paper cites Vocal cord augmentation with autogenous fat.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Vocal cord augmentation with autogenous fat

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.220186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.220186Z digest=sha256:05ef5b4825845ea509c93300891629878783080d2023a0a773902fe5f738ce82

Observation 6b01ec0d-1268-4c79-8737-fb9ee07b3ad1 · outbound

This paper cites Audio and visually augmented teleconferencing.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Audio and visually augmented teleconferencing

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.225307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.225307Z digest=sha256:599acaf3165475a74f30c6fecc4c510fcf6194c8023f5d9d39c52f6a4732aebb

Observation ef72fd8c-b6d2-4a92-9412-4f156063c2f0 · outbound

This paper cites Ackerman, and Debby Hindus.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Ackerman, and Debby Hindus

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.230099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.230099Z digest=sha256:e1cc7ca07454626fddb14d8582bc082e495547cd7961e09c36bef82ab9b0acff

Observation d30e73a8-294e-48d7-a27e-2f431bf18d8e · outbound

This paper cites Can the lombard effect be used to improve low voice intensity in parkinson’s disease? European Journal of Disorders of Communication , 27(2):121–127, 1992.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Can the lombard effect be used to improve low voice intensity in parkinson’s disease? European Journal of Disorders of Communication , 27(2):121–127, 1992

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.235099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.235099Z digest=sha256:617f33a893fc5b71d7a96b24a4a398571e8029398688299bbe0232caa9a81c35

Observation 32902113-038e-4a72-b45f-6e7cbd72e8a6 · outbound

This paper cites Speech perception, localization, and lateralization with bilateral cochlear implants.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Speech perception, localization, and lateralization with bilateral cochlear implants

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.239690Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.239690Z digest=sha256:9c003d78e85bfa1d86fc080e0d4d374b95630ee2486daa9be1a5ed4a7eb8e5ae

Observation 7e7c8182-2a5d-41a7-baab-7d5730cb3611 · outbound

This paper cites Improving performance in noise for hearing aids and cochlear implants using coherent modulation filtering.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Improving performance in noise for hearing aids and cochlear implants using coherent modulation filtering

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.244655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.244655Z digest=sha256:3ef780577066abfbcde2e469fe512a3091d0478f30dd0f03206fe8490d314207

Observation 4d824749-df9e-4229-9065-ec4133a6bfbc · outbound

This paper cites Speech recognition for a humanoid with motor noise utilizing missing feature theory.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Speech recognition for a humanoid with motor noise utilizing missing feature theory

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.249717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.249717Z digest=sha256:981ecaa2b6d044455dd483880382abc4c262b22e33e2f18bb3b343de8227b171

Observation 80b8c125-affd-40cb-b290-e20e686ed734 · outbound

This paper cites Augmented reality audio for mobile and wearable appliances.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Augmented reality audio for mobile and wearable appliances

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.254610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.254610Z digest=sha256:cd64ed2102a3229db07e9dcbf6345304f33925a9d09536f315b218155c14900e

Observation 61b3996c-e4c7-4e28-bf19-f9ae3f4af0e4 · outbound

This paper cites Power supply noise in analog audio class d amplifiers.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Power supply noise in analog audio class d amplifiers

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.259505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.259505Z digest=sha256:0e87262c43374e5d01be127716681ad6ada0aa1a9196e6423c943a12fca8ee53

Observation b7b396fc-ee80-4798-a315-eb60b9c05f85 · outbound

This paper cites Deep convolutional neural network based medical image classification for disease diagnosis.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Deep convolutional neural network based medical image classification for disease diagnosis

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.264379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.264379Z digest=sha256:b5f2de149ec57742d6593de228ca72344691a75d712154ad8e4e2151a00bce7a

Observation 25e9b7e2-f5c9-4c1f-be9a-6c3398ab068f · outbound

This paper cites Going deep in medical image analysis: concepts, methods, challenges, and future directions.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Going deep in medical image analysis: concepts, methods, challenges, and future directions

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.269219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.269219Z digest=sha256:d1643333af6d940639dde9dc113694cc3b49f39da6bae18f6ad7738c2a345d5d

Observation ea5177f1-14c0-444e-98ed-48c19828a46d · outbound

This paper cites Adversarial differentiable data augmentation for autonomous systems.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Adversarial differentiable data augmentation for autonomous systems

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.274303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.274303Z digest=sha256:1f1e33af03909f2acbcdf2a6b3ac3ab0c571f01337acec4e29ec038635051000

Observation 11a311c2-d345-4bf3-a521-b44d9567cfa4 · outbound

This paper cites Data augmentation for deep learning based semantic segmentation and crop-weed classification in agricultural robotics.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Data augmentation for deep learning based semantic segmentation and crop-weed classification in agricultural robotics

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.279774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.279774Z digest=sha256:82703cad8924dee70b00b84732de86c1461adfd32932689efb8af7212b3f68fd

Observation f338ce1e-4f3a-4453-9a37-0b07c3278872 · outbound

This paper cites DAGAM: Data Augmentation with Generation And Modification.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey DAGAM: Data Augmentation with Generation And Modification

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.284791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.284791Z digest=sha256:52dd732e71b27d367ebdf836c49c6dde00168acc96038625d43de2a5c5eaa1ce

Observation 7ae1d42d-27a9-4f11-a1f6-dd6f5df6df43 · outbound

This paper cites Generative AI in Agriculture: Creating Image Datasets Using DALL.E's Advanced Large Language Model Capabilities.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Generative AI in Agriculture: Creating Image Datasets Using DALL.E's Advanced Large Language Model Capabilities

Reference 80

Resolution
verified exact
local_arxiv, observed 2026-08-10T04:36:40.551209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T04:36:37.290650Z digest=sha256:c5b43f56efe2635c5214bacd48c2e207752e71e6b814cb6a50366379a2c1c0d3

Observation 66cda6a2-0cf3-4af4-8763-b9000de2769f · outbound

This paper cites Synthetic data generation for tabular health records: A systematic review.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Synthetic data generation for tabular health records: A systematic review

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.296356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.296356Z digest=sha256:0ec3ce7f42b9a75a63af4e649fdbfa2ea9785bbd24a1da7409379a15623229bd

Observation ef86f27d-84d5-4c4d-bfe1-1b3f042ebc67 · outbound

This paper cites Data augmentation for medical imaging: A systematic literature review.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Data augmentation for medical imaging: A systematic literature review

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.301684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.301684Z digest=sha256:bf1dcdafcb248b912cdff02483627fd73b6db89d64f4cd3cf2b8e154c262fdcc

Observation c848dc83-e0e6-4821-8644-be026470ab82 · outbound

This paper cites Multi-modal llms in agriculture: A comprehensive review.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Multi-modal llms in agriculture: A comprehensive review

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.306903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.306903Z digest=sha256:80e909f6cdfd8bff17fe97528a54cf3009817b1c75b41f59f034976e7c946a67

Observation 1859460f-82db-41c6-bd72-383f27b9a4bf · outbound

This paper cites Research on data augmentation for image classification based on convolution neural networks.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Research on data augmentation for image classification based on convolution neural networks

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.312065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.312065Z digest=sha256:19d4ab394588b84b8ee673638e5012ae65a202f20b3d070a6be46e38bb7fd806

Observation 7cf4a101-8fd4-4295-8906-671ef30ae74d · outbound

This paper cites Data augmentation and generative machine learning on the cloud platform.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Data augmentation and generative machine learning on the cloud platform

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.317193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.317193Z digest=sha256:20232b0d87f2ac59801bc4b1b1d834543a51204616d20bc0fe9bc39ffe9bf210

Observation f99bdc6d-2f1f-4985-aa55-124ffe2e1be6 · outbound

This paper cites Data augmentation using deep generative models for embedding based speaker recognition.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Data augmentation using deep generative models for embedding based speaker recognition

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.322050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.322050Z digest=sha256:f937474fa4c5ce079726466bc2de447929913502b06016e7afdd074465359478

Observation 93ce9540-afb3-4093-bf4a-d7df71200239 · outbound

This paper cites Statistical Augmentation of a Chinese Machine-Readable Dictionary.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Statistical Augmentation of a Chinese Machine-Readable Dictionary

Reference 87

Resolution
verified exact
local_arxiv, observed 2026-08-10T04:36:40.524795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T04:36:37.326820Z digest=sha256:0a84e3c0ce7eb365fa8db7f5502d8c43ed0443db5b46933d84be474d9741063b

Observation 25744ac4-66a5-4abf-bf79-8ed70b2a9c03 · outbound

This paper cites A comparison of id3 and backpropagation for english text-to-speech mapping.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey A comparison of id3 and backpropagation for english text-to-speech mapping

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.332059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.332059Z digest=sha256:d89ef86906dc9083058a57551d959abcf29dcaffb2a6874596b8c848241146a9

Observation 5b0ba7af-0f83-4446-9faf-6b1d4923bf61 · outbound

This paper cites an unresolved cited work.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Unresolved cited work

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.336892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.336892Z digest=sha256:ff244196ec34d64661be27836949a6cf8140dccaea2b0bc23f80fed06186ae0d

Observation 5ef8e37c-f0ff-4371-8ae0-70c261709fb3 · outbound

This paper cites Srl-aco: A text augmentation framework based on semantic role labeling and ant colony optimization.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Srl-aco: A text augmentation framework based on semantic role labeling and ant colony optimization

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.341637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.341637Z digest=sha256:ba8f9e8460e16160badc9dae3558470d361fa308a1666779ff43f2e29ac87ccf

Observation 24e63926-3e4d-468b-a93e-1c35a2c781f9 · outbound

This paper cites From theories on styles to their transfer in text: Bridging the gap with a hierarchical survey.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey From theories on styles to their transfer in text: Bridging the gap with a hierarchical survey

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.346759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.346759Z digest=sha256:d37dc956128f06f56c98dbded4a9ce98ab7b26f6e01f189279ff4d3bf80685b5

Observation 8662a7b1-9908-44b4-b1d8-d66254b59665 · outbound

This paper cites Summary of chatgpt-related research and perspective towards the future of large language models.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Summary of chatgpt-related research and perspective towards the future of large language models

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.351795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.351795Z digest=sha256:0a843419ae415c1fb15e7d4256dbb6fe183d479e604322e3cc770f9994f71f5d

Observation 25ef9681-2fbe-4811-8d35-af8604763111 · outbound

This paper cites Survey on deep neural networks in speech and vision systems.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Survey on deep neural networks in speech and vision systems

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.357623Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.357623Z digest=sha256:2cf19a2b34e0636f71aa3425a3eab2c95b14b605c4f3a0c19cb9fe13a8916967

Observation 7fa32c4f-8dbe-4dab-a294-3ee124a44298 · outbound

This paper cites an unresolved cited work.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Unresolved cited work

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.363135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.363135Z digest=sha256:f20fb5e1c03a35133ed83dbf0ddd2a1be5c8f8e7d01a5e385980c6f5e3e79f72

Observation abd27d9c-19ff-481e-8596-2c3c65e3ebca · outbound

This paper cites Automated building damage assessment and large-scale mapping by integrating satellite imagery, gis, and deep learning.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Automated building damage assessment and large-scale mapping by integrating satellite imagery, gis, and deep learning

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.368046Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.368046Z digest=sha256:3a033af72aa6bc3f0c16a92e12537f6455d85e61959a244b129862d63a99d88f

Observation e117fe4c-cbf5-48d1-9a64-f181e685b8ac · outbound

This paper cites Machine learning for advanced emission monitoring and reduction strategies in fossil fuel power plants.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey Machine learning for advanced emission monitoring and reduction strategies in fossil fuel power plants

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.373656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.373656Z digest=sha256:8d1437fe66f0ae8a38a5a13f85b21c17a406bcca7ad9e3c427ea38430a691552

Observation 8a93cebc-465f-4130-8bcd-1fd05bad4371 · outbound

This paper cites A review on large language models: Architectures, applications, taxonomies, open issues and challenges.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey A review on large language models: Architectures, applications, taxonomies, open issues and challenges

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.378456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.378456Z digest=sha256:6e83c34a2e2ea0ca555f9c2f1bcec52bf240874086fc9e48167fe100335b221d

Observation c7d05b4c-179d-4d19-aedf-b3ae5ddff980 · outbound

This paper cites A comparison on data augmentation methods based on deep learning for audio classification.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey A comparison on data augmentation methods based on deep learning for audio classification

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.383521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.383521Z digest=sha256:2f4ff0f278d50c3caa04c68188947e36c53985960656aa136dbd32721ed1c9a1

Observation 6fcc3640-d25c-4f29-83d7-16ef18effa2c · outbound

This paper cites A study on data augmentation in voice anti-spoofing.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey A study on data augmentation in voice anti-spoofing

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.388689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.388689Z digest=sha256:a8adbcc601037d16d18c7e91147432b259d753461085c7ddcc5013d64c8b321b

Observation 052846a0-90c6-4626-aa5d-9913de8dbf8d · outbound

This paper cites On the analysis of data augmentation methods for spectral imaged based heart sound classification using convolutional neural networks.

Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey On the analysis of data augmentation methods for spectral imaged based heart sound classification using convolutional neural networks

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-10T04:36:37.393483Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:36:37.393483Z digest=sha256:b75648a0f8971e0de6689813e40866740e7fe1ea6ff34604a9a86ae704dc9673

Pith citing papers

Observation 763b5aaf-b75e-4ada-8887-4dfad38ffd6c · inbound

SafeTrans: LLM-assisted Transpilation from C to Rust cites this paper.

SafeTrans: LLM-assisted Transpilation from C to Rust Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-22T14:14:55.616435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T14:12:55.504993Z digest=sha256:921628ef0ae3d452177c00107a610f59f96e797a0bd36980ce37d067cfe8b1ea

Observation 80ddc295-914b-43f5-864b-888679f4acd2 · inbound

Breaking the Barriers of Text-Hungry and Audio-Deficient AI cites this paper.

Breaking the Barriers of Text-Hungry and Audio-Deficient AI Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey

Reference 112

Resolution
unresolved
no resolver link, observed 2026-08-07T11:28:56.015260Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:28:56.015260Z digest=sha256:b30baabb815f436f659dd5a42a2e2f216bd2685d40bf888497de3ff13e4c7932

Observation 73b99071-00fa-46f8-a614-09d889d6ee33 · inbound

ORBIT: Guided Agentic Orchestration for Autonomous C-to-Rust Transpilation cites this paper.

ORBIT: Guided Agentic Orchestration for Autonomous C-to-Rust Transpilation Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:46:05.073472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T15:19:09.020005Z digest=sha256:8b367d6b91b5e6ba688ea7efdc8032b1feb9d92751beb41610100150d744f87e