Pith. sign in

Paper Citation Record · LEDGER

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation

As of 12 August 2026, this Paper Citation Record lists 71 of 71 outbound references and 3 inbound Pith citation observations for arXiv:2501.04155.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.04155 v1

Coverage vector

measured 71 of 71 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T21:45:46.148987Z

measured 74 of 74 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-14T21:28:37.680681Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

71 of 71 outbound references displayed

  • verified exact0
  • verified fuzzy41
  • unresolved29
  • parse uncertain1
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation e0def824-6a1a-4cca-a2cd-8c401bd92ee2 · outbound

This paper cites an unresolved cited work.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:45:47.698235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.674255Z digest=sha256:9d6820452dbdac11a315fee02ecc7014bf10846b31860ceb04d95ae68d4399ef

Observation 16a9f92a-8eaf-46fc-a474-aa8f9bf4390d · outbound

This paper cites an unresolved cited work.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:45:47.679392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.679645Z digest=sha256:dd0e1975e614e6073f9bbaffc8e6f714adcc337d5a1ff799dd9f3d5e6ed6b914

Observation fa811d3f-279e-405b-87ec-5b725f5caf2c · outbound

This paper cites Eureka: Evaluating and Understanding Large Foundation Models.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Eureka: Evaluating and Understanding Large Foundation Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:45.685534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:45.685534Z digest=sha256:59d79ffbf7f24629034bd0d32957dc6747b1a21384501d2b108c7a0d71c53f71

Observation 6b99cb32-86f6-4adc-b1f8-35f0b0144e5b · outbound

This paper cites Introduction to Natural Language Pro- cessing.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Introduction to Natural Language Pro- cessing

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.653822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.691150Z digest=sha256:7be3dee12a4b330f76d74dbf45338280fa6abf20448ebbb924015162ebb39e95

Observation f90a3c62-72a6-41df-bf1e-1a24d80661fe · outbound

This paper cites Conceptual 12m: Pushing web-scale image-text pre- training to recognize long-tail visual concepts.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Conceptual 12m: Pushing web-scale image-text pre- training to recognize long-tail visual concepts

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.636329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.703431Z digest=sha256:7bc272c29d07c7c4b843f2d066b8dcb55ee09fad4bb64658e08351f1e2979fa4

Observation d597aca8-aa56-4094-9aa8-b3306e2aed41 · outbound

This paper cites Sharegpt4v: Improving large multi-modal models with better captions,.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Sharegpt4v: Improving large multi-modal models with better captions,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.621454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.708945Z digest=sha256:8a3136a4910316f123c75013cbe623c69e64f3b7ce8a96582a1885143a3aea4d

Observation 2f826d38-c6e4-4022-a905-0f2c1dde5257 · outbound

This paper cites AlpaGasus: Training A Better Alpaca with Fewer Data.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation AlpaGasus: Training A Better Alpaca with Fewer Data

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:45.714805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:45.714805Z digest=sha256:f8c34286a15e046108e0617b81d7682aadb19cf92dec5c65f096cd6dcc73ca20

Observation f57deb5f-8094-49ef-aa8e-8c43bd611be7 · outbound

This paper cites Selection via Proxy: Efficient Data Selection for Deep Learning.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Selection via Proxy: Efficient Data Selection for Deep Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:45.721082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:45.721082Z digest=sha256:78d753abbccbb78ea6f3679794495a6da59b39ae458cb2d465105cd9349f0804

Observation 3f0ef15c-d78f-4b4b-8729-b705e4a8e8c2 · outbound

This paper cites A survey on in-context learning, 2024.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation A survey on in-context learning, 2024

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.597592Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.725412Z digest=sha256:664901d486cc5c1b9d1eeb04c1baca6beb45943bbceba8d1e93172e81a4f036d

Observation b88e0666-15ca-46fa-9cea-85783dd9ac06 · outbound

This paper cites The Llama 3 Herd of Models.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation The Llama 3 Herd of Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:45.729541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:45.729541Z digest=sha256:2eb8165752ae691c11effadc5e62267028635a31e4695e02f683f89328b00bba

Observation a7f45de2-aae3-463c-b563-90662e2ae71b · outbound

This paper cites Tinystories: How small can language models be and still speak coherent english?, 2023.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Tinystories: How small can language models be and still speak coherent english?, 2023

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.579297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.733543Z digest=sha256:bd3b476bd9e4b29ca665610de3eeaadd7c45852167ace204c1cfde94f4f5d97c

Observation 91ce6dd0-283b-4568-9cc0-10540dec98d6 · outbound

This paper cites Data curation via joint example selection further accelerates multimodal learning.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Data curation via joint example selection further accelerates multimodal learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:45.737944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:45.737944Z digest=sha256:cf3f5888fe0ad07abdde2b2452aece1392a565f4b1916187e22b89874ea09d6b

Observation a6175318-33b7-43c8-89f7-70956a29f416 · outbound

This paper cites Improving clip training with language rewrites.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Improving clip training with language rewrites

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.538874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.743101Z digest=sha256:536d58987dc2ee8b44fa304c67a1c304c39060b8ed14481c78b9e1c0c2ad70a1

Observation 91bceeee-3b9b-485b-ada7-45b350f679a6 · outbound

This paper cites Data fil- tering networks, 2023.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Data fil- tering networks, 2023

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.520855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.747389Z digest=sha256:843df0f4efa224e32467d7c4f7cdf18d699e4852780f3771b6c7c74f35bff13d

Observation 480ff99d-ec12-41ec-8a86-7267c4fe9fef · outbound

This paper cites Blink: Multimodal large language models can see but not perceive.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Blink: Multimodal large language models can see but not perceive

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.492428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.752067Z digest=sha256:e2be00bcb8c8275426d6e07003db943cd75146b1a82dc9827d5a529ad4e161db

Observation d06e20d5-c06c-417c-9eb6-8ad535de0467 · outbound

This paper cites Dat- acomp: In search of the next generation of multimodal datasets.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Dat- acomp: In search of the next generation of multimodal datasets

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.475310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.757274Z digest=sha256:cefaeb5531c65335d2d13d1309054a503cbd577465249ae741a04cf77312da45

Observation f3951993-f15d-49da-8fe6-8f24fb16161f · outbound

This paper cites Textbooks are all you need, 2023.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Textbooks are all you need, 2023

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.453986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.762665Z digest=sha256:95ad44baaf9af020d6a75070aca946fdc24deb44db5cb4935b38a8694e316150

Observation 9a90dad8-22f3-4995-9be3-23a89620a890 · outbound

This paper cites Statistical Methods for Speech Recogni- tion.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Statistical Methods for Speech Recogni- tion

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.433242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.766460Z digest=sha256:88fb2b0f65fad4bfab8a83d8364eb1bfdd93aa7f8323609b7be09a0e4df303a0

Observation 2d18b204-ead3-48a5-987f-7d9df402fbe2 · outbound

This paper cites Data-efficient contrastive self-supervised learning: Most beneficial exam- ples for supervised learning contribute the least.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Data-efficient contrastive self-supervised learning: Most beneficial exam- ples for supervised learning contribute the least

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.418100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.771697Z digest=sha256:730ecd6a599581f9f26695ee910520ce46eda2248b04c04b20841d258c1ff910

Observation 26c40066-ca40-4111-ba37-05090a8ab588 · outbound

This paper cites Data-efficient contrastive language-image pre- training: Prioritizing data quality over quantity, 2024.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Data-efficient contrastive language-image pre- training: Prioritizing data quality over quantity, 2024

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.399558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.906802Z digest=sha256:412dd562e3f7dbe75ef29d21e760e7b083b1147e3e579b54ddee2e9c68bbe367

Observation 9fc42298-ffae-4ff3-b920-8f8db743d20e · outbound

This paper cites What’s ”up” with vision-language models? investigating their strug- gle with spatial reasoning, 2023.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation What’s ”up” with vision-language models? investigating their strug- gle with spatial reasoning, 2023

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.383268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.911771Z digest=sha256:4356e9c6d8c456f739fd0ba26a09888babc6b4d8d0f1726ed12de9523635bb78

Observation d9d34882-3d13-45ae-91ff-e3fc602c97ed · outbound

This paper cites Not all sam- ples are created equal: Deep learning with importance sam- pling.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Not all sam- ples are created equal: Deep learning with importance sam- pling

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.360864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.915930Z digest=sha256:6844c785a09fd9778b4ffc75d8ba732075c0756424fa112d3edb30e2c2588637

Observation 80fbdd5b-f870-4798-9d68-55de9976f92d · outbound

This paper cites A diagram is worth a dozen images.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation A diagram is worth a dozen images

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:45.920097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:45.920097Z digest=sha256:2060d99da8c6f824501886295d4a3efec7848eedf308bbfc03889f103536a8f6

Observation 94ca5eeb-8653-4765-8e0f-abcd00236f8e · outbound

This paper cites Grad-match: Gradient matching based data subset selection for efficient deep model training.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Grad-match: Gradient matching based data subset selection for efficient deep model training

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.309657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.928292Z digest=sha256:17880bbedc60af79045e919e76be344df8baef79693bf147dbea37392e0a9f6d

Observation cd7275d1-fe30-4ccc-adeb-5b68524de13b · outbound

This paper cites Revisit large-scale image-caption data in pre-training multi- modal foundation models, 2024.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Revisit large-scale image-caption data in pre-training multi- modal foundation models, 2024

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.290284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.932593Z digest=sha256:08af76d9f134a5388ce87cf0b62b8486ebcd19a971b959ff9f669e2c401e99bf

Observation c6bc482d-75a7-4721-8036-7e202af722ea · outbound

This paper cites Veclip: Improving clip training via visual-enriched captions,.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Veclip: Improving clip training via visual-enriched captions,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:45.936558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:45.936558Z digest=sha256:24197135abf3f5a28a38c042fccc9cdcbe3456ba65d7b7b0ec67de40ae9c68da

Observation 7dd9a7a5-e192-4923-9288-c98bfaceeb94 · outbound

This paper cites M$^3$IT: A Large-Scale Dataset towards Multi-Modal Multilingual Instruction Tuning.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation M$^3$IT: A Large-Scale Dataset towards Multi-Modal Multilingual Instruction Tuning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:45.940806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:45.940806Z digest=sha256:835d69e4a668cbe0150fc320e80035218154b98c4ccd995bc882162025a549e1

Observation f6ccc2c0-cf5c-4081-b7b4-db4d30d78697 · outbound

This paper cites Textbooks are all you need ii: phi-1.5 technical report, 2023.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Textbooks are all you need ii: phi-1.5 technical report, 2023

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:45.945149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:45.945149Z digest=sha256:ab20c4a050d3ba71dd01a374af9286825a1256483e98096ee59698d0b5d6b726

Observation a9b8a560-b0a3-4fe4-bfe3-246c2a309289 · outbound

This paper cites Visual instruction tuning, 2023.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Visual instruction tuning, 2023

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.246927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.948900Z digest=sha256:854571dd381222afe64d5efee31d908efa3edabc3ca87181f7c49808e51bea03

Observation 6d8ee901-ecfc-4de0-818e-394d550b92a5 · outbound

This paper cites T-MARS: Improving visual representations by circumventing text feature learning.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation T-MARS: Improving visual representations by circumventing text feature learning

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.232888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.952814Z digest=sha256:03e530c36e217f0757bbc6a6f681bf1f6960360250fe0c531bb2fdd91192d726

Observation 81b1b235-bad7-4ea9-9c48-ef461e92a2e8 · outbound

This paper cites When Less is More: Investigating Data Pruning for Pretraining LLMs at Scale.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation When Less is More: Investigating Data Pruning for Pretraining LLMs at Scale

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:45.956267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:45.956267Z digest=sha256:712cd736b237c09b38908530ce05c55d2fe8354ebacb4758693aee43025385fb

Observation 97b13408-1408-497c-bbcd-e2ec1173d76f · outbound

This paper cites Joty, and Enamul Hoque.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Joty, and Enamul Hoque

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.217798Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.960156Z digest=sha256:1358f60538cf426bd4f2f7fff966338ebff530f0dda651c660371503d96742fd

Observation 585c7a07-f294-4e00-ba51-6e49d2fc65f0 · outbound

This paper cites Chartinstruct: Instruction tuning for chart comprehension and reasoning,.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Chartinstruct: Instruction tuning for chart comprehension and reasoning,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.184369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.969503Z digest=sha256:a07ae92f8610ba243555d4f1610afe74e0b9e8349bfdb5498d8a61050f719945

Observation 5f923886-9885-455c-a8c6-9fe06caed7df · outbound

This paper cites Mmiu: Multimodal multi- image understanding for evaluating large vision-language models, 2024.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Mmiu: Multimodal multi- image understanding for evaluating large vision-language models, 2024

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.168981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.974012Z digest=sha256:2b3cb64e54e4ba4e1ef59fa5ce7e5f69a302f7fbdedfb75a2b786c671a382712

Observation 375f069d-e6e4-4116-adcf-8aabd8afa165 · outbound

This paper cites Coresets for data-efficient training of machine learning mod- els.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Coresets for data-efficient training of machine learning mod- els

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.147986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.978496Z digest=sha256:142d9076513cd7aac408efa8a30fd349e137e4582e8b14ee6af16c0de3afd654

Observation d6d30720-6c6f-4300-a80a-8500b3f2ce19 · outbound

This paper cites Orca 2: Teaching small language models how to reason, 2023.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Orca 2: Teaching small language models how to reason, 2023

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.128426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.982471Z digest=sha256:e866b6a968a3666a52769a5d9788a395e4c2fc447ccd0851ad719e58ce425472

Observation 6772d19d-d32f-4bf5-a53f-9cf2e2f4baf9 · outbound

This paper cites Agentinstruct: Toward generative teaching with agentic flows, 2024.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Agentinstruct: Toward generative teaching with agentic flows, 2024

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.111690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.987055Z digest=sha256:95a854536bf5baafe23a6b36cd712dc8825e96106ce4a74fd914de2fc46e1d20

Observation 7d9775f9-a91c-48ae-94b0-4db88dd951a6 · outbound

This paper cites Orca: Progressive learning from complex explanation traces of gpt- 4, 2023.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Orca: Progressive learning from complex explanation traces of gpt- 4, 2023

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.086437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.991377Z digest=sha256:0585732632becf99ca09e9a14eb32d03d22093332a0c4bf270b4558da1d01cd7

Observation 966db625-5b0a-4672-ba1c-ccbe55accc66 · outbound

This paper cites Improving multimodal datasets with image captioning.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Improving multimodal datasets with image captioning

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.060679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.995650Z digest=sha256:87ded80c7e93eb6908edccbd3153e43e743a3f0af3728a4498790cdfa5affe6f

Observation 0d1b369f-d3c2-455c-9aa6-90b60dcc1cdb · outbound

This paper cites GPT-4 Technical Report.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation GPT-4 Technical Report

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:46.000150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:46.000150Z digest=sha256:f82d3c51b31f78cb8ddeafc31fd3dba1b2e1871886c568312b727ee39cd10c41

Observation b37b0661-1dbe-4c0b-a23b-623d3d5e449d · outbound

This paper cites Deep learning on a data diet: Finding important ex- amples early in training.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Deep learning on a data diet: Finding important ex- amples early in training

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:46.004308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:46.004308Z digest=sha256:a40d772e1d8ed2df390bd9e716e3429abc05fcdb6db96bb73548922edc993a0b

Observation 7441493e-9901-42cd-8f78-b9741ec5c742 · outbound

This paper cites Adaptive second order coresets for data-efficient machine learning.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Adaptive second order coresets for data-efficient machine learning

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.028134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:46.008764Z digest=sha256:21c4ef54dc2dda2f048a29abb4ccabbc9181952e29770f3e38279e3582c2a9f6

Observation 0c9f476e-2b9f-404c-bc01-aeda92c41599 · outbound

This paper cites Learning transferable visual models from natural language supervision, 2021.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Learning transferable visual models from natural language supervision, 2021

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:46.012853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:46.012853Z digest=sha256:522798b27010b7e88dd5787a3817099f4d026a6fb7a877d3d4983af961dd8fd0

Observation e09306b9-615d-45fb-96bf-a95e914fc6e9 · outbound

This paper cites FuseCap: Leveraging Large Language Models for Enriched Fused Image Captions.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation FuseCap: Leveraging Large Language Models for Enriched Fused Image Captions

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:46.017427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:46.017427Z digest=sha256:0f1a39eff85c0fd752585d712b11cd812446ea1a668c9f1ba4e8eed3a37ba929

Observation 1fe657d6-43b0-4aa6-98ff-41a1eadeca3e · outbound

This paper cites Is a caption worth a thousand im- ages? a study on representation learning.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Is a caption worth a thousand im- ages? a study on representation learning

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:46.992347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:46.022417Z digest=sha256:f6fba49eccadd7eaa474eccc3403764bd29b318be935bb2f46ff833f52787b9d

Observation 39e4111d-ce49-4428-b350-7cedebdd518d · outbound

This paper cites Laion-5b: An open large-scale dataset for training next generation image-text models.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Laion-5b: An open large-scale dataset for training next generation image-text models

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:46.965511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:46.027100Z digest=sha256:9b41f0430be4a90df968ac4ea1c96bdf468ac44a2c20259968654a2e2c83a4b8

Observation 36f9edc2-e3cb-4aa2-a14b-42e62cb9f5b1 · outbound

This paper cites Conceptual captions: A cleaned, hypernymed, im- age alt-text dataset for automatic image captioning.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Conceptual captions: A cleaned, hypernymed, im- age alt-text dataset for automatic image captioning

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:46.943332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:46.031188Z digest=sha256:0fb600fea22d7493e75f2d4c31e7ee11df6d76608c82451311282dad6bec70c7

Observation d94aea82-d231-4228-bf57-7325fe8e7948 · outbound

This paper cites Math- llava: Bootstrapping mathematical reasoning for multimodal large language models, 2024.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Math- llava: Bootstrapping mathematical reasoning for multimodal large language models, 2024

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:46.915219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:46.035329Z digest=sha256:458b8fcb92ccadd13928104eb2483c58421bfd11450bd9130b1ae15043e5e0ab

Observation f96b1b11-bae4-4952-b324-f9d0c1fd5603 · outbound

This paper cites Chatgpt-4 vision struggles with radiologic image interpretation.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Chatgpt-4 vision struggles with radiologic image interpretation

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:46.896538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:46.039698Z digest=sha256:f047001ccd5f9413192d246fbdeb1769e4183b8a4f19be8c3dd7e530ae7f9b22

Observation 8d4114e7-e92e-4315-919c-15eae6eea1b7 · outbound

This paper cites Dataset Cartography: Mapping and Diagnosing Datasets with Training Dynamics.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Dataset Cartography: Mapping and Diagnosing Datasets with Training Dynamics

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:46.044678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:46.044678Z digest=sha256:826298dd437725b10f67d4cdbb0b8138356392c38d4c50e98cd49e7f2ff3f26d

Observation f7827bac-5b2f-49c4-9ae3-25caf3d7bdb2 · outbound

This paper cites an unresolved cited work.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:45:46.872048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:46.049655Z digest=sha256:05ba2cb88a7a781f62de3905f991e047e3cd08ed7046595c09cb2754222ee5fe

Observation 780bbae5-20fd-4acc-98d1-5ea42e275b6a · outbound

This paper cites An Empirical Study of Example Forgetting during Deep Neural Network Learning.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation An Empirical Study of Example Forgetting during Deep Neural Network Learning

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:46.054263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:46.054263Z digest=sha256:43e4cc9d2232b7e2b61671373b2228332bc98b68a3e12cf5ad46442db8b92a63

Observation a1565f15-2808-4542-a8ff-29620c750f04 · outbound

This paper cites Dynamic data selection for efficient ssl via coarse-to-fine re- finement.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Dynamic data selection for efficient ssl via coarse-to-fine re- finement

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:46.841057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:46.058675Z digest=sha256:654004d88a627c2d41b5746b5eafc236456e18c28aadf438d7b4cbe8913c90e7

Observation 352f8531-7729-42b3-ac28-45ac41f85937 · outbound

This paper cites Show and tell: Lessons learned from the 2015 mscoco image captioning challenge.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Show and tell: Lessons learned from the 2015 mscoco image captioning challenge

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:46.063180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:46.063180Z digest=sha256:0f7fb5d836134c3d2a7ad559655b4e4a8181096f8487e8e07c1171719e766feb

Observation 1715ed42-84bd-4452-97d9-33cd6d7d81a4 · outbound

This paper cites Is a picture worth a thou- sand words? delving into spatial reasoning for vision lan- guage models.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Is a picture worth a thou- sand words? delving into spatial reasoning for vision lan- guage models

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:46.818213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:46.068355Z digest=sha256:f4851ab6580d2fe03da43c5bee1e49f3f1cf3717e61c0fce4a0389dd87aff97b

Observation 082548be-0f79-4cb3-b979-4b187a08643c · outbound

This paper cites Decoding Data Quality via Synthetic Corruptions: Embedding-guided Pruning of Code Data.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Decoding Data Quality via Synthetic Corruptions: Embedding-guided Pruning of Code Data

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:46.075086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:46.075086Z digest=sha256:53860b9dd94a8aa4fb521e2485e6cd6f72c3b2fdcbe8807e3b19ea0d6972b264

Observation 0728e6dd-9bae-4dee-8845-0769751b2e0b · outbound

This paper cites Capsfu- sion: Rethinking image-text data at scale, 2024.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Capsfu- sion: Rethinking image-text data at scale, 2024

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:46.803460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:46.081145Z digest=sha256:000b89a994b516ef61222575eb088cf03294b937a40e86d9e31f103ab6bb3a23

Observation 4c2bfe0e-b3a2-465f-98a9-51acff9f3453 · outbound

This paper cites Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for ex- pert agi.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for ex- pert agi

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:46.788371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:46.086087Z digest=sha256:bc270e1a56191a097d422b0ea998d80572a364cf67b6620f9301930239a17f69

Observation 42f49a25-c3fb-495b-876c-e260f90e1e37 · outbound

This paper cites Multimodal self-instruct: Synthetic abstract image and visual reasoning instruction using language model, 2024.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Multimodal self-instruct: Synthetic abstract image and visual reasoning instruction using language model, 2024

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:46.771221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:46.091233Z digest=sha256:f4b307723e52bc506866cd28d317b37b1507b18d070815274798f1595d60b582

Observation 338dfdf7-4491-485b-b8a4-aa512523bbea · outbound

This paper cites Lima: Less is more for alignment.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Lima: Less is more for alignment

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:46.755567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:46.096926Z digest=sha256:50c8a65be920f8cb125130b24e95a048a6382dc2400e2efa5dccef5fbc3150aa

Observation 0384b6f5-82c5-4c52-82d0-a9ce40c7e24e · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:46.102791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:46.102791Z digest=sha256:c1ec0c3d1084aedf4e5b6f3b950bfe940483d8ed63276096216bd3bae65902e7

Observation fca5074e-2878-4323-b264-106069c62193 · outbound

This paper cites an unresolved cited work.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Unresolved cited work

Reference 64

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:45:46.737796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:46.109989Z digest=sha256:2cfc5612da3b20d37d7a043e02e403d30bb1fe298f66aeac1059bba02f089688

Observation 7d031cdb-4b33-4c4b-9006-b195c8611540 · outbound

This paper cites an unresolved cited work.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Unresolved cited work

Reference 65

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:45:46.718986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:46.115378Z digest=sha256:da3b7f5c9dd0c4a19c27efa6cb93f3b271bccb9b0701c5bf3525a7d3509be3bf

Observation 3622dcb6-4e75-4d0a-a05b-a258c69a8ee3 · outbound

This paper cites an unresolved cited work.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Unresolved cited work

Reference 66

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:45:46.701823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:46.120785Z digest=sha256:bf86dc16ed6521ce2590747879bf1032bb06c2eeba7064272dba80fb7d527f4e

Observation 2a2cd15e-85ae-4343-99ba-77929c6c5cd2 · outbound

This paper cites an unresolved cited work.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Unresolved cited work

Reference 67

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:45:46.683180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:46.125900Z digest=sha256:57ecf9b398d54e3ca93b9040285e13ab71d4b2f3842c56ae4954441ad38c4a14

Observation 59f8e91d-9f74-4057-80de-d888344960c2 · outbound

This paper cites Q": The generated question (include options if it’s multiple-choice). -.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Q": The generated question (include options if it’s multiple-choice). -

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:46.446901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:46.131769Z digest=sha256:996266673ec9836bef7e1fc05d2767ea2d5148af27fcf1b7e7dbb611048bc65b

Observation 842fba42-ce9e-4cd6-a8e7-53cf0c9ec6a4 · outbound

This paper cites an unresolved cited work.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Unresolved cited work

Reference 69

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:45:46.428751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:46.137605Z digest=sha256:691ab23964e641fe8628ccfbe975a40c61a85334b8e5076160243be3b208b633

Observation cbbbe3ac-33bd-45f5-9780-c2d150fa0176 · outbound

This paper cites an unresolved cited work.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Unresolved cited work

Reference 70

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:45:46.409958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:46.142195Z digest=sha256:0931dbd9d9582208ff18c3884b87b97e084790f8a63966dd8bd6d517d87c1c3e

Observation 363f0474-8c9a-4ed0-a547-9ffa06d47600 · outbound

This paper cites an unresolved cited work.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Unresolved cited work

Reference 71

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:45:46.393032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:46.148987Z digest=sha256:295a62224a6f724fe8a297a4b61da950e2f58e0d40cdddf65ba64e542d98176b

Observation 26f1a52b-f90d-4bdc-b54b-ae663d6e4b1a · outbound

This paper cites an unresolved cited work.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Unresolved cited work

Reference 251

Resolution
parse uncertain
raw_fallback, observed 2026-08-10T21:45:47.331313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.923926Z digest=sha256:f0443e0bb0983da31ee13ed91ca3f6349cab318bdd9236ee3afea7981065f91c

Observation bc4665e9-da03-44f8-9961-72f23475c1bb · outbound

This paper cites 3, 4, 5, 6.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation 3, 4, 5, 6

Reference 2279

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.201313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.964422Z digest=sha256:b15b257d56a630b09a6ccf0e28a234969e8bffb9e1261f95f20d04cd3ef4dcac

Pith citing papers

Observation ac973bde-508f-42a9-8b61-26faac0367b8 · inbound

20/20 Vision Language Models: A Prescription for Better VLMs through Data Curation Alone cites this paper.

20/20 Vision Language Models: A Prescription for Better VLMs through Data Curation Alone MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-13T02:57:09.442144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-13T02:52:43.674969Z digest=sha256:73b8c99fa7a323daad96e32e23c0228798455f2e0e3ceb31f34c0ec128b95cab

Observation e8383d39-89ac-468f-83f8-6624a3821a07 · inbound

20/20 Vision Language Models: A Prescription for Better VLMs through Data Curation Alone cites this paper.

20/20 Vision Language Models: A Prescription for Better VLMs through Data Curation Alone MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:29:28.681673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-14T21:28:37.680681Z digest=sha256:3123eacf2b8d55a3c50b3842b85fb573b1bfe0871d8300b96ad3c27e6127f62f

Observation f93a6f5e-3930-4dc7-ad66-56309254a29c · inbound

Unlocking UML Class Diagram Understanding in Vision Language Models cites this paper.

Unlocking UML Class Diagram Understanding in Vision Language Models MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:42:02.618815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-13T01:39:50.027264Z digest=sha256:52f56f8f43eef7334d010f6953dc8ade35034d9d1963b874e98a679dc6537f1e