Pith. sign in

Paper Citation Record · LEDGER

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation

As of 12 August 2026, this Paper Citation Record lists 71 of 71 outbound references and 3 inbound Pith citation observations for arXiv:2501.04155.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.04155 v1

Coverage vector

measured 71 of 71 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T21:45:46.148987Z

measured 74 of 74 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-14T21:28:37.680681Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

71 of 71 outbound references displayed

  • verified exact0
  • verified fuzzy41
  • unresolved29
  • parse uncertain1
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation e0def824-6a1a-4cca-a2cd-8c401bd92ee2 · outbound

This paper cites an unresolved cited work.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:45:47.698235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.674255Z digest=sha256:cb26df362a9c98d493f1254e4f22005b10bae18d4eaca0d4ed6afab63867741c

Observation 16a9f92a-8eaf-46fc-a474-aa8f9bf4390d · outbound

This paper cites an unresolved cited work.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:45:47.679392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.679645Z digest=sha256:410cefc444c465b8dc7ddac660dbe9d7b154e9c62a2924298f5925e703bf1761

Observation fa811d3f-279e-405b-87ec-5b725f5caf2c · outbound

This paper cites Eureka: Evaluating and Understanding Large Foundation Models.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Eureka: Evaluating and Understanding Large Foundation Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:45.685534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:45.685534Z digest=sha256:1474af03dc3c173dadbeef78d9551b6e04508dd45d8a8246ba31e4826557b3c5

Observation 6b99cb32-86f6-4adc-b1f8-35f0b0144e5b · outbound

This paper cites Introduction to Natural Language Pro- cessing.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Introduction to Natural Language Pro- cessing

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.653822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.691150Z digest=sha256:2bad43315579c8d13d8503afb5a48a102b3a6613b3f60fe318bfd97fd82408d5

Observation f90a3c62-72a6-41df-bf1e-1a24d80661fe · outbound

This paper cites Conceptual 12m: Pushing web-scale image-text pre- training to recognize long-tail visual concepts.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Conceptual 12m: Pushing web-scale image-text pre- training to recognize long-tail visual concepts

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.636329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.703431Z digest=sha256:884bd53f9a93886dae13583a58ba9817ec113e7ab4fdfaa14d18b54a6f0df6fa

Observation d597aca8-aa56-4094-9aa8-b3306e2aed41 · outbound

This paper cites Sharegpt4v: Improving large multi-modal models with better captions,.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Sharegpt4v: Improving large multi-modal models with better captions,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.621454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.708945Z digest=sha256:2a9ad5a5037d48122c2fffecdab8ab182390d2f0bddc443394fd0cf767e8635d

Observation 2f826d38-c6e4-4022-a905-0f2c1dde5257 · outbound

This paper cites AlpaGasus: Training A Better Alpaca with Fewer Data.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation AlpaGasus: Training A Better Alpaca with Fewer Data

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:45.714805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:45.714805Z digest=sha256:44921e80241da95ac55376cebce83fa59edcb9d51532fa17415e588190839509

Observation f57deb5f-8094-49ef-aa8e-8c43bd611be7 · outbound

This paper cites Selection via Proxy: Efficient Data Selection for Deep Learning.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Selection via Proxy: Efficient Data Selection for Deep Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:45.721082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:45.721082Z digest=sha256:7c9cd7328a50fe982b5ef68185a7b0840ccf1ae2a744c6bbf17b6231952d204e

Observation 3f0ef15c-d78f-4b4b-8729-b705e4a8e8c2 · outbound

This paper cites A survey on in-context learning, 2024.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation A survey on in-context learning, 2024

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.597592Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.725412Z digest=sha256:ec35d43ea80b38227d74e5a56f5a2d71944c3d3f683de9373a01c1df8305b1dd

Observation b88e0666-15ca-46fa-9cea-85783dd9ac06 · outbound

This paper cites The Llama 3 Herd of Models.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation The Llama 3 Herd of Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:45.729541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:45.729541Z digest=sha256:f194e72620047f58dc25fc797c3a7604d72f5468562fca5e6093bce9809e545f

Observation a7f45de2-aae3-463c-b563-90662e2ae71b · outbound

This paper cites Tinystories: How small can language models be and still speak coherent english?, 2023.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Tinystories: How small can language models be and still speak coherent english?, 2023

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.579297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.733543Z digest=sha256:f1c340f415d58545053b06532795dbaa640ea64c5bbb204d91b4f6b5958a9aa9

Observation 91ce6dd0-283b-4568-9cc0-10540dec98d6 · outbound

This paper cites Data curation via joint example selection further accelerates multimodal learning.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Data curation via joint example selection further accelerates multimodal learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:45.737944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:45.737944Z digest=sha256:35b07985067c088d1b852add61e8648eb582867b63ea81cef436417360ce7141

Observation a6175318-33b7-43c8-89f7-70956a29f416 · outbound

This paper cites Improving clip training with language rewrites.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Improving clip training with language rewrites

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.538874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.743101Z digest=sha256:24bcb664699595340eace3c52642288345376b4c66969b9411e8d5da334e5a9e

Observation 91bceeee-3b9b-485b-ada7-45b350f679a6 · outbound

This paper cites Data fil- tering networks, 2023.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Data fil- tering networks, 2023

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.520855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.747389Z digest=sha256:b3b388a622e34119f5b53d3eafb3e67510ade8cd187028727712b7827ee7f3ba

Observation 480ff99d-ec12-41ec-8a86-7267c4fe9fef · outbound

This paper cites Blink: Multimodal large language models can see but not perceive.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Blink: Multimodal large language models can see but not perceive

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.492428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.752067Z digest=sha256:5212a6ad6af42f6517f28c1b4469549eece025f6224917a84d9f6f7a362ea98c

Observation d06e20d5-c06c-417c-9eb6-8ad535de0467 · outbound

This paper cites Dat- acomp: In search of the next generation of multimodal datasets.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Dat- acomp: In search of the next generation of multimodal datasets

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.475310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.757274Z digest=sha256:db705823e1f010dff91df78be2b10f64b2fe98c052e3c6271862e76a418844ea

Observation f3951993-f15d-49da-8fe6-8f24fb16161f · outbound

This paper cites Textbooks are all you need, 2023.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Textbooks are all you need, 2023

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.453986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.762665Z digest=sha256:ed8110b077af8130a80e160cec654028d28ef34f20c513a3588cf910d93059f7

Observation 9a90dad8-22f3-4995-9be3-23a89620a890 · outbound

This paper cites Statistical Methods for Speech Recogni- tion.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Statistical Methods for Speech Recogni- tion

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.433242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.766460Z digest=sha256:ac4466b318a7cfb75c2c5bb4f26af71e952282784d3ee871ca8680b78a9c8346

Observation 2d18b204-ead3-48a5-987f-7d9df402fbe2 · outbound

This paper cites Data-efficient contrastive self-supervised learning: Most beneficial exam- ples for supervised learning contribute the least.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Data-efficient contrastive self-supervised learning: Most beneficial exam- ples for supervised learning contribute the least

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.418100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.771697Z digest=sha256:2ca7681a8281fe436b0ad9b4f6b0c98b26e0e867eaf20f7503f6b573caa476cf

Observation 26c40066-ca40-4111-ba37-05090a8ab588 · outbound

This paper cites Data-efficient contrastive language-image pre- training: Prioritizing data quality over quantity, 2024.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Data-efficient contrastive language-image pre- training: Prioritizing data quality over quantity, 2024

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.399558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.906802Z digest=sha256:f41a5f177ce8653df070d838fc1be873aa52618687e5333d94da6fee253e64e0

Observation 9fc42298-ffae-4ff3-b920-8f8db743d20e · outbound

This paper cites What’s ”up” with vision-language models? investigating their strug- gle with spatial reasoning, 2023.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation What’s ”up” with vision-language models? investigating their strug- gle with spatial reasoning, 2023

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.383268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.911771Z digest=sha256:66bccee72c2b73256c4453ed89bf54f8ad1e152ffead7eee3ed3082e68e22a04

Observation d9d34882-3d13-45ae-91ff-e3fc602c97ed · outbound

This paper cites Not all sam- ples are created equal: Deep learning with importance sam- pling.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Not all sam- ples are created equal: Deep learning with importance sam- pling

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.360864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.915930Z digest=sha256:1e6937ba9da6b5e64f9c07d05eb2b53406b225ce09659d2916154609ee4f617f

Observation 80fbdd5b-f870-4798-9d68-55de9976f92d · outbound

This paper cites A diagram is worth a dozen images.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation A diagram is worth a dozen images

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:45.920097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:45.920097Z digest=sha256:eb0ae1824a08f4aa75ebace8db3cb5b9d0bcaf89d742434fb4edde4da3a16449

Observation 94ca5eeb-8653-4765-8e0f-abcd00236f8e · outbound

This paper cites Grad-match: Gradient matching based data subset selection for efficient deep model training.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Grad-match: Gradient matching based data subset selection for efficient deep model training

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.309657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.928292Z digest=sha256:1a8b53ad8fb4c579b921c65a79fcda6fc65c8d4b4763580f06c387039dfa279e

Observation cd7275d1-fe30-4ccc-adeb-5b68524de13b · outbound

This paper cites Revisit large-scale image-caption data in pre-training multi- modal foundation models, 2024.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Revisit large-scale image-caption data in pre-training multi- modal foundation models, 2024

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.290284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.932593Z digest=sha256:567dfb45e6b944e0f45e3c5d09be54405a1d80e44f9a7a5d43830470f5be4a15

Observation c6bc482d-75a7-4721-8036-7e202af722ea · outbound

This paper cites Veclip: Improving clip training via visual-enriched captions,.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Veclip: Improving clip training via visual-enriched captions,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:45.936558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:45.936558Z digest=sha256:91a65ec11f78c7b9609cc9c33a19eae2f069fcd2801bf396c0b8626553d72674

Observation 7dd9a7a5-e192-4923-9288-c98bfaceeb94 · outbound

This paper cites M$^3$IT: A Large-Scale Dataset towards Multi-Modal Multilingual Instruction Tuning.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation M$^3$IT: A Large-Scale Dataset towards Multi-Modal Multilingual Instruction Tuning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:45.940806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:45.940806Z digest=sha256:efbf1ca65cfe4123b3b71bc5080614645d5939dc1b0953db158fbdee15c6e603

Observation f6ccc2c0-cf5c-4081-b7b4-db4d30d78697 · outbound

This paper cites Textbooks are all you need ii: phi-1.5 technical report, 2023.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Textbooks are all you need ii: phi-1.5 technical report, 2023

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:45.945149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:45.945149Z digest=sha256:41c30554a945f94265d872ff82143a9a025c749510338190ef01b166f94d1ce3

Observation a9b8a560-b0a3-4fe4-bfe3-246c2a309289 · outbound

This paper cites Visual instruction tuning, 2023.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Visual instruction tuning, 2023

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.246927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.948900Z digest=sha256:228f0e30e4125cc8f132bea967097101d8a6add0b01acfa36a69e4ca8ca6e124

Observation 6d8ee901-ecfc-4de0-818e-394d550b92a5 · outbound

This paper cites T-MARS: Improving visual representations by circumventing text feature learning.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation T-MARS: Improving visual representations by circumventing text feature learning

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.232888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.952814Z digest=sha256:710d5d33171f08470a5a6325f05ffc33cc25fc02b6cf15108db33e533c66ef93

Observation 81b1b235-bad7-4ea9-9c48-ef461e92a2e8 · outbound

This paper cites When Less is More: Investigating Data Pruning for Pretraining LLMs at Scale.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation When Less is More: Investigating Data Pruning for Pretraining LLMs at Scale

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:45.956267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:45.956267Z digest=sha256:c627e08e5778e927e113b02d7d4c2beac6deae12b721266b1a2217b62af26dc2

Observation 97b13408-1408-497c-bbcd-e2ec1173d76f · outbound

This paper cites Joty, and Enamul Hoque.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Joty, and Enamul Hoque

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.217798Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.960156Z digest=sha256:b21d940a91a0fdc91ccbd23c7d7a88cf6a2759186ff43453b91fc3d2beb1e06e

Observation 585c7a07-f294-4e00-ba51-6e49d2fc65f0 · outbound

This paper cites Chartinstruct: Instruction tuning for chart comprehension and reasoning,.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Chartinstruct: Instruction tuning for chart comprehension and reasoning,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.184369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.969503Z digest=sha256:9c9e982c8929e756ca4f938efa645043cf68b0cd234f21081f13ccefbc64aaf6

Observation 5f923886-9885-455c-a8c6-9fe06caed7df · outbound

This paper cites Mmiu: Multimodal multi- image understanding for evaluating large vision-language models, 2024.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Mmiu: Multimodal multi- image understanding for evaluating large vision-language models, 2024

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.168981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.974012Z digest=sha256:2311ae4af6d42ff5ca8907b6abdba65f2011f15fff3ae205838dc46b2f00d5fb

Observation 375f069d-e6e4-4116-adcf-8aabd8afa165 · outbound

This paper cites Coresets for data-efficient training of machine learning mod- els.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Coresets for data-efficient training of machine learning mod- els

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.147986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.978496Z digest=sha256:c2b6f166d521d592d04b730d8f8c66dfeeaf2d8d4e4741f13e29decf31c6edad

Observation d6d30720-6c6f-4300-a80a-8500b3f2ce19 · outbound

This paper cites Orca 2: Teaching small language models how to reason, 2023.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Orca 2: Teaching small language models how to reason, 2023

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.128426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.982471Z digest=sha256:676a4974a3b76196c663d14f0a95410311c99dcaf585153c2474ff020eaf88a5

Observation 6772d19d-d32f-4bf5-a53f-9cf2e2f4baf9 · outbound

This paper cites Agentinstruct: Toward generative teaching with agentic flows, 2024.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Agentinstruct: Toward generative teaching with agentic flows, 2024

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.111690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.987055Z digest=sha256:f66aca1cd5f1cc2e14d840d37fc8324af085341aefb3d1a5cf54e20b4c3d388a

Observation 7d9775f9-a91c-48ae-94b0-4db88dd951a6 · outbound

This paper cites Orca: Progressive learning from complex explanation traces of gpt- 4, 2023.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Orca: Progressive learning from complex explanation traces of gpt- 4, 2023

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.086437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.991377Z digest=sha256:0f8ca2074562f0a4850a89f7311ec70a5f1d8fa1807956dbfd3302a2b6021f5f

Observation 966db625-5b0a-4672-ba1c-ccbe55accc66 · outbound

This paper cites Improving multimodal datasets with image captioning.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Improving multimodal datasets with image captioning

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.060679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.995650Z digest=sha256:c8e52da631a122a6e424c41d60302cb1b6a6768b7ed08e06318c4beef2b0d85a

Observation 0d1b369f-d3c2-455c-9aa6-90b60dcc1cdb · outbound

This paper cites GPT-4 Technical Report.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation GPT-4 Technical Report

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:46.000150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:46.000150Z digest=sha256:a091e0c7afee48dacdb259076a4eae12e7de200efc9fed1044e7c22feb06fd17

Observation b37b0661-1dbe-4c0b-a23b-623d3d5e449d · outbound

This paper cites Deep learning on a data diet: Finding important ex- amples early in training.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Deep learning on a data diet: Finding important ex- amples early in training

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:46.004308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:46.004308Z digest=sha256:11e439110d4c167833ea272b6079a2979946da74025212e7a32b8f8f90dcf30b

Observation 7441493e-9901-42cd-8f78-b9741ec5c742 · outbound

This paper cites Adaptive second order coresets for data-efficient machine learning.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Adaptive second order coresets for data-efficient machine learning

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.028134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:46.008764Z digest=sha256:b9e99a577125731cd68e65e0dc60c77bfef07ead27e732fe07f369009ebda83a

Observation 0c9f476e-2b9f-404c-bc01-aeda92c41599 · outbound

This paper cites Learning transferable visual models from natural language supervision, 2021.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Learning transferable visual models from natural language supervision, 2021

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:46.012853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:46.012853Z digest=sha256:579ce152a57ecfe6316c30ef15db06e0bd0e458126d74c0c501f0a241c7bae92

Observation e09306b9-615d-45fb-96bf-a95e914fc6e9 · outbound

This paper cites FuseCap: Leveraging Large Language Models for Enriched Fused Image Captions.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation FuseCap: Leveraging Large Language Models for Enriched Fused Image Captions

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:46.017427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:46.017427Z digest=sha256:1cc92064d3b7002be9677065cc3da91d6073c9fb176b6ffc1771a323db3feb70

Observation 1fe657d6-43b0-4aa6-98ff-41a1eadeca3e · outbound

This paper cites Is a caption worth a thousand im- ages? a study on representation learning.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Is a caption worth a thousand im- ages? a study on representation learning

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:46.992347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:46.022417Z digest=sha256:3db0e34ad7ed3b13d2a2138b83c68cd090ae098414b6c00e6769402754ed427a

Observation 39e4111d-ce49-4428-b350-7cedebdd518d · outbound

This paper cites Laion-5b: An open large-scale dataset for training next generation image-text models.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Laion-5b: An open large-scale dataset for training next generation image-text models

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:46.965511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:46.027100Z digest=sha256:89e39fa39cd37fe15424a307b55de95ee9b1891bd7067ad8f37c1bb0190aa334

Observation 36f9edc2-e3cb-4aa2-a14b-42e62cb9f5b1 · outbound

This paper cites Conceptual captions: A cleaned, hypernymed, im- age alt-text dataset for automatic image captioning.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Conceptual captions: A cleaned, hypernymed, im- age alt-text dataset for automatic image captioning

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:46.943332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:46.031188Z digest=sha256:87d9633eae48b7df925e6e8dca0e6a49f09387cdd934dde2f4ca48ac5aef5874

Observation d94aea82-d231-4228-bf57-7325fe8e7948 · outbound

This paper cites Math- llava: Bootstrapping mathematical reasoning for multimodal large language models, 2024.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Math- llava: Bootstrapping mathematical reasoning for multimodal large language models, 2024

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:46.915219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:46.035329Z digest=sha256:fe8a0e8515d752193dd57a6d4bc96022f7981bd02a0a51d937df977ee5e125d5

Observation f96b1b11-bae4-4952-b324-f9d0c1fd5603 · outbound

This paper cites Chatgpt-4 vision struggles with radiologic image interpretation.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Chatgpt-4 vision struggles with radiologic image interpretation

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:46.896538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:46.039698Z digest=sha256:29726b18adb917bc1cb0a0bd1d0c74655f035607908cafd3958b16963760cbd2

Observation 8d4114e7-e92e-4315-919c-15eae6eea1b7 · outbound

This paper cites Dataset Cartography: Mapping and Diagnosing Datasets with Training Dynamics.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Dataset Cartography: Mapping and Diagnosing Datasets with Training Dynamics

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:46.044678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:46.044678Z digest=sha256:a74c430c07b967a7b3b104d4ef3485cbb6348fdde9c52b137ba02376fc735e77

Observation f7827bac-5b2f-49c4-9ae3-25caf3d7bdb2 · outbound

This paper cites an unresolved cited work.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:45:46.872048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:46.049655Z digest=sha256:7df31646f5e279c422ff635c998b06db0336b828b59059b1673772abe0ee8d5b

Observation 780bbae5-20fd-4acc-98d1-5ea42e275b6a · outbound

This paper cites An Empirical Study of Example Forgetting during Deep Neural Network Learning.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation An Empirical Study of Example Forgetting during Deep Neural Network Learning

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:46.054263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:46.054263Z digest=sha256:04f908c5fcbe417e97219c5e4b4f2040debc27502240aab07a8b7dbfbc2ede25

Observation a1565f15-2808-4542-a8ff-29620c750f04 · outbound

This paper cites Dynamic data selection for efficient ssl via coarse-to-fine re- finement.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Dynamic data selection for efficient ssl via coarse-to-fine re- finement

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:46.841057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:46.058675Z digest=sha256:b9f6cc0876fdecd3ecfc1c6696dc2fbce5ecf8b897719ce8e5c2bdc06452a13b

Observation 352f8531-7729-42b3-ac28-45ac41f85937 · outbound

This paper cites Show and tell: Lessons learned from the 2015 mscoco image captioning challenge.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Show and tell: Lessons learned from the 2015 mscoco image captioning challenge

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:46.063180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:46.063180Z digest=sha256:f40dc9c8472c4059ef610772c01e7afccf58de2e8e4cf077d534795bb3589947

Observation 1715ed42-84bd-4452-97d9-33cd6d7d81a4 · outbound

This paper cites Is a picture worth a thou- sand words? delving into spatial reasoning for vision lan- guage models.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Is a picture worth a thou- sand words? delving into spatial reasoning for vision lan- guage models

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:46.818213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:46.068355Z digest=sha256:f79fba003bbcf0af8ab24ef4dc1573d667cd4edf7d3b79c069c0536d40ec0754

Observation 082548be-0f79-4cb3-b979-4b187a08643c · outbound

This paper cites Decoding Data Quality via Synthetic Corruptions: Embedding-guided Pruning of Code Data.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Decoding Data Quality via Synthetic Corruptions: Embedding-guided Pruning of Code Data

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:46.075086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:46.075086Z digest=sha256:bb4b77d789f1782ca01dcc2decae37fd46475fdf6cbae7445c036196126c0278

Observation 0728e6dd-9bae-4dee-8845-0769751b2e0b · outbound

This paper cites Capsfu- sion: Rethinking image-text data at scale, 2024.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Capsfu- sion: Rethinking image-text data at scale, 2024

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:46.803460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:46.081145Z digest=sha256:1d511efd5aefd97e0b27c3cbe083f2fd5fdf01a4b061db13ff5e8c47415692f2

Observation 4c2bfe0e-b3a2-465f-98a9-51acff9f3453 · outbound

This paper cites Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for ex- pert agi.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for ex- pert agi

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:46.788371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:46.086087Z digest=sha256:32a09055bf91e3218fb4ad0cece860be1d7fa1f232aa4aea773a7a661d1b4511

Observation 42f49a25-c3fb-495b-876c-e260f90e1e37 · outbound

This paper cites Multimodal self-instruct: Synthetic abstract image and visual reasoning instruction using language model, 2024.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Multimodal self-instruct: Synthetic abstract image and visual reasoning instruction using language model, 2024

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:46.771221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:46.091233Z digest=sha256:3b5409b5eac4d1dd064e308c9fcd0b59063af4e26f5d3c5eb593d5ce492bdc0c

Observation 338dfdf7-4491-485b-b8a4-aa512523bbea · outbound

This paper cites Lima: Less is more for alignment.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Lima: Less is more for alignment

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:46.755567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:46.096926Z digest=sha256:9f08c263dd3c11cdd083256294acd25f124732a958a863d7b9fbae43a4bc6131

Observation 0384b6f5-82c5-4c52-82d0-a9ce40c7e24e · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-10T21:45:46.102791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:45:46.102791Z digest=sha256:cc18b9bb67c56eed4d04ccdb768e29fbf5d478f0b81cc31793d7532129302602

Observation fca5074e-2878-4323-b264-106069c62193 · outbound

This paper cites an unresolved cited work.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Unresolved cited work

Reference 64

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:45:46.737796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:46.109989Z digest=sha256:0665df12a794786857e1bed7205e2ac9e6345660c07c64d850359f798d0a9d6c

Observation 7d031cdb-4b33-4c4b-9006-b195c8611540 · outbound

This paper cites an unresolved cited work.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Unresolved cited work

Reference 65

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:45:46.718986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:46.115378Z digest=sha256:5cb8d41b7c2dd4421555e0a23ac76060ed4c93a11442dd147c0afbd7debc3af5

Observation 3622dcb6-4e75-4d0a-a05b-a258c69a8ee3 · outbound

This paper cites an unresolved cited work.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Unresolved cited work

Reference 66

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:45:46.701823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:46.120785Z digest=sha256:21bcdc54f0bca284d3fc28035e34772b3186b2e32b442d12337d6e187afc0558

Observation 2a2cd15e-85ae-4343-99ba-77929c6c5cd2 · outbound

This paper cites an unresolved cited work.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Unresolved cited work

Reference 67

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:45:46.683180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:46.125900Z digest=sha256:aeda1e62b4db68adec0e5fc56f3bc25bbc489065cb47abca8780ffd531471814

Observation 59f8e91d-9f74-4057-80de-d888344960c2 · outbound

This paper cites Q": The generated question (include options if it’s multiple-choice). -.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Q": The generated question (include options if it’s multiple-choice). -

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:46.446901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:46.131769Z digest=sha256:386bbbaee0d90cc0eb0cc1962aace50e74480cb45ce7c3282b73e4444aece40e

Observation 842fba42-ce9e-4cd6-a8e7-53cf0c9ec6a4 · outbound

This paper cites an unresolved cited work.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Unresolved cited work

Reference 69

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:45:46.428751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:46.137605Z digest=sha256:c21844ddcc87503011acae4cd76ebb5a27436330a562545ad0227b3995c38a6c

Observation cbbbe3ac-33bd-45f5-9780-c2d150fa0176 · outbound

This paper cites an unresolved cited work.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Unresolved cited work

Reference 70

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:45:46.409958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:46.142195Z digest=sha256:fdb7456452f24904db478aaa8bd4a15a619c0ba2162806d718b08c31999b15d4

Observation 363f0474-8c9a-4ed0-a547-9ffa06d47600 · outbound

This paper cites an unresolved cited work.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Unresolved cited work

Reference 71

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:45:46.393032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:46.148987Z digest=sha256:8ab77bdc4e9c517d3aec237e8d701c9a27f9b82bda6a8a7a0361e717ecd386e1

Observation 26f1a52b-f90d-4bdc-b54b-ae663d6e4b1a · outbound

This paper cites an unresolved cited work.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation Unresolved cited work

Reference 251

Resolution
parse uncertain
raw_fallback, observed 2026-08-10T21:45:47.331313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.923926Z digest=sha256:ede6a749c3fb953b6c016cb1160cd824415e8071c597338212991577e269b35f

Observation bc4665e9-da03-44f8-9961-72f23475c1bb · outbound

This paper cites 3, 4, 5, 6.

MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation 3, 4, 5, 6

Reference 2279

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:45:47.201313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:45:45.964422Z digest=sha256:012bbdc93505c174f61f629f50745d4860046433068587e32edbeba237dc726d

Pith citing papers

Observation ac973bde-508f-42a9-8b61-26faac0367b8 · inbound

20/20 Vision Language Models: A Prescription for Better VLMs through Data Curation Alone cites this paper.

20/20 Vision Language Models: A Prescription for Better VLMs through Data Curation Alone MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-13T02:57:09.442144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-13T02:52:43.674969Z digest=sha256:0eef8ee584e185eb4813813ee2d9809b61328feb16d3680d1fe34b5b09310e97

Observation e8383d39-89ac-468f-83f8-6624a3821a07 · inbound

20/20 Vision Language Models: A Prescription for Better VLMs through Data Curation Alone cites this paper.

20/20 Vision Language Models: A Prescription for Better VLMs through Data Curation Alone MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:29:28.681673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-14T21:28:37.680681Z digest=sha256:4d5a847c88885b30ded77d626e66208654d36891a995dc843d15bb2f8125547c

Observation f93a6f5e-3930-4dc7-ad66-56309254a29c · inbound

Unlocking UML Class Diagram Understanding in Vision Language Models cites this paper.

Unlocking UML Class Diagram Understanding in Vision Language Models MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:42:02.618815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-13T01:39:50.027264Z digest=sha256:54f0ff2b37a13e3dc69056f07ef6fd15f3e5a75b8a391ef8f1443782ccacc306