Pith. sign in

Paper Citation Record · LEDGER

Bridging the Data Provenance Gap Across Text, Speech and Video

As of 18 August 2026, this Paper Citation Record lists 100 of 294 outbound references and 4 inbound Pith citation observations for arXiv:2412.17847.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.17847 v2

Coverage vector

measured 100 of 294 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T12:18:39.408235Z

measured 104 of 104 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T05:22:55.579720Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T10:29:46.922814Z

Reference resolution

100 of 294 outbound references displayed

  • verified exact3
  • verified fuzzy0
  • unresolved95
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4d2e697d-9363-4731-9070-bf871ededd9e · outbound

This paper cites Untitled review,.

Bridging the Data Provenance Gap Across Text, Speech and Video Untitled review,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:37.126845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:37.126845Z digest=sha256:17e5839c24e32e23305d1afd3f5411d07c9185f88130202ac2e4512fd7eda3d9

Observation 35b7cf81-d240-4fbd-ada6-a818d742b7b9 · outbound

This paper cites On the measurement of inequality,.

Bridging the Data Provenance Gap Across Text, Speech and Video On the measurement of inequality,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:37.134025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:37.134025Z digest=sha256:b65b52b37ea6a5846854c04610dc113623c65f3be7f929a022d6b1bd2e28019d

Observation e2282deb-125e-45b2-882f-16665140cebb · outbound

This paper cites A survey of video datasets for human action and activity recognition,.

Bridging the Data Provenance Gap Across Text, Speech and Video A survey of video datasets for human action and activity recognition,

Reference 3

Resolution
verified exact
doi, observed 2026-08-11T12:18:43.954046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:18:37.139721Z digest=sha256:8676ecb315511e07c26f2d321b994ae76c65485a697a3cf0c5dfff2978301410

Observation 451848a5-994d-49c4-9fff-d7657153b0eb · outbound

This paper cites No Classification without Representation: Assessing Geodiversity Issues in Open Data Sets for the Developing World.

Bridging the Data Provenance Gap Across Text, Speech and Video No Classification without Representation: Assessing Geodiversity Issues in Open Data Sets for the Developing World

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:37.431142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:37.431142Z digest=sha256:a00a65bc07c5e7dd0a7dd47d4a1c3e4cc165891b18179ad65fd0b4aee0d11802

Observation e33bdddd-86c1-47be-b88e-69b800095733 · outbound

This paper cites Playing hard exploration games by watching youtube,.

Bridging the Data Provenance Gap Across Text, Speech and Video Playing hard exploration games by watching youtube,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:37.435498Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:37.435498Z digest=sha256:38f6cf83912628241882196d94302d11feda0cf957d349d65b146b38f7c3123e

Observation bb59c224-4697-4f53-8ab6-5172831b101f · outbound

This paper cites Data statements for natural language processing: Toward mitigating system bias and enabling better science,.

Bridging the Data Provenance Gap Across Text, Speech and Video Data statements for natural language processing: Toward mitigating system bias and enabling better science,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:37.439558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:37.439558Z digest=sha256:5d162f853dd1e78e737b52e2071bb7008c549e80ab75d6dbaecae91d3a175ed0

Observation ffe633ed-dd3b-45da-b569-e1fad83c52ea · outbound

This paper cites Gender shades: Intersectional accuracy disparities in commer- cial gender classification,.

Bridging the Data Provenance Gap Across Text, Speech and Video Gender shades: Intersectional accuracy disparities in commer- cial gender classification,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:37.449771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:37.449771Z digest=sha256:d736ae747661b8408b4b27ad5f62a6239262c0c5bae891f9226e59f595a802ff

Observation 4bad7737-d010-4bdb-8af7-b4f872455d9b · outbound

This paper cites Common Voice: A Massively-Multilingual Speech Corpus.

Bridging the Data Provenance Gap Across Text, Speech and Video Common Voice: A Massively-Multilingual Speech Corpus

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:37.455322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:37.455322Z digest=sha256:0f62ee6ad605cef7ed7d7a9b2b84776f68f022db4159a5890bd1f120bb412ab9

Observation 1c45452e-f139-45a0-bf04-d93d8c0f3bd4 · outbound

This paper cites Does object recognition work for everyone?.

Bridging the Data Provenance Gap Across Text, Speech and Video Does object recognition work for everyone?

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:37.460876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:37.460876Z digest=sha256:39dffe9bf1d47fc1c5a9c6fb558ff8546e8dc7a250933ff7bb538c30247f76ed

Observation 5ff71237-162e-436d-9b39-b03f1d114e73 · outbound

This paper cites Visual to text: Survey of image and video captioning,.

Bridging the Data Provenance Gap Across Text, Speech and Video Visual to text: Survey of image and video captioning,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:37.465688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:37.465688Z digest=sha256:7d3e8ef2512f6a257690f5c82a5738a93ed991930d1bdd49281407206b34ee78

Observation f35a663c-9573-485b-9e10-9abecab58c56 · outbound

This paper cites Mundane content on social media: Creation, circulation, and the copyright problem,.

Bridging the Data Provenance Gap Across Text, Speech and Video Mundane content on social media: Creation, circulation, and the copyright problem,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:37.475425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:37.475425Z digest=sha256:99c35d21c46ce10b3c5c29c41fd0ecb28ad52808bafcce73d72e25820a17f594

Observation 1fc08473-de3a-4765-ac48-f08e88e3bc29 · outbound

This paper cites Model cards for model reporting,.

Bridging the Data Provenance Gap Across Text, Speech and Video Model cards for model reporting,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:37.480576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:37.480576Z digest=sha256:9252663db424d025f21eceabd9920ddb9039f444c40496757a9ae0dfd5c364f9

Observation 344370fa-0a08-4f67-b9e7-9e12a29c9f31 · outbound

This paper cites Moments in time dataset: One million videos for event understanding,.

Bridging the Data Provenance Gap Across Text, Speech and Video Moments in time dataset: One million videos for event understanding,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:37.485834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:37.485834Z digest=sha256:57417a01ca3fc48f01d3371ec7fada8d705d5a40cb43534fe5aed732c8276ac9

Observation 1e5c51d5-f155-4514-8a01-f5e99c5516ff · outbound

This paper cites Social data: Biases, methodological pitfalls, and ethical boundaries,.

Bridging the Data Provenance Gap Across Text, Speech and Video Social data: Biases, methodological pitfalls, and ethical boundaries,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:37.490744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:37.490744Z digest=sha256:7aa8b6f7e50f32675010af2b21912eda9999487f38a9ea19c92e2411b0824e64

Observation 3946237d-fc0f-40da-b651-d5dfc4a3a7a2 · outbound

This paper cites Common voice: A massively-multilingual speech corpus,.

Bridging the Data Provenance Gap Across Text, Speech and Video Common voice: A massively-multilingual speech corpus,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:37.495107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:37.495107Z digest=sha256:8dbeaa7d077d015167e12451666b74ed09e35a3f5ae5649051ca238d88b03f04

Observation b88d587a-ed5b-44b9-ba0a-88b339426238 · outbound

This paper cites Language models are few-shot learners,.

Bridging the Data Provenance Gap Across Text, Speech and Video Language models are few-shot learners,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:37.499991Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:37.499991Z digest=sha256:29b723e3a0a8134628ba4e11c0771440c689a827d188f75a613fdb820ab4ec6a

Observation a6dac15b-1494-4175-8e2c-69d289565a6f · outbound

This paper cites Semantic visual navigation by watching youtube videos,.

Bridging the Data Provenance Gap Across Text, Speech and Video Semantic visual navigation by watching youtube videos,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:37.504919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:37.504919Z digest=sha256:5e5fde9ce1a7ccfe394aabed0a89f6e85c623e02b5d256f7c487deaf13e861d4

Observation ce6f8933-2302-4e48-b68f-0f587ae271c6 · outbound

This paper cites The Pile: An 800GB Dataset of Diverse Text for Language Modeling.

Bridging the Data Provenance Gap Across Text, Speech and Video The Pile: An 800GB Dataset of Diverse Text for Language Modeling

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:37.680509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:37.680509Z digest=sha256:ca88bde341989f46a0ceef205b5719b05439e16e281b5fe151e5db7afe1694b1

Observation c30bb912-7912-4b60-b6fe-6fba66ac48df · outbound

This paper cites Henighan, J.

Bridging the Data Provenance Gap Across Text, Speech and Video Henighan, J

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:37.779146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:37.779146Z digest=sha256:4a4bea2c7302d9405ba564e40c9f5af8f357ae86d6f0bf143bbef50276c04998

Observation 0c752b47-712b-423e-b07d-ea61403fc7af · outbound

This paper cites The State and Fate of Linguistic Diversity and Inclusion in the NLP World.

Bridging the Data Provenance Gap Across Text, Speech and Video The State and Fate of Linguistic Diversity and Inclusion in the NLP World

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:37.791548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:37.791548Z digest=sha256:655c6356388cb89e90c6972636f1091841a81ea4bd76ef15ea8b6b27734f2920

Observation 74dbc32e-d3be-4898-bfd6-fa3f21ed6852 · outbound

This paper cites Scaling Laws for Neural Language Models.

Bridging the Data Provenance Gap Across Text, Speech and Video Scaling Laws for Neural Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:37.797181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:37.797181Z digest=sha256:da5100bd4763239c03d9fc69c5191c5d750509bd3296dfcb06b43828c5f79a50

Observation 5d6e7218-edcc-455c-891e-f25bd84cb20a · outbound

This paper cites AI4Bharat-IndicNLP Corpus: Monolingual Corpora and Word Embeddings for Indic Languages.

Bridging the Data Provenance Gap Across Text, Speech and Video AI4Bharat-IndicNLP Corpus: Monolingual Corpora and Word Embeddings for Indic Languages

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:37.801569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:37.801569Z digest=sha256:6ceb3771d90e502279ce4a563338a6c3e3b6ffa996dc18f3fd4e167f0b23c91e

Observation b6405bc6-1edc-4336-8eb5-d9bdd853938c · outbound

This paper cites Beyond “i agree.

Bridging the Data Provenance Gap Across Text, Speech and Video Beyond “i agree

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:37.806778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:37.806778Z digest=sha256:b2e9f60c05ffb2861256f5ba9ca3489c49d61473e12199a2ad5adf4e5b5d36e0

Observation f8998262-70d3-47b0-8c23-cd235167b500 · outbound

This paper cites The new legal landscape for text mining and machine learning,.

Bridging the Data Provenance Gap Across Text, Speech and Video The new legal landscape for text mining and machine learning,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:37.811826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:37.811826Z digest=sha256:91779651f985f2ccc7aa54230818f35c50834d8bead5b078c18a7aea09e9507a

Observation 2d69d7a3-54c0-452e-a900-87ece0af7b7f · outbound

This paper cites A Comprehensive Study of Deep Video Action Recognition.

Bridging the Data Provenance Gap Across Text, Speech and Video A Comprehensive Study of Deep Video Action Recognition

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:37.818453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:37.818453Z digest=sha256:c9c07299e09a4b99be357327f7d244ca73865d601095749840131c81b9678067

Observation f66b7838-838b-40fe-a4b6-321720bcc393 · outbound

This paper cites Masakhaner: Named entity recognition for african languages,.

Bridging the Data Provenance Gap Across Text, Speech and Video Masakhaner: Named entity recognition for african languages,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:37.823677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:37.823677Z digest=sha256:7c9ee5e71c338e0bac74aa97c78c44b6b8686e151ee2e5295493ff294e9e8fc2

Observation 893e0746-134a-4274-b4b5-0297d839abf1 · outbound

This paper cites How might we create better benchmarks for speech recognition?.

Bridging the Data Provenance Gap Across Text, Speech and Video How might we create better benchmarks for speech recognition?

Reference 32

Resolution
verified exact
doi, observed 2026-08-11T12:18:43.839640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:18:37.828242Z digest=sha256:55aaddc6937d1d675b5b070e53b35c1342a30aeab85c4d1981dd2b93f96b4bcc

Observation dc9a17e6-7fc7-4c7d-a0bd-87032b8376a0 · outbound

This paper cites XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale.

Bridging the Data Provenance Gap Across Text, Speech and Video XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:37.833298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:37.833298Z digest=sha256:0994f77c1405a1c342013b322e901bc9dc539c9447da87de269d22e460b7aebf

Observation b110d844-2909-4230-986f-87078a386f56 · outbound

This paper cites Addressing "Documentation Debt" in Machine Learning Research: A Retrospective Datasheet for BookCorpus.

Bridging the Data Provenance Gap Across Text, Speech and Video Addressing "Documentation Debt" in Machine Learning Research: A Retrospective Datasheet for BookCorpus

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:37.869550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:37.869550Z digest=sha256:19eb27038f28e1b7ef1b3494ec5c5679dd05294f0b8c4d40f517f1b0ebe88a65

Observation 12001a36-45c2-481e-ba93-b4a44767bdbd · outbound

This paper cites Multimodal datasets: misogyny, pornography, and malignant stereotypes.

Bridging the Data Provenance Gap Across Text, Speech and Video Multimodal datasets: misogyny, pornography, and malignant stereotypes

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:37.938469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:37.938469Z digest=sha256:e62e78c55bdce65c499be6d2c4c8c1acb7c11f8af74bc348781b4bbf7f50d24c

Observation 37cc4280-6d2c-41bd-b2a2-2ca514b35e48 · outbound

This paper cites Quality at a Glance: An Audit of Web-Crawled Multilingual Datasets.

Bridging the Data Provenance Gap Across Text, Speech and Video Quality at a Glance: An Audit of Web-Crawled Multilingual Datasets

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.016129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.016129Z digest=sha256:99652635e52d031132f27364f0e2684101e2e7ade6da3e43a318c9f4f26303ed

Observation f7eaaa3e-eb0a-40ca-8030-aea1d999d7f2 · outbound

This paper cites Documenting large webtext corpora: A case study on the colossal clean crawled corpus,.

Bridging the Data Provenance Gap Across Text, Speech and Video Documenting large webtext corpora: A case study on the colossal clean crawled corpus,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.021819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.021819Z digest=sha256:bf7530b0bc753c6afe7d6dd77186a34287ddd3da91ecfadb840420e773a78b4b

Observation b194b82b-4595-4f52-8e28-41de7c5e4670 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Bridging the Data Provenance Gap Across Text, Speech and Video An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.026230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.026230Z digest=sha256:ae0c1d7b363bb2ef696af92b3e33f058cfa6ee147e9f92e5e6520d9cf1a2ac99

Observation cedb07ac-4798-4960-8d60-e70b1e5b5d63 · outbound

This paper cites Datasheets for datasets,.

Bridging the Data Provenance Gap Across Text, Speech and Video Datasheets for datasets,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.031666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.031666Z digest=sha256:cb1f0c09d02f083a745820c5dccbfd9be8cef697e4f512aa82e29a71d53bc396

Observation 5a7a7463-4d51-46a6-b8b1-c9a9dd798a04 · outbound

This paper cites What's in the Box? A Preliminary Analysis of Undesirable Content in the Common Crawl Corpus.

Bridging the Data Provenance Gap Across Text, Speech and Video What's in the Box? A Preliminary Analysis of Undesirable Content in the Common Crawl Corpus

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.036804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.036804Z digest=sha256:2371919c967dafe5cea2a13ba0729a99d6a7022b4bfee1e3c40507ccce59f76b

Observation 8be59f2d-bd70-4abc-b42d-9b35b1ab51e4 · outbound

This paper cites Understanding Gender and Racial Disparities in Image Recognition Models.

Bridging the Data Provenance Gap Across Text, Speech and Video Understanding Gender and Racial Disparities in Image Recognition Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.041579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.041579Z digest=sha256:384d2e9343ffbb3eaf1497625a34cfb0cc306e11784c498b1ced19cb97a85bf2

Observation 7370843e-9c09-490f-b391-7a6650342d81 · outbound

This paper cites Automatic speech recognition: A survey,.

Bridging the Data Provenance Gap Across Text, Speech and Video Automatic speech recognition: A survey,

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.047296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.047296Z digest=sha256:50157387b17682537ab7b6ad793f0f4351a1e2a3786b55fb876f27a8bd547f9e

Observation 037abbb7-10d4-4be6-a9f4-41c338d56874 · outbound

This paper cites Spoken Moments: Learning Joint Audio-Visual Representations from Video Descriptions.

Bridging the Data Provenance Gap Across Text, Speech and Video Spoken Moments: Learning Joint Audio-Visual Representations from Video Descriptions

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-08-11T12:18:43.759925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-11T12:18:38.051968Z digest=sha256:04a5b66944d1001ae19644b14b3e3fc4b97909ffa21d91d86285362536c46152

Observation 6acb6b60-e698-4cc1-aa47-1a6b478bff4b · outbound

This paper cites Data and its (dis) contents: A survey of dataset development and use in machine learning research,.

Bridging the Data Provenance Gap Across Text, Speech and Video Data and its (dis) contents: A survey of dataset development and use in machine learning research,

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.057796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.057796Z digest=sha256:7eb08835ee7fed43c58c6443f3c697d43b9e2483895588c54597650e4136bf3a

Observation 5e73f1e9-ceb0-42db-b2ca-6c2dedee9f3e · outbound

This paper cites Learning Transferable Visual Models From Natural Language Supervision.

Bridging the Data Provenance Gap Across Text, Speech and Video Learning Transferable Visual Models From Natural Language Supervision

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.062678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.062678Z digest=sha256:0fc1cfad4bd486d593878c4336ce626af1ee6ccf71920080fd67036e5024ce10

Observation c9e0dc0a-1865-4c99-91b1-2d059aa0d449 · outbound

This paper cites Changing the world by changing the data,.

Bridging the Data Provenance Gap Across Text, Speech and Video Changing the world by changing the data,

Reference 46

Resolution
malformed identifier
no resolver link, observed 2026-08-11T12:18:38.068300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.068300Z digest=sha256:3caa0419fe142ef7afbfa29bce0a4248db7988a554d26d77b8ee839f254327fa

Observation cc104248-65a2-4751-853e-28500ea7957b · outbound

This paper cites “Everyone wants to do the model work, not the data work.

Bridging the Data Provenance Gap Across Text, Speech and Video “Everyone wants to do the model work, not the data work

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.072970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.072970Z digest=sha256:e84dc36d412733a21c63990e68e96c906f2c2429dd69ead67c7e8719baa646ea

Observation daded205-9e1e-489b-871a-2538c80fa3e7 · outbound

This paper cites Multitask Prompted Training Enables Zero-Shot Task Generalization.

Bridging the Data Provenance Gap Across Text, Speech and Video Multitask Prompted Training Enables Zero-Shot Task Generalization

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.077796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.077796Z digest=sha256:daec4c56f13651a80a5fa87513c21fec4e8fcabeed363abc6f9b0bce41ca5236

Observation 19c4b6b2-b8c7-4325-8ed4-009c3e3e5a45 · outbound

This paper cites Finetuned language models are zero-shot learners,.

Bridging the Data Provenance Gap Across Text, Speech and Video Finetuned language models are zero-shot learners,

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.099319Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.099319Z digest=sha256:52f0341745054eb0af61eb1d877088084647524b5192493b9af2294c813c9854

Observation 39a725a8-afcb-408c-b5ad-3d565b4afd2d · outbound

This paper cites Challenges in detoxifying language models,.

Bridging the Data Provenance Gap Across Text, Speech and Video Challenges in detoxifying language models,

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.203075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.203075Z digest=sha256:8f3a2eafc95f8d811cbde3564a894030e57f9e1b367624eca29c55f094257d4e

Observation c076c99c-161f-476b-bf7a-2bacfa00ff6b · outbound

This paper cites Detoxifying language models risks marginalizing minority voices,.

Bridging the Data Provenance Gap Across Text, Speech and Video Detoxifying language models risks marginalizing minority voices,

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.301572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.301572Z digest=sha256:6de5afaeb97c7f48261479ababfbb310c05d2c876615510f23798d70f7ea94f1

Observation eef597fb-cc84-4f52-991d-9ca45741d2c1 · outbound

This paper cites Masader: Metadata sourcing for arabic text and speech data resources,.

Bridging the Data Provenance Gap Across Text, Speech and Video Masader: Metadata sourcing for arabic text and speech data resources,

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.374350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.374350Z digest=sha256:73004077d2311ab5780f92c8bc4f16077a90cb2e4b7fc19dd1f2caa37f7874cf

Observation cba3522f-5ce5-4a0f-a0c3-9477e0333ea7 · outbound

This paper cites Quantifying Memorization Across Neural Language Models.

Bridging the Data Provenance Gap Across Text, Speech and Video Quantifying Memorization Across Neural Language Models

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.379014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.379014Z digest=sha256:0f087f6ed4613e7cffdf0c208a2df420e3927590a395a95de71ad35d7fd48028

Observation b615bad5-4b4a-4209-84a0-bee73cb7cbea · outbound

This paper cites CLAP: Learning Audio Concepts From Natural Language Supervision.

Bridging the Data Provenance Gap Across Text, Speech and Video CLAP: Learning Audio Concepts From Natural Language Supervision

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.384193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.384193Z digest=sha256:093d808ee6322a4331309ff309c0752de1d9d4cc305c342a0504d4b97d36bc9e

Observation c584ab02-e0d1-475b-84d3-fa4ffcd9b150 · outbound

This paper cites Dataset geography: Mapping language data to language users,.

Bridging the Data Provenance Gap Across Text, Speech and Video Dataset geography: Mapping language data to language users,

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.388736Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.388736Z digest=sha256:15c2c12148950bb01732e61ed62d60043fbc3bc5be42284d4082326128dd499e

Observation b2401b8d-e086-47c9-a2ce-876bd273dec2 · outbound

This paper cites The flores-101 evaluation benchmark for low- resource and multilingual machine translation,.

Bridging the Data Provenance Gap Across Text, Speech and Video The flores-101 evaluation benchmark for low- resource and multilingual machine translation,

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.392967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.392967Z digest=sha256:42df1186bf7666116fb1830cd7c02fb4ed13ba2d45c04a7accd6d397b4926b73

Observation 764ed178-a112-4e73-b527-f671461daa62 · outbound

This paper cites Training Compute-Optimal Large Language Models.

Bridging the Data Provenance Gap Across Text, Speech and Video Training Compute-Optimal Large Language Models

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.396913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.396913Z digest=sha256:40fee8ab57f06aa3bdafb69717def8b7195a469c5ac8a316f87395ee89b4b5d3

Observation ba5e3d54-2a89-4a45-a117-3b53c3249014 · outbound

This paper cites Leakage and the Reproducibility Crisis in ML-based Science.

Bridging the Data Provenance Gap Across Text, Speech and Video Leakage and the Reproducibility Crisis in ML-based Science

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.401723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.401723Z digest=sha256:91172eb7dbf31ae8eff0c4ded90c149a0a059f50b3b22a8cc7cb037c59ae4c24

Observation 0c506884-047f-4cca-9a0c-7812bb72e2b1 · outbound

This paper cites Quality at a glance: An audit of web-crawled multilingual datasets,.

Bridging the Data Provenance Gap Across Text, Speech and Video Quality at a glance: An audit of web-crawled multilingual datasets,

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.406325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.406325Z digest=sha256:eacea46f50f799443eea1fe37224b2fbfa4ff10dcadf185580549d5e87ecf3b5

Observation 8ce539a6-b5a1-45b2-b993-7ce4b9130ac2 · outbound

This paper cites The bigscience roots corpus: A 1.6tb composite multilingual dataset,.

Bridging the Data Provenance Gap Across Text, Speech and Video The bigscience roots corpus: A 1.6tb composite multilingual dataset,

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.411146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.411146Z digest=sha256:bd855967dac0be57352c1bdd94801a8380de126ae3c72a29bf61941b26dc8763

Observation baf44963-345d-440f-b9a3-a1ed92b6e504 · outbound

This paper cites McMillan-Major, Z.

Bridging the Data Provenance Gap Across Text, Speech and Video McMillan-Major, Z

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.415788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.415788Z digest=sha256:a8acbf6a3bcc0d049f7fec3123fc31cc4bd8939277a0f890e39d8db2a1b3214c

Observation 4ff7621d-871e-4d76-bf55-ada84cf18fe6 · outbound

This paper cites Documenting Geographically and Contextually Diverse Data Sources: The BigScience Catalogue of Language Data and Resources.

Bridging the Data Provenance Gap Across Text, Speech and Video Documenting Geographically and Contextually Diverse Data Sources: The BigScience Catalogue of Language Data and Resources

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.425838Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.425838Z digest=sha256:611a8ff5558eb1a4a7e4d667472a81308e7f6375c1905055b40bed15f36dcd23

Observation 203b6d24-b31f-4bad-bd5c-3ae18f6ccb3e · outbound

This paper cites Video Captioning: a comparative review of where we are and which could be the route.

Bridging the Data Provenance Gap Across Text, Speech and Video Video Captioning: a comparative review of where we are and which could be the route

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.430624Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.430624Z digest=sha256:c6c8cc18d6e4f2b771e4eb21250d1bfa07343b4efe7a7e4914cc937bdaf9a4bb

Observation ece03659-e309-47c9-bd25-9187a2f22a40 · outbound

This paper cites Training language models to follow instructions with human feedback.

Bridging the Data Provenance Gap Across Text, Speech and Video Training language models to follow instructions with human feedback

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.435910Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.435910Z digest=sha256:e51b17e45cf9d5d48c2c30592fae4f2334a48c9152d2ebf1119ac2d1d764bfad

Observation eeb9ff7f-ebfb-4a92-b1c9-59ba47d99210 · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

Bridging the Data Provenance Gap Across Text, Speech and Video Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.441839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.441839Z digest=sha256:db164709da5a060999e9d741bd25b433bafba0dc0ffbe95b0f6df2789573c5ab

Observation bfcda661-d96c-45b1-8a7c-9f9e218c7b27 · outbound

This paper cites Red-Teaming the Stable Diffusion Safety Filter.

Bridging the Data Provenance Gap Across Text, Speech and Video Red-Teaming the Stable Diffusion Safety Filter

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.446393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.446393Z digest=sha256:da23ec2ac18a4672b84a3470f516a1813be58b46b9d7961dbc3cfd2545adbd4d

Observation 620b7cc8-ab03-427c-8fb9-dc438dc81a84 · outbound

This paper cites Make-A-Video: Text-to-Video Generation without Text-Video Data.

Bridging the Data Provenance Gap Across Text, Speech and Video Make-A-Video: Text-to-Video Generation without Text-Video Data

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.451457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.451457Z digest=sha256:556046f69331c491cfe5fc9fcac03d75128c1a38783cbcdde0761a4e03c42631

Observation b04f1218-2a03-428d-aee9-26b3fed62131 · outbound

This paper cites Bigssl: Exploring the frontier of large-scale semi- supervised learning for automatic speech recognition,.

Bridging the Data Provenance Gap Across Text, Speech and Video Bigssl: Exploring the frontier of large-scale semi- supervised learning for automatic speech recognition,

Reference 68

Resolution
malformed identifier
no resolver link, observed 2026-08-11T12:18:38.455743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.455743Z digest=sha256:b824c8a4ae9335f6f9864d3585670d31808804e99fbe11a67e29f8f66b7df16c

Observation 04c8d0e3-cb37-4f94-96ac-568ca63cdc54 · outbound

This paper cites Survey of video object detection algorithms based on deep learning,.

Bridging the Data Provenance Gap Across Text, Speech and Video Survey of video object detection algorithms based on deep learning,

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.550049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.550049Z digest=sha256:4d3cc8680380972fc23a2fc4f43a32ba8d4c4d9dd9cfa9524c45a3f596b16611

Observation d9e123a1-2b13-4e5c-9d06-3c722eafefbf · outbound

This paper cites Scaling laws for generative mixed-modal language models,.

Bridging the Data Provenance Gap Across Text, Speech and Video Scaling laws for generative mixed-modal language models,

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.616320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.616320Z digest=sha256:a7eb2fcc1e1c4b761ad3ec192c49bd0919f4defaab15c76f3dbf4f2aad4c8c57

Observation b367f587-fa50-47ed-bc01-c0d2b9d1fc8d · outbound

This paper cites Into the LAIONs Den: Investigating Hate in Multimodal Datasets.

Bridging the Data Provenance Gap Across Text, Speech and Video Into the LAIONs Den: Investigating Hate in Multimodal Datasets

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.759641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.759641Z digest=sha256:5d609afe5f03807d91d4e0a0f8f54001d0a47d02bba80a6a2bf55bcc736bb486

Observation f31d00a1-0f37-4769-92c6-c1a21d9d20f0 · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

Bridging the Data Provenance Gap Across Text, Speech and Video Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.831652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.831652Z digest=sha256:c6349898864c68d5571b376916730834586d1c420348ac6e82a3424f3bba8732

Observation 72403e49-9387-489e-a2db-bb198f835fdf · outbound

This paper cites Bommasani, K.

Bridging the Data Provenance Gap Across Text, Speech and Video Bommasani, K

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.836900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.836900Z digest=sha256:4d5e898d085aca32ebe70da5f93bb647f667ad3730afef785aed2bafad1192b3

Observation ce871711-4c48-4f9b-b65a-a3446da89ae4 · outbound

This paper cites Quantifying mem- orization across neural language models,.

Bridging the Data Provenance Gap Across Text, Speech and Video Quantifying mem- orization across neural language models,

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.849686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.849686Z digest=sha256:318329f521f02156616cb5978ee3683d2fe5fe37b9307743a1bc4ab60cd821b7

Observation a42caaa4-cdcb-45c4-9ead-d19002074711 · outbound

This paper cites Extracting training data from diffusion models,.

Bridging the Data Provenance Gap Across Text, Speech and Video Extracting training data from diffusion models,

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.855948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.855948Z digest=sha256:08ce9b2bc00c87271dbc2fec070adbcadc761f3f50fbcddd19d4913cadfed8f0

Observation 54fe29a2-755f-4047-99b8-789c979b5006 · outbound

This paper cites an unresolved cited work.

Bridging the Data Provenance Gap Across Text, Speech and Video Unresolved cited work

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.860032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.860032Z digest=sha256:34dab17b28f91783a943d4d67238125909529e5933f1286787618b3b2eeaf387

Observation f815bbd4-4f46-493a-bb30-8184f486d2d7 · outbound

This paper cites Gender bias in hiring: An analysis of the impact of amazon’s recruiting algorithm,.

Bridging the Data Provenance Gap Across Text, Speech and Video Gender bias in hiring: An analysis of the impact of amazon’s recruiting algorithm,

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.864990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.864990Z digest=sha256:3ab4f34dc1301ca3ca7147fdd38a19731f1e0b270d08f883b28968c4c27a301f

Observation 2ed54f6d-6903-4bf5-b6ef-d471d05808f1 · outbound

This paper cites Can language models be instructed to protect personal information?.

Bridging the Data Provenance Gap Across Text, Speech and Video Can language models be instructed to protect personal information?

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.869440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.869440Z digest=sha256:51531527fc76814710c323cfcee110ca3737aa0ce190d6c1244d0182983a9bf7

Observation 6a7cf425-41ba-471c-a819-b8dd29c195ea · outbound

This paper cites Dialect corpora from youtube,.

Bridging the Data Provenance Gap Across Text, Speech and Video Dialect corpora from youtube,

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.873717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.873717Z digest=sha256:efe0b57d893dc54866c0b682c009d2e147f46c4b081954e595bda58f727d8ab9

Observation 3608ee7b-d191-4736-a7ab-13f3ac2ec008 · outbound

This paper cites Ai image training dataset found to include child sexual abuse imagery,.

Bridging the Data Provenance Gap Across Text, Speech and Video Ai image training dataset found to include child sexual abuse imagery,

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.877961Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.877961Z digest=sha256:0ea717bed2ba783f0df932923b0270b5c94d554826cc7b0bc7e1561d1f3a2ef5

Observation ac8b0b81-05d2-44d2-80ab-8b9b75a26d3c · outbound

This paper cites What’s in my big data?.

Bridging the Data Provenance Gap Across Text, Speech and Video What’s in my big data?

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.882071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.882071Z digest=sha256:7a34beb11422329c8dddce4d63de56e21874473a8aaed5b87fdc97b3086ea855

Observation a7580bb9-13e2-45ef-a91d-4514cebb8114 · outbound

This paper cites Structure and Content-Guided Video Synthesis with Diffusion Models.

Bridging the Data Provenance Gap Across Text, Speech and Video Structure and Content-Guided Video Synthesis with Diffusion Models

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.886382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.886382Z digest=sha256:fe1a4f6a20b16d04c4c752e0221392aae5e991e48df02feb5397672ea8160868

Observation 2a39e77d-7742-41f5-89f1-707f3d4021a5 · outbound

This paper cites Datacomp: In search of the next generation of multi- modal datasets,.

Bridging the Data Provenance Gap Across Text, Speech and Video Datacomp: In search of the next generation of multi- modal datasets,

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:38.973189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:38.973189Z digest=sha256:ce964c989366fda0320e0fe41020aa2bb1b62ea050c1823de695c3efab40bb90

Observation d0cf6703-313f-457b-8b67-e58098545fd6 · outbound

This paper cites Foundation Models and Fair Use.

Bridging the Data Provenance Gap Across Text, Speech and Video Foundation Models and Fair Use

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:39.033787Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:39.033787Z digest=sha256:fe5bb0d72c1752d8851c97d345a2455930a91bd76d9e77837814e27c0b9f4c35

Observation 317b730d-7fe8-4779-9215-d5d98a852ac5 · outbound

This paper cites Understanding Catastrophic Forgetting in Language Models via Implicit Inference.

Bridging the Data Provenance Gap Across Text, Speech and Video Understanding Catastrophic Forgetting in Language Models via Implicit Inference

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:39.147383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:39.147383Z digest=sha256:38a96462ec7aa02590c092b245c9c6fc2f881466df5511f7fa7adc34b31ad763

Observation 483f2c46-0785-4b07-8c37-a8775c89a0a9 · outbound

This paper cites Harnessing large-language models to generate private synthetic text.

Bridging the Data Provenance Gap Across Text, Speech and Video Harnessing large-language models to generate private synthetic text

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:39.153956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:39.153956Z digest=sha256:050ba3e447ef2053d503773dd9bc99499ee4b19d8ec06c1b0a29e990cd6dadeb

Observation 7aa2bb08-689a-4eee-a193-db704d874c8d · outbound

This paper cites Platypus: Quick, cheap, and powerful refinement of llms,.

Bridging the Data Provenance Gap Across Text, Speech and Video Platypus: Quick, cheap, and powerful refinement of llms,

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:39.158999Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:39.158999Z digest=sha256:6ecd8090df0c39acee364ef88b6c60f7be266fd5448d7ddc228d0d33d8e2cdb9

Observation ba48ea07-989c-47a4-aee0-07531fc70397 · outbound

This paper cites Talkin' 'Bout AI Generation: Copyright and the Generative-AI Supply Chain.

Bridging the Data Provenance Gap Across Text, Speech and Video Talkin' 'Bout AI Generation: Copyright and the Generative-AI Supply Chain

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:39.164488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:39.164488Z digest=sha256:e46e3c7eb44160d1709fca1abb0f6ec443f2b3b63a9d681053edd6940873efb8

Observation 7e49434b-f6f6-430c-af41-0691e3e4fcdc · outbound

This paper cites Yodas: Youtube- oriented dataset for audio and speech,.

Bridging the Data Provenance Gap Across Text, Speech and Video Yodas: Youtube- oriented dataset for audio and speech,

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:39.170432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:39.170432Z digest=sha256:98986f489c91941b9234ed885ad18139e559e7adce72f07fb8df58cd7f488609

Observation d11f82cf-ddcf-4ed4-ab4a-5c78826f1023 · outbound

This paper cites an unresolved cited work.

Bridging the Data Provenance Gap Across Text, Speech and Video Unresolved cited work

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:39.176280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:39.176280Z digest=sha256:ca6a58fac119bed855a1d6ab3a3e6ba66e57a8244c3ea3fc78b5abad9e97f8e9

Observation d8fd51e7-8672-4417-a2aa-442e460dd04e · outbound

This paper cites Visual instruction tuning,.

Bridging the Data Provenance Gap Across Text, Speech and Video Visual instruction tuning,

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:39.181162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:39.181162Z digest=sha256:8d9f3bcb644c503d28916b9435feeb282741ff686da26e66bd5e1aafe116b37d

Observation de67c4b4-c154-4763-9730-cf52a8102e86 · outbound

This paper cites A Pretrainer's Guide to Training Data: Measuring the Effects of Data Age, Domain Coverage, Quality, & Toxicity.

Bridging the Data Provenance Gap Across Text, Speech and Video A Pretrainer's Guide to Training Data: Measuring the Effects of Data Age, Domain Coverage, Quality, & Toxicity

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:39.186213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:39.186213Z digest=sha256:da375af1085cf7e885ec3b4ce1ff44ea4f05f094644496cffaf77b9b7c1bc309

Observation fdb8e334-5406-470b-95e3-22b98326fd19 · outbound

This paper cites Discit ergo est: Training data provenance and fair use,.

Bridging the Data Provenance Gap Across Text, Speech and Video Discit ergo est: Training data provenance and fair use,

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:39.191010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:39.191010Z digest=sha256:450499b4a029567ce2a80339f13b508a9239b89296725bd03b0974e1c06442a6

Observation a9b0fc45-814d-4f19-b5e5-308d7dc85e21 · outbound

This paper cites Mahari, L.

Bridging the Data Provenance Gap Across Text, Speech and Video Mahari, L

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:39.195797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:39.195797Z digest=sha256:c2ea20842969d6e0d4fbde931ecd1900254a778569074aa45a894259eb29d614

Observation 53ee1090-4624-49ae-9838-0e175b389bef · outbound

This paper cites When Less is More: Investigating Data Pruning for Pretraining LLMs at Scale.

Bridging the Data Provenance Gap Across Text, Speech and Video When Less is More: Investigating Data Pruning for Pretraining LLMs at Scale

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:39.200726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:39.200726Z digest=sha256:866dae4d36a771016a9b78e857dfc607d6b343bef612bbdd66c5ac988b0e0423

Observation 741641ef-285f-4674-9d8e-3baea912e8ec · outbound

This paper cites Silo language models: Isolating legal risk in a nonparametric datastore,.

Bridging the Data Provenance Gap Across Text, Speech and Video Silo language models: Isolating legal risk in a nonparametric datastore,

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:39.220163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:39.220163Z digest=sha256:4812967eec8b1b13c1e575995f0919cfdcfe4fa15c86ea07c3653166518050f8

Observation 86a62542-018e-444d-bf1a-c8f8262d6ac3 · outbound

This paper cites Licensed to learn: Mitigating copyright infringement liability of generative ai systems through contracts,.

Bridging the Data Provenance Gap Across Text, Speech and Video Licensed to learn: Mitigating copyright infringement liability of generative ai systems through contracts,

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:39.369149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:39.369149Z digest=sha256:8b7b944be0520f805315801a5f2e9600d30522c0d952b70f7367c05c8b964c83

Observation 46cfb70d-acf1-4ec1-a188-4a4217883aa6 · outbound

This paper cites Crosslingual generalization through multitask finetuning,.

Bridging the Data Provenance Gap Across Text, Speech and Video Crosslingual generalization through multitask finetuning,

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:39.375354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:39.375354Z digest=sha256:edb7980bab47dd5624baaaf6de50c7fa1bd7253b5ba45cf01191389532d61ee7

Observation 1475fb1c-0e1b-4d11-8334-90ac6fdb744d · outbound

This paper cites The RefinedWeb dataset for falcon LLM: Outperforming curated corpora with web data, and web data only,.

Bridging the Data Provenance Gap Across Text, Speech and Video The RefinedWeb dataset for falcon LLM: Outperforming curated corpora with web data, and web data only,

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:39.380955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:39.380955Z digest=sha256:914c629624622cf1186ca870af32500ab74fb3fa7f25f8c46f15da650eb55cd7

Observation f924aa00-fc87-41a6-93f1-51e58f558f97 · outbound

This paper cites Reproducing whisper-style training using an open-source toolkit and publicly available data,.

Bridging the Data Provenance Gap Across Text, Speech and Video Reproducing whisper-style training using an open-source toolkit and publicly available data,

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:39.385524Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:39.385524Z digest=sha256:798936c545ef6930621854ec2aa3a6dbc8a76a187f78b7ae6c100ca594613002

Observation 3e1ca41a-b4bd-4842-b4f9-97955803c8b8 · outbound

This paper cites The Casual Conversations v2 Dataset.

Bridging the Data Provenance Gap Across Text, Speech and Video The Casual Conversations v2 Dataset

Reference 101

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:39.389698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:39.389698Z digest=sha256:9a3c185fd6d817fe590ffdcb739f7505ddeb48ac3d144d2194740b4e9b1eccaf

Observation 68096db9-cb6e-44ed-b8c4-2a61e4ca4d73 · outbound

This paper cites Goodtriever: Adaptive Toxicity Mitigation with Retrieval-augmented Models.

Bridging the Data Provenance Gap Across Text, Speech and Video Goodtriever: Adaptive Toxicity Mitigation with Retrieval-augmented Models

Reference 102

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:39.394434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:39.394434Z digest=sha256:0b65e8f68f8976de6feb933fb1618279a3a67823b51aabad9639a036d3ada61d

Observation f666d717-3996-42c5-82a0-8677c9420887 · outbound

This paper cites End-to-end speech recognition: A survey,.

Bridging the Data Provenance Gap Across Text, Speech and Video End-to-end speech recognition: A survey,

Reference 103

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:39.399058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:39.399058Z digest=sha256:669163d0a282994834b13f0c5fc67f06be384e69359a2a51b8c9f1a14c0eb8e2

Observation 7befe848-6fbc-43f0-931b-1f71b3d37d24 · outbound

This paper cites Robust speech recognition via large-scale weak supervision,.

Bridging the Data Provenance Gap Across Text, Speech and Video Robust speech recognition via large-scale weak supervision,

Reference 104

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:39.403513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:39.403513Z digest=sha256:5159f0fabac666a2d46e526a2caf5acfd2a2557ff19d33bd6d62e411469f6930

Observation 08285e19-b154-4f80-a186-dafb12fe9f95 · outbound

This paper cites Direct Preference Optimization: Your Language Model is Secretly a Reward Model.

Bridging the Data Provenance Gap Across Text, Speech and Video Direct Preference Optimization: Your Language Model is Secretly a Reward Model

Reference 105

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:39.408235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:18:39.408235Z digest=sha256:0b88cd23331438f4cd246758caa187c4b0d30fc67a6fff3a4d6d46be2f5a97d4

Pith citing papers

Observation 2ea33814-c125-4f59-b3eb-30b7ab1d5d0f · inbound

The Leaderboard Illusion cites this paper.

The Leaderboard Illusion Bridging the Data Provenance Gap Across Text, Speech and Video

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-16T05:22:55.579720Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T05:22:55.579720Z digest=sha256:1172aae9df88723fe7e0a7bdc55b26cde4bfd0cb7a555b56a08962c406de5aff

Observation a93a6730-803a-4979-8f1e-eeefa0299a95 · inbound

TEDI: Trustworthy and Ethical Dataset Indicators to Analyze and Compare Dataset Documentation cites this paper.

TEDI: Trustworthy and Ethical Dataset Indicators to Analyze and Compare Dataset Documentation Bridging the Data Provenance Gap Across Text, Speech and Video

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T14:42:06.940110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:42:06.940110Z digest=sha256:bca7f817446205969a4cf7a11803badf5382acd5cba475c176abf8566eb10509

Observation f731eed2-b730-4270-841a-088f1f5382dd · inbound

The Common Pile v0.1: An 8TB Dataset of Public Domain and Openly Licensed Text cites this paper.

The Common Pile v0.1: An 8TB Dataset of Public Domain and Openly Licensed Text Bridging the Data Provenance Gap Across Text, Speech and Video

Reference 109

Resolution
verified exact
local_arxiv, observed 2026-08-07T10:29:46.930466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T10:29:44.548165Z digest=sha256:2f49bee1145a89413febd505c3209c9a66cdbb049ac5f779b16b6de45b8acdbc

Observation ba0306ed-e07e-42f3-bcf8-09d27c104048 · inbound

BRoverbs -- Measuring how much LLMs understand Portuguese proverbs cites this paper.

BRoverbs -- Measuring how much LLMs understand Portuguese proverbs Bridging the Data Provenance Gap Across Text, Speech and Video

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-04T19:58:54.302390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T19:58:54.302390Z digest=sha256:319b9d7fa7412d79d1daa92020f7fecaacb471666dfdf9e0ed11d516d2fa59c6