Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T00:48:30.581939Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 0 inbound Pith citation observations for arXiv:2506.12623.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T00:48:30.581939Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
13 of 13 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 34ea7804-2841-406c-a928-502d210f5a08 · outbound
MS4UI: A Dataset for Multi-modal Summarization of User Interface Instructional Videos BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and Comprehension
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf0cc22b-2fe0-4413-a07d-e2ec1215f176 · outbound
MS4UI: A Dataset for Multi-modal Summarization of User Interface Instructional Videos InProceedings of the 2020 Con- ference on Empirical Methods in Natural Language Processing (EMNLP), pages 4707–4716, Online
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 795be711-6f3f-4c84-a376-8019c03f01d3 · outbound
MS4UI: A Dataset for Multi-modal Summarization of User Interface Instructional Videos MMSum: A Dataset for Multimodal Summarization and Thumbnail Generation of Videos
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2de4cfa6-e2a6-4e58-af28-867351195999 · outbound
MS4UI: A Dataset for Multi-modal Summarization of User Interface Instructional Videos How2: A Large-scale Dataset for Multimodal Language Understanding
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f5296aa-02e4-4b02-892e-3f33da3e84d6 · outbound
MS4UI: A Dataset for Multi-modal Summarization of User Interface Instructional Videos Very Deep Convolutional Networks for Large-Scale Image Recognition
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70c46022-23c8-4b93-8fac-770eac918e80 · outbound
MS4UI: A Dataset for Multi-modal Summarization of User Interface Instructional Videos InProceed- ings of the IEEE conference on computer vision and pattern recognition, pages 1059–1067
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7607221a-21b7-4272-9467-33ef23071a16 · outbound
MS4UI: A Dataset for Multi-modal Summarization of User Interface Instructional Videos InComputer Vision–ECCV 2014: 13th European Conference, Zurich, Switzerland, September 6-12, 2014, Proceedings, Part VII 13, pages 505–520
Reference 2014
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4a6f7f72-e090-462a-9b56-96a4427c9d89 · outbound
MS4UI: A Dataset for Multi-modal Summarization of User Interface Instructional Videos Abstractive Text Summarization Using Sequence-to-Sequence RNNs and Beyond
Reference 2016
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 264c2662-35b9-4326-84c6-0d60e4ab4ab3 · outbound
MS4UI: A Dataset for Multi-modal Summarization of User Interface Instructional Videos Deep Communicating Agents for Abstractive Summarization
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63b842e4-0c12-47d4-98b7-1c8586ae3834 · outbound
MS4UI: A Dataset for Multi-modal Summarization of User Interface Instructional Videos InPro- ceedings of the 2019 CHI Conference on Human Factors in Computing Systems, pages 1–10
Reference 2019
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8063f4d8-c1ec-4ac4-89b5-c4169252c7e1 · outbound
MS4UI: A Dataset for Multi-modal Summarization of User Interface Instructional Videos Multi-modal Summarization for Video-containing Documents
Reference 2020
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 76ee0f58-4157-4dd3-b9f3-e554262b00db · outbound
MS4UI: A Dataset for Multi-modal Summarization of User Interface Instructional Videos bert2BERT: Towards Reusable Pretrained Language Models
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a5eb272a-6f4d-4fad-a5ed-e63d7755510e · outbound
MS4UI: A Dataset for Multi-modal Summarization of User Interface Instructional Videos InFindings of the Association for Computa- tional Linguistics: EACL 2023, pages 880–894
Reference 2023
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
No inbound Pith citation observations are available.