Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-14T06:06:44.499410Z
Paper Citation Record · LEDGER
As of 23 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 0 inbound Pith citation observations for arXiv:1909.00121.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-14T06:06:44.499410Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
49 of 49 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 2bf5375e-92a8-4163-a8f5-7407e3aa4d51 · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Jointly modeling embedding and translation to bridge video and language,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 45fc1d89-3cda-42f7-bf65-301f02a1ebf7 · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Semantic compositional networks for visual captioning,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 2b1ddf43-83e2-4730-9772-a92f983559b5 · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Video captioning with attention-based lstm and semantic con- sistency,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation f4aabf7f-b95d-4151-b267-2eb570373754 · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Reinforced video caption- ing with entailment rewards,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 1e834d07-ab52-453a-85e2-d8b6ec16bff1 · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Sequence to sequence - video to text,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation f3233cf2-84c9-4507-b149-75f910086577 · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Long-term recurrent convolutional networks for visual recognition and description,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation f53b4f61-b125-4143-bf2d-c9724a9c1aef · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Sched- uled sampling for sequence prediction with recurrent neural networks,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 67c81ae3-31ae-48fe-a7bf-d9bf47c1e59a · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Learning phrase representations using RNN encoder-decoder for statistical machine translation,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 4fd4f53c-1b94-4fc5-aba9-a729b557d6c7 · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Show and tell: A neural image caption generator,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation cb1bf66e-d6ee-447b-bf2d-b63a4b6a32ae · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Explain Images with Multimodal Recurrent Neural Networks
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df5dcd4b-217b-4dbd-85a6-61e5b2909dfd · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Neural machine translation by jointly learning to align and translate,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 91dbd382-25e7-48f9-be37-709c84993af0 · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Multiple Object Recognition with Visual Attention
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ae46ee7-4b37-48b7-913e-e88a60d45c14 · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Image captioning with semantic attention,
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d34eebab-bd07-4d29-a23b-792ff1fec0d8 · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Bottom-up and top-down attention for image captioning and visual question answering,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 3a0c0e74-f73f-4923-b536-b50c519b0f59 · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Self-critical sequence training for image captioning,
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa7fc737-65e4-4b90-b103-bcd77cfd67fe · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Exploring visual rela- tionship for image captioning,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 2e5aa482-c04a-420c-a51c-0b21576930d3 · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Multimodal trans- former with multi-view visual representation for image captioning,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 1847ea75-d5b4-4889-b646-6753a39f076e · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Meshed-Memory Transformer for Image Captioning,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation c1a30203-1804-4196-8ca6-6d41170cc6d2 · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Controllable video captioning with pos se- quence guidance based on gated fusion network,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation d8e9073e-77d7-4ae2-a664-d4a625e3eb38 · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Memory-attended recurrent network for video captioning,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 48e0d08f-a4fd-47aa-bb92-b808172a3563 · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Joint syntax representation learning and visual cue translation for video captioning,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation c31d7597-7d58-4392-9dde-eef20fdd0ac9 · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Spatio-temporal dynamics and semantic attribute en- riched visual encoding for video captioning,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 84f23a4c-6298-4dd9-975a-12a79570c1d6 · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Syntax-aware action targeting for video captioning,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 49460937-d7f6-4039-89fa-aa9af015551d · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Video paragraph captioning using hierarchical recurrent neural networks,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 2e952d92-56e7-4c12-a685-67829cbce697 · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Show, attend and tell: Neural image caption generation with visual attention,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation e6aadcce-ce67-4871-a23d-e0759bf01ea1 · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Top-down visual saliency guided by captions,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 49ca774d-f854-4622-abe0-fe67f999202a · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Less is more: Picking informative frames for video captioning,
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8515a68-3524-4354-b627-1aa5f0693935 · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Watch, lis- ten, and describe: Globally and locally aligned cross- modal attentions for video captioning,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 0a684f27-ffac-40bf-9eeb-9eba672329c2 · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Multi-task video captioning with video and entailment generation,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 4bd62083-3c82-4e17-bbb5-7cfd5d85da14 · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Videobert: A joint model for video and language representation learning,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 6673947a-22c5-4b3e-810f-861a7f7e52da · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 852d4678-8910-40d1-bfe8-8e40f8d737f0 · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Learning to compose topic-aware mixture of experts for zero-shot video captioning,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 15b7ff8e-88c6-457e-940a-8a2608881faf · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Spatio-temporal graph for video captioning with knowledge distillation,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation c25fddfe-95c7-40f0-9ee0-0851630838b0 · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling A learning algorithm for continually running fully recurrent neural networks,
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ce565dc-561f-45aa-ae63-3a51ff2864ef · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling How (not) to Train your Generative Model: Scheduled Sampling, Likelihood, Adversary?
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7bd7a385-4d82-4455-a3e1-c122c55be88a · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Professor forcing: A new algorithm for training recurrent networks,
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation d5b35215-b44d-4e26-8667-029397fcd84a · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Object relational graph with teacher-recommended learning for video captioning,
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 0742d209-7878-47c3-b576-dfda851ecb33 · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Simple statistical gradient-following al- gorithms for connectionist reinforcement learning,
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 29169e2d-e5e2-4c8a-b759-ff6f39eb21b5 · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Finding structure in time,
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e1e6673-c21f-4a57-af06-7ed13417c9d8 · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Long short-term memory,
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be883842-f70e-4905-99ab-7376c54ef632 · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Google’s Neural Machine Translation System: Bridging the Gap between Human and Machine Translation,
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation c9ed91e9-932d-40c4-a1b1-77058d54bb73 · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Youtube2text: Recognizing and describing arbitrary ac- tivities using semantic hierarchies and zero-shot recog- nition,
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 8c284eec-572c-4d33-873c-9c5288f84a8c · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Collecting highly parallel data for paraphrase evaluation,
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation ce02e931-d919-47aa-b30c-6a23d5efad9c · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling MSR-VTT: A large video description dataset for bridging video and language,
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 7898fb66-7880-4e2a-b8f3-61abe7fbb891 · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Aggregated residual transformations for deep neural networks,
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a15bb6f7-ae9b-42c0-893d-c2d3b4aa3292 · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling ECO: efficient convolutional network for online video understanding,
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19848132-ef73-4833-8e6b-cd6681b5d31a · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Sibnet: Sibling convolutional encoder for video captioning,
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab3864ee-2408-4cb9-a5d9-f08b58e07811 · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Multi-label classification: An overview,
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 59132fb8-4aa9-4310-bee9-ac53d26ba833 · outbound
A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Neural Machine Translation by Jointly Learning to Align and Translate
Reference 2015
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.