Pith. sign in

Paper Citation Record · LEDGER

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling

As of 23 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 0 inbound Pith citation observations for arXiv:1909.00121.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1909.00121 v3

Coverage vector

measured 49 of 49 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-14T06:06:44.499410Z

measured 49 of 49 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

49 of 49 outbound references displayed

  • verified exact2
  • verified fuzzy33
  • unresolved14
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2bf5375e-92a8-4163-a8f5-7407e3aa4d51 · outbound

This paper cites Jointly modeling embedding and translation to bridge video and language,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Jointly modeling embedding and translation to bridge video and language,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T06:06:45.011263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T06:06:44.361502Z digest=sha256:a7ff56ed7c0129de6cac2af0e5002a190919994cb19920121072918569a55a62

Observation 45fc1d89-3cda-42f7-bf65-301f02a1ebf7 · outbound

This paper cites Semantic compositional networks for visual captioning,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Semantic compositional networks for visual captioning,

Reference 2

Resolution
verified exact
doi, observed 2026-08-14T06:06:44.554569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T06:06:44.365257Z digest=sha256:b3008b55e0ae5274cb3746cd81d0eebfdf4c1abdae98956d84bc58d1cf2c3302

Observation 2b1ddf43-83e2-4730-9772-a92f983559b5 · outbound

This paper cites Video captioning with attention-based lstm and semantic con- sistency,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Video captioning with attention-based lstm and semantic con- sistency,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T06:06:45.004263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T06:06:44.368152Z digest=sha256:81a60fa693ae8e244830e3efdbccccc48d9b393d3860b3a1202bc20e0bd11cff

Observation f4aabf7f-b95d-4151-b267-2eb570373754 · outbound

This paper cites Reinforced video caption- ing with entailment rewards,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Reinforced video caption- ing with entailment rewards,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T06:06:44.997168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T06:06:44.371070Z digest=sha256:376e4b1eb52a499d92fd31276bacd6e7e80ab4f27a3b9db86f210ba1322bb57c

Observation 1e834d07-ab52-453a-85e2-d8b6ec16bff1 · outbound

This paper cites Sequence to sequence - video to text,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Sequence to sequence - video to text,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T06:06:44.989948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T06:06:44.374158Z digest=sha256:2b838694cb402b483dbec6b02f07bbe212f4190f804fc0df6f779db5c9f2c49d

Observation f3233cf2-84c9-4507-b149-75f910086577 · outbound

This paper cites Long-term recurrent convolutional networks for visual recognition and description,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Long-term recurrent convolutional networks for visual recognition and description,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T06:06:44.982795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T06:06:44.377091Z digest=sha256:5ad84a073e78384f6b91b24d3a17605c37ce270b7c917f9d40536198ce1b6748

Observation f53b4f61-b125-4143-bf2d-c9724a9c1aef · outbound

This paper cites Sched- uled sampling for sequence prediction with recurrent neural networks,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Sched- uled sampling for sequence prediction with recurrent neural networks,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T06:06:44.975135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T06:06:44.380358Z digest=sha256:29c5187dc08552de794b80880e2d2605652609eb3ed95333831ead2c33555a96

Observation 67c81ae3-31ae-48fe-a7bf-d9bf47c1e59a · outbound

This paper cites Learning phrase representations using RNN encoder-decoder for statistical machine translation,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Learning phrase representations using RNN encoder-decoder for statistical machine translation,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T06:06:44.967431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T06:06:44.383158Z digest=sha256:8e9b01bb4d5be02618309e695a125405c71f3212167fdf87608491b6a65c9b8e

Observation 4fd4f53c-1b94-4fc5-aba9-a729b557d6c7 · outbound

This paper cites Show and tell: A neural image caption generator,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Show and tell: A neural image caption generator,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T06:06:44.959992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T06:06:44.385751Z digest=sha256:8f673ce143cc3ecc6a3a9a2b647e9b14ca6da9f81ba76c9b41c37b9750ca9611

Observation cb1bf66e-d6ee-447b-bf2d-b63a4b6a32ae · outbound

This paper cites Explain Images with Multimodal Recurrent Neural Networks.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Explain Images with Multimodal Recurrent Neural Networks

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-14T06:06:44.388985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T06:06:44.388985Z digest=sha256:0cbddec1f00c8c2272bf9c7b270878f88b1f314edeb66dcabdc1ee2a8743952f

Observation df5dcd4b-217b-4dbd-85a6-61e5b2909dfd · outbound

This paper cites Neural machine translation by jointly learning to align and translate,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Neural machine translation by jointly learning to align and translate,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T06:06:44.952548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T06:06:44.392701Z digest=sha256:ac388fd277cb0952a69862632fef43c4131b3a2e43186294631c6bf6bd2de5c4

Observation 91dbd382-25e7-48f9-be37-709c84993af0 · outbound

This paper cites Multiple Object Recognition with Visual Attention.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Multiple Object Recognition with Visual Attention

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-14T06:06:44.398896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T06:06:44.398896Z digest=sha256:7290465602997471356a083b3f0376dfc5062780d90ec8c22756ba32c6402b48

Observation 5ae46ee7-4b37-48b7-913e-e88a60d45c14 · outbound

This paper cites Image captioning with semantic attention,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Image captioning with semantic attention,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-14T06:06:44.401874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T06:06:44.401874Z digest=sha256:1c98927b8b68f94b0df6480b68025f00e3484bf2c8a81b20de34add478927714

Observation d34eebab-bd07-4d29-a23b-792ff1fec0d8 · outbound

This paper cites Bottom-up and top-down attention for image captioning and visual question answering,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Bottom-up and top-down attention for image captioning and visual question answering,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T06:06:44.945047Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T06:06:44.405099Z digest=sha256:cda31d94ecad34ac2ba1a717025c7708e7fd735ea1841e69507d64c478d7f474

Observation 3a0c0e74-f73f-4923-b536-b50c519b0f59 · outbound

This paper cites Self-critical sequence training for image captioning,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Self-critical sequence training for image captioning,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-14T06:06:44.408048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T06:06:44.408048Z digest=sha256:db8f339d69b0a9795809f4b19357cff92df31ccbd6f29a2badda0e9147971ccf

Observation fa7fc737-65e4-4b90-b103-bcd77cfd67fe · outbound

This paper cites Exploring visual rela- tionship for image captioning,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Exploring visual rela- tionship for image captioning,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T06:06:44.937156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T06:06:44.411115Z digest=sha256:983d8052bf29f9a66e624f03f401b3f101cc57a1534bb27f413b4779cb7f8d5e

Observation 2e5aa482-c04a-420c-a51c-0b21576930d3 · outbound

This paper cites Multimodal trans- former with multi-view visual representation for image captioning,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Multimodal trans- former with multi-view visual representation for image captioning,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T06:06:44.929080Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T06:06:44.413770Z digest=sha256:3453ec9eae1727606a095d4b40bdb36a9e8aab56161de9c36d66503996424d40

Observation 1847ea75-d5b4-4889-b646-6753a39f076e · outbound

This paper cites Meshed-Memory Transformer for Image Captioning,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Meshed-Memory Transformer for Image Captioning,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T06:06:44.920957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T06:06:44.416423Z digest=sha256:d4a2f25b61b665286b78f9d362873302f3cc24c952687c29122787c724137268

Observation c1a30203-1804-4196-8ca6-6d41170cc6d2 · outbound

This paper cites Controllable video captioning with pos se- quence guidance based on gated fusion network,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Controllable video captioning with pos se- quence guidance based on gated fusion network,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T06:06:44.912353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T06:06:44.419220Z digest=sha256:9c409201a7548cf834cf7456d06034360dfe842398de95e2f44c47c498362ffa

Observation d8e9073e-77d7-4ae2-a664-d4a625e3eb38 · outbound

This paper cites Memory-attended recurrent network for video captioning,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Memory-attended recurrent network for video captioning,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T06:06:44.903361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T06:06:44.421849Z digest=sha256:89505efc4ea4756b7b137cd4db03b9c1a05f2ea82d1cf86a73c8bfddde0ee241

Observation 48e0d08f-a4fd-47aa-bb92-b808172a3563 · outbound

This paper cites Joint syntax representation learning and visual cue translation for video captioning,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Joint syntax representation learning and visual cue translation for video captioning,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T06:06:44.895332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T06:06:44.424398Z digest=sha256:f429e33d7e8ac4f309e8de067afb78e2859aef8966c4fac6aa3d14c06b19e149

Observation c31d7597-7d58-4392-9dde-eef20fdd0ac9 · outbound

This paper cites Spatio-temporal dynamics and semantic attribute en- riched visual encoding for video captioning,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Spatio-temporal dynamics and semantic attribute en- riched visual encoding for video captioning,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T06:06:44.887461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T06:06:44.427108Z digest=sha256:43e05739574269677985d07c511bc9d8dbb64679880a906f904ab11a9653b501

Observation 84f23a4c-6298-4dd9-975a-12a79570c1d6 · outbound

This paper cites Syntax-aware action targeting for video captioning,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Syntax-aware action targeting for video captioning,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T06:06:44.879459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T06:06:44.429769Z digest=sha256:206ba8fc0be023f07d5a2529c0df151314c435c582598ab963e7e1b15de65e27

Observation 49460937-d7f6-4039-89fa-aa9af015551d · outbound

This paper cites Video paragraph captioning using hierarchical recurrent neural networks,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Video paragraph captioning using hierarchical recurrent neural networks,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T06:06:44.870474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T06:06:44.432514Z digest=sha256:91318478cd9c93d6f75f668b019ced0115660560e2beacbde13a0a4952bf80f0

Observation 2e952d92-56e7-4c12-a685-67829cbce697 · outbound

This paper cites Show, attend and tell: Neural image caption generation with visual attention,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Show, attend and tell: Neural image caption generation with visual attention,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T06:06:44.861077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T06:06:44.435006Z digest=sha256:a5ce4ee22d8aa9c1eb5cd3b1e87c3cf0957655d9e5371379dfa6d5d22e9b45a5

Observation e6aadcce-ce67-4871-a23d-e0759bf01ea1 · outbound

This paper cites Top-down visual saliency guided by captions,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Top-down visual saliency guided by captions,

Reference 26

Resolution
verified exact
doi, observed 2026-08-14T06:06:44.537169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T06:06:44.437755Z digest=sha256:20c279991f67afc6d811e2da542ecc065f366fa7ff546600881d9f412be3327a

Observation 49ca774d-f854-4622-abe0-fe67f999202a · outbound

This paper cites Less is more: Picking informative frames for video captioning,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Less is more: Picking informative frames for video captioning,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-14T06:06:44.440408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T06:06:44.440408Z digest=sha256:3f85fde28ca407a33fad75cf908997e40033c194bd499c65649450a363087552

Observation d8515a68-3524-4354-b627-1aa5f0693935 · outbound

This paper cites Watch, lis- ten, and describe: Globally and locally aligned cross- modal attentions for video captioning,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Watch, lis- ten, and describe: Globally and locally aligned cross- modal attentions for video captioning,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T06:06:44.851609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T06:06:44.443883Z digest=sha256:c0a927c59c7fd07e92d1d37ba877a60b89f23c50217e4895dfde4193f9438a63

Observation 0a684f27-ffac-40bf-9eeb-9eba672329c2 · outbound

This paper cites Multi-task video captioning with video and entailment generation,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Multi-task video captioning with video and entailment generation,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T06:06:44.842776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T06:06:44.446627Z digest=sha256:7f0a9bf27a97ab4ddc2175fb716bc53efa5be99f1a03a00a83e329ee342c7968

Observation 4bd62083-3c82-4e17-bbb5-7cfd5d85da14 · outbound

This paper cites Videobert: A joint model for video and language representation learning,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Videobert: A joint model for video and language representation learning,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T06:06:44.833573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T06:06:44.449222Z digest=sha256:cf1eb2f6f6acd68b9b5f197064ebec0ab1003089753ba84c200b3394237268d4

Observation 6673947a-22c5-4b3e-810f-861a7f7e52da · outbound

This paper cites BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-14T06:06:44.451806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T06:06:44.451806Z digest=sha256:2e4002dc1909e0e27ff069a7e9a1908bb804ed0be797e9d062b052d643e5b9cb

Observation 852d4678-8910-40d1-bfe8-8e40f8d737f0 · outbound

This paper cites Learning to compose topic-aware mixture of experts for zero-shot video captioning,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Learning to compose topic-aware mixture of experts for zero-shot video captioning,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T06:06:44.824579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T06:06:44.454636Z digest=sha256:2961bc7703d24ca6bec198eca7fd210d387c453c796858a9826c8dfade515dd4

Observation 15b7ff8e-88c6-457e-940a-8a2608881faf · outbound

This paper cites Spatio-temporal graph for video captioning with knowledge distillation,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Spatio-temporal graph for video captioning with knowledge distillation,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T06:06:44.816173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T06:06:44.457037Z digest=sha256:96ae6a2d41fa7ec74069a9aadf2f5c0e25b1a06cd36b6dee95b0616e965087b9

Observation c25fddfe-95c7-40f0-9ee0-0851630838b0 · outbound

This paper cites A learning algorithm for continually running fully recurrent neural networks,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling A learning algorithm for continually running fully recurrent neural networks,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-14T06:06:44.459692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T06:06:44.459692Z digest=sha256:3c047aee89e85ab9ab72f5ee7fd82ea316bf87c4e6618305eff45b37e4950af7

Observation 3ce565dc-561f-45aa-ae63-3a51ff2864ef · outbound

This paper cites How (not) to Train your Generative Model: Scheduled Sampling, Likelihood, Adversary?.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling How (not) to Train your Generative Model: Scheduled Sampling, Likelihood, Adversary?

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-14T06:06:44.462514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T06:06:44.462514Z digest=sha256:2ab7e89a3683529059acd747bd2004a6565bd625bbfb063fe398a873b23fb3c3

Observation 7bd7a385-4d82-4455-a3e1-c122c55be88a · outbound

This paper cites Professor forcing: A new algorithm for training recurrent networks,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Professor forcing: A new algorithm for training recurrent networks,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T06:06:44.804616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T06:06:44.465779Z digest=sha256:996fbb814b5c5d1f19e6b2777486b33f57a9269b4ea736629a2c3b6cac9f7961

Observation d5b35215-b44d-4e26-8667-029397fcd84a · outbound

This paper cites Object relational graph with teacher-recommended learning for video captioning,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Object relational graph with teacher-recommended learning for video captioning,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T06:06:44.796622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T06:06:44.468478Z digest=sha256:9f7147a3a5f3e048ba795c55f52f04f18c0964d5d66621bd1f35e604b11175e5

Observation 0742d209-7878-47c3-b576-dfda851ecb33 · outbound

This paper cites Simple statistical gradient-following al- gorithms for connectionist reinforcement learning,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Simple statistical gradient-following al- gorithms for connectionist reinforcement learning,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T06:06:44.788172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T06:06:44.471067Z digest=sha256:90ea498c7ae8e194482722779df49f41a80213fcceb1dde4dcee638308cf8cca

Observation 29169e2d-e5e2-4c8a-b759-ff6f39eb21b5 · outbound

This paper cites Finding structure in time,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Finding structure in time,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-14T06:06:44.473774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T06:06:44.473774Z digest=sha256:40265d10d44a9369f8b2dc1b459cb4816f9250bffd4cbad4ce84e066f2b60341

Observation 7e1e6673-c21f-4a57-af06-7ed13417c9d8 · outbound

This paper cites Long short-term memory,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Long short-term memory,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-14T06:06:44.476382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T06:06:44.476382Z digest=sha256:459cf697b06379b67be680bb424ab5bc80a90973309515eba3c0e1bdd3c7fcb1

Observation be883842-f70e-4905-99ab-7376c54ef632 · outbound

This paper cites Google’s Neural Machine Translation System: Bridging the Gap between Human and Machine Translation,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Google’s Neural Machine Translation System: Bridging the Gap between Human and Machine Translation,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T06:06:44.769694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T06:06:44.479107Z digest=sha256:0ae3cdae301d7639fc87411abf1c27a4c980088f09a4147236cc8f229c302bc7

Observation c9ed91e9-932d-40c4-a1b1-77058d54bb73 · outbound

This paper cites Youtube2text: Recognizing and describing arbitrary ac- tivities using semantic hierarchies and zero-shot recog- nition,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Youtube2text: Recognizing and describing arbitrary ac- tivities using semantic hierarchies and zero-shot recog- nition,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T06:06:44.761231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T06:06:44.483067Z digest=sha256:78fcd0b21aa34dd7cbcaad3ead7b5cc36a17e145142140722c560e0411044bd9

Observation 8c284eec-572c-4d33-873c-9c5288f84a8c · outbound

This paper cites Collecting highly parallel data for paraphrase evaluation,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Collecting highly parallel data for paraphrase evaluation,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T06:06:44.752786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T06:06:44.485659Z digest=sha256:fc3f536df31b859512d84fc4d8eee7f00860ba9b1c3bb8f5be04cfb872021f15

Observation ce02e931-d919-47aa-b30c-6a23d5efad9c · outbound

This paper cites MSR-VTT: A large video description dataset for bridging video and language,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling MSR-VTT: A large video description dataset for bridging video and language,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T06:06:44.744212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T06:06:44.488668Z digest=sha256:90ff455a62a032a4e17fdd319e1f86303731b1e19eda3e8b8a0d14436cab77ad

Observation 7898fb66-7880-4e2a-b8f3-61abe7fbb891 · outbound

This paper cites Aggregated residual transformations for deep neural networks,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Aggregated residual transformations for deep neural networks,

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-14T06:06:44.491259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T06:06:44.491259Z digest=sha256:dcc94f4be36bb31e8392474ae01b04971b0236284732569f77703565bc077f15

Observation a15bb6f7-ae9b-42c0-893d-c2d3b4aa3292 · outbound

This paper cites ECO: efficient convolutional network for online video understanding,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling ECO: efficient convolutional network for online video understanding,

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-14T06:06:44.494009Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T06:06:44.494009Z digest=sha256:b5d61712347678889947f43eaccca86e432d2bb08a1c051608beea72148616ff

Observation 19848132-ef73-4833-8e6b-cd6681b5d31a · outbound

This paper cites Sibnet: Sibling convolutional encoder for video captioning,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Sibnet: Sibling convolutional encoder for video captioning,

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-14T06:06:44.496761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T06:06:44.496761Z digest=sha256:bab96bbead4db90dd61eb960c2211a210ab2cec342ce01075b8bdfb2cf8c750e

Observation ab3864ee-2408-4cb9-a5d9-f08b58e07811 · outbound

This paper cites Multi-label classification: An overview,.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Multi-label classification: An overview,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T06:06:44.735290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-14T06:06:44.499410Z digest=sha256:ff4dc7b6b5b1ba19e7ed2d8c39da6a79ae560590c8696d9197804d2b84a96130

Observation 59132fb8-4aa9-4310-bee9-ac53d26ba833 · outbound

This paper cites Neural Machine Translation by Jointly Learning to Align and Translate.

A Semantics-Assisted Video Captioning Model Trained with Scheduled Sampling Neural Machine Translation by Jointly Learning to Align and Translate

Reference 2015

Resolution
unresolved
no resolver link, observed 2026-08-14T06:06:44.395375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T06:06:44.395375Z digest=sha256:b2d32eac622010195db7ee1c4df7e14941ec6e7b6d5b68dc8bc4b6d839bba440

Pith citing papers

No inbound Pith citation observations are available.