Pith. sign in

Paper Citation Record · LEDGER

Transferable Representation Learning in Vision-and-Language Navigation

As of 15 August 2026, this Paper Citation Record lists 46 of 46 outbound references and 0 inbound Pith citation observations for arXiv:1908.03409.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1908.03409 v2

Coverage vector

measured 46 of 46 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-14T14:19:11.407540Z

measured 46 of 46 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

46 of 46 outbound references displayed

  • verified exact1
  • verified fuzzy36
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4ca3a9ef-a873-45fe-87d3-6477a2f6686a · outbound

This paper cites On Evaluation of Embodied Navigation Agents.

Transferable Representation Learning in Vision-and-Language Navigation On Evaluation of Embodied Navigation Agents

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-14T14:19:11.119168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:19:11.119168Z digest=sha256:6c7220b46b694ed929e405ad6023d5962ee3869c0d384669f250f05919ab7080

Observation d66bc0d8-8957-4a79-9cab-3931c2707c47 · outbound

This paper cites Vision-and- Language Navigation: Interpreting visually-grounded navigation instructions in real environments.

Transferable Representation Learning in Vision-and-Language Navigation Vision-and- Language Navigation: Interpreting visually-grounded navigation instructions in real environments

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:19:12.471471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:19:11.126454Z digest=sha256:21d9756b33dfe8c975498ddf2bcde0a7b1406adb9a79ef54b6ee27517ee1bc89

Observation faa70d35-0050-45e0-b60c-b64025140b7c · outbound

This paper cites Antol, A.

Transferable Representation Learning in Vision-and-Language Navigation Antol, A

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:19:12.450905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:19:11.133722Z digest=sha256:3e082c2e12b51b3aedf106b83d33f937d61994068b28cb6b1726e77ab0cd6fe2

Observation 341838b6-d8bf-4161-9697-baeacd275943 · outbound

This paper cites A framework for behavioural cloning.

Transferable Representation Learning in Vision-and-Language Navigation A framework for behavioural cloning

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:19:12.433784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:19:11.139657Z digest=sha256:feca86fc73bd171a33141bbee73faf7e39ef77ed69c4a31432dabf6e6f3b55bb

Observation 3c21a7ec-8f2a-4a78-ae6a-c970bd309726 · outbound

This paper cites Pre- diction, cognition and the brain.

Transferable Representation Learning in Vision-and-Language Navigation Pre- diction, cognition and the brain

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:19:12.416315Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:19:11.146609Z digest=sha256:d932b701d54c2135ae62798fcce08375575b205c65fad009a06fd89d9afb7b16

Observation 4a9b3ebd-e9bb-4e38-b8b7-169898bcaa54 · outbound

This paper cites Matterport3D: Learning from RGB-D data in indoor environments.

Transferable Representation Learning in Vision-and-Language Navigation Matterport3D: Learning from RGB-D data in indoor environments

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:19:12.397683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:19:11.153427Z digest=sha256:9598890a61391b5a6fbe906165833196c4a3ce631f9431f588087f4a7267cb15

Observation 8685371d-43af-43de-9e4c-86891f969140 · outbound

This paper cites Fol- lowing formulaic map instructions in a street simu- lation environment.

Transferable Representation Learning in Vision-and-Language Navigation Fol- lowing formulaic map instructions in a street simu- lation environment

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:19:12.375342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:19:11.160600Z digest=sha256:874cfa3110146ed6e48cf0495c8d77db7980f6111b4fe901507297c97a098e50

Observation 670e568d-83dc-4ca6-8c15-1beffc4d5c88 · outbound

This paper cites Learning Transferable Policies for Monocular Reactive MAV Control.

Transferable Representation Learning in Vision-and-Language Navigation Learning Transferable Policies for Monocular Reactive MAV Control

Reference 8

Resolution
metadata mismatch
local_arxiv, observed 2026-08-14T14:19:11.606364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:19:11.166962Z digest=sha256:b1001be1f318310ea0a056fa1af10a051c1a17fabc39b07d7b68d9f113d78198

Observation 3e21dc95-b895-426c-bbab-56cd59ecfb7b · outbound

This paper cites Moura, Devi Parikh, and Dhruv Batra.

Transferable Representation Learning in Vision-and-Language Navigation Moura, Devi Parikh, and Dhruv Batra

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:19:12.351647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:19:11.174068Z digest=sha256:1b9d1da6c8ee097ea092f9fd65cb3969e17f41c679881a4ff70df59b38f49691

Observation 35e0a59a-86dd-4076-b338-1b3feefee9df · outbound

This paper cites Talk the Walk: Navigating New York City through Grounded Dialogue.

Transferable Representation Learning in Vision-and-Language Navigation Talk the Walk: Navigating New York City through Grounded Dialogue

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-14T14:19:11.179655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:19:11.179655Z digest=sha256:a310d0ee9ef7340639a6835df9e9514fbe72d9777faf311e60c3db3ecfab88a9

Observation ed7fa6fa-8428-446b-9280-2dffbf9a8f4e · outbound

This paper cites Optimal perceived timing: Integrating sensory information with dynamically updated expectations.

Transferable Representation Learning in Vision-and-Language Navigation Optimal perceived timing: Integrating sensory information with dynamically updated expectations

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:19:12.331171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:19:11.186018Z digest=sha256:06bb37fb2a35cbad71cb1225a6b2e6cda7b31f90d43edda64ac608c9df02e32f

Observation 3f152cc0-e59c-46cd-9336-b9ab3210b861 · outbound

This paper cites Donahue, L.

Transferable Representation Learning in Vision-and-Language Navigation Donahue, L

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:19:12.307427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:19:11.193586Z digest=sha256:8a66dba0260c6e90a9cdda66e7dd571b3fa7e23131da73734a0d2e44e976de41

Observation 1614c3ca-37e9-4d97-ba8a-d8a4eeb9f32a · outbound

This paper cites an unresolved cited work.

Transferable Representation Learning in Vision-and-Language Navigation Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-14T14:19:12.282670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:19:11.200210Z digest=sha256:62972e3e48f94daef2bf0220fffa1f97b06fc333d79c6562bbe29f18c6ef272b

Observation 60a95a36-8e3e-4354-86ec-2c58d5c5b8c4 · outbound

This paper cites Speaker-follower models for vision- and-language navigation.

Transferable Representation Learning in Vision-and-Language Navigation Speaker-follower models for vision- and-language navigation

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:19:12.262343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:19:11.208250Z digest=sha256:b783737b4997382ff05346e1ad61e795143f3b8797bba9e86d9486b5c9c0ef3f

Observation 28331094-57a3-47da-81f5-a17999152111 · outbound

This paper cites End-to-End Retrieval in Continuous Space.

Transferable Representation Learning in Vision-and-Language Navigation End-to-End Retrieval in Continuous Space

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-14T14:19:11.214599Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:19:11.214599Z digest=sha256:2a3517da265881ba896b9a289244fc58db43c2d14ddf2b1a68dd3d845ea7632e

Observation 1b861484-879b-464d-8fa7-29174e7b1386 · outbound

This paper cites Girshick, Jeff Donahue, Trevor Darrell, and Jitendra Malik.

Transferable Representation Learning in Vision-and-Language Navigation Girshick, Jeff Donahue, Trevor Darrell, and Jitendra Malik

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:19:12.242161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:19:11.221766Z digest=sha256:74c7f91153ff053eed852dbbef25e68ab3ae01282e8603de119ee31c7f0907c5

Observation 9cf15d99-9a0a-4597-8a6b-72b8acf382ec · outbound

This paper cites Deep residual learning for image recognition.

Transferable Representation Learning in Vision-and-Language Navigation Deep residual learning for image recognition

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:19:12.224239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:19:11.227571Z digest=sha256:21d7e8e1e3dd9e10ecfa187dd2cd20ebfc19cddb9e84dddc9764197d91be6e52

Observation 8a58c6b0-cd30-4cd2-9ffd-f7268985f628 · outbound

This paper cites Howard, Nicholas Roy, Anthony Stentz, and Matthew R.

Transferable Representation Learning in Vision-and-Language Navigation Howard, Nicholas Roy, Anthony Stentz, and Matthew R

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:19:12.200570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:19:11.233120Z digest=sha256:7a3a773d947b2a3813998fdb2f5683107e79598e13e897c7bc66100188e0ae65

Observation e96c25a2-f218-439a-955e-b3c3e3896ebf · outbound

This paper cites Learning To Follow Directions in Street View.

Transferable Representation Learning in Vision-and-Language Navigation Learning To Follow Directions in Street View

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-08-14T14:19:11.536817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:19:11.239882Z digest=sha256:13519c8ff3f9bc84666b43360c93be8f16c535f591d0ebbfc18114f419eabd4f

Observation 805b5679-27db-4032-8fef-eaa01907d0d5 · outbound

This paper cites Long short- term memory.

Transferable Representation Learning in Vision-and-Language Navigation Long short- term memory

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:19:12.179454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:19:11.245632Z digest=sha256:033911d759b6426673fe03361b8f7de533ceb6fe344101343a5be7ad1190cd45

Observation 58d8c99a-14a4-4376-856c-866ac19d34fe · outbound

This paper cites Segmentation from natural language expressions.

Transferable Representation Learning in Vision-and-Language Navigation Segmentation from natural language expressions

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:19:12.162181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:19:11.253116Z digest=sha256:7bf9dfba41a0b19946e85263be0ba809a76f7f5fcd88ab5e87df66a795aea364

Observation 3b75a289-9de2-48e8-9e21-68f81480624c · outbound

This paper cites Multi-modal discriminative model for vision-and-language navigation.

Transferable Representation Learning in Vision-and-Language Navigation Multi-modal discriminative model for vision-and-language navigation

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:19:12.141081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:19:11.258985Z digest=sha256:5e130c1607d056988d3581a33bb79ca663b983267182b87cb5c1f7f38d043059

Observation 4d6f2199-80b8-43b3-a450-5c62fd86f94a · outbound

This paper cites Deep visual-semantic alignments for generating image descriptions.

Transferable Representation Learning in Vision-and-Language Navigation Deep visual-semantic alignments for generating image descriptions

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:19:12.123816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:19:11.265041Z digest=sha256:733b9a6ec8d275613fba2970d028ecd4fb876d04dd8953a9dc76a1ce85a53716

Observation 1452f6ec-fa82-474d-aff0-775007d17ada · outbound

This paper cites Self-monitoring navigation agent via auxiliary progress estimation.

Transferable Representation Learning in Vision-and-Language Navigation Self-monitoring navigation agent via auxiliary progress estimation

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:19:12.104517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:19:11.271858Z digest=sha256:09c36b79a8192658e22567af5349461134bd20c201310b375403b126dd7bb9da

Observation e358715e-22f6-433a-b221-d529ec1aa216 · outbound

This paper cites The regretful agent: Heuristic- aided navigation through progress estimation.

Transferable Representation Learning in Vision-and-Language Navigation The regretful agent: Heuristic- aided navigation through progress estimation

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:19:12.082224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:19:11.276925Z digest=sha256:716caac3e7557879187f64eabd8b21296a752f5e4fc73ead0982f86ab65a3655

Observation 9760fe42-5c10-47d3-971f-0f5899d66a71 · outbound

This paper cites Walk the talk: Connecting language, knowl- edge, action in route instructions.

Transferable Representation Learning in Vision-and-Language Navigation Walk the talk: Connecting language, knowl- edge, action in route instructions

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:19:12.060487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:19:11.283362Z digest=sha256:096ed5dd7eebabc2393cec78d79d56ec729aa1a3e713a06d3839b40718d661f8

Observation 50e99ed7-a15f-4679-bd94-1ff291b03a35 · outbound

This paper cites Yuille, and Kevin Murphy.

Transferable Representation Learning in Vision-and-Language Navigation Yuille, and Kevin Murphy

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:19:12.040851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:19:11.288760Z digest=sha256:a3555f869387e2b02fb43c643cf1d821802ca0187401a9c1350f6207dd56d381

Observation 8e186d58-e3d8-4e78-a584-92ddcc3345c3 · outbound

This paper cites Grounded language learning: Where robotics and NLP meet.

Transferable Representation Learning in Vision-and-Language Navigation Grounded language learning: Where robotics and NLP meet

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:19:12.023588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:19:11.293814Z digest=sha256:bacd18108c538b3a26d15265e4f49d0c89042427ca5cacf215d96a73159f6342

Observation 475ffa7a-c893-4ff8-b491-190bbf236cdb · outbound

This paper cites Learning to nav- igate in cities without a map.

Transferable Representation Learning in Vision-and-Language Navigation Learning to nav- igate in cities without a map

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:19:12.001906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:19:11.299826Z digest=sha256:33a2bdd2896e9224ca9cf41f2f7330b9d0133736b05a92fd2060ce2c58cf91e5

Observation a0e41960-6f16-42a4-97df-dbf7d3ba4bf1 · outbound

This paper cites GloVe: Global vectors for word represen- tation.

Transferable Representation Learning in Vision-and-Language Navigation GloVe: Global vectors for word represen- tation

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:19:11.967093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:19:11.304689Z digest=sha256:183ea6f241b537ec14bb7f151945c1cfd185da01bc5cc4245bdd115025e4d1c5

Observation 90192a8a-397b-49c3-8809-c9283596ee65 · outbound

This paper cites Berg, and Li Fei-Fei.

Transferable Representation Learning in Vision-and-Language Navigation Berg, and Li Fei-Fei

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:19:11.943264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:19:11.309896Z digest=sha256:a1805b89c25ec12affcca03eceed80facf80025a28746fa24b57c62858cdaf96

Observation bf93a4c8-8411-4a43-8f6a-bc0844882ee8 · outbound

This paper cites an unresolved cited work.

Transferable Representation Learning in Vision-and-Language Navigation Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-14T14:19:11.919833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:19:11.315769Z digest=sha256:db54f60654ab142d0f5bdb367730932a52de599ee6b404c5e58ac5f5ae07c328

Observation 1e0c0de9-8e88-4c90-8513-39941aa2cbeb · outbound

This paper cites A survey of available corpora for building data-driven dialogue systems: The journal version.

Transferable Representation Learning in Vision-and-Language Navigation A survey of available corpora for building data-driven dialogue systems: The journal version

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:19:11.891699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:19:11.320991Z digest=sha256:bfd666da35cba8e408811b0c0af5c7c2aa57722824d623802dbb5d3cd59e45c9

Observation 2bb91e95-dcfe-413c-bb98-0b9a6267fe2d · outbound

This paper cites Learning to navigate unseen environments: Back translation with environmental dropout.

Transferable Representation Learning in Vision-and-Language Navigation Learning to navigate unseen environments: Back translation with environmental dropout

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:19:11.869767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:19:11.326538Z digest=sha256:d4cd26d314908476ca1f345aee1c50c65b8856c097ba12d929a3c05c16f37af8

Observation 339a8912-e055-4e5a-90e4-4168db9edcae · outbound

This paper cites Visual represen- tations for semantic target driven navigation.

Transferable Representation Learning in Vision-and-Language Navigation Visual represen- tations for semantic target driven navigation

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:19:11.845846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:19:11.331732Z digest=sha256:caee3d98e634a6551b82e187e2409b22c96e0141598a7bb0dcfd3f44723ae28c

Observation 544fac11-47d1-473e-acbd-dbbf55c5285f · outbound

This paper cites Representation Learning with Contrastive Predictive Coding.

Transferable Representation Learning in Vision-and-Language Navigation Representation Learning with Contrastive Predictive Coding

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-14T14:19:11.337497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:19:11.337497Z digest=sha256:3fd38ad2523c0ec4ad93477e9f7690781daeb1be44b89f2b13b9e8bd8226d1b6

Observation 033f22e3-9ceb-4a3a-96f6-e8a5b7685841 · outbound

This paper cites Show and tell: A neural image caption generator.

Transferable Representation Learning in Vision-and-Language Navigation Show and tell: A neural image caption generator

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:19:11.826135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:19:11.344785Z digest=sha256:d718e9b04b2f6b5557ae810f49951704f3d10a0f5b56fa1153bc5ec51dc8c002

Observation f643b55d-115a-4579-9068-7af0d9fc9592 · outbound

This paper cites Video captioning via hierar- chical reinforcement learning.

Transferable Representation Learning in Vision-and-Language Navigation Video captioning via hierar- chical reinforcement learning

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:19:11.805430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:19:11.351938Z digest=sha256:9e0cdd8f96cdacec7f3d82c5968a1a0929d014dbc0b76f254d8507a27c20b3b5

Observation 80eefb52-59c7-4a15-8bac-d8dfbb4092ae · outbound

This paper cites Reinforced Cross-Modal Matching and Self-Supervised Imitation Learning for Vision-Language Navigation.

Transferable Representation Learning in Vision-and-Language Navigation Reinforced Cross-Modal Matching and Self-Supervised Imitation Learning for Vision-Language Navigation

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-14T14:19:11.359505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:19:11.359505Z digest=sha256:7bd7d9bae7160d52a0fd5ea062040b4ab85df236562da37290e407e8145625ec

Observation c9486b45-6581-403c-b3dc-9d2cda7a1abb · outbound

This paper cites Look before you leap: Bridg- ing model-free and model-based reinforcement learn- ing for planned-ahead vision-and-language navigation.

Transferable Representation Learning in Vision-and-Language Navigation Look before you leap: Bridg- ing model-free and model-based reinforcement learn- ing for planned-ahead vision-and-language navigation

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:19:11.779931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:19:11.366180Z digest=sha256:4a446033c11d4049425c1169af26b34588e75234029a6aa7d4d494615d0e6449

Observation 3c844af7-b456-47ab-93b7-48fec306f542 · outbound

This paper cites Williams.

Transferable Representation Learning in Vision-and-Language Navigation Williams

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:19:11.755950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:19:11.372679Z digest=sha256:3e956bf067569fdf92709495b2acf057b20bf832bd72123fbe938e780dc111c6

Observation 49345936-bec0-49c5-853b-f808c47ba4b6 · outbound

This paper cites Courville, Ruslan Salakhutdinov, Richard S.

Transferable Representation Learning in Vision-and-Language Navigation Courville, Ruslan Salakhutdinov, Richard S

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:19:11.726654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:19:11.379857Z digest=sha256:e71fe11e23610ece0a4c7d5f26c2d062a92b4cd3d79e3ef555e3b39093c5a8d3

Observation 9553e5d2-ba09-4a6c-9bbf-6a047de59cdf · outbound

This paper cites an unresolved cited work.

Transferable Representation Learning in Vision-and-Language Navigation Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-14T14:19:11.708296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:19:11.385627Z digest=sha256:3059160f8f2b09963586a5f74159ec23353ba0bb0ce53b6154449efb3bb87984

Observation 27ab8297-92be-493d-82eb-b4ab2aa7f09e · outbound

This paper cites Video paragraph captioning using hierarchical recurrent neural networks.

Transferable Representation Learning in Vision-and-Language Navigation Video paragraph captioning using hierarchical recurrent neural networks

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:19:11.690262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:19:11.393470Z digest=sha256:3c2b33f2492a10ee9628c67abd56729fe82527b992b7f4d845ce9620f8c04811

Observation 3d2baa88-dbf4-46a3-b6ab-5aa573943c44 · outbound

This paper cites Zeiler and Rob Fergus.

Transferable Representation Learning in Vision-and-Language Navigation Zeiler and Rob Fergus

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:19:11.668772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:19:11.399683Z digest=sha256:8677220c88bbec9fcd0a656ad3bf7bd9eaf21b3c60d25499155414af60238fbb

Observation 5223a639-628c-4df9-adea-5a34e3bb547f · outbound

This paper cites Lim, Abhinav Gupta, Li Fei-Fei, and Ali Farhadi.

Transferable Representation Learning in Vision-and-Language Navigation Lim, Abhinav Gupta, Li Fei-Fei, and Ali Farhadi

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:19:11.646883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:19:11.407540Z digest=sha256:71eadfc95daefe389dec278acd3cf505c6594c89e4db3878ca7c1203f0682513

Pith citing papers

No inbound Pith citation observations are available.