Pith. sign in

Paper Citation Record · LEDGER

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization

As of 7 August 2026, this Paper Citation Record lists 77 of 77 outbound references and 0 inbound Pith citation observations for arXiv:2507.10894.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.10894 v1

Coverage vector

measured 77 of 77 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:26:43.437116Z

measured 77 of 77 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

77 of 77 outbound references displayed

  • verified exact0
  • verified fuzzy62
  • unresolved14
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 067f3fef-cfd5-4b9b-914a-1a43d89c144b · outbound

This paper cites A survey of embodied ai: From simulators to research tasks,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization A survey of embodied ai: From simulators to research tasks,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.494005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:37.385842Z digest=sha256:cced827acda33a78048cad76605aee272612fa15188584c1deca4a44b3f57324

Observation a740e2b8-e30e-4bd1-8417-2c33ac260644 · outbound

This paper cites Behavior-1k: A benchmark for embodied ai with 1,000 everyday activities and realistic simulation,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Behavior-1k: A benchmark for embodied ai with 1,000 everyday activities and realistic simulation,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.485004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:37.447109Z digest=sha256:6e959d5049494984fefb6ac41cf0e948554f534ae775a7b5a73528171ed4c4cf

Observation c84f1d60-1240-4862-8b95-285523181373 · outbound

This paper cites Aligning Cyber Space with Physical World: A Comprehensive Survey on Embodied AI.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Aligning Cyber Space with Physical World: A Comprehensive Survey on Embodied AI

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T17:26:37.507750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:26:37.507750Z digest=sha256:293a216ee1b3667428cac977cc6ca20a429e3eb18c4f0c7672f9a1bb56baa815

Observation d95ca566-acea-41bd-a67c-56d81f7d35c0 · outbound

This paper cites Vision-and-language navigation: Interpreting visually-grounded navigation instructions in real environments,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Vision-and-language navigation: Interpreting visually-grounded navigation instructions in real environments,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.476143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:37.597645Z digest=sha256:32f204191883ae15e6d62d684448474758b0bb3fbd0954acfcd98350314d947b

Observation 08615390-332c-426b-a2cc-25b031ee047f · outbound

This paper cites Room-across-room: Multilingual vision-and-language navigation with dense spatiotemporal grounding,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Room-across-room: Multilingual vision-and-language navigation with dense spatiotemporal grounding,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.466733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:37.642304Z digest=sha256:2be780e5b9577b45d2e236b808500ee33b4b7c9de1e2208275d4e76512ef6298

Observation 320f098a-a661-45fd-9ce5-b5d5716e3908 · outbound

This paper cites Talk2nav: Long-range vision-and-language navigation with dual attention and spatial memory,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Talk2nav: Long-range vision-and-language navigation with dual attention and spatial memory,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.457258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:37.732340Z digest=sha256:945853cbf2e649352adb0dc973e2c254a510f484853655f0acd9c59b4e15d406

Observation 3fd76607-44a9-47e8-b314-11137f882b18 · outbound

This paper cites Reverie: Remote embodied visual referring expression in real indoor environments,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Reverie: Remote embodied visual referring expression in real indoor environments,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.447485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:37.821615Z digest=sha256:e7cc80b6b8509973a5a2875524de360db361b61a7cd33571ac10bb48441df244

Observation 79b91aaf-2067-483d-9252-0e4951e56f13 · outbound

This paper cites Speaker-follower models for vision-and-language nav- igation,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Speaker-follower models for vision-and-language nav- igation,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.438579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:37.926692Z digest=sha256:9d8fabf3b2df365a39a2892dfb65b7a9674b1375890d68f1803a1e86fb1707e9

Observation e7b3cd13-b28a-4c39-ac02-363b48101623 · outbound

This paper cites Learning to navigate unseen environments: Back trans- lation with environmental dropout,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Learning to navigate unseen environments: Back trans- lation with environmental dropout,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.428820Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:38.019873Z digest=sha256:8ce36d6e7a8325df37915b2cb9c2712617be3c0fc15cf3ab28ea46e8d7298af3

Observation 2f3b20a6-0499-4de2-ae3e-cef0c369a98f · outbound

This paper cites Visual landmark selection for generating grounded and interpretable navigation instructions,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Visual landmark selection for generating grounded and interpretable navigation instructions,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.420060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:38.083529Z digest=sha256:65c62f2272119df5236e4bad7f19539255be98157b4714041fc03d25a13e0ee9

Observation 86a5f0cd-6322-460f-b3e6-7fc111e24d61 · outbound

This paper cites Crossmap transformer: A crossmodal masked path transformer using double back-translation for vision-and-language navigation,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Crossmap transformer: A crossmodal masked path transformer using double back-translation for vision-and-language navigation,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.410397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:38.155998Z digest=sha256:c7885d2cb5d39cacb77aaff98bda9e4afb2469354e9a7cd311383c390229e142

Observation f84da2ac-b995-4950-8876-1ebac4b38a3f · outbound

This paper cites Improved speaker and navigator for vision-and-language navigation,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Improved speaker and navigator for vision-and-language navigation,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.401744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:38.239153Z digest=sha256:aaa2704749bb1ae556d4f590e28430b50d04cba45dd12716179748c47479f7d0

Observation 0c5fa32d-cc31-4fe4-a8e0-f6298bde3b1b · outbound

This paper cites Res-sts: Referring expression speaker via self-training with scorer for goal-oriented vision-language navigation,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Res-sts: Referring expression speaker via self-training with scorer for goal-oriented vision-language navigation,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.392673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:38.295300Z digest=sha256:e8a69f7c94e4451e5e6964db77bde7d60095e15b1cd1ee389d59f4027a800afa

Observation a0c9bbcc-6905-443c-a24a-1dcc90d9ed0e · outbound

This paper cites Touchdown: Natural language navigation and spatial reasoning in visual street environments,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Touchdown: Natural language navigation and spatial reasoning in visual street environments,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.381772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:38.367875Z digest=sha256:372670b9ec47a881078c169a608cd5b39425c84388ec2ca7fdefbababf948429

Observation 6a7ef5d2-e715-423b-8567-3ba1fa884db6 · outbound

This paper cites Vision-language navigation with self-supervised auxil- iary reasoning tasks,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Vision-language navigation with self-supervised auxil- iary reasoning tasks,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.372815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:38.452079Z digest=sha256:33750842cd0369740d476b2c61a04810d34717599068959d36997c10531315fb

Observation 1eb139da-2ee3-49ec-85ea-c8e94ab8b33b · outbound

This paper cites Towards navigation by reasoning over spatial config- urations,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Towards navigation by reasoning over spatial config- urations,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.363650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:38.554838Z digest=sha256:8fed2afdd26ed9cacc4b00975f718b5c27d7baf1eb6dfb24837f2698aec366ad

Observation 34f7b84e-9f1c-4bf3-ac1d-e560385cd547 · outbound

This paper cites Language-guided navigation via cross-modal ground- ing and alternate adversarial learning,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Language-guided navigation via cross-modal ground- ing and alternate adversarial learning,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.354831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:38.609153Z digest=sha256:161e0dc7f85d64770ede7dca397c04312e16d56595eed0c2b510a8dcd8e18df1

Observation 21b5f186-52d0-4f3a-990a-1bcc4428e6e2 · outbound

This paper cites A dual semantic-aware recurrent global-adaptive net- work for vision-and-language navigation,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization A dual semantic-aware recurrent global-adaptive net- work for vision-and-language navigation,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.345561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:38.716252Z digest=sha256:a6530a29cb102e4f6a17b96204bf9ec2cd612a95a636d7850dc533b9bf32733d

Observation 007a28bd-c66b-4057-9f25-ab84b9db3c2a · outbound

This paper cites Vision-and-language navigation via causal learning,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Vision-and-language navigation via causal learning,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.335692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:38.796758Z digest=sha256:cc87cd8cedcf5c24c9592af74be4809c6b44469a285292489128824018a28ee1

Observation bcacd41b-15ed-4c7e-8901-c55c2688cf0f · outbound

This paper cites Waypoint models for instruction-guided navigation in continuous environments,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Waypoint models for instruction-guided navigation in continuous environments,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.327313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:38.871862Z digest=sha256:ba1b360f9418f44f354bff62e4d06dc96faf6d0bf7d9c449e8a6f09851d06cc4

Observation eec50d35-4578-4d91-bbad-ed303647c806 · outbound

This paper cites Bridging the gap between learning in discrete and continuous environments for vision-and-language navigation,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Bridging the gap between learning in discrete and continuous environments for vision-and-language navigation,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.317948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:38.964356Z digest=sha256:eb6c1b965b8af425a6ee5ace63760357a4c94b1ff128db731d90ec105758e462

Observation fcf4d2f5-53ea-4e97-923c-2184a576ffa8 · outbound

This paper cites Instruction-aligned hierarchical waypoint planner for vision-and-language navigation in continuous environments,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Instruction-aligned hierarchical waypoint planner for vision-and-language navigation in continuous environments,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.308510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:39.031354Z digest=sha256:a054cac1a14a9e0ad1cecbbd63b0b0347c69f43bbcacdb829d49f832e9c992af

Observation 8ddd9a38-05e6-472b-add2-8ea5728405b9 · outbound

This paper cites Improving vision-and-language navigation with image-text pairs from the web,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Improving vision-and-language navigation with image-text pairs from the web,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.299440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:39.088275Z digest=sha256:f0b344f7a2de0baa454c3b4edfe76d39813575eba11fca8b01adf43ec1c1d66c

Observation b6dd2d05-fc7a-4002-85c8-3741ea750657 · outbound

This paper cites Vln bert: A recurrent vision-and-language bert for navigation,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Vln bert: A recurrent vision-and-language bert for navigation,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.290467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:39.181868Z digest=sha256:47df58d1601337a8f3f095bfbe9db726e8c351965a37e3323ded1d36134851d8

Observation 8f05ce81-8683-43c7-bc49-de2b6baf70ba · outbound

This paper cites Think global, act local: Dual-scale graph transformer for vision-and-language navigation,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Think global, act local: Dual-scale graph transformer for vision-and-language navigation,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.281913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:39.255109Z digest=sha256:b3cfdad51bf2b1329f92014f2f5c2eb9457083b61f88b2baced9cb215cda728e

Observation fcbc33e1-5ace-47b3-81b0-fc401782c2ed · outbound

This paper cites Multimodal evolutionary encoder for continuous vision- language navigation,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Multimodal evolutionary encoder for continuous vision- language navigation,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.273145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:39.371677Z digest=sha256:35218f0557a98e58c1c1fdcd0cf74f6116bae47daaea38e1adb48f1a52bbfa0b

Observation 7884320a-c9f4-469a-9b2e-152643a064d4 · outbound

This paper cites Lana: A language-capable navigator for instruction following and generation,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Lana: A language-capable navigator for instruction following and generation,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.264496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:39.423682Z digest=sha256:6c166d5dbd1753173048d477d613c49aff14cbc88d5e7c89a3ce1cd3e7de2b14

Observation 630df405-98ab-4d04-88f1-866f562a9a4d · outbound

This paper cites Pasts: Progress-aware spatio-temporal transformer speaker for vision-and-language navigation,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Pasts: Progress-aware spatio-temporal transformer speaker for vision-and-language navigation,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.255999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:39.509361Z digest=sha256:31d1e387a36146207517c153bf6b5714cf6227eaeef8387587540f9715e10ef9

Observation 8a48d74d-8bc1-4388-ad64-42f79fd1f07b · outbound

This paper cites Envedit: Environment editing for vision-and-language navigation,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Envedit: Environment editing for vision-and-language navigation,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.245998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:39.600011Z digest=sha256:24323cc8b47cdf03b4eda644369fe1f98a93fa1ff730ecd1cff9f5604e9ce404

Observation eb44402d-c6aa-4704-8cc9-b4d15dea0189 · outbound

This paper cites Less is more: Generating grounded navigation instruc- tions from landmarks,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Less is more: Generating grounded navigation instruc- tions from landmarks,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.236709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:39.657739Z digest=sha256:8910543b2dcf74ce07d16ce70bae7d3c2ade95d55a575e77952943bd51116d14

Observation 81ba1bbc-87b9-4838-a7af-147d1a0540c9 · outbound

This paper cites Learning vision-and-language navigation from youtube videos,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Learning vision-and-language navigation from youtube videos,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.226926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:39.749105Z digest=sha256:3d1bb236710b554d1ac5be330865a69258dbed2c5b2012824ce4ce1a5d6283b5

Observation 61f4f532-c790-4ee5-9f6a-486991bddce8 · outbound

This paper cites Video captioning: A comparative review of where we are and which could be the route,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Video captioning: A comparative review of where we are and which could be the route,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.217843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:39.825425Z digest=sha256:ab5536271019d0061827b21ec45d9deb25ef82535cbd353de3fad3c137336fe3

Observation 22ef3e08-ccdf-4734-8ce8-8e62679309e9 · outbound

This paper cites A review of deep learning for video captioning,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization A review of deep learning for video captioning,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.209487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:39.922378Z digest=sha256:059bd7e1da45acf5d59c426c2a0ff0a1a27d5712cad0c792947a8ae9b0aa93cd

Observation d2458f56-b1f6-4903-8b55-56e5a6224c4d · outbound

This paper cites A survey of video datasets for grounded event understanding,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization A survey of video datasets for grounded event understanding,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.201062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:40.001084Z digest=sha256:548ee10029d44c8eadbdfda7ab6d4f911203ccc019ea58e3f0049d655ca67e91

Observation 77cd7376-8182-465e-9944-167429902d73 · outbound

This paper cites Dual-stream recurrent neural network for video caption- ing,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Dual-stream recurrent neural network for video caption- ing,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.191684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:40.140493Z digest=sha256:15e5c8a253e49e68589f33da61bb4e678cb077e4b3b5011be1db79740c1105e6

Observation 8d03ee06-26ed-4750-b4cc-300ca7c23b15 · outbound

This paper cites Video captioning using global-local representation,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Video captioning using global-local representation,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.182805Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:40.207516Z digest=sha256:1adcade24bf114c301a448a4bc3ff66a56f995811fcaff52b10ce2a4f4668524

Observation 0458ea46-934a-4639-8938-dc4f54258e01 · outbound

This paper cites Evcap: Element-aware video captioning,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Evcap: Element-aware video captioning,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.173422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:40.274303Z digest=sha256:96e3bba3eaf8e72f03b8aa18726aac1424faeec7efb6e6306714b828781951e9

Observation b26fa6d3-8e11-4c3e-9537-a920ae3bceae · outbound

This paper cites InternVideo2.5: Empowering Video MLLMs with Long and Rich Context Modeling.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization InternVideo2.5: Empowering Video MLLMs with Long and Rich Context Modeling

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T17:26:40.342191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:26:40.342191Z digest=sha256:ae414961e2dcd1051b8f57abed0f33ba829a70bc2bab476392ce3574d3c0b13e

Observation 6aae20aa-65e1-4b81-8d3c-1412558d63f6 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T17:26:40.424787Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:26:40.424787Z digest=sha256:ce724ef15697b0e412506bf4f1d13692db1f0697878b7d3a71cc867f9a12fc18

Observation 5a0f4ef5-bca5-4426-8e87-71b62743ce9c · outbound

This paper cites Msr-vtt: A large video description dataset for bridging video and language,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Msr-vtt: A large video description dataset for bridging video and language,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.164482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:40.514568Z digest=sha256:139c83c23c62b99d935379858ed594aa082fd19f9e197112bfe87089cedb989e

Observation 0098bd94-98bd-4d4f-b0c4-c85abfd7fd19 · outbound

This paper cites Frozen in time: A joint video and image encoder for end-to-end retrieval,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Frozen in time: A joint video and image encoder for end-to-end retrieval,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.154235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:40.611316Z digest=sha256:271914b647640a8d265f16b91e450c955eba74768d20f7fa5afbefe6be6c0dce

Observation 91305db8-ed73-4a43-ae18-476114721c4e · outbound

This paper cites Beyond the nav-graph: Vision-and-language navigation in continuous environments,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Beyond the nav-graph: Vision-and-language navigation in continuous environments,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.144410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:40.722428Z digest=sha256:766317eb6c77f825212abed8a9090743d4ea8f6fdc4f77ba5957b0f3f0b735fc

Observation 990bc053-7431-47a2-ba82-b380648b8fa7 · outbound

This paper cites Deep residual learning for image recognition,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Deep residual learning for image recognition,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.136022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:40.776309Z digest=sha256:8e09074e9a732264c8cb532bc95e33003f5fb6f2528893982be512ab68957a3c

Observation 1fffbdb2-5fb0-4ea9-802f-e52533dc242c · outbound

This paper cites Object recognition from local scale-invariant features,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Object recognition from local scale-invariant features,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.127067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:40.864850Z digest=sha256:5f1f3058815697856940c23477815acf438a61fe8cc710beb334e67e14f3a3e3

Observation 2debacc8-7b06-45ae-b108-0611d651b798 · outbound

This paper cites Fast approximate nearest neighbors with au- tomatic algorithm configuration,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Fast approximate nearest neighbors with au- tomatic algorithm configuration,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.115716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:40.960158Z digest=sha256:e43a907647b19b0deaa64ec2904b072acedb60bbb8275baf93e2033764fdc96e

Observation d8d7c4b6-bd4f-43ac-a3c4-e7d6a716bff1 · outbound

This paper cites Revisiting weakly supervised pre-training of visual perception models,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Revisiting weakly supervised pre-training of visual perception models,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.105728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:41.044904Z digest=sha256:d1f47ca8eef0b7fed9340e75d84f1a32ee6681eae3d718c502ada0fb52badff9

Observation 23f5e418-644f-47af-af40-1efbbb7e09b8 · outbound

This paper cites BLIP-2: Bootstrapping language-image pre-training with frozen image encoders and large language models,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization BLIP-2: Bootstrapping language-image pre-training with frozen image encoders and large language models,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.096238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:41.136000Z digest=sha256:e60e22af653a60f922b8a6e3fa8031812cacda9c5be92d5a97703806bb02a69e

Observation c4b53ea2-908e-4d6e-ac65-f4e5e6974249 · outbound

This paper cites End-to-end object detection with transformers,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization End-to-end object detection with transformers,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.085885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:41.246340Z digest=sha256:e671a242c2b979664b3aa094976385d0a9c1af3618d4d57d16aa4153a774c6ec

Observation 14a3a940-618e-426e-b011-b094ce65ca34 · outbound

This paper cites Visual instruction tuning,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Visual instruction tuning,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.077443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:41.321346Z digest=sha256:8978bc010f8eb0adfbce090c27b9dae8004eeee814677ef1a88b4f489398389f

Observation 27eb50c0-19e7-4ef8-b2cd-281d96d8e02f · outbound

This paper cites Bleu: a method for automatic evaluation of machine translation,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Bleu: a method for automatic evaluation of machine translation,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.067482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:41.381224Z digest=sha256:1b6feb7734d31b2fd4fab9a0c036f136acd68f53e57f167337ba8da404305e5c

Observation 7e4715a2-04d0-41fb-abd4-b97a82fd66a9 · outbound

This paper cites Learning transferable visual models from natural language supervision,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Learning transferable visual models from natural language supervision,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.057996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:41.520663Z digest=sha256:b86c839d65269680beb470badf5fc8242c36d35f20b42ee5b9a5d4ac7e0d195d

Observation fcc4dadf-e352-4213-a471-636f6c12b421 · outbound

This paper cites Cutting the gordian knot: The moving-average type–token ratio (mattr),.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Cutting the gordian knot: The moving-average type–token ratio (mattr),

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T17:26:41.645320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:26:41.645320Z digest=sha256:69b49aa318865fe1629c902abef46786875bad4ca1a04ecff6dd4cd5473cc03c

Observation 9d0a2613-97d5-4adc-93a8-0ae22ebcc604 · outbound

This paper cites Locally typical sampling,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Locally typical sampling,

Reference 53

Resolution
malformed identifier
no resolver link, observed 2026-08-06T17:26:41.755589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:26:41.755589Z digest=sha256:db61a80a3b7ec8361196b0a20be45cb6dc920199aa4b52bf7556c02ce7417260

Observation f0e708ec-b3a9-45e2-ae65-c66232e35414 · outbound

This paper cites Texygen: A benchmarking platform for text generation models,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Texygen: A benchmarking platform for text generation models,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.048817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:41.876239Z digest=sha256:0da4c557c943f9cfa305be6bca35ab09259ee9442f4b60b4aa870682065e5cea

Observation 15627557-607d-42bc-9114-4b9486b1ce47 · outbound

This paper cites Standardizing the measurement of text diversity: A tool and a comparative analysis of scores,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Standardizing the measurement of text diversity: A tool and a comparative analysis of scores,

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T17:26:42.017758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:26:42.017758Z digest=sha256:31d5c8c83ee3567fa88ee949b3d0a1c1f7cee0b3e3583a05f8da110a30e7c046

Observation 8c2ba964-e269-4872-b6ac-2d45b22fb411 · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization DINOv2: Learning Robust Visual Features without Supervision

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T17:26:42.152760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:26:42.152760Z digest=sha256:365c1fc0d58e04cfdfe4eecf839527cae85bd98714bc23035ab3c977c5b7d625

Observation a6530019-9259-4309-80ff-38276acef15a · outbound

This paper cites Masked autoencoders are scalable vision learners,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Masked autoencoders are scalable vision learners,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.038618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:42.420639Z digest=sha256:752037df3f90febeb8268b6074019130ed6d8c3c3b26c6c37321674f71d3c88a

Observation 897009f2-c011-48d7-b139-16b23f4927cb · outbound

This paper cites Places: A 10 million image database for scene recognition,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Places: A 10 million image database for scene recognition,

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.029351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:42.589056Z digest=sha256:1dbdaf0c21f7f378ddfd392c514d530f91b3a61cb931b7d781ccf89895aac8a0

Observation 8f2f6ac3-43d4-4d5c-adb6-cb66239809e6 · outbound

This paper cites Llava-next: Improved reasoning, ocr, and world knowledge,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Llava-next: Improved reasoning, ocr, and world knowledge,

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.020039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:42.755630Z digest=sha256:80cf390db235036909e9d2063331b1241d129a217a2f2e8354d2b38826a538fb

Observation c6cccb69-2e66-428c-a754-1c8b3e39b82d · outbound

This paper cites Qwen2.5-VL Technical Report.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Qwen2.5-VL Technical Report

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T17:26:42.870226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:26:42.870226Z digest=sha256:12bef7b534586f4636268f81ec2adc89e905f9250bb2c3265c9db0e55fcb0864

Observation 4452d2f4-3dbb-4912-beb9-97b5addd3fc4 · outbound

This paper cites GPT-4o mini: advancing cost-efficient in- telligence.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization GPT-4o mini: advancing cost-efficient in- telligence

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.011684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:42.873716Z digest=sha256:828a6dd27730226ca556cc65546e43ab88591637a5876f5d9fc928bebb417569

Observation 4ab11380-5c49-4707-aab2-06129999fad6 · outbound

This paper cites The Llama 3 Herd of Models.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization The Llama 3 Herd of Models

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-06T17:26:42.877278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:26:42.877278Z digest=sha256:72655d7b15969384e171dfdd05df4839506157c645af4abd66ce10f2451dcbcd

Observation a8c8dc48-1b17-4bd8-b79c-91866c6a70e0 · outbound

This paper cites Gemma 2: Improving Open Language Models at a Practical Size.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Gemma 2: Improving Open Language Models at a Practical Size

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T17:26:42.892226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:26:42.892226Z digest=sha256:af16bd8e1645108bda48617484f20ce673ef4010142663ce186ca873784747a8

Observation b0bdb53a-ba0f-41b8-9619-687c91c69ae3 · outbound

This paper cites Qwen2.5 Technical Report.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Qwen2.5 Technical Report

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-06T17:26:42.996218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:26:42.996218Z digest=sha256:0e1e76e18a96f3808f4f7075c93018c34492cb11c206935b601aa33facc44505

Observation b7bdd802-c3b5-401a-92a6-815bd1315e18 · outbound

This paper cites Openclip,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Openclip,

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T17:26:43.092511Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:26:43.092511Z digest=sha256:987686e391b279415ea8521c0a889c4cc6b4a3769bdc0d1c6c4e1262f7fd8233

Observation fee5a68b-2a12-4e76-b3fd-6f2feac95f21 · outbound

This paper cites Torchmetrics - measuring reproducibility in pytorch,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Torchmetrics - measuring reproducibility in pytorch,

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-06T17:26:43.191390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:26:43.191390Z digest=sha256:8bedc44011715ece8b769234cc5fbac7b5eb72808d83e4e75d4ecde13afeb6ff

Observation e7acea5c-bc75-4730-9fc6-307e4ca84cfd · outbound

This paper cites Coca: Contrastive captioners are image-text foundation models,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Coca: Contrastive captioners are image-text foundation models,

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:44.001587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:43.275308Z digest=sha256:1ebb7c02f9232a71c03cf31ad158d8816df94b97a11cf71fe7e98fa5c2ea567e

Observation f81e374e-5483-48e2-a264-9e627404a796 · outbound

This paper cites Llava-next: A strong zero-shot video understanding model,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Llava-next: A strong zero-shot video understanding model,

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:43.992269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:43.314145Z digest=sha256:9c8ab58f0cf284f2c76760738ef0753f455acb058f89d74c249d66fff0155cfb

Observation b1854ce4-7a63-4ad5-92b7-736cd657a4f0 · outbound

This paper cites Habitat-matterport 3d dataset (HM3d): 1000 large-scale 3d environments for embodied AI,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Habitat-matterport 3d dataset (HM3d): 1000 large-scale 3d environments for embodied AI,

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:43.982841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:43.350446Z digest=sha256:6758c758230b012cb58de6f9615477ac6fc0e32d3acecb8d7498bb3975b443cb

Observation 657da7f3-f228-4c2a-9e96-5dd84ce52051 · outbound

This paper cites Scenenet rgb-d: Can 5m synthetic images beat generic imagenet pre-training on indoor segmentation?.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Scenenet rgb-d: Can 5m synthetic images beat generic imagenet pre-training on indoor segmentation?

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:43.973049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:43.380318Z digest=sha256:ae804d942e938babb2eaca3736496a6f3cd61c5d88ab8848ef7bf82a73e286fd

Observation 50046585-1b74-4c97-a40c-c8877d309220 · outbound

This paper cites GRUtopia: Dream General Robots in a City at Scale.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization GRUtopia: Dream General Robots in a City at Scale

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-06T17:26:43.401732Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:26:43.401732Z digest=sha256:425314eab0e2e86d881719fad12c80bb522900bd1477fe585ae7b81ab961cdd6

Observation 0a01de21-e13d-40f3-9ec3-33838189d519 · outbound

This paper cites Sun3d: A database of big spaces reconstructed using sfm and object labels,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Sun3d: A database of big spaces reconstructed using sfm and object labels,

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:43.927862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:43.407659Z digest=sha256:356e4b4d9fadc57dd273b90494510a7c77917a28649bd10d003074ec3854ab6c

Observation 505e524b-fc7c-4f03-a10f-e591e1571320 · outbound

This paper cites Scannet: Richly-annotated 3d reconstructions of indoor scenes,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Scannet: Richly-annotated 3d reconstructions of indoor scenes,

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:43.849072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:43.413279Z digest=sha256:95dadc328c4d023d148cf90f8dd562d1688c20485e2c0173ac7d0cd7c94b5597

Observation a268142b-4b52-423d-a890-b6594a577936 · outbound

This paper cites A benchmark for the evaluation of rgb-d slam systems,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization A benchmark for the evaluation of rgb-d slam systems,

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:43.798243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:43.419507Z digest=sha256:cc53431c224a7bfafbba8b2d132813ac9acec49160cf8b8b6fa15571e83da187

Observation 6467cdaf-05aa-47e0-a458-28270e1a30ac · outbound

This paper cites Learning to navigate the energy landscape,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Learning to navigate the energy landscape,

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:43.748638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:43.425274Z digest=sha256:dce3287dfbb4747381bb6f821f46866aee4cab0d42972520978bc55cbc96b25e

Observation 224e5921-e4fc-47e8-811d-6c2248a5c419 · outbound

This paper cites DIODE: A Dense Indoor and Outdoor DEpth Dataset.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization DIODE: A Dense Indoor and Outdoor DEpth Dataset

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-06T17:26:43.431393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:26:43.431393Z digest=sha256:a74fa70dc1797677c2d050d6e37b281cd09322c4a486fc610e7ffa923cda1581

Observation 3672e4eb-aa87-4be4-964d-ee67f5298424 · outbound

This paper cites Vision meets robotics: The kitti dataset,.

NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization Vision meets robotics: The kitti dataset,

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:43.696891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:43.437116Z digest=sha256:a01c5f017cce4696f741d6363600218fde187023f304070865c9e9208e257dd5

Pith citing papers

No inbound Pith citation observations are available.