Pith. sign in

Paper Citation Record · LEDGER

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation

As of 22 August 2026, this Paper Citation Record lists 90 of 90 outbound references and 0 inbound Pith citation observations for arXiv:2605.26500.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.26500 v1

Coverage vector

measured 90 of 90 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-29T18:21:18.475105Z

measured 90 of 90 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

90 of 90 outbound references displayed

  • verified exact2
  • verified fuzzy0
  • unresolved88
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 8308f7a8-7ac7-4cc0-8bad-2fd3e86cbf5b · outbound

This paper cites Bevbert: Multimodal map pre-training for language-guided navigation.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Bevbert: Multimodal map pre-training for language-guided navigation

Reference 1

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:cb2e5c0307ce45cea7378164bf708de797eb864ba54a3b991ce63e5c9d636f7f

Observation d1200a19-6589-4e2a-9a89-8604a3e59b31 · outbound

This paper cites Etpnav: Evolving topo- logical planning for vision-language navigation in continu- ous environments.IEEE TPAMI, 2024.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Etpnav: Evolving topo- logical planning for vision-language navigation in continu- ous environments.IEEE TPAMI, 2024

Reference 2

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:a28267a3e6414bfba8b3ce91032f2ba718aca85b524eb08caecf308fa93e4e0f

Observation 0cc48786-e186-45f1-8784-b01d20276ce6 · outbound

This paper cites Reid, Stephen Gould, and Anton van den Hengel.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Reid, Stephen Gould, and Anton van den Hengel

Reference 3

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:63e38490e8a22691d11a562da29bb679ff16a3cc9375802143d688275539f441

Observation 3a18e11a-f9f9-436a-9984-213991c494ac · outbound

This paper cites 3d semantic parsing of large-scale indoor spaces.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation 3d semantic parsing of large-scale indoor spaces

Reference 4

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:d94bfd3d8e0ff7887dfeecd78393f53e375674d189aaa17c81be2ce212e5ea1e

Observation 80cf65e6-00b1-4b33-b1ed-c8b9e1ce3fb3 · outbound

This paper cites Matterport3d: Learning from rgb-d data in indoor environments.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Matterport3d: Learning from rgb-d data in indoor environments

Reference 5

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:49e24f5ce8c06179c06d80b9fb021fc8bb9850cb261e4f5f3f5e58f297ffcd84

Observation 9af4a82a-0af9-4010-bc5c-af23ac89eb72 · outbound

This paper cites Object goal naviga- tion using goal-oriented semantic exploration.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Object goal naviga- tion using goal-oriented semantic exploration

Reference 6

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:ba8bf45e0e3f5516ae9ad9192feae664b1d710bc4c7293630b24a9b047dd9a9a

Observation c03b04bf-92e2-43cb-a32a-b54b0494a38b · outbound

This paper cites Neural topological slam for vi- sual navigation.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Neural topological slam for vi- sual navigation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:22aa097491be9760ca0f313cb8731903ece7f0f07b6ad35d9bfdc0d3173924cc

Observation c0d74a32-f383-4c4c-b8b5-40a21fa71cc6 · outbound

This paper cites pixelsplat: 3d gaussian splats from image pairs for scalable generalizable 3d reconstruction.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation pixelsplat: 3d gaussian splats from image pairs for scalable generalizable 3d reconstruction

Reference 8

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:e8274f9fb88ac521abb79bab16755f9e838350314cf5cddc5798338af5a1d257

Observation 9949e942-3d08-47f5-a9cf-d6ff7d6b75de · outbound

This paper cites A Survey on 3D Gaussian Splatting.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation A Survey on 3D Gaussian Splatting

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-06-29T18:23:50.613597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:03bb773824f784e930ef9cacf97256eb922c8277336a1592e5651e45e54c8108

Observation 56b6b7d8-b64b-440a-b350-07222e87fa29 · outbound

This paper cites Learning active camera for multi-object navigation.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Learning active camera for multi-object navigation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:567e2932e9ad1ff89921fe2a8b159edfbb34821ba220eaa534634adcccc56829

Observation e455f297-c7f9-4504-9df4-0062c87b4672 · outbound

This paper cites Weakly- supervised multi-granularity map learning for vision-and- language navigation.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Weakly- supervised multi-granularity map learning for vision-and- language navigation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:dfc8d7fd41d739e3e73d6788c1abb392265c8956b30de50716453c0e4d848e96

Observation 42cfafe3-28e8-4bef-bc93-28cfb175c4c7 · outbound

This paper cites History aware multimodal transformer for vision-and-language navigation.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation History aware multimodal transformer for vision-and-language navigation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:22d433ca36f0d4eb67444539a946bc3b8cec6437a260c6b072ac2c783b959b5c

Observation 246297b1-ea35-4fab-8f66-324d4b526fff · outbound

This paper cites Think global, act local: Dual-scale graph transformer for vision-and-language navi- gation.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Think global, act local: Dual-scale graph transformer for vision-and-language navi- gation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:1dc765397d7ec0f74b1136b6adabe46f1ce8b13036fbdb2acfe449d697a0934c

Observation a9b3c51b-8ac0-490a-8cef-6fba3b4f323a · outbound

This paper cites Text-to-3d using gaussian splatting.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Text-to-3d using gaussian splatting

Reference 14

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:57b422e21d995bb569c8676b16c20f476443040d3271eda80a7eea9508fa2902

Observation c6251885-d403-4b7c-a52d-380537f6b5d7 · outbound

This paper cites Scannet: Richly-annotated 3d reconstructions of indoor scenes.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Scannet: Richly-annotated 3d reconstructions of indoor scenes

Reference 15

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:468bd2bc063181e8199c97244d6f1909caf65bd4a73d1ddb1d7a44d04707f953

Observation ae774b99-5541-4ecb-979e-581a0ee720d1 · outbound

This paper cites Evolving graphical planner: Contextual global planning for vision-and-language navigation.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Evolving graphical planner: Contextual global planning for vision-and-language navigation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:ead943bd59239e91212e7e83f8b3008f32a6f4c8436ab7b7f028f5a154df24c6

Observation 78ccdae6-8676-4590-ac4f-e564730ac90b · outbound

This paper cites Unconstrained scene generation with locally conditioned radiance fields.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Unconstrained scene generation with locally conditioned radiance fields

Reference 17

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:ff4e242e5dcc1716c192942efee7befee5d81e53b29f648449c730734258cfac

Observation f02c6fcd-b2ae-4c26-81a7-d6191afec65b · outbound

This paper cites Reinforcement learning with neural ra- diance fields.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Reinforcement learning with neural ra- diance fields

Reference 18

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:8a42475fc55732e9ac9bc7ccb0abaa2ec678e59ac60e531fc9e54970f525eaa9

Observation 9aff3430-2c9e-497c-8423-57b5df345ef1 · outbound

This paper cites Evidential active recognition: Intelligent and prudent open-world embodied perception.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Evidential active recognition: Intelligent and prudent open-world embodied perception

Reference 19

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:dfebad0356c6e708712ebea90f82c26a35ad198351e57ae231b6a3e99d5cc7a0

Observation 58df7574-bc5c-43f9-af38-5ebfc4e789ec · outbound

This paper cites Navi- gation instruction generation with bev perception and large language models.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Navi- gation instruction generation with bev perception and large language models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:5dca206dd73f7c39596b30c6b0bc9c478a78d1b2022d3edadc88efabfe5e742f

Observation 029f7ce4-a075-4cba-928d-0c6ade433dda · outbound

This paper cites Scene map-based prompt tuning for navigation instruction genera- tion.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Scene map-based prompt tuning for navigation instruction genera- tion

Reference 21

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:83815c25680541e85622567d2049252065ca874f7572ae7c616b61d850d85dbe

Observation 33592cfc-f5f1-428c-959e-2e46888f0f3b · outbound

This paper cites Speaker-follower models for vision-and-language naviga- tion.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Speaker-follower models for vision-and-language naviga- tion

Reference 22

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:ae3f786ed4eb8e602717c2b6c1c469c16daf52d5fcae445914573d73213d40c1

Observation 8d70fb94-3df5-43dc-bed6-970689175c5e · outbound

This paper cites Panoptic nerf: 3d-to-2d label transfer for panoptic urban scene segmentation.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Panoptic nerf: 3d-to-2d label transfer for panoptic urban scene segmentation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:f0f96d1b4b6c9e2d6577d4589d6a061f5048dce65709c855aea267148ad8a1ad

Observation bd70849e-6ad8-4f2b-9ed4-1277a228bb4a · outbound

This paper cites Dynamic view synthesis from dynamic monocular video.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Dynamic view synthesis from dynamic monocular video

Reference 24

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:fda00d1a9768d116b2b13838e21fe8f23e72820d517d813d94c6a20d62f58a9c

Observation aefe504e-a80b-4f10-9608-def9dfc64b04 · outbound

This paper cites Room-object entity prompting and reasoning for embodied referring expression.IEEE TPAMI, 46(2):994– 1010, 2023.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Room-object entity prompting and reasoning for embodied referring expression.IEEE TPAMI, 46(2):994– 1010, 2023

Reference 25

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:3a3e14a95d4dbbfd627ce5e4b402297229b30cdc7a0caf7b5653559e3e8ace6b

Observation 32b62103-2834-4ab0-956d-9a36f6d9c8cc · outbound

This paper cites Cross-modal map learning for vision and language navigation.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Cross-modal map learning for vision and language navigation

Reference 26

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:cb9c85f347dd1075b4e994f1fd37e0933bb150a492c4e158456e06651e929d76

Observation 474fa3a9-43c1-4f5d-b37e-25a59a16b1b3 · outbound

This paper cites Airbert: In-domain pretrain- ing for vision-and-language navigation.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Airbert: In-domain pretrain- ing for vision-and-language navigation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:ce7d0c680bbb27a479454ff806ce3800b13b4bc8e9a0ac28ee21462b7cd73433

Observation 65ac717c-b9f2-43d4-90a4-a6056b44874b · outbound

This paper cites Multi-view reconstruction via sfm-guided monocular depth estimation.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Multi-view reconstruction via sfm-guided monocular depth estimation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:fc3cbc1f8038fbc1f18cdd466a17efdd4f61d2429ab9196a2e05ed6788694b53

Observation 80c66e76-f22f-4749-930e-0abb550a1d04 · outbound

This paper cites Language and visual entity relationship graph for agent navigation.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Language and visual entity relationship graph for agent navigation

Reference 29

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:03038befe63c33c2ab813ea9a77da72d3091b8cc5244007ebab3bef62c29d5e3

Observation 4d8af440-16e8-4388-b74a-01534cc687d5 · outbound

This paper cites Vln bert: A recurrent vision- and-language bert for navigation.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Vln bert: A recurrent vision- and-language bert for navigation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:7835f0140499f01a93d60b681bfe1ebbbb05c521428633d3103a3d95b4acd801

Observation 35b53c0e-25a1-43ac-bcc8-cbb71c506d44 · outbound

This paper cites Learning navigational visual representations with semantic map super- vision.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Learning navigational visual representations with semantic map super- vision

Reference 31

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:910c2315a81a4d47c96f25bbee004d35cc979d663ca7bb7830c6cc75f7da98e4

Observation e50662a6-048d-42ec-97c2-367a34933241 · outbound

This paper cites Stay on the path: Instruction fidelity in vision-and-language navigation.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Stay on the path: Instruction fidelity in vision-and-language navigation

Reference 32

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:cea8d50cb365c774b5ffd75b2026c484b31778b3117e95dc835f4c16c2169ba7

Observation cf836c5a-bcdd-4519-933f-8500cc73a67b · outbound

This paper cites Hifi4g: High-fidelity human performance rendering via compact gaussian splatting.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Hifi4g: High-fidelity human performance rendering via compact gaussian splatting

Reference 33

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:d683823e8aa8fdabfae23a8ca299bc741c300eeb99149e17fd8d5fe158e8caa0

Observation c785602c-dd0e-4d42-a87d-4413146abe0d · outbound

This paper cites Deformation and correspondence aware un- supervised synthetic-to-real scene flow estimation for point clouds.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Deformation and correspondence aware un- supervised synthetic-to-real scene flow estimation for point clouds

Reference 34

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:44baf94855b93b0d062af98d5e3e0e0522a0b875c879092a94e9feba4b81c3db

Observation a922c9d9-a6b0-4fa8-8987-3e8eba12391c · outbound

This paper cites Splatam: Splat track & map 3d gaussians for dense rgb-d slam.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Splatam: Splat track & map 3d gaussians for dense rgb-d slam

Reference 35

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:dfa3390e205a701f58b330e847ece453f2db543c12d37d7294a90f9bf0519537

Observation f8bc45f8-fa7f-400d-b7ae-3a85fd0a5c2f · outbound

This paper cites Bert: Pre-training of deep bidirectional trans- formers for language understanding.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Bert: Pre-training of deep bidirectional trans- formers for language understanding

Reference 36

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:2480bc672955cb6342fdfc45fc60ed399aa98377b559cc03e18b24fc036e9a2f

Observation 709cad3e-611a-4da0-90dc-304fe95a99b5 · outbound

This paper cites 3d gaussian splatting for real-time radiance field rendering.ACM TOG, 42(4), 2023.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation 3d gaussian splatting for real-time radiance field rendering.ACM TOG, 42(4), 2023

Reference 37

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:976a97efbb2ead019f4e4db1e984df8bdb3ec59a4829cf03d6d09099e18b37fd

Observation 82bb6540-9bce-4e69-9f82-7c04c53ef4fb · outbound

This paper cites Kingma and Jimmy Ba.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Kingma and Jimmy Ba

Reference 38

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:28cbb100d3de4fcbbeb8341f87e27250763058ee6d3367412857360dee938bdc

Observation 5bd35366-ce58-45db-809e-c35f6c48048a · outbound

This paper cites Controllable navigation in- struction generation with chain of thought prompting.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Controllable navigation in- struction generation with chain of thought prompting

Reference 39

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:7e14e17021165034f19afe3bdcd622041879511bba4a542e9a074f647f16876e

Observation c33b45f8-8656-4582-892a-cdc48a610aa5 · outbound

This paper cites Renderable neural radiance map for visual navigation.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Renderable neural radiance map for visual navigation

Reference 40

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:80162c9a25926e061ee82fd96cfebf43a36e5b28e12be5fa923668eeb5bce1ed

Observation 9fa95f73-feb0-4dc3-af68-7503cc07569e · outbound

This paper cites Envedit: Environment editing for vision-and-language navigation.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Envedit: Environment editing for vision-and-language navigation

Reference 41

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:988388bf62e0f200e77afa4f0c79e6ef19e046e6169f195c1c24cfe2fe6e474f

Observation 70857a6a-9260-47ef-9b44-386f046fe834 · outbound

This paper cites 3d neural scene representations for visuomotor control.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation 3d neural scene representations for visuomotor control

Reference 42

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:d293763a871acb2d5281347d0179196569dd884f0de8523ab102fb8559068568

Observation 20315b3c-d6e2-4a61-9cd9-4869e654e450 · outbound

This paper cites V oxformer: Sparse voxel transformer for camera- based 3d semantic scene completion.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation V oxformer: Sparse voxel transformer for camera- based 3d semantic scene completion

Reference 43

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:c9dcb9bace2981ad67bda835313f006e2b35a300a264b8154c9fecffb9f2835f

Observation b34ea2b8-1cd4-4fb2-9dce-a54d292d414b · outbound

This paper cites Luciddreamer: Towards high- fidelity text-to-3d generation via interval score matching.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Luciddreamer: Towards high- fidelity text-to-3d generation via interval score matching

Reference 44

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:fe2d3b7e0dcab81cb6cc98dc0e7555f6c9d2296a51e2b652e1e893d2578bcad9

Observation 2c627795-43d4-40c0-ab2c-b744af895fb7 · outbound

This paper cites Scene-intuitive agent for remote embodied visual grounding.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Scene-intuitive agent for remote embodied visual grounding

Reference 45

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:1e444233dc1711c24b8e8e30549f895fb3fd37eee5b45820e2a924e8ca2e8aac

Observation 6b3bd8e0-03ee-4d2c-9b11-9c203080b681 · outbound

This paper cites Vision-language naviga- tion with random environmental mixup.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Vision-language naviga- tion with random environmental mixup

Reference 46

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:575e7fa73944bc06ee391f5da30e3fbb847024f8b4b004684486702ac4caf0fb

Observation d7045f83-e1cb-4a08-a7af-2e01f07d5c0d · outbound

This paper cites Bird’s-eye-view scene graph for vision-language navigation.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Bird’s-eye-view scene graph for vision-language navigation

Reference 47

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:1129bc27f60666ac4ca8f1cb074d55ac0760cb8753f9444c66336bcf94a0b242

Observation bd6b843a-914a-4fa2-bf0e-e4950741b64b · outbound

This paper cites Vision-language nav- igation with energy-based policy.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Vision-language nav- igation with energy-based policy

Reference 48

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:817cb9b457365d1132544b2a9294fff9a1f7b921c1c4e5b4c1e1d18da6486c09

Observation 23df8aac-b14d-4c8c-a0a6-beeb56760cbf · outbound

This paper cites V olumetric envi- ronment representation for vision-language navigation.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation V olumetric envi- ronment representation for vision-language navigation

Reference 49

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:88e297398fb6ae518db652ad9a0c93a835c9a3f5031f6ffab8877d49a4cddfaa

Observation ad3f74c1-5a1a-4055-bd99-629401d412ff · outbound

This paper cites Editing conditional radiance fields.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Editing conditional radiance fields

Reference 50

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:a8d01d84ccc68d09d8ac647f2e06ec24e166752e58e730ee98d7c9dadffe1d3f

Observation f56b006f-6ebd-4220-b806-288f41409ade · outbound

This paper cites Srinivasan, Matthew Tancik, Jonathan T.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Srinivasan, Matthew Tancik, Jonathan T

Reference 51

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:5e262239d0cfcecb8c0cdeeadea79f2258820316161cb2caa0634c2603a95788

Observation a88d7427-87b6-44a6-b7ca-3512bb74b80d · outbound

This paper cites Soat: A scene-and object-aware transformer for vision-and-language navigation.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Soat: A scene-and object-aware transformer for vision-and-language navigation

Reference 52

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:b4f1e5b459f6e21c2ecce63806e7fe0dae8eadf6cdf3ed0d9845346da478e07b

Observation c398176c-2e08-46b9-81f8-6d547f40db7d · outbound

This paper cites Seeing the un-scene: Learning amodal semantic maps for room navigation.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Seeing the un-scene: Learning amodal semantic maps for room navigation

Reference 53

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:6a592effcf8c375080e0f83da1c2351474ae6e0a8899bd835285b4a9e9bf7b21

Observation 33345515-4949-4edb-a7d3-b637efb77c90 · outbound

This paper cites Neural map: Structured memory for deep reinforcement learning.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Neural map: Structured memory for deep reinforcement learning

Reference 54

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:a4977cc32b05d2dd9393c03c076e4c3c0ce4c605893de8308a577bfa5b147d8a

Observation b99a6570-a09b-49f7-b73b-709fbc69ec4a · outbound

This paper cites D-nerf: Neural radiance fields for dynamic scenes.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation D-nerf: Neural radiance fields for dynamic scenes

Reference 55

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:799a99ffa320c5168fb042e1f76ab7e7edc27dfb8a2d553a62efd7822641f01d

Observation 9c7128cf-e41c-4023-b886-9cb5aa394734 · outbound

This paper cites Reverie: Remote embodied visual referring expres- sion in real indoor environments.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Reverie: Remote embodied visual referring expres- sion in real indoor environments

Reference 56

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:541503af39e819f03a2bdb8897279b7a251a3f04de6fac9358a7ce8a33abc224

Observation 781744d9-02b4-42ba-b02b-8074e3608030 · outbound

This paper cites Hop: history-and-order aware pre- training for vision-and-language navigation.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Hop: history-and-order aware pre- training for vision-and-language navigation

Reference 57

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:b00f27240d8d98e11cfecdb4c0c66e947129ffa1fe36c30b6247365decb6a112

Observation 5f85be6f-b2fb-42b9-bbb3-6ed40583551f · outbound

This paper cites Holis- tic lstm for pedestrian trajectory prediction.IEEE TIP, 30: 3229–3239, 2021.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Holis- tic lstm for pedestrian trajectory prediction.IEEE TIP, 30: 3229–3239, 2021

Reference 58

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:9de9e0e40c2f8f5b6406bd9e228267cf3f80979884feb20da1f202d728b73ba7

Observation 8823f0b0-d875-4d42-aea0-79d9e0e8598a · outbound

This paper cites Learn- ing transferable visual models from natural language super- vision.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Learn- ing transferable visual models from natural language super- vision

Reference 59

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:eeec961a293ed8e62c3152720288995e23e1ff4718a1c4cb7beca04f62389a2b

Observation dc4408a4-7e89-4ab0-a42c-fe11de67f07f · outbound

This paper cites Occupancy anticipation for efficient exploration and navigation.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Occupancy anticipation for efficient exploration and navigation

Reference 60

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:90f47a19f345d50a7c828a3f0f16f497e2acda4dc739742e883388db81cb262d

Observation 8664526a-3c59-4511-8b11-b1315763b04e · outbound

This paper cites SAM 2: Segment Anything in Images and Videos.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation SAM 2: Segment Anything in Images and Videos

Reference 61

Resolution
verified exact
local_arxiv, observed 2026-06-29T18:23:50.610305Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:6605ced4f81e772a26b2e17ee24595a839c31e11050ea3c92605c9ed875c62bb

Observation 5cfb5b44-428a-4d87-9d8d-2101a184cb3b · outbound

This paper cites Gordon, and Drew Bagnell.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Gordon, and Drew Bagnell

Reference 62

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:e18d9e2fe79dacc7256bea66705eee9f1929b83ad1706dbec29779a560a88649

Observation 42f07a43-6742-4dcc-bfba-56ab4f0d9a02 · outbound

This paper cites Toward open set recogni- tion.IEEE TPAMI, 35(7):1757–1772, 2012.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Toward open set recogni- tion.IEEE TPAMI, 35(7):1757–1772, 2012

Reference 63

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:2f08037c9b9fe73eea8f57cd9aac5ac9a9836f1a38ed5d0d0de6b06def6fbeb1

Observation 90dccd76-63fa-4c82-b738-344a62707d90 · outbound

This paper cites Language embedded 3d gaussians for open- vocabulary scene understanding.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Language embedded 3d gaussians for open- vocabulary scene understanding

Reference 64

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:52b3315974c44060f13edd7ee10af86b236195d934a7d4eb1680c9cec9d6a311

Observation 8596cd39-fece-494a-b34c-7313c69a5ed5 · outbound

This paper cites Snerl: Semantic-aware neural radiance fields for reinforcement learning.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Snerl: Semantic-aware neural radiance fields for reinforcement learning

Reference 65

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:3326643809e465a1b0f8bcc756d677678b468c26a80da56b5ebac2d6b0e5525e

Observation 3d3b2322-b260-4706-85ae-4024d7ced8d9 · outbound

This paper cites Sequence to sequence learning with neural networks.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Sequence to sequence learning with neural networks

Reference 66

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:f08f86be7f937069bbaead83142d66351b9b704b799f1bc44a7b8e0c2de12c69

Observation 7f0a9d19-d824-48f6-9410-c1952e86f7bf · outbound

This paper cites Learning to nav- igate unseen environments: Back translation with environ- mental dropout.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Learning to nav- igate unseen environments: Back translation with environ- mental dropout

Reference 67

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:440965bec64bc2091577d2c82fa11f64d784f5e595f74b84f1e2210d7bc1a7cd

Observation 28020d93-299d-44a5-9908-c66dce421096 · outbound

This paper cites Active visual information gathering for vision-language navigation.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Active visual information gathering for vision-language navigation

Reference 68

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:57c75d75fd7e8bc072aa1f4a190b6b677ae29afd7c4f9b5df0706717251a60bc

Observation 27c03164-09c3-4f42-863b-cee9c4fe1b79 · outbound

This paper cites Structured scene memory for vision- language navigation.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Structured scene memory for vision- language navigation

Reference 69

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:88bf7cbcba76f93c56e266d3fe72dd77b1b390eb141c9319ab89f417b5e2c2a2

Observation eecf4612-81ae-4214-9b57-a20b6e9adf0f · outbound

This paper cites Towards versatile embodied navigation.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Towards versatile embodied navigation

Reference 70

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:cb93aafa231a94208945e49fe4e5e772aa13fadea39d0468ce1c449ee10318d5

Observation de14de4f-d171-4176-b7d3-9234a04e97b7 · outbound

This paper cites Counterfactual cycle-consistent learn- ing for instruction following and generation in vision- language navigation.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Counterfactual cycle-consistent learn- ing for instruction following and generation in vision- language navigation

Reference 71

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:7ca270a88346f1059f54d7d191a2baf270621afdd81f67b9309d1da5fe5b06d3

Observation 16c96448-9905-4303-a345-94128a03914e · outbound

This paper cites Dreamwalker: Mental planning for continuous vision-language navigation.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Dreamwalker: Mental planning for continuous vision-language navigation

Reference 72

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:d9e6cb7a2875886b00739268de4de4b717a0996a1778854908081ace52788f11

Observation 58412a53-9722-4e8c-88ed-98cf3965b4b3 · outbound

This paper cites Active perception for visual-language navigation.IJCV, 131(3):607–625, 2023.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Active perception for visual-language navigation.IJCV, 131(3):607–625, 2023

Reference 73

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:fcc9cff9532287ca847168475dccd03f65f6293ad82b2e358a3db524e6eff73b

Observation 26840858-5036-46d7-84d0-f067b17809a6 · outbound

This paper cites Reinforced cross-modal matching and self- supervised imitation learning for vision-language navigation.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Reinforced cross-modal matching and self- supervised imitation learning for vision-language navigation

Reference 74

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:7edbec25682fd2e2970a44046f28d553a1c1d783145716390bf5544fb48b0f3a

Observation 8ae485e1-8fe1-4ba2-8bbc-20f9861ca2ad · outbound

This paper cites Lana: A language-capable navigator for instruction follow- ing and generation.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Lana: A language-capable navigator for instruction follow- ing and generation

Reference 75

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:193d6d0d8a08b900f30dbf7df3b1636626897959bcf9fa3c31729bda3eeab2ca

Observation 67ecf29e-d35e-4f3f-9a52-8bd828fa678e · outbound

This paper cites Image quality assessment: from error visibility to structural similarity.IEEE TIP, 13(4):600–612, 2004.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Image quality assessment: from error visibility to structural similarity.IEEE TIP, 13(4):600–612, 2004

Reference 76

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:03e3dc60e9c492cc1e34e517efb3627f5c435818d02748768fc7cf19d0509a30

Observation 2878fbc9-d2b6-41df-a9d5-fcea5e0df168 · outbound

This paper cites Gridmm: Grid memory map for vision-and- language navigation.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Gridmm: Grid memory map for vision-and- language navigation

Reference 77

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:d4a6b0b2afd980dc581c0dee9005457cfc8c8722e4235114e548c0080b7d6cb3

Observation 66eeba34-0bdd-457e-87e2-f286de6ee3fa · outbound

This paper cites Lookahead exploration with neural radiance representation for continuous vision- language navigation.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Lookahead exploration with neural radiance representation for continuous vision- language navigation

Reference 78

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:95b82dc10cdf1ce5a88b61eb5998fa1a62c0e19c534d82e00f42bcbc1a61fb03

Observation 43ab28b0-3bd5-42fd-88c7-96d4a7b34eb8 · outbound

This paper cites Vector-decomposed disentanglement for domain- invariant object detection.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Vector-decomposed disentanglement for domain- invariant object detection

Reference 79

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:cb0b3f889aa661d37900d76b098eaedc99b4640b0333fe109d1e9d396e7b0b3a

Observation 02964bbf-4474-4f88-8b11-938472e68ea6 · outbound

This paper cites 4k4d: Real-time 4d view synthesis at 4k resolution.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation 4k4d: Real-time 4d view synthesis at 4k resolution

Reference 80

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:866587e9c15d98f26fbe5fe0181315e0b7126e1df34de8f3a12eb200cfabb02b

Observation 345f004a-7b5b-4e44-adbe-0898ba2ee4e0 · outbound

This paper cites Gaussian grouping: Segment and edit anything in 3d scenes.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Gaussian grouping: Segment and edit anything in 3d scenes

Reference 81

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:b47e5a8632149c038593dea32cd5c4847a4eec8a01f5b2fdd0d43dce47ee83ce

Observation 7d9a1962-663c-4dc7-94a5-731125395c51 · outbound

This paper cites Compositional scene representation learning via reconstruc- tion: A survey.IEEE TPAMI, 45(10):11540–11560, 2023.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Compositional scene representation learning via reconstruc- tion: A survey.IEEE TPAMI, 45(10):11540–11560, 2023

Reference 82

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:618dab648b0b14a16e3324a1912afc0f2e2613b633dfbb1ea43ba349d5016ec4

Observation b617152c-97e2-4703-b3d5-9b728202cffd · outbound

This paper cites Target- driven structured transformer planner for vision-language navigation.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Target- driven structured transformer planner for vision-language navigation

Reference 83

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:bcefff8bec73be3153befe6eb9e9c4184de407033fbce8a55c460f36297f7fee

Observation b6f58505-6981-4da5-8908-8547c64f6d86 · outbound

This paper cites In-place scene labelling and understanding with implicit scene representation.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation In-place scene labelling and understanding with implicit scene representation

Reference 84

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:b53799f840c0f1311bd3bb62c4c39c46f8b93926cdc7e9e987e3faa98f1295e5

Observation f0277453-a614-4d23-9874-cfe84e38f9b3 · outbound

This paper cites Empowering embodied visual tracking with visual foundation models and offline rl.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Empowering embodied visual tracking with visual foundation models and offline rl

Reference 85

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:2ef1de395532b04eaa32792d4c0d7b04869156140840588e741d54e21c5cbd12

Observation ee42e9a5-f17c-469f-af8e-10b5de0c6b55 · outbound

This paper cites Unrealzoo: Enriching photo- realistic virtual worlds for embodied ai.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Unrealzoo: Enriching photo- realistic virtual worlds for embodied ai

Reference 86

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:5c70db5dd72a03468e53c919c651710f2c8238a19ef9105119304d97edd8a3a1

Observation 30cb1c68-12db-47e0-b261-e4f1077228ba · outbound

This paper cites Migc: Multi-instance generation controller for text-to-image synthesis.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Migc: Multi-instance generation controller for text-to-image synthesis

Reference 87

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:b3d66a73edbbc9c93d37ba4da99adc4ce927f87add5ff68af9f71a0e4c7edc26

Observation 983f84d5-90b9-4cd7-ace3-33628d0ffa48 · outbound

This paper cites Hugs: Holistic urban 3d scene understanding via gaus- sian splatting.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Hugs: Holistic urban 3d scene understanding via gaus- sian splatting

Reference 88

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:36434618dbfea6fb9ca1890c5faa602725d8609ed4324f50661b9329de9bbe1d

Observation cbd7edbc-8b35-420a-b1b3-0dccb378d5d2 · outbound

This paper cites Feature 3dgs: Supercharging 3d gaussian splatting to enable distilled feature fields.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Feature 3dgs: Supercharging 3d gaussian splatting to enable distilled feature fields

Reference 89

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:042d6b4d1e8249e8a321baf06b9e6ec93be37447a3ece3690496879098b28662

Observation 45430512-0a20-42cd-b94f-4544038764b2 · outbound

This paper cites Vision-language navigation with self-supervised auxiliary reasoning tasks.

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation Vision-language navigation with self-supervised auxiliary reasoning tasks

Reference 90

Resolution
unresolved
no resolver link, observed 2026-06-29T18:21:18.475105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T18:21:18.475105Z digest=sha256:62866c58809a3f77b7afa5b9a8c44efe5b4e16f84dc1e666febc9dad8bad7129

Pith citing papers

No inbound Pith citation observations are available.