Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:02:57.418588Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 86 of 86 outbound references and 1 inbound Pith citation observation for arXiv:2505.16663.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:02:57.418588Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-08T08:43:54.882467Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-11T20:31:12.651566Z
86 of 86 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 5459cc4b-f01f-431d-811e-70fa256f774c · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Flamingo: a visual language model for few-shot learning.Advances in neural information processing systems, 35:23716–23736, 2022
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c700b45-057a-4f36-b138-e8092009fdce · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation BEVBert: Multimodal Map Pre-training for Language-guided Navigation
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12a93b6a-2de0-4a3b-acf1-c122a90bddb8 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation On Evaluation of Embodied Navigation Agents
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c0f15c6-d55d-4869-ab98-77d9bdbd8d80 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Vision-and-language navigation: Interpreting visually-grounded navigation instructions in real environments
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c087e12-cdb5-48df-bb8d-3038b7c04819 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Vision-and-language navigation: Interpreting visually-grounded navigation instructions in real environments
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 269d6139-cfc6-4f78-92ba-cf68da220b74 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Scanqa: 3d question answering for spatial scene understanding
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35418d20-2502-485e-a488-1c4afad4e2b0 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Curriculum learning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 803644f5-5be0-438a-80f2-527e9f4bd7e2 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Coyo-700m: Image-text pair dataset.https://github.com/kakaobrain/coyo-dataset, 2022
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b8ba055-c9ed-4467-9bd8-23b8b5020ee4 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Matterport3D: Learning from RGB-D Data in Indoor Environments
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fca35d50-3a88-4a5c-97ec-fc333d96aade · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Conceptual 12M: Pushing web-scale image-text pre-training to recognize long-tail visual concepts
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21515a60-5e7f-4b5e-a449-37f90ce34a6b · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Spatialvlm: Endowing vision-language models with spatial reasoning capabilities
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b0cfe7a-0f67-4cca-bc1d-f10d121e586e · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Scanrefer: 3d object localization in rgb-d scans using natural language
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f2a2215-a244-4ee0-a580-1a532f346492 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation MapGPT: Map-Guided Prompting with Adaptive Path Planning for Vision-and-Language Navigation
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7c8a17c-4d9a-43c2-adf4-dc0f4a8ade47 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation History aware multimodal transformer for vision-and-language navigation
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b9f94454-ca63-4994-813e-df88e8cea3a0 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Think global, act local: Dual-scale graph transformer for vision-and-language navigation
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1357fe32-c10f-44be-bd3b-b8bce7efd881 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Microsoft COCO Captions: Data Collection and Evaluation Server
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1477e4b-6647-4c5e-a97a-220f3527a3d4 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Scan2cap: Context-aware dense captioning in rgb-d scans
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 293d8a3a-991d-4e9b-addf-4ada9f1f93d3 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Gonzalez, Ion Stoica, and Eric P
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a60419e-8b42-47bf-b579-7db8a7d0aba5 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Objaverse: A universe of annotated 3d objects
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 34b713a6-811b-4cee-9c1b-c4b1df398b0a · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Imagenet: A large-scale hierarchical image database
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01a671c2-f7c8-4748-8754-eabc42c1e184 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Eva: Exploring the limits of masked visual representation learning at scale
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f8c0a50-ba21-4e3a-bbcb-6f25178918bf · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation 3d-front: 3d furnished rooms with layouts and semantics
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c3e6fa09-42b1-48d6-a78b-9321040f646b · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Adaptive zone-aware hierarchical planner for vision-language navigation
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8a37d701-4536-4bbc-95f5-269a18e675c3 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation RoomTour3D: Geometry-Aware Video-Instruction Tuning for Embodied Navigation
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04d85396-d76d-4cc5-b220-5ce55cc2b4be · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Multimodal fusion and vision-language models: A survey for robot vision.arXiv preprint arXiv:2504.02477, 2025
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7bb2342d-17dc-43eb-af21-240b2d6ea61a · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Towards learning a generic agent for vision-and-language navigation via pre-training
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 945ffac9-c13b-4e9f-a43a-5590ceba39b7 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation A recurrent vision- and-language bert for navigation
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ebe4c6f3-69ba-438e-ab30-7be0acdd0fa5 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation 3d-llm: Injecting the 3d world into large language models.NeurIPS, 2023
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a22bf88d-733b-4dec-b1a1-caf9e420c007 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation An Embodied Generalist Agent in 3D World
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a393671c-a2b7-4a59-a3e5-ecbc5f59fc4a · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Clip2point: Transfer clip to point cloud classification with image-depth pre-training
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fcb7fd77-45ec-4480-a47d-e56f92ab226f · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Autonomous multi-view navigation via deep reinforcement learning
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 43932677-81df-4f63-805e-e20b81389fa8 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Meta-explore: Ex- ploratory hierarchical vision-and-language navigation using scene object spectrum grounding
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation af25ac1b-7f06-467f-806b-2c9379fe1297 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Sceneverse: Scaling 3d vision-language learning for grounded scene understanding
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bdcf0f27-bce6-4e8f-a9f5-070a8529295c · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Sceneverse: Scaling 3d vision-language learning for grounded scene understanding
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a954ca17-1345-4920-bd16-b85d97a4781f · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Scaling up visual and vision-language representation learning with noisy text supervision
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0e30d26-ca0d-46fe-9392-3190ab0fa939 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Two Heads are Better Than One: Test-time Scaling of Multi-agent Collaborative Reasoning
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d20e5b75-bab7-4ef4-9a68-037ff399b8a2 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Segment anything
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4bc65567-79f4-4e10-94cc-e9eb979b0e76 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation AI2-THOR: An Interactive 3D Environment for Visual AI
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 78900ee0-bd7e-4085-9648-01d4a981d002 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Room-across-room: Multi- lingual vision-and-language navigation with dense spatiotemporal grounding
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7eb2eab1-cb88-453f-842c-e7bec29be4ea · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation LLaVA-OneVision: Easy Visual Task Transfer
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb9ee419-1261-45a5-b399-c5dfc8e61bb8 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Improving vision-and-language navigation by generating future-view image semantics
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dffb12bc-4341-45b1-9b24-967f6df9bf54 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Robust Navigation with Language Pretraining and Stochastic Sampling
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 513df580-b561-427d-95e7-6f93dc7e7ac1 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Bevformer: Learning bird’s-eye-view representation from multi-camera images via spatiotemporal trans- formers
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b5be0337-f724-48f5-88e7-cc05dbf5631a · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation NavCoT: Boosting LLM-Based Vision-and-Language Navigation via Learning Disentangled Reasoning
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eea42581-97ad-4948-bbf0-ed86b64c3015 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Multi-modal situated reasoning in 3d scenes.Advances in Neural Information Processing Systems, 37:140903–140936, 2024
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6b4cd09e-bffe-469e-bfa5-600ef92525bf · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Visual instruction tuning
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8811b094-3f3b-4112-9130-55fdf1125a18 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation OpenShape: Scaling Up 3D Shape Representation Towards Open-World Understanding
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b30f639-bbbb-4e5b-988e-95098dab123c · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation V olumetric environment representation for vision-language navigation
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d876afb-d19d-47d0-8fee-60ffbdbfc0b0 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Bird’s-eye-view scene graph for vision-language navigation
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e538cdb-bffd-470c-9f79-f7afbafbb1fe · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Discuss Before Moving: Visual Language Navigation via Multi-expert Discussions
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39a03453-711b-451a-9ffe-65d72760c4bf · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Scalable 3D Captioning with Pretrained Models
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3037b647-cbac-4793-bddc-87e095cc4d53 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation SQA3D: Situated Question Answering in 3D Scenes
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a34b8f07-52a7-4db5-91ca-7ef4906dd1ca · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Reverie: Remote embodied visual referring expression in real indoor environments
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 51a4544e-e0be-467d-9cd5-b42f849b1292 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Hop: history-and-order aware pre-training for vision-and-language navigation
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 958ec0a8-02da-4144-99e3-c054ce7e8077 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Vln-petl: Parameter-efficient transfer learning for vision-and- language navigation
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 47a8587d-5d42-479a-a22c-05576ebd4d3e · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Learning transferable visual models from natural language supervision
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd7a9f9c-1ecf-4663-ae8a-5e18124ef717 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Learning transferable visual models from natural language supervision
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 302d6f8d-62f5-45df-91b4-7d7f05f7f5e7 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Habitat: A platform for embodied ai research
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b3ee603e-a914-4208-9194-6aeaa7e779f5 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Semantic scene completion from a single depth image
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66e45730-928f-43b4-b4fd-a3a493cb5495 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Learning to navigate unseen environments: Back translation with environmental dropout
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 076ef3af-9947-4f6e-b23d-fc4488f3a57b · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Vision-and-dialog navigation
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d045a787-7259-4359-afa1-8b656c83d47e · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation LLaMA: Open and Efficient Foundation Language Models
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa6e9eed-72ef-419c-a719-81667f724bcd · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation AutoML-Agent: A Multi-Agent LLM Framework for Full-Pipeline AutoML
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation afd67ae2-3246-4ddc-8d81-79f2a81fb221 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Vision-and- language navigation via causal learning
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 26942c5b-6d93-49d4-af6d-3aa246475cad · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Reinforced cross-modal matching and self-supervised imitation learning for vision-language navigation
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 976f3e18-4513-4f5e-8e84-88fb1b95bcdc · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Lavie: High-quality video generation with cascaded latent diffusion models.International Journal of Computer Vision, 133(5):3059–3078, 2025
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27644782-d99f-41e0-a703-7a8c054e9aad · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Bootstrapping Language-Guided Navigation Learning with Self-Refining Data Flywheel
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e025e17-6005-4031-aaf7-5bfc60d352e3 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Florence-2: Advancing a unified representation for a variety of vision tasks
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1477c5e8-bb0c-4272-badf-9fd57709601c · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Pointllm: Empower- ing large language models to understand point clouds
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 422e44f8-b0a7-4e5e-99c4-8d8709d1186e · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation ULIP-2: Towards Scalable Multimodal Pre-training for 3D Understanding
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 106edb86-96d0-450b-83fc-9fbad3d7fc30 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Qwen2.5 Technical Report
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 902165a0-6a26-4d05-ae92-bfeb755b600c · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation 3D-GRAND: A Million-Scale Dataset for 3D-LLMs with Better Grounding and Less Hallucination
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eea5d638-352e-4d7a-8c91-2a6825f9a949 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Point-bert: Pre-training 3d point cloud transformers with masked point modeling
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 131b30e5-3f36-454d-af67-ee9aac0ad72e · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Mlink: Linking black-box models from multiple domains for collaborative inference.IEEE Transactions on Pattern Analysis and Machine Intelligence, 45(10):12085–12097, 2023
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f6907c81-9cf6-4ae2-881f-95d685a14dfc · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Building Cooperative Embodied Agents Modularly with Large Language Models
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 678618de-b9a8-4866-abe3-7a1395ba1550 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Agent journey beyond rgb: Unveiling hybrid semantic-spatial environmental representations for vision-and-language navigation.arXiv preprint arXiv:2412.06465, 2024
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf6a3967-dc8c-4814-90ad-81b871184152 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Meta-Transformer: A Unified Framework for Multimodal Learning
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5cadbb87-d677-46db-b1a3-8baf2fd7fd94 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Embodied-R: Collaborative Framework for Activating Embodied Spatial Reasoning in Foundation Models via Reinforcement Learning
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 17b84147-d9cf-4f34-80d0-cad22a36a9f8 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Towards Learning a Generalist Model for Embodied Navigation
Reference 79
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f462e98-b543-4ac2-a80d-5ded14db671b · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Navgpt-2: Unleashing navigational reasoning capability for large vision-language models
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c0940b1-eda1-4b47-a4c7-ed158560e9fa · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation SAME: Learning Generic Language-Guided Visual Navigation with State-Adaptive Mixture of Experts
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2404d4a-54b4-47ff-aaef-83b712f96928 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46cea761-e513-4aba-9b82-c784db7a1807 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Soon: Scenario oriented object navigation with graph-based exploration
Reference 83
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 78805592-6493-44e1-a65a-4e3a85ed4905 · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation Pointclip v2: Prompting clip and gpt for powerful 3d open-world learning
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b5278fdf-86ba-41b1-946f-243abdc0dc2a · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation How would you interpret this3D point cloud?
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4a9188a2-cffd-42bb-8cfb-ea7d60013eeb · outbound
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation We express our deep respect for the contributions of the developers and researchers who have made these models and datasets available
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c130f979-6912-4589-845e-75488568ee75 · inbound
Cross-Modal Navigation with Multi-Agent Reinforcement Learning CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.