{"record_type":"pith_number_record","schema_url":"https://pith.science/schemas/pith-number/v1.json","pith_number":"pith:2025:HJV2XG5FJCGD2G4AWW2SPRLM7M","short_pith_number":"pith:HJV2XG5F","schema_version":"1.0","canonical_sha256":"3a6bab9ba5488c3d1b80b5b527c56cfb233e869461c1bd0779b2775c6eb1350e","source":{"kind":"arxiv","id":"2502.14254","version":2},"attestation_state":"computed","paper":{"title":"Mem2Ego: Empowering Vision-Language Models with Global-to-Ego Memory for Long-Horizon Embodied Navigation","license":"http://creativecommons.org/licenses/by-nc-sa/4.0/","headline":"","cross_cats":["cs.AI"],"primary_cat":"cs.RO","authors_text":"Atia Hamidizadeh, David Gamaliel Arcos Bravo, Guowei Huang, Haoping Xu, Hongjian Gu, Jianye Hao, Lingfeng Zhang, Matin Aghaei, Mohammad Ali Alomrani, Raika Karimi, Tongtong Cao, Weichao Qiu, Xingyue Quan, Yaochen Hu, Yingxue Zhang, Yuecheng Liu, Yuzheng Zhuang, Zhanguang Zhang, Zhanpeng Zhang","submitted_at":"2025-02-20T04:41:40Z","abstract_excerpt":"Recent advancements in Large Language Models (LLMs) and Vision-Language Models (VLMs) have made them powerful tools in embodied navigation, enabling agents to leverage commonsense and spatial reasoning for efficient exploration in unfamiliar environments. Existing LLM-based approaches convert global memory, such as semantic or topological maps, into language descriptions to guide navigation. While this improves efficiency and reduces redundant exploration, the loss of geometric information in language-based representations hinders spatial reasoning, especially in intricate environments. To add"},"verification_status":{"content_addressed":true,"pith_receipt":true,"author_attested":false,"weak_author_claims":0,"strong_author_claims":0,"externally_anchored":false,"storage_verified":false,"citation_signatures":0,"replication_records":0,"graph_snapshot":true,"references_resolved":false,"formal_links_present":false},"canonical_record":{"source":{"id":"2502.14254","kind":"arxiv","version":2},"metadata":{"license":"http://creativecommons.org/licenses/by-nc-sa/4.0/","primary_cat":"cs.RO","submitted_at":"2025-02-20T04:41:40Z","cross_cats_sorted":["cs.AI"],"title_canon_sha256":"d4cc2827e871e7d0b7f6cbc1d0987077ef92e2cd2613eadf978543df684e4fbe","abstract_canon_sha256":"44afcbba418084303c58d7500404062c84c000bc288a6b62bf0131f42a80238e"},"schema_version":"1.0"},"receipt":{"kind":"pith_receipt","key_id":"pith-v1-2026-05","algorithm":"ed25519","signed_at":"2026-07-05T11:19:24.764144Z","signature_b64":"zN+AiDeYyPXxy9x7AimJuGOxlzgk0zywebWPfPO3+/Rdl9NsFVbLLZkXJ20so01y91lSIpPCo/1jhVvR4A3JDQ==","signed_message":"canonical_sha256_bytes","builder_version":"pith-number-builder-2026-05-17-v1","receipt_version":"0.3","canonical_sha256":"3a6bab9ba5488c3d1b80b5b527c56cfb233e869461c1bd0779b2775c6eb1350e","last_reissued_at":"2026-07-05T11:19:24.763675Z","signature_status":"signed_v1","first_computed_at":"2026-07-05T11:19:24.763675Z","public_key_fingerprint":"8d4b5ee74e4693bcd1df2446408b0d54"},"graph_snapshot":{"paper":{"title":"Mem2Ego: Empowering Vision-Language Models with Global-to-Ego Memory for Long-Horizon Embodied Navigation","license":"http://creativecommons.org/licenses/by-nc-sa/4.0/","headline":"","cross_cats":["cs.AI"],"primary_cat":"cs.RO","authors_text":"Atia Hamidizadeh, David Gamaliel Arcos Bravo, Guowei Huang, Haoping Xu, Hongjian Gu, Jianye Hao, Lingfeng Zhang, Matin Aghaei, Mohammad Ali Alomrani, Raika Karimi, Tongtong Cao, Weichao Qiu, Xingyue Quan, Yaochen Hu, Yingxue Zhang, Yuecheng Liu, Yuzheng Zhuang, Zhanguang Zhang, Zhanpeng Zhang","submitted_at":"2025-02-20T04:41:40Z","abstract_excerpt":"Recent advancements in Large Language Models (LLMs) and Vision-Language Models (VLMs) have made them powerful tools in embodied navigation, enabling agents to leverage commonsense and spatial reasoning for efficient exploration in unfamiliar environments. Existing LLM-based approaches convert global memory, such as semantic or topological maps, into language descriptions to guide navigation. While this improves efficiency and reduces redundant exploration, the loss of geometric information in language-based representations hinders spatial reasoning, especially in intricate environments. To add"},"claims":{"count":0,"items":[],"snapshot_sha256":"258153158e38e3291e3d48162225fcdb2d5a3ed65a07baac614ab91432fd4f57"},"source":{"id":"2502.14254","kind":"arxiv","version":2},"verdict":{"id":null,"model_set":{},"created_at":null,"strongest_claim":"","one_line_summary":"","pipeline_version":null,"weakest_assumption":"","pith_extraction_headline":""},"integrity":{"clean":true,"summary":{"advisory":0,"critical":0,"by_detector":{},"informational":0},"endpoint":"/pith/2502.14254/integrity.json","findings":[],"available":true,"detectors_run":[],"snapshot_sha256":"c28c3603d3b5d939e8dc4c7e95fa8dfce3d595e45f758748cecf8e644a296938"},"references":{"count":0,"sample":[],"resolved_work":0,"snapshot_sha256":"258153158e38e3291e3d48162225fcdb2d5a3ed65a07baac614ab91432fd4f57","internal_anchors":0},"formal_canon":{"evidence_count":0,"snapshot_sha256":"258153158e38e3291e3d48162225fcdb2d5a3ed65a07baac614ab91432fd4f57"},"author_claims":{"count":0,"strong_count":0,"snapshot_sha256":"258153158e38e3291e3d48162225fcdb2d5a3ed65a07baac614ab91432fd4f57"},"builder_version":"pith-number-builder-2026-05-17-v1"},"aliases":[{"alias_kind":"arxiv","alias_value":"2502.14254","created_at":"2026-07-05T11:19:24.763731+00:00"},{"alias_kind":"arxiv_version","alias_value":"2502.14254v2","created_at":"2026-07-05T11:19:24.763731+00:00"},{"alias_kind":"doi","alias_value":"10.48550/arxiv.2502.14254","created_at":"2026-07-05T11:19:24.763731+00:00"},{"alias_kind":"pith_short_12","alias_value":"HJV2XG5FJCGD","created_at":"2026-07-05T11:19:24.763731+00:00"},{"alias_kind":"pith_short_16","alias_value":"HJV2XG5FJCGD2G4A","created_at":"2026-07-05T11:19:24.763731+00:00"},{"alias_kind":"pith_short_8","alias_value":"HJV2XG5F","created_at":"2026-07-05T11:19:24.763731+00:00"}],"events":[],"event_summary":{},"paper_claims":[],"inbound_citations":{"count":5,"internal_anchor_count":0,"sample":[{"citing_arxiv_id":"2606.28760","citing_title":"Vision-Language Models for Deployable Social Robot Navigation: Bridging Semantic Reasoning and Low-Level Control","ref_index":53,"is_internal_anchor":false},{"citing_arxiv_id":"2606.29774","citing_title":"Analytic Concept-Centric Memory for Agentic Embodied Manipulation","ref_index":35,"is_internal_anchor":false},{"citing_arxiv_id":"2605.18729","citing_title":"Robo-Cortex: A Self-Evolving Embodied Agent via Dual-Grain Cognitive Memory and Autonomous Knowledge Induction","ref_index":44,"is_internal_anchor":false},{"citing_arxiv_id":"2604.08232","citing_title":"HiRO-Nav: Hybrid ReasOning Enables Efficient Embodied Navigation","ref_index":48,"is_internal_anchor":false},{"citing_arxiv_id":"2604.07705","citing_title":"Vision-Language Navigation for Aerial Robots: Towards the Era of Large Language Models","ref_index":152,"is_internal_anchor":false}]},"formal_canon":{"evidence_count":0,"sample":[],"anchors":[]},"links":{"html":"https://pith.science/pith/HJV2XG5FJCGD2G4AWW2SPRLM7M","json":"https://pith.science/pith/HJV2XG5FJCGD2G4AWW2SPRLM7M.json","graph_json":"https://pith.science/api/pith-number/HJV2XG5FJCGD2G4AWW2SPRLM7M/graph.json","events_json":"https://pith.science/api/pith-number/HJV2XG5FJCGD2G4AWW2SPRLM7M/events.json","paper":"https://pith.science/paper/HJV2XG5F"},"agent_actions":{"view_html":"https://pith.science/pith/HJV2XG5FJCGD2G4AWW2SPRLM7M","download_json":"https://pith.science/pith/HJV2XG5FJCGD2G4AWW2SPRLM7M.json","view_paper":"https://pith.science/paper/HJV2XG5F","resolve_alias":"https://pith.science/api/pith-number/resolve?arxiv=2502.14254&json=true","fetch_graph":"https://pith.science/api/pith-number/HJV2XG5FJCGD2G4AWW2SPRLM7M/graph.json","fetch_events":"https://pith.science/api/pith-number/HJV2XG5FJCGD2G4AWW2SPRLM7M/events.json","actions":{"anchor_timestamp":"https://pith.science/pith/HJV2XG5FJCGD2G4AWW2SPRLM7M/action/timestamp_anchor","attest_storage":"https://pith.science/pith/HJV2XG5FJCGD2G4AWW2SPRLM7M/action/storage_attestation","attest_author":"https://pith.science/pith/HJV2XG5FJCGD2G4AWW2SPRLM7M/action/author_attestation","sign_citation":"https://pith.science/pith/HJV2XG5FJCGD2G4AWW2SPRLM7M/action/citation_signature","submit_replication":"https://pith.science/pith/HJV2XG5FJCGD2G4AWW2SPRLM7M/action/replication_record"}},"created_at":"2026-07-05T11:19:24.763731+00:00","updated_at":"2026-07-05T11:19:24.763731+00:00"}