Pith. sign in

REVIEW 7 cited by

RayFronts: Open-Set Semantic Ray Frontiers for Online Scene Understanding and Exploration

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2504.06994 v1 pith:WBRZBDEC submitted 2025-04-09 cs.RO cs.AIcs.CVcs.LG

classification cs.ROcs.AIcs.CVcs.LG
keywords beyond-rangerayfrontsmappingonlinesemanticopen-setsearchsemantics
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Open-set semantic mapping is crucial for open-world robots. Current mapping approaches either are limited by the depth range or only map beyond-range entities in constrained settings, where overall they fail to combine within-range and beyond-range observations. Furthermore, these methods make a trade-off between fine-grained semantics and efficiency. We introduce RayFronts, a unified representation that enables both dense and beyond-range efficient semantic mapping. RayFronts encodes task-agnostic open-set semantics to both in-range voxels and beyond-range rays encoded at map boundaries, empowering the robot to reduce search volumes significantly and make informed decisions both within & beyond sensory range, while running at 8.84 Hz on an Orin AGX. Benchmarking the within-range semantics shows that RayFronts's fine-grained image encoding provides 1.34x zero-shot 3D semantic segmentation performance while improving throughput by 16.5x. Traditionally, online mapping performance is entangled with other system components, complicating evaluation. We propose a planner-agnostic evaluation framework that captures the utility for online beyond-range search and exploration, and show RayFronts reduces search volume 2.2x more efficiently than the closest online baselines.

Discussion (0). Sign in to comment.

Forward citations

Cited by 7 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. SignScene: Visual Sign Grounding for Mapless Navigation

    cs.RO 2026-02 conditional novelty 6.0 of 10

    A sign-centric abstract map lets a vision-language model turn navigational sign instructions into correct paths 88.6% of the time across nine environment types.

  2. RADSeg: Unleashing Parameter and Compute Efficient Zero-Shot Open-Vocabulary Segmentation Using Agglomerative Models

    cs.CV 2025-11 unverdicted novelty 6.0 of 10

    RADSeg adapts the RADIO model with targeted enhancements to deliver 6-30% higher mIoU in zero-shot OVSS while using 2.5x fewer parameters and running 3.95x faster than prior large-model combinations.

  3. Don't Fool Me Twice: Adapting to Adversity in the Wild with Experience-Driven Reasoning

    cs.RO 2026-05 conditional novelty 5.0 of 10

    By detecting trajectory disturbances, attributing them to visual causes with a VLM, and fitting a few-shot spatial disturbance model, robots build personalized danger libraries that improve later navigation.

  4. G-DRAGON: Geospatial Reasoning and Dynamic Planning for Retrieval-Augmented Outdoor Navigation

    cs.RO 2026-05 unverdicted novelty 5.0 of 10

    G-DRAGON framework maps language commands to OSM coordinates via lightweight LLM for global planning and uses frontier exploration for local targets, outperforming baselines in simulation and completing real UGV perso...

  5. FUS3DMaps: Scalable and Accurate Open-Vocabulary Semantic Mapping by 3D Fusion of Voxel- and Instance-Level Layers

    cs.RO 2026-05 unverdicted novelty 5.0 of 10

    FUS3DMaps fuses voxel- and instance-level open-vocabulary layers inside a shared 3D voxel map to improve both layers and enable scalable accurate semantic mapping.

  6. Don't Fool Me Twice: Adapting to Adversity in the Wild with Experience-Driven Reasoning

    cs.RO 2026-05 unverdicted novelty 4.0 of 10

    A robotics framework combines VLMs and kernel regression for online learning from embodiment-specific disturbances to enable better adaptation in unstructured environments.

  7. PLAF: Pixel-wise Language-Aligned Feature Extraction for Efficient 3D Scene Understanding

    cs.CV 2026-04 unverdicted novelty 4.0 of 10

    PLAF introduces a 2D pixel-wise language-aligned feature extractor paired with a redundancy-reducing storage scheme that supports accurate open-vocabulary 3D scene understanding.

Pith tools