FlowPilot combines anchored flow matching for multimodal action pre-training with human-in-the-loop preference learning to improve long-horizon monocular sidewalk navigation, reporting 42% success in simulation and reduced interruptions in real-world tests.
Navdp: Learning sim-to-real navigation diffusion policy with privileged information guidance
10 Pith papers cite this work. Polarity classification is still indexing.
citation-role summary
citation-polarity summary
fields
cs.RO 10years
2026 10verdicts
UNVERDICTED 10roles
background 1polarities
background 1representative citing papers
A zero-shot unified agent for VLN-CE, ObjectNav, EQA and Aerial-VLN on wheeled, quadruped, humanoid and UAV platforms that translates language and vision inputs into actions via MLLMs plus TDM and SCB mechanisms, matching trained foundation models on multiple benchmarks.
NavOL collects expert trajectory labels online from a global planner during policy rollouts in simulation to train a diffusion navigation policy, mitigating distribution shift and improving performance on visual navigation tasks.
RoamFlow applies MeanFlow to predict average velocity fields for one-step action policies in image-goal navigation, trained via expert imitation followed by RL refinement.
SwarmFly is a modular MATLAB simulator for UAV swarms with leader-follower, decentralized, heterogeneous relay, and heterogeneous speed modes plus plugins for faults and analysis.
A training-free fusion layer enables stale VLM selections to improve a real-time planner's trajectory scoring for urban sidewalk navigation, yielding 30% ADE reduction in challenging scenarios.
StereoNav reaches new benchmark highs on R2R-CE and RxR-CE and improves real-robot reliability by supplying persistent target-location priors and stereo-derived geometry that stay stable under lighting changes and blur.
VISTA conditions visual navigation policies on action history and DINOv3 features to achieve scale-aware zero-shot deployment, reporting 100% goal accuracy and 95% checkpoint success in real-world outdoor, forest, and office tests.
AgenticDiffusion proposes a multi-view UAV navigation framework using language-guided reasoning, open-vocabulary grounding, vision-based diffusion planning, and NMPC, reporting 80% mission success across 40 real-world trials.
GN0 curates GN-Matrix dataset, builds 3DGS simulator and GN-Bench, and trains BAE model via supervised learning plus DAgger and RL to unify VLN tasks and outperform prior methods on GN-Bench and VLN-CE.
citing papers explorer
-
From Imitation to Alignment: Human-Preference Flow Policies for Long-Horizon Sidewalk Navigation
FlowPilot combines anchored flow matching for multimodal action pre-training with human-in-the-loop preference learning to improve long-horizon monocular sidewalk navigation, reporting 42% success in simulation and reduced interruptions in real-world tests.
-
Uni-LaViRA: Language-Vision-Robot Actions Translation for Unified Embodied Navigation
A zero-shot unified agent for VLN-CE, ObjectNav, EQA and Aerial-VLN on wheeled, quadruped, humanoid and UAV platforms that translates language and vision inputs into actions via MLLMs plus TDM and SCB mechanisms, matching trained foundation models on multiple benchmarks.
-
NavOL: Navigation Policy with Online Imitation Learning
NavOL collects expert trajectory labels online from a global planner during policy rollouts in simulation to train a diffusion navigation policy, mitigating distribution shift and improving performance on visual navigation tasks.
-
RoamFlow: Reinforcement-Aligned One-Step Action MeanFlow Policy for Image-Goal Navigation
RoamFlow applies MeanFlow to predict average velocity fields for one-step action policies in image-goal navigation, trained via expert imitation followed by RL refinement.
-
SwarmFly: A simulation platform for UAV swarm experiment design and validation
SwarmFly is a modular MATLAB simulator for UAV swarms with leader-follower, decentralized, heterogeneous relay, and heterogeneous speed modes plus plugins for faults and analysis.
-
Slow Brain, Fast Planner: Latency-Resilient VLM-Augmented Urban Navigation
A training-free fusion layer enables stale VLM selections to improve a real-time planner's trajectory scoring for urban sidewalk navigation, yielding 30% ADE reduction in challenging scenarios.
-
What Limits Vision-and-Language Navigation ?
StereoNav reaches new benchmark highs on R2R-CE and RxR-CE and improves real-robot reliability by supplying persistent target-location priors and stereo-derived geometry that stay stable under lighting changes and blur.
-
VISTA: Scale-Aware Visual Navigation via Action History Conditioning
VISTA conditions visual navigation policies on action history and DINOv3 features to achieve scale-aware zero-shot deployment, reporting 100% goal accuracy and 95% checkpoint success in real-world outdoor, forest, and office tests.
-
AgenticDiffusion: Agentic Diffusion-based Path Planning for Vision-Based UAV Navigation
AgenticDiffusion proposes a multi-view UAV navigation framework using language-guided reasoning, open-vocabulary grounding, vision-based diffusion planning, and NMPC, reporting 80% mission success across 40 real-world trials.
-
GN0: Toward a Unified Paradigm for Generation, Evaluation, and Policy Learning in Visual-Language Navigation
GN0 curates GN-Matrix dataset, builds 3DGS simulator and GN-Bench, and trains BAE model via supervised learning plus DAgger and RL to unify VLN tasks and outperform prior methods on GN-Bench and VLN-CE.