REVIEW 22 cited by
MultiPath: Multiple Probabilistic Anchor Trajectory Hypotheses for Behavior Prediction
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Predicting human behavior is a difficult and crucial task required for motion planning. It is challenging in large part due to the highly uncertain and multi-modal set of possible outcomes in real-world domains such as autonomous driving. Beyond single MAP trajectory prediction, obtaining an accurate probability distribution of the future is an area of active interest. We present MultiPath, which leverages a fixed set of future state-sequence anchors that correspond to modes of the trajectory distribution. At inference, our model predicts a discrete distribution over the anchors and, for each anchor, regresses offsets from anchor waypoints along with uncertainties, yielding a Gaussian mixture at each time step. Our model is efficient, requiring only one forward inference pass to obtain multi-modal future distributions, and the output is parametric, allowing compact communication and analytical probabilistic queries. We show on several datasets that our model achieves more accurate predictions, and compared to sampling baselines, does so with an order of magnitude fewer trajectories.
Forward citations
Cited by 22 Pith papers
-
ILNet: Trajectory Prediction with Inverse Learning Attention for Enhancing Intention Capture
ILNet reports top INTERACTION joint metrics and strong Argoverse marginal metrics using inverse temporal attention plus learned dynamic anchor refinement.
-
SRefiner: Soft-Braid Attention for Multi-Agent Trajectory Refinement
SRefiner improves multi-agent trajectory prediction accuracy by using soft-braid attention that encodes closeness and motion at nearest trajectory and lane points.
-
Autoregressive Meta-Actions for Unified Controllable Trajectory Generation
Frame-level meta-actions, predicted and injected at every time step in an autoregressive trajectory model, improve alignment between high-level driving decisions and generated motion.
-
Robust Planning for Autonomous Driving via Mixed Adversarial Diffusion Predictions
The authors mix normal and adversarially biased diffusion predictions under expected cost, and report a closed-loop score of 86.6 versus 83.5 for the best baseline in three adversarial driving scenarios.
-
Distilling Multi-modal Large Language Models for Autonomous Driving
DiMA jointly trains a vision-only planner with an LLM and uses auxiliary language, reconstruction, and scene editing tasks to improve planning on nuScenes while dropping the LLM at inference.
-
Resonance: Learning to Predict Social-Aware Pedestrian Trajectories as Co-Vibrations
A trajectory prediction model that decomposes forecasts into a linear base, a self-sourced vibration, and a social resonance vibration, achieving strong benchmark results with an interpretable decomposition.
-
A Dynamic Scene Interaction Reasoning Framework for Scene-level Lane-Change Intention and Trajectory Prediction of Multiple Interacting Vehicles
A dynamic graph-attention model jointly predicts every nearby vehicle’s lane-change intention and trajectory, cutting trajectory error by up to ~53% and improving scene coherence on NGSIM and highD.
-
Pivot-Centric Trajectory Prediction: Bridging Long Horizons via Dynamical Guidance
PCTP predicts multiple hierarchical "pivot" waypoints first and then refines the segments between them, improving long-horizon trajectory prediction on Argoverse I and II.
-
TGRIP: A Text-Guided Approach to Vehicle Instance Prediction in Autonomous Driving
Auxiliary CLIP-derived BEV semantic supervision during training improves nuScenes end-to-end vehicle instance prediction over a geometric-only baseline, with the semantic head removed at inference.
-
Adaptive Output Steps: FlexiSteps Network for Dynamic Trajectory Prediction
FSN dynamically selects the number of future trajectory steps to predict, using a learned classifier plus a Fréchet-distance-based score, claiming improved accuracy and efficiency on two driving benchmarks.
-
Foresight in Motion: Reinforcing Trajectory Prediction with Reward Heuristics
FiM predicts future trajectories by first learning a reward distribution over a grid world via inverse reinforcement learning, then rolling out intention plans that condition a Mamba-enhanced trajectory decoder.
-
RoCA: Robust Cross-Domain End-to-End Autonomous Driving
RoCA, a Gaussian-process codebook over ego and agent tokens, improves cross-domain generalization and adaptation of end-to-end autonomous driving models without extra inference cost.
-
CogAD: Cognitive-Hierarchy Guided End-to-End Autonomous Driving
CogAD reports state-of-the-art open-loop and closed-loop planning results by combining hierarchical scene-to-instance perception with intent-to-trajectory planning and dual-level uncertainty.
-
HAMF: A Hybrid Attention-Mamba Framework for Joint Scene Context Understanding and Future Motion Representation Learning
HAMF feeds learnable future motion tokens into the scene encoder alongside road and agent tokens, then uses a Mamba decoder to output six diverse trajectories, achieving competitive Argoverse 2 results with 3.0M parameters.
-
Direct Preference Optimization-Enhanced Multi-Guided Diffusion Model for Traffic Scenario Generation
MuDi-Pro fine-tunes a multi-guided diffusion transformer with DPO using guidance-score preferences to improve controllability of traffic scenario generation on nuScenes.
-
GaussianAD: Gaussian-Centric End-to-End Autonomous Driving
GaussianAD uses sparse 3D semantic Gaussians as the intermediate representation for camera-only end-to-end driving, adding Gaussian flow prediction and future-scene supervision to achieve strong open-loop planning res...
-
A Spatio-temporal Continuous Network for Stochastic 3D Human Motion Prediction
STCN predicts stochastic 3D human futures by combining a spatio-temporal continuous network with anchor-based Gaussian mixture sampling.
-
Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections
A hierarchical RL agent with a goal-conditioned collision prediction module achieves 94.7% success and 3.3% collisions in SMARTS intersection tasks, outperforming flat RL baselines.
-
Learning Soft Driving Constraints from Vectorized Scene Embeddings while Imitating Expert Trajectories
A motion planner that learns soft driving constraints from vectorized scene embeddings improves closed-loop safety and interpretability over a reward-only imitation-learning baseline.
-
Map-Free Trajectory Prediction with Map Distillation and Hierarchical Encoding
MFTP distills HD-map priors into a map-free trajectory predictor and reports state-of-the-art minADE, minFDE, and MR on Argoverse among the compared map-free methods.
-
PhysVarMix: Physics-Informed Variational Mixture Model for Multi-Modal Trajectory Prediction
PhysVarMix is a claimed physics-informed variational mixture model for multimodal trajectory prediction, but its latent variable never affects the predicted trajectories, so it reduces to a standard mixture density ne...
-
Generative AI for Autonomous Driving: Frontiers and Opportunities
A comprehensive, structured survey of generative AI for autonomous driving, covering model families, sensor modalities, real-world applications, and open research challenges.
Discussion (0). Continue with ORCID to comment.