Pith. sign in

REVIEW 2 cited by

A Scenario-Based Platform for Testing Autonomous Vehicle Behavior Prediction Models in Simulation

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2110.14870 v2 pith:LIAWEVVG submitted 2021-10-28 cs.AI

classification cs.AI
keywords predictionplatformscenariostestingbehaviormodelsautonomousbehaviors
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Behavior prediction remains one of the most challenging tasks in the autonomous vehicle (AV) software stack. Forecasting the future trajectories of nearby agents plays a critical role in ensuring road safety, as it equips AVs with the necessary information to plan safe routes of travel. However, these prediction models are data-driven and trained on data collected in real life that may not represent the full range of scenarios an AV can encounter. Hence, it is important that these prediction models are extensively tested in various test scenarios involving interactive behaviors prior to deployment. To support this need, we present a simulation-based testing platform which supports (1) intuitive scenario modeling with a probabilistic programming language called Scenic, (2) specifying a multi-objective evaluation metric with a partial priority ordering, (3) falsification of the provided metric, and (4) parallelization of simulations for scalable testing. As a part of the platform, we provide a library of 25 Scenic programs that model challenging test scenarios involving interactive traffic participant behaviors. We demonstrate the effectiveness and the scalability of our platform by testing a trained behavior prediction model and searching for failure scenarios.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Unsupervised Discovery of Failure Taxonomies from Deployment Logs

    cs.RO 2025-06 conditional novelty 6.0 of 10

    An unsupervised pipeline converts robot failure videos into natural language explanations, clusters them into recurring failure types, and uses those types to guide data collection and runtime monitoring.

  2. Bayesian Optimization applied for accelerated Virtual Validation of the Autonomous Driving Function

    cs.RO 2025-07 conditional novelty 4.0 of 10

    A Bayesian optimization framework finds critical scenarios for an MPC motion planner using one to two orders of magnitude fewer simulations than full-factorial testing.

Pith tools