Pith. sign in

REVIEW 17 cited by

HUGSIM: A Real-Time, Photo-Realistic and Closed-Loop Simulator for Autonomous Driving

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2412.01718 v1 pith:6XFCMGNN submitted 2024-12-02 cs.CV cs.RO

HUGSIM: A Real-Time, Photo-Realistic and Closed-Loop Simulator for Autonomous Driving

classification cs.CV cs.RO
keywords closed-loopautonomousdrivinghugsimalgorithmsrenderingscenariosbenchmark
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
Share X Bluesky LinkedIn Reddit HN
read the original abstract

In the past few decades, autonomous driving algorithms have made significant progress in perception, planning, and control. However, evaluating individual components does not fully reflect the performance of entire systems, highlighting the need for more holistic assessment methods. This motivates the development of HUGSIM, a closed-loop, photo-realistic, and real-time simulator for evaluating autonomous driving algorithms. We achieve this by lifting captured 2D RGB images into the 3D space via 3D Gaussian Splatting, improving the rendering quality for closed-loop scenarios, and building the closed-loop environment. In terms of rendering, We tackle challenges of novel view synthesis in closed-loop scenarios, including viewpoint extrapolation and 360-degree vehicle rendering. Beyond novel view synthesis, HUGSIM further enables the full closed simulation loop, dynamically updating the ego and actor states and observations based on control commands. Moreover, HUGSIM offers a comprehensive benchmark across more than 70 sequences from KITTI-360, Waymo, nuScenes, and PandaSet, along with over 400 varying scenarios, providing a fair and realistic evaluation platform for existing autonomous driving algorithms. HUGSIM not only serves as an intuitive evaluation benchmark but also unlocks the potential for fine-tuning autonomous driving algorithms in a photorealistic closed-loop setting.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 17 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. MDrive: Benchmarking Closed-Loop Cooperative Driving for End-to-End Multi-agent Systems

    cs.RO 2026-05 unverdicted novelty 7.0

    MDrive benchmark shows multi-agent cooperative driving systems generally outperform single-agent ones in closed-loop settings but perception sharing does not always improve planning and negotiation can harm performanc...

  2. Fail2Drive: Benchmarking Closed-Loop Driving Generalization

    cs.RO 2026-04 conditional novelty 7.0

    Fail2Drive is the first paired-route benchmark for closed-loop generalization in CARLA, showing an average 22.8% success-rate drop on shifted scenarios and revealing failure modes such as ignoring visible LiDAR objects.

  3. Robust 4D Driving Scene Reconstruction from Imperfect Visual Priors

    cs.CV 2026-07 conditional novelty 6.0

    A self-correcting Gaussian scene graph uses semantic attention and adaptive topology updates to reconstruct dynamic driving scenes from noisy video-only priors.

  4. TerraTransfer: Learning End-to-End Driving Policies Without Expert Demonstrations

    cs.CV 2026-06 conditional novelty 6.0

    Self-play in a vector simulator trains a driving teacher; a vision student aligned by action and low-rank structural losses drives end-to-end without expert trajectory supervision.

  5. TASE: Truncation-Aware Semantic Embeddings for 3D Scene Understanding and Editing

    cs.CV 2026-06 unverdicted novelty 6.0

    TASE introduces truncation-aware semantic embeddings from 2D features for controllable text-driven 3D scene editing with added multi-view consistency via equivariance loss and diffusion finetuning.

  6. P2GS: Physical Prior-guided Gaussian Splatting for Photometrically Consistent Urban Reconstruction

    cs.CV 2026-05 unverdicted novelty 6.0

    P2GS jointly decomposes LDR images into a view-invariant linear HDR radiance field, per-view exposure scales, and tone-mapping functions without HDR supervision to enforce photometric consistency in urban Gaussian Splatting.

  7. BridgeSim: Unveiling the OL-CL Gap in End-to-End Autonomous Driving

    cs.RO 2026-04 unverdicted novelty 6.0

    The primary OL-CL gap in end-to-end autonomous driving arises from objective mismatch creating structural inability to model reactive behaviors, which a test-time adaptation method can mitigate.

  8. DriveLaW:Unifying Planning and Video Generation in a Latent Driving World

    cs.CV 2025-12 unverdicted novelty 6.0

    DriveLaW unifies video world modeling and trajectory planning by injecting video-generator latents into a diffusion planner, achieving SOTA video prediction and a new record on the NAVSIM planning benchmark.

  9. SimScale: Learning to Drive via Real-World Simulation at Scale

    cs.CV 2025-11 conditional novelty 6.0

    SimScale synthesizes unseen driving states from real logs via neural rendering and reactive environments, generates pseudo-expert trajectories, and shows that co-training on real plus simulated data improves planning ...

  10. Zero-Human Demonstration End-to-end Autonomous Driving with Trajectory Scorer

    cs.RO 2025-10 conditional novelty 6.0

    A reward-only offline RL method for trajectory planning in end-to-end autonomous driving achieves state-of-the-art on Navhard and competitive closed-loop HUGSIM performance without imitation learning.

  11. OmniNWM: Omniscient Driving Navigation World Models

    cs.CV 2025-10 conditional novelty 6.0

    OmniNWM jointly generates long panoramic multi-modal driving videos, controls them precisely via normalized Plücker ray-maps, and derives dense driving rewards from generated 3D occupancy.

  12. VERDI: VLM-Embedded Reasoning for Autonomous Driving

    cs.RO 2025-05 conditional novelty 6.0

    VERDI aligns perception, prediction, and planning outputs of end-to-end AD models with VLM-generated text features at training time to embed structured reasoning, yielding up to 11% better l2 distance and 10% higher n...

  13. Robust 4D Driving Scene Reconstruction from Imperfect Visual Priors

    cs.CV 2026-07 unverdicted novelty 5.5

    AGG reconstructs 4D driving scenes from noisy in-the-wild video by decoupling camera/background updates from agents and adaptively fixing the scene-graph topology.

  14. TerraTransfer: Learning End-to-End Driving Policies Without Expert Demonstrations

    cs.CV 2026-06 unverdicted novelty 5.0

    TerraTransfer decouples self-play policy pretraining from vision alignment via KL divergence and low-rank loss to produce end-to-end driving policies without expert demonstrations, matching prior methods on closed-loo...

  15. NVIDIA OmniDreams: Real-Time Generative World Model for Closed-Loop Autonomous Vehicle Simulation

    cs.CV 2026-06 conditional novelty 5.0

    OmniDreams turns the Cosmos video-diffusion model into a real-time, action-conditioned driving simulator and shows that its internal representations can be fine-tuned into a competitive driving policy.

  16. RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework

    cs.CV 2026-04 unverdicted novelty 5.0

    RAD-2 uses a diffusion generator and RL discriminator to cut collision rates by 56% in closed-loop autonomous driving planning.

  17. NVIDIA OmniDreams: Real-Time Generative World Model for Closed-Loop Autonomous Vehicle Simulation

    cs.CV 2026-06 unverdicted novelty 4.0

    OmniDreams is a real-time generative world model mid- and post-trained from the Cosmos diffusion model on 21k hours of driving data to autoregressively generate action-conditioned videos for closed-loop AV simulation.