Pith. sign in

REVIEW 19 cited by

iGibson 2.0: Object-Centric Simulation for Robot Learning of Everyday Household Tasks

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2108.03272 v4 pith:SXPN6A4G submitted 2021-08-06 cs.RO cs.AIcs.CVcs.LG

classification cs.ROcs.AIcs.CVcs.LG
keywords igibsontaskssimulationstateslearninglogicrobotcollect
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

Recent research in embodied AI has been boosted by the use of simulation environments to develop and train robot learning approaches. However, the use of simulation has skewed the attention to tasks that only require what robotics simulators can simulate: motion and physical contact. We present iGibson 2.0, an open-source simulation environment that supports the simulation of a more diverse set of household tasks through three key innovations. First, iGibson 2.0 supports object states, including temperature, wetness level, cleanliness level, and toggled and sliced states, necessary to cover a wider range of tasks. Second, iGibson 2.0 implements a set of predicate logic functions that map the simulator states to logic states like Cooked or Soaked. Additionally, given a logic state, iGibson 2.0 can sample valid physical states that satisfy it. This functionality can generate potentially infinite instances of tasks with minimal effort from the users. The sampling mechanism allows our scenes to be more densely populated with small objects in semantically meaningful locations. Third, iGibson 2.0 includes a virtual reality (VR) interface to immerse humans in its scenes to collect demonstrations. As a result, we can collect demonstrations from humans on these new types of tasks, and use them for imitation learning. We evaluate the new capabilities of iGibson 2.0 to enable robot learning of novel tasks, in the hope of demonstrating the potential of this new simulator to support new research in embodied AI. iGibson 2.0 and its new dataset are publicly available at http://svl.stanford.edu/igibson/.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 19 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. EgoExoBench: A Benchmark for First- and Third-person View Video Understanding in MLLMs

    cs.CV 2025-07 conditional novelty 7.0 of 10

    EgoExoBench introduces 7,330 multiple-choice questions across 11 tasks that measure how well multimodal AI models align, associate, and temporally reason across paired first-person and third-person videos; the best mo...

  2. BioVLN: A Simulation Platform for Visual Language Navigation in Biomedical Laboratories

    cs.RO 2026-07 conditional novelty 6.0 of 10

    BioVLN introduces a three-zone operational envelope for biomedical lab navigation, with 47 scenes and benchmarks showing multi-point operation-area goals raise success to 83–92% while cutting unsafe proximity.

  3. DISCOVERSE: Efficient Robot Simulation in Complex High-Fidelity Environments

    cs.RO 2025-07 conditional novelty 6.0 of 10

    A simulation framework that combines 3D Gaussian Splatting with MuJoCo reports improved zero-shot transfer of manipulation policies from simulation to real robots.

  4. StarDojo: Benchmarking Open-Ended Behaviors of Agentic Multimodal LLMs in Production-Living Simulations with Stardew Valley

    cs.AI 2025-07 conditional novelty 6.0 of 10

    StarDojo is a 1,000-task benchmark in Stardew Valley combining production and social activities, and the best tested MLLM (GPT-4.1) achieves only 12.7% success on its 100-task subset.

  5. DIPO: Dual-State Images Controlled Articulated Object Generation Powered by Diverse Data

    cs.CV 2025-05 conditional novelty 6.0 of 10

    DIPO generates articulated 3D objects from a closed and an open image, and the new PM-X dataset improves generalization to complex objects.

  6. Subtle Risks, Critical Failures: A Framework for Diagnosing Physical Safety of LLMs for Embodied Decision Making

    cs.AI 2025-05 conditional novelty 6.0 of 10

    A new PDDL benchmark and modular test suite shows that LLMs reject clear safety violations but fail to plan around subtle physical hazards.

  7. MetaScenes: Towards Automated Replica Creation for Real-world 3D Scans

    cs.CV 2025-05 conditional novelty 6.0 of 10

    MetaScenes converts 706 ScanNet scenes into simulatable 3D replicas with 15,366 objects and ranked candidate assets, and introduces Scan2Sim for automated asset replacement.

  8. Episodic Novelty Through Temporal Distance

    cs.LG 2025-01 conditional novelty 6.0 of 10

    An episodic intrinsic reward based on a contrastively learned temporal distance quasimetric improves exploration in sparse-reward Contextual MDPs.

  9. AAM-SEALS: Developing Aerial-Aquatic Manipulators in SEa, Air, and Land Simulator

    cs.RO 2024-12 conditional novelty 6.0 of 10

    A simulator integrates particle-based water, aerial flight, manipulation, and animal models to train aerial-aquatic robots across air and water.

  10. Locate n' Rotate: Two-stage Openable Part Detection with Foundation Model Priors

    cs.CV 2024-12 conditional novelty 6.0 of 10

    MOPD improves openable part detection and motion parameter prediction by fusing perceptual grouping and geometric priors from foundation models into a two-decoder transformer with a motion-aware optimal transport matc...

  11. ClevrSkills: Compositional Language and Visual Reasoning in Robotics

    cs.RO 2024-11 conditional novelty 6.0 of 10

    ClevrSkills provides a 33-task, 330k-trajectory benchmark showing that vision-language robot policies struggle to compose base manipulation skills into novel long-horizon tasks.

  12. Physics-informed Neural Motion Planning via Domain Decomposition in Large Environments

    cs.RO 2025-06 conditional novelty 5.0 of 10

    A domain-decomposed neural field predicts cost-to-go as a latent-space distance, enabling physics-informed motion planning in large and real-world environments.

  13. Physics-informed Temporal Difference Metric Learning for Robot Motion Planning

    cs.RO 2025-05 conditional novelty 5.0 of 10

    A self-supervised neural planner combines Eikonal, temporal difference, obstacle alignment, and causality losses with a learned L1 and L-infinity metric, improving success rates and generalization in robot motion planning.

  14. InfiniteWorld: A Unified Scalable Simulation Framework for General Visual-Language Robot Interaction

    cs.RO 2024-12 conditional novelty 5.0 of 10

    InfiniteWorld presents an Isaac Sim based simulator with unified assets and four benchmarks, including scene graph exploration and social mobile manipulation, but reports zero success on the main social task.

  15. Perception, Reason, Think, and Plan: A Survey on Large Multimodal Reasoning Models

    cs.CV 2025-05 conditional novelty 4.0 of 10

    A survey organizes multimodal reasoning research into a staged roadmap and proposes native large multimodal reasoning models that unify perception, generation, and agentic planning.

  16. Demonstrating DVS: Dynamic Virtual-Real Simulation Platform for Mobile Robotic Tasks

    cs.RO 2025-04 conditional novelty 4.0 of 10

    DVS is a virtual-real simulation platform that combines dynamic pedestrian modeling, optical motion capture, and ROS communication for closed-loop mobile robot research.

  17. Toward Generalizable Evaluation in the LLM Era: A Survey Beyond Benchmarks

    cs.CL 2025-04 conditional novelty 4.0 of 10

    A survey that frames the core problem of LLM evaluation as 'evaluation generalization': finite test sets cannot scale with unbounded model capabilities, and proposes two transitions in evaluation design.

  18. A Survey of Robotic Navigation and Manipulation with Physics Simulators in the Era of Embodied AI

    cs.RO 2025-05 conditional novelty 2.0 of 10

    A review of navigation and manipulation simulators, datasets, and methods, framed around the sim-to-real gap.

  19. A Survey of State of the Art Large Vision Language Models: Alignment, Benchmark, Evaluations and Challenges

    cs.CV 2025-01 reject novelty 2.0 of 10

    A survey that catalogs large vision-language models, their alignment methods, benchmarks, and challenges, but is compromised by inconsistent counts and misclassified entries.

Pith tools