{"record_type":"pith_number_record","schema_url":"https://pith.science/schemas/pith-number/v1.json","pith_number":"pith:2023:6VMGCJ4T6T2N7654RVQSEU76VZ","short_pith_number":"pith:6VMGCJ4T","schema_version":"1.0","canonical_sha256":"f558612793f4f4dffbbc8d612253feae4b3c5643e2e6805d8dfc88a1f75d3689","source":{"kind":"arxiv","id":"2301.00493","version":1},"attestation_state":"computed","paper":{"title":"Argoverse 2: Next Generation Datasets for Self-Driving Perception and Forecasting","license":"http://arxiv.org/licenses/nonexclusive-distrib/1.0/","headline":"Argoverse 2 releases three large datasets to support new research in self-driving perception and forecasting.","cross_cats":["cs.AI","cs.LG","cs.RO"],"primary_cat":"cs.CV","authors_text":"Andrew Hartnett, Benjamin Wilson, Bowen Pan, Deva Ramanan, Jagjeet Singh, James Hays, Jhony Kaesemodel Pontes, John Lambert, Peter Carr, Ratnesh Kumar, Siddhesh Khandelwal, Tanmay Agarwal, William Qi","submitted_at":"2023-01-02T00:36:22Z","abstract_excerpt":"We introduce Argoverse 2 (AV2) - a collection of three datasets for perception and forecasting research in the self-driving domain. The annotated Sensor Dataset contains 1,000 sequences of multimodal data, encompassing high-resolution imagery from seven ring cameras, and two stereo cameras in addition to lidar point clouds, and 6-DOF map-aligned pose. Sequences contain 3D cuboid annotations for 26 object categories, all of which are sufficiently-sampled to support training and evaluation of 3D perception models. The Lidar Dataset contains 20,000 sequences of unlabeled lidar point clouds and ma"},"verification_status":{"content_addressed":true,"pith_receipt":true,"author_attested":false,"weak_author_claims":0,"strong_author_claims":0,"externally_anchored":false,"storage_verified":false,"citation_signatures":0,"replication_records":0,"graph_snapshot":true,"references_resolved":true,"formal_links_present":false},"canonical_record":{"source":{"id":"2301.00493","kind":"arxiv","version":1},"metadata":{"license":"http://arxiv.org/licenses/nonexclusive-distrib/1.0/","primary_cat":"cs.CV","submitted_at":"2023-01-02T00:36:22Z","cross_cats_sorted":["cs.AI","cs.LG","cs.RO"],"title_canon_sha256":"5c3fdb0e957ac924d7ec1520c41f1b32a330fb198bf3f3d998c1939feeb50f17","abstract_canon_sha256":"d272b78f0db95ded10d197eff9463d8ca1586ed60993f83e8772f5ec62d97eee"},"schema_version":"1.0"},"receipt":{"kind":"pith_receipt","key_id":"pith-v1-2026-05","algorithm":"ed25519","signed_at":"2026-07-05T05:29:46.939665Z","signature_b64":"dpWRH8PqCDkDpgVn99RW7MOAjl32kbRa3diS+A6N6ktvpoeg3fkr5SMNBDSzzY5hcWE3sIGUxiibDwRv9W8hDA==","signed_message":"canonical_sha256_bytes","builder_version":"pith-number-builder-2026-05-17-v1","receipt_version":"0.3","canonical_sha256":"f558612793f4f4dffbbc8d612253feae4b3c5643e2e6805d8dfc88a1f75d3689","last_reissued_at":"2026-07-05T05:29:46.939095Z","signature_status":"signed_v1","first_computed_at":"2026-07-05T05:29:46.939095Z","public_key_fingerprint":"8d4b5ee74e4693bcd1df2446408b0d54"},"graph_snapshot":{"paper":{"title":"Argoverse 2: Next Generation Datasets for Self-Driving Perception and Forecasting","license":"http://arxiv.org/licenses/nonexclusive-distrib/1.0/","headline":"Argoverse 2 releases three large datasets to support new research in self-driving perception and forecasting.","cross_cats":["cs.AI","cs.LG","cs.RO"],"primary_cat":"cs.CV","authors_text":"Andrew Hartnett, Benjamin Wilson, Bowen Pan, Deva Ramanan, Jagjeet Singh, James Hays, Jhony Kaesemodel Pontes, John Lambert, Peter Carr, Ratnesh Kumar, Siddhesh Khandelwal, Tanmay Agarwal, William Qi","submitted_at":"2023-01-02T00:36:22Z","abstract_excerpt":"We introduce Argoverse 2 (AV2) - a collection of three datasets for perception and forecasting research in the self-driving domain. The annotated Sensor Dataset contains 1,000 sequences of multimodal data, encompassing high-resolution imagery from seven ring cameras, and two stereo cameras in addition to lidar point clouds, and 6-DOF map-aligned pose. Sequences contain 3D cuboid annotations for 26 object categories, all of which are sufficiently-sampled to support training and evaluation of 3D perception models. The Lidar Dataset contains 20,000 sequences of unlabeled lidar point clouds and ma"},"claims":{"count":4,"items":[{"kind":"strongest_claim","text":"We believe these datasets will support new and existing machine learning research problems in ways that existing datasets do not.","source":"verdict.strongest_claim","status":"machine_extracted","claim_id":"C1","attestation":"unclaimed"},{"kind":"weakest_assumption","text":"That the provided annotations are accurate enough and the selected scenarios sufficiently representative to drive meaningful improvements in deployed self-driving systems.","source":"verdict.weakest_assumption","status":"machine_extracted","claim_id":"C2","attestation":"unclaimed"},{"kind":"one_line_summary","text":"Argoverse 2 introduces three new datasets with annotated sensor data, massive lidar collections, and challenging motion forecasting scenarios for autonomous driving research.","source":"verdict.one_line_summary","status":"machine_extracted","claim_id":"C3","attestation":"unclaimed"},{"kind":"headline","text":"Argoverse 2 releases three large datasets to support new research in self-driving perception and forecasting.","source":"verdict.pith_extraction.headline","status":"machine_extracted","claim_id":"C4","attestation":"unclaimed"}],"snapshot_sha256":"c1aebfb621b3bb1bbe9c690e6b34ff203451091304b61e6561bfb938ea30506c"},"source":{"id":"2301.00493","kind":"arxiv","version":1},"verdict":{"id":"077689b3-084e-429b-bf06-0e62bc5ebfd6","model_set":{"reader":"grok-4.3"},"created_at":"2026-05-12T20:10:13.850741Z","strongest_claim":"We believe these datasets will support new and existing machine learning research problems in ways that existing datasets do not.","one_line_summary":"Argoverse 2 introduces three new datasets with annotated sensor data, massive lidar collections, and challenging motion forecasting scenarios for autonomous driving research.","pipeline_version":"pith-pipeline@v0.9.0","weakest_assumption":"That the provided annotations are accurate enough and the selected scenarios sufficiently representative to drive meaningful improvements in deployed self-driving systems.","pith_extraction_headline":"Argoverse 2 releases three large datasets to support new research in self-driving perception and forecasting."},"integrity":{"clean":true,"summary":{"advisory":0,"critical":0,"by_detector":{},"informational":0},"endpoint":"/pith/2301.00493/integrity.json","findings":[],"available":true,"detectors_run":[],"snapshot_sha256":"c28c3603d3b5d939e8dc4c7e95fa8dfce3d595e45f758748cecf8e644a296938"},"references":{"count":54,"sample":[{"doi":"","year":2019,"title":"SemanticKITTI: A dataset for semantic scene understanding of lidar sequences","work_id":"27a8f474-426a-44fe-84d3-baa5116bce2d","ref_index":1,"cited_arxiv_id":"","is_internal_anchor":false},{"doi":"","year":2020,"title":"Range conditioned dilated convolutions for scale invariant 3d object detection","work_id":"99a248a9-df28-4e84-9041-34ce5b65aced","ref_index":2,"cited_arxiv_id":"","is_internal_anchor":false},{"doi":"","year":2005,"title":"Language Models are Few-Shot Learners","work_id":"214732c0-2edd-44a0-af9e-28184a2b8279","ref_index":3,"cited_arxiv_id":"2005.14165","is_internal_anchor":true},{"doi":"","year":2020,"title":"Lang, Sourabh V ora, Venice Erin Liong, Qiang Xu, Anush Krishnan, Yu Pan, Giancarlo Baldan, and Oscar Beijbom","work_id":"09126d7a-6b02-4972-b45c-7ff87220ed10","ref_index":4,"cited_arxiv_id":"","is_internal_anchor":false},{"doi":"","year":2021,"title":"To the point: Efﬁcient 3d object detection in the range image with graph convolution kernels","work_id":"a733b164-6e16-483c-ac28-24feabff3117","ref_index":5,"cited_arxiv_id":"","is_internal_anchor":false}],"resolved_work":54,"snapshot_sha256":"8b49227c34dae52f88f7ca1bb866e3007c8c1293c2ac943d038693752d37418f","internal_anchors":2},"formal_canon":{"evidence_count":0,"snapshot_sha256":"258153158e38e3291e3d48162225fcdb2d5a3ed65a07baac614ab91432fd4f57"},"author_claims":{"count":0,"strong_count":0,"snapshot_sha256":"258153158e38e3291e3d48162225fcdb2d5a3ed65a07baac614ab91432fd4f57"},"builder_version":"pith-number-builder-2026-05-17-v1"},"aliases":[{"alias_kind":"arxiv","alias_value":"2301.00493","created_at":"2026-07-05T05:29:46.939158+00:00"},{"alias_kind":"arxiv_version","alias_value":"2301.00493v1","created_at":"2026-07-05T05:29:46.939158+00:00"},{"alias_kind":"doi","alias_value":"10.48550/arxiv.2301.00493","created_at":"2026-07-05T05:29:46.939158+00:00"},{"alias_kind":"pith_short_12","alias_value":"6VMGCJ4T6T2N","created_at":"2026-07-05T05:29:46.939158+00:00"},{"alias_kind":"pith_short_16","alias_value":"6VMGCJ4T6T2N7654","created_at":"2026-07-05T05:29:46.939158+00:00"},{"alias_kind":"pith_short_8","alias_value":"6VMGCJ4T","created_at":"2026-07-05T05:29:46.939158+00:00"}],"events":[],"event_summary":{},"paper_claims":[],"inbound_citations":{"count":66,"internal_anchor_count":66,"sample":[{"citing_arxiv_id":"2607.05705","citing_title":"IMR: Iterative Mode-World Weighted Regression for Multi-Agent Trajectory Prediction","ref_index":7,"is_internal_anchor":true},{"citing_arxiv_id":"2607.06838","citing_title":"WildCity: A Real-World City-Scale Testbed for Rendering, Simulation, and Spatial Intelligence","ref_index":39,"is_internal_anchor":true},{"citing_arxiv_id":"2607.07103","citing_title":"A knowledge-augmented dataset of high-risk driving scenarios with LLM annotations for autonomous driving","ref_index":2,"is_internal_anchor":true},{"citing_arxiv_id":"2606.26424","citing_title":"Rethinking Training & Inference for Forecasting: Linking Winner-Take-All back to GMMs","ref_index":34,"is_internal_anchor":true},{"citing_arxiv_id":"2606.27317","citing_title":"OctoSense: Self-Supervised Learning for Multimodal Robot Perception","ref_index":43,"is_internal_anchor":true},{"citing_arxiv_id":"2606.22617","citing_title":"OmniSpace: Efficient Geometry Awareness for Autonomous Vehicles MLLMs","ref_index":65,"is_internal_anchor":true},{"citing_arxiv_id":"2606.21344","citing_title":"Mind the Noise: Sensitivity of Transformer-based Interaction-Aware Trajectory Prediction Models to Noisy Data","ref_index":9,"is_internal_anchor":true},{"citing_arxiv_id":"2606.20725","citing_title":"D2HDMap: Non-visible Driveline Map Prior for Online Vectorized HD Map Prediction","ref_index":32,"is_internal_anchor":true},{"citing_arxiv_id":"2606.17080","citing_title":"HRDX: A Large-Scale Vector HD-Map Dataset","ref_index":28,"is_internal_anchor":true},{"citing_arxiv_id":"2606.19370","citing_title":"Human-like autonomy emerges from self-play and a pinch of human data","ref_index":37,"is_internal_anchor":true},{"citing_arxiv_id":"2606.11874","citing_title":"AutoMine Solution for AV2 2026 Scenario Mining Challenge","ref_index":6,"is_internal_anchor":true},{"citing_arxiv_id":"2606.11739","citing_title":"Multi-View In-Cabin Monitoring System for Public Transport Vehicles","ref_index":14,"is_internal_anchor":true},{"citing_arxiv_id":"2606.10641","citing_title":"CAMASA: A CAM-based Dataset from the MASA Living Lab","ref_index":3,"is_internal_anchor":true},{"citing_arxiv_id":"2606.11120","citing_title":"Monte Carlo Pass Search: Using Trajectory Generation for 3D Counterfactual Pass Evaluation in Football","ref_index":29,"is_internal_anchor":true},{"citing_arxiv_id":"2606.09882","citing_title":"WHU-Infra3D: A Full-stack Multi-modal Dataset and Benchmark for 3D Roadside Infrastructure Inventory","ref_index":9,"is_internal_anchor":true},{"citing_arxiv_id":"2606.02379","citing_title":"Honey, I Shrunk the Arc de Triomphe!","ref_index":38,"is_internal_anchor":true},{"citing_arxiv_id":"2605.31572","citing_title":"nuReasoning: A Reasoning-Centric Dataset and Benchmark for Long-Tail Autonomous Driving","ref_index":70,"is_internal_anchor":true},{"citing_arxiv_id":"2606.31844","citing_title":"Bridging Local Observation and Global Simulation in Closed-Loop Traffic Modeling","ref_index":32,"is_internal_anchor":true},{"citing_arxiv_id":"2606.31814","citing_title":"Generative Lane Topology Reasoning via Autoregressive Model with Geometry Prior","ref_index":48,"is_internal_anchor":true},{"citing_arxiv_id":"2604.24119","citing_title":"TopoHR: Hierarchical Centerline Representation for Cyclic Topology Reasoning in Driving Scenes with Point-to-Instance Relations","ref_index":29,"is_internal_anchor":true},{"citing_arxiv_id":"2605.24037","citing_title":"Mode-as-Sequence: Translating Multimodal Motion Prediction into Unified Sequential Mode Modeling","ref_index":2,"is_internal_anchor":true},{"citing_arxiv_id":"2606.29716","citing_title":"AerialMetric: Benchmarking and Adapting UAV Monocular Metric Depth Estimation in the Real World","ref_index":69,"is_internal_anchor":true},{"citing_arxiv_id":"2606.02379","citing_title":"Honey, I Shrunk the Arc de Triomphe!","ref_index":38,"is_internal_anchor":true},{"citing_arxiv_id":"2605.28552","citing_title":"Modeling Vehicle-Type-Specific Pedestrian Crash Avoidance Behavior in Safety-Critical Interactions Using Smooth-Mamba Deep Reinforcement Learning","ref_index":10,"is_internal_anchor":true},{"citing_arxiv_id":"2605.30561","citing_title":"VLM3: Vision Language Models Are Native 3D Learners","ref_index":17,"is_internal_anchor":true}]},"formal_canon":{"evidence_count":0,"sample":[],"anchors":[]},"links":{"html":"https://pith.science/pith/6VMGCJ4T6T2N7654RVQSEU76VZ","json":"https://pith.science/pith/6VMGCJ4T6T2N7654RVQSEU76VZ.json","graph_json":"https://pith.science/api/pith-number/6VMGCJ4T6T2N7654RVQSEU76VZ/graph.json","events_json":"https://pith.science/api/pith-number/6VMGCJ4T6T2N7654RVQSEU76VZ/events.json","paper":"https://pith.science/paper/6VMGCJ4T"},"agent_actions":{"view_html":"https://pith.science/pith/6VMGCJ4T6T2N7654RVQSEU76VZ","download_json":"https://pith.science/pith/6VMGCJ4T6T2N7654RVQSEU76VZ.json","view_paper":"https://pith.science/paper/6VMGCJ4T","resolve_alias":"https://pith.science/api/pith-number/resolve?arxiv=2301.00493&json=true","fetch_graph":"https://pith.science/api/pith-number/6VMGCJ4T6T2N7654RVQSEU76VZ/graph.json","fetch_events":"https://pith.science/api/pith-number/6VMGCJ4T6T2N7654RVQSEU76VZ/events.json","actions":{"anchor_timestamp":"https://pith.science/pith/6VMGCJ4T6T2N7654RVQSEU76VZ/action/timestamp_anchor","attest_storage":"https://pith.science/pith/6VMGCJ4T6T2N7654RVQSEU76VZ/action/storage_attestation","attest_author":"https://pith.science/pith/6VMGCJ4T6T2N7654RVQSEU76VZ/action/author_attestation","sign_citation":"https://pith.science/pith/6VMGCJ4T6T2N7654RVQSEU76VZ/action/citation_signature","submit_replication":"https://pith.science/pith/6VMGCJ4T6T2N7654RVQSEU76VZ/action/replication_record"}},"created_at":"2026-07-05T05:29:46.939158+00:00","updated_at":"2026-07-05T05:29:46.939158+00:00"}