Pith. sign in

Paper Citation Record · LEDGER

Spatially Visual Perception for End-to-End Robotic Learning

As of 17 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 1 inbound Pith citation observation for arXiv:2411.17458.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.17458 v1

Coverage vector

measured 54 of 54 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T12:11:06.173763Z

measured 55 of 55 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:53:41.125211Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T13:53:48.384530Z

Reference resolution

54 of 54 outbound references displayed

  • verified exact0
  • verified fuzzy38
  • unresolved16
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 05292ab4-7592-4685-bbe2-7d00d23bbcda · outbound

This paper cites ReMEmbR: Building and Reasoning Over Long-Horizon Spatio-Temporal Memory for Robot Navigation.

Spatially Visual Perception for End-to-End Robotic Learning ReMEmbR: Building and Reasoning Over Long-Horizon Spatio-Temporal Memory for Robot Navigation

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T12:11:05.998750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:11:05.998750Z digest=sha256:3b463b359cf33d12262ab3eff8106b61b6d2d3ac6b8748edd33a06c35e5fb559

Observation 5da46c21-4592-4bde-9125-4497ac67d6ae · outbound

This paper cites Apollo: An open autonomous driving platform.

Spatially Visual Perception for End-to-End Robotic Learning Apollo: An open autonomous driving platform

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:11:06.692850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:11:06.002889Z digest=sha256:880c1c3b45f30999a0122d72c16609874de7a04e49d9f5b20bc35ca7b3840cfc

Observation e6bb57c4-c001-4563-8d8d-52f56dc1e29a · outbound

This paper cites Midas v3.1 – a model zoo for robust monocular relative depth estimation,.

Spatially Visual Perception for End-to-End Robotic Learning Midas v3.1 – a model zoo for robust monocular relative depth estimation,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T12:11:06.006478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:11:06.006478Z digest=sha256:5c7debc94f8e196bba946c05208acdb419285595cb7eee96d52b1c9b8a249267

Observation b6f76a4f-7d49-499b-9a86-ae9874b24064 · outbound

This paper cites Jackel, Mathew Monfort, Urs Muller, Jiakai Zhang, Xin Zhang, Jake Zhao, and Karol Zieba.

Spatially Visual Perception for End-to-End Robotic Learning Jackel, Mathew Monfort, Urs Muller, Jiakai Zhang, Xin Zhang, Jake Zhao, and Karol Zieba

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:11:06.677483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:11:06.009591Z digest=sha256:b0f920d42de5306f5e5c230d914d0198b22249157e924470d1ea842890e6846b

Observation d8a7978b-24fd-4db2-85ba-6588255f7bd3 · outbound

This paper cites Emerg- ing properties in self-supervised vision transformers, 2021.

Spatially Visual Perception for End-to-End Robotic Learning Emerg- ing properties in self-supervised vision transformers, 2021

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:11:06.666409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:11:06.014148Z digest=sha256:6e44ff34e324c6970da7c82557c5e4e63812be90e3b0c1583188bdacc1fb4944

Observation bd12b86e-0516-4d2c-87bb-c0eae77808f3 · outbound

This paper cites Diffusion policy: Visuomotor policy learning via action dif- fusion, 2024.

Spatially Visual Perception for End-to-End Robotic Learning Diffusion policy: Visuomotor policy learning via action dif- fusion, 2024

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:11:06.656143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:11:06.017559Z digest=sha256:ac5c0ff392b90a13243c6ee37af7f294bcde68a0b646ca19d80652e857589c51

Observation 3de98456-9ac8-4106-83a1-595aa2294fb7 · outbound

This paper cites Benchmarking robustness of 3d object detection to common corruptions in autonomous driving, 2023.

Spatially Visual Perception for End-to-End Robotic Learning Benchmarking robustness of 3d object detection to common corruptions in autonomous driving, 2023

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:11:06.646607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:11:06.020995Z digest=sha256:290ad6b0e0f55cbbf728bbe8bcafc763795ad827fcca9980f360617749212efc

Observation 55821bf4-f2ee-4649-a53c-4f9923024e3c · outbound

This paper cites Scene memory transformer for embodied agents in long-horizon tasks.

Spatially Visual Perception for End-to-End Robotic Learning Scene memory transformer for embodied agents in long-horizon tasks

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:11:06.636071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:11:06.024570Z digest=sha256:57a01759f42dc2c665de4474dcf98cfa63804f98b1b78d5803029a4e9d27ccce

Observation 930ebf98-1653-42c2-bc59-d760aa38b90d · outbound

This paper cites Zhao, and Chelsea Finn.

Spatially Visual Perception for End-to-End Robotic Learning Zhao, and Chelsea Finn

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:11:06.625829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:11:06.027615Z digest=sha256:93e3679c4c5b0505b2a7694fe0f0b88d921b6783dab4b56d84eb2f271d17e38f

Observation 0b024a8a-8a48-4f46-baa8-2d3aa87ee571 · outbound

This paper cites Paschalidis.

Spatially Visual Perception for End-to-End Robotic Learning Paschalidis

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:11:06.616123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:11:06.031366Z digest=sha256:f825afa88ce6e4ae4707b9b797b021bd294314b03738ee8500dbb991654d0d8a

Observation 51faa333-791a-4e2e-8b0e-5ec08dfd631a · outbound

This paper cites Deep residual learning for image recognition, 2015.

Spatially Visual Perception for End-to-End Robotic Learning Deep residual learning for image recognition, 2015

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T12:11:06.034373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:11:06.034373Z digest=sha256:5978c5cc37eded5d0e7d7f304b95ece3cb5e18b758dc4a392f0cae21e7201b14

Observation 40c0f03f-4c7b-44eb-a316-9940aa962930 · outbound

This paper cites Benchmarking neu- ral network robustness to common corruptions and perturba- tions, 2019.

Spatially Visual Perception for End-to-End Robotic Learning Benchmarking neu- ral network robustness to common corruptions and perturba- tions, 2019

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:11:06.600588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:11:06.037463Z digest=sha256:e97edff53ec802fc2cff05d774ebd7174ca89a8b322cc5acc8e5d5a0b9f6f5d6

Observation 4bbfd44e-1f49-4642-8712-e3777da4631a · outbound

This paper cites Cubuk, Barret Zoph, Justin Gilmer, and Balaji Lakshminarayanan.

Spatially Visual Perception for End-to-End Robotic Learning Cubuk, Barret Zoph, Justin Gilmer, and Balaji Lakshminarayanan

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:11:06.590619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:11:06.040735Z digest=sha256:3d6a4fa085c605594b23cb154d3d86730ecf2c3353d9e595fbcab3b441efb05d

Observation 8b252dcc-41b9-4d61-aeb0-fe07db93735b · outbound

This paper cites Denoising diffu- sion probabilistic models, 2020.

Spatially Visual Perception for End-to-End Robotic Learning Denoising diffu- sion probabilistic models, 2020

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T12:11:06.043979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:11:06.043979Z digest=sha256:2e194dee55b7c0a5e314345ea51d25d04ac6c7dcb0d771ccb176639fe224a068

Observation 2f0e422d-2544-4b7f-987d-62580a4f7be8 · outbound

This paper cites Denoising diffu- sion probabilistic models, 2020.

Spatially Visual Perception for End-to-End Robotic Learning Denoising diffu- sion probabilistic models, 2020

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:11:06.574467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:11:06.048481Z digest=sha256:e81e712fdf468f6333a27457b197cebd55b4f59c403f1cb12aa2818e5a24f82a

Observation 13ee5704-8452-448f-bb21-00dc28be95f0 · outbound

This paper cites Imitation with spatial-temporal heatmap: 2nd place solution for nuplan challenge, 2023.

Spatially Visual Perception for End-to-End Robotic Learning Imitation with spatial-temporal heatmap: 2nd place solution for nuplan challenge, 2023

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:11:06.564541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:11:06.051348Z digest=sha256:fa85f3ceb8d55f383111fe57fc5c1c517b5c98dc01ca25f3dccbe02cab47a57d

Observation baa3a2d1-e2b5-4e0e-82c3-76731d9a76ba · outbound

This paper cites Rekep: Spatio-temporal reasoning of rela- tional keypoint constraints for robotic manipulation, 2024.

Spatially Visual Perception for End-to-End Robotic Learning Rekep: Spatio-temporal reasoning of rela- tional keypoint constraints for robotic manipulation, 2024

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:11:06.554356Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:11:06.054277Z digest=sha256:f03b720e486c4b30cf18d479edcb5c5be4fbffd5b14030b8f29c148a9dd9d8a2

Observation a7024e10-5360-46b7-b313-c8c4902faff2 · outbound

This paper cites an unresolved cited work.

Spatially Visual Perception for End-to-End Robotic Learning Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-12T12:11:06.543429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:11:06.057152Z digest=sha256:c814d31aa6d5866e7308e7af7a29b3ce4c212624d128183fd3d93ea45fa4e5a0

Observation 1edb32ea-cb36-4614-b2f7-609917775d55 · outbound

This paper cites FORTRESS: Feature optimization and robustness techniques for 3d object detection systems.

Spatially Visual Perception for End-to-End Robotic Learning FORTRESS: Feature optimization and robustness techniques for 3d object detection systems

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:11:06.533148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:11:06.060238Z digest=sha256:3e4f46fd378db7b601ac445ae37f7c9677d94c98c696f503643cafef3c09f504

Observation 2af1f14e-5e7b-49a7-8372-0623b9baa02f · outbound

This paper cites an unresolved cited work.

Spatially Visual Perception for End-to-End Robotic Learning Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-12T12:11:06.523206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:11:06.063227Z digest=sha256:6a28312a0b8987b86da9d3acaf9f817800fef6d3cfcb4c5677eb99175cf7e474

Observation 1dd02c3a-31aa-40e7-86d1-167eef805a80 · outbound

This paper cites 3d gaussian splatting for real-time radiance field rendering.

Spatially Visual Perception for End-to-End Robotic Learning 3d gaussian splatting for real-time radiance field rendering

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T12:11:06.067182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:11:06.067182Z digest=sha256:38789a1b5764238af38f09bf68cf5145be0fb183e6335b092323973ea22a1ed4

Observation 26f973a4-f821-4f13-95f1-b477ec9c2418 · outbound

This paper cites Domain adaptive imitation learning, 2020.

Spatially Visual Perception for End-to-End Robotic Learning Domain adaptive imitation learning, 2020

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:11:06.507570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:11:06.070759Z digest=sha256:ea7af72f89aae005db0cdfaa734abd96974d154c516f8f4d55167d5f308d318e

Observation bee77bd3-b4a6-4ca6-80de-c1a596f63857 · outbound

This paper cites Berg, Wan-Yen Lo, Piotr Doll ´ar, and Ross Girshick.

Spatially Visual Perception for End-to-End Robotic Learning Berg, Wan-Yen Lo, Piotr Doll ´ar, and Ross Girshick

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:11:06.496777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:11:06.074201Z digest=sha256:d82a43a4ec60e8c05e17599bb07f864209df2b8b881316cf0a666da5d3ebd414

Observation ab7c5d8c-e958-4ae4-8bc1-5cb1f5e24b9b · outbound

This paper cites End-to-end planning of au- tonomous driving in industry and academia: 2022-2023,.

Spatially Visual Perception for End-to-End Robotic Learning End-to-end planning of au- tonomous driving in industry and academia: 2022-2023,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:11:06.487292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:11:06.077155Z digest=sha256:8af2775e6c7967a290a2cf7097a85172c27d7ba1a4fc1ac66c6d981516e6be7e

Observation 2d89ebcf-d984-4897-af65-70aa664c2e76 · outbound

This paper cites Okami: Teaching hu- manoid robots manipulation skills through single video imi- tation.

Spatially Visual Perception for End-to-End Robotic Learning Okami: Teaching hu- manoid robots manipulation skills through single video imi- tation

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:11:06.476798Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:11:06.079956Z digest=sha256:2ca58c3b09223caebf2fbf771afe202e707ab3f639c813f395b1269dcf995575

Observation 7f950add-244f-4605-a44e-ddd0f0c18ba5 · outbound

This paper cites Robust visual imi- tation learning with inverse dynamics representations, 2023.

Spatially Visual Perception for End-to-End Robotic Learning Robust visual imi- tation learning with inverse dynamics representations, 2023

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:11:06.467262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:11:06.083229Z digest=sha256:ab9114417abed0522abf8f1537447ae507f44aa866d9887a85437c7c17b9ba08

Observation bf545937-1fd5-43ae-9db4-bb40a6a04ee2 · outbound

This paper cites Alvarez, Sanja Fidler, Chen Feng, and Anima Anandkumar.

Spatially Visual Perception for End-to-End Robotic Learning Alvarez, Sanja Fidler, Chen Feng, and Anima Anandkumar

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:11:06.455765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:11:06.086198Z digest=sha256:23236e8e488cec04ba5b7994a42873f5e69b6fb2e6f1f68e2ffcc27e47dcb2a6

Observation 3b31211a-e08b-447b-a758-1a862bd0df1e · outbound

This paper cites Data Scaling Laws in Imitation Learning for Robotic Manipulation.

Spatially Visual Perception for End-to-End Robotic Learning Data Scaling Laws in Imitation Learning for Robotic Manipulation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T12:11:06.089934Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:11:06.089934Z digest=sha256:3bd2ac774908c4392df8501255733fb8ec60022732518ccb9049eb8f5e659b3f

Observation 923ed01d-a193-488c-8e22-e285877e3d52 · outbound

This paper cites Feature pyramid networks for object detection, 2017.

Spatially Visual Perception for End-to-End Robotic Learning Feature pyramid networks for object detection, 2017

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-12T12:11:06.093513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:11:06.093513Z digest=sha256:add97fa37b6c73a71320e86731bb1850dd6667870ef36d86e5c9d537e1bca8c0

Observation 8219f36a-4d8e-4438-9df8-8743b6a37c88 · outbound

This paper cites Occupancy prediction-guided neural planner for autonomous driving,.

Spatially Visual Perception for End-to-End Robotic Learning Occupancy prediction-guided neural planner for autonomous driving,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:11:06.439946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:11:06.097408Z digest=sha256:717c02a0edba6572646cbad2222c97a18bcea1649e9c91de455be158217441c8

Observation 45b45b28-2c97-4a96-a4fa-d222d0dca406 · outbound

This paper cites Robust imitation learning from corrupted demonstrations, 2022.

Spatially Visual Perception for End-to-End Robotic Learning Robust imitation learning from corrupted demonstrations, 2022

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:11:06.430049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:11:06.100566Z digest=sha256:128e4ba48e2ae63b62ab4c77222ed8bdb45edef4d8be899cf3a285a26f2deff0

Observation 208e5e7a-f2fc-48cc-abd2-d40e2eb94c7a · outbound

This paper cites The practice of mass produc- tion autonomous driving.

Spatially Visual Perception for End-to-End Robotic Learning The practice of mass produc- tion autonomous driving

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:11:06.420462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:11:06.103451Z digest=sha256:b40b9c18c911743d2d66060286648fd2afa88b3806cab7d0ea8125c054aa3a95

Observation 5548b18a-b12c-4441-8b5b-80f165ad3c81 · outbound

This paper cites Point- voxel cnn for efficient 3d deep learning, 2019.

Spatially Visual Perception for End-to-End Robotic Learning Point- voxel cnn for efficient 3d deep learning, 2019

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:11:06.410182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:11:06.106579Z digest=sha256:d70c429f1bfa994cfbbc8fbdb17b170ff04b942a25ded7f45d5a04b1b629d267

Observation 1b4aa277-7ef5-4661-8c79-eb96b0d9cb51 · outbound

This paper cites Ecker, Matthias Bethge, and Wieland Brendel.

Spatially Visual Perception for End-to-End Robotic Learning Ecker, Matthias Bethge, and Wieland Brendel

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:11:06.400366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:11:06.109482Z digest=sha256:366083b6d5199491bb51c21105348252f1f5d86c6febf0d5d47a2b107e14442b

Observation 9ecf87b2-dc2a-4059-80bd-47c763f9644b · outbound

This paper cites Srinivasan, Matthew Tancik, Jonathan T.

Spatially Visual Perception for End-to-End Robotic Learning Srinivasan, Matthew Tancik, Jonathan T

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T12:11:06.112441Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:11:06.112441Z digest=sha256:3a3e19e55b67641d40d56c2ee485ea541c38ae1f8d837c279c8a4b4317073621

Observation 8ca51b71-4b15-480f-9234-5b83ff81272e · outbound

This paper cites Dinov2: Learning robust visual features with- out supervision, 2024.

Spatially Visual Perception for End-to-End Robotic Learning Dinov2: Learning robust visual features with- out supervision, 2024

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:11:06.384151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:11:06.115494Z digest=sha256:7d47f73d2bee51e418fc324623d6f0e97533ab46f828f8eec973e170cf182f60

Observation a19c879d-4659-4872-8fb4-9c0999b1d237 · outbound

This paper cites Ro- bust multimodal vehicle detection in foggy weather using complementary lidar and radar signals.

Spatially Visual Perception for End-to-End Robotic Learning Ro- bust multimodal vehicle detection in foggy weather using complementary lidar and radar signals

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:11:06.374004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:11:06.118747Z digest=sha256:e5d3f2a6dbe7d734b382586ee5e80c541b103761775160b83674612089e18109

Observation 7c988782-7f13-4341-8b15-d479b02de059 · outbound

This paper cites Learning transferable visual models from natural language supervision, 2021.

Spatially Visual Perception for End-to-End Robotic Learning Learning transferable visual models from natural language supervision, 2021

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-12T12:11:06.121684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:11:06.121684Z digest=sha256:7eb9163064ef9915e7a2a3256afaf4f23a3419495bf5060ecbe32329b06730e8

Observation 78bd8131-3d91-417c-bd42-774f5acaa08f · outbound

This paper cites 3d-outdet: A fast and memory efficient outlier detector for 3d lidar point clouds in adverse weather.

Spatially Visual Perception for End-to-End Robotic Learning 3d-outdet: A fast and memory efficient outlier detector for 3d lidar point clouds in adverse weather

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:11:06.359139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:11:06.124719Z digest=sha256:1da54571ea373845604a877d1557a109b3a6e0d010e1b1e80fb1b1fbfff10a72

Observation 6c7b082a-71f5-4fc6-bf87-3632672f0aa2 · outbound

This paper cites Sam 2: Segment anything in images and videos,.

Spatially Visual Perception for End-to-End Robotic Learning Sam 2: Segment anything in images and videos,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-12T12:11:06.128514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:11:06.128514Z digest=sha256:192cb1fc5c6f655f6edceb1936849033f9ee0f9356c2669ce126043664ed2df9

Observation 759924a2-2e49-432b-a265-18631f4bf98b · outbound

This paper cites Raychaudhuri, Sujoy Paul, Jeroen van Baar, and Amit K.

Spatially Visual Perception for End-to-End Robotic Learning Raychaudhuri, Sujoy Paul, Jeroen van Baar, and Amit K

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:11:06.344235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:11:06.132259Z digest=sha256:3747f7e71ed9c6a7d6427ca3125b5cf2f2160a06699461604f787bea9a2e9764

Observation a08cf6cc-4ce0-4319-9f5c-7bb9f2cdf61e · outbound

This paper cites Cliport: What and where pathways for robotic manipulation, 2021.

Spatially Visual Perception for End-to-End Robotic Learning Cliport: What and where pathways for robotic manipulation, 2021

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:11:06.334713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:11:06.135178Z digest=sha256:7be2fea8f4973ab8653ce6bd950f64681fce1f5b9f528fba077b5cbcff799b87

Observation 98987e00-98c4-483c-9311-dd3bc9cf4d6c · outbound

This paper cites Robust imitation learning from noisy demon- strations.

Spatially Visual Perception for End-to-End Robotic Learning Robust imitation learning from noisy demon- strations

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:11:06.324554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:11:06.138077Z digest=sha256:89f765269e00b8562c3064f38e056a383fb6829e3748ecc38618d80f34eba1be

Observation 11a403a0-0469-451e-87a4-1b5ccf31885c · outbound

This paper cites an unresolved cited work.

Spatially Visual Perception for End-to-End Robotic Learning Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-12T12:11:06.314659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:11:06.141063Z digest=sha256:3fa2c945aa2837a298a05894be5e938806e09a8bad8d6a14aebba7a2f77474a1

Observation d671a01d-b05f-4cfd-a051-793d4b4ba993 · outbound

This paper cites Attention is all you need.

Spatially Visual Perception for End-to-End Robotic Learning Attention is all you need

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-12T12:11:06.144398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:11:06.144398Z digest=sha256:78372fd04c82d3d12e3a175dcb6755a05b42f7f91104d64837033b9fb02d2310

Observation 85d13d6e-ef95-447a-86fe-7e4ccf771520 · outbound

This paper cites 4seasons: A cross-season dataset for multi-weather slam in autonomous driving, 2020.

Spatially Visual Perception for End-to-End Robotic Learning 4seasons: A cross-season dataset for multi-weather slam in autonomous driving, 2020

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:11:06.299863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:11:06.147285Z digest=sha256:010fb3ff4fb12b20103e0d738275e6afdf0c2c29012f24ed56302c09239a6aed

Observation e829b718-02ee-4af4-b1a2-fdf55325b12e · outbound

This paper cites Generalized robot learn- ing framework, 2024.

Spatially Visual Perception for End-to-End Robotic Learning Generalized robot learn- ing framework, 2024

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:11:06.289786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:11:06.150126Z digest=sha256:285405b23094281ac70db0ff1e2ee9d4c5a507e2f0a46a16fb24f81e025a9253

Observation b84897ce-0550-4e87-a5a9-d01a2272e809 · outbound

This paper cites Depth anything: Unleashing the power of large-scale unlabeled data, 2024.

Spatially Visual Perception for End-to-End Robotic Learning Depth anything: Unleashing the power of large-scale unlabeled data, 2024

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-12T12:11:06.153216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:11:06.153216Z digest=sha256:5a0a5a303e61a770d375152e999e2ffbf5802d914ada205e13e737b2b7d906fd

Observation 9c2553ae-fa63-4d74-980a-45aa5e05a6ea · outbound

This paper cites Depth any- thing v2, 2024.

Spatially Visual Perception for End-to-End Robotic Learning Depth any- thing v2, 2024

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:11:06.274032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:11:06.156316Z digest=sha256:41e9a4baad8c238be64f1785325354ed4b69cc227dd375c3294b6b4c0f41d0f4

Observation 686a6ab6-2083-45c7-939b-b4dd5bb029ec · outbound

This paper cites Primedepth: Efficient monocular depth estimation with a sta- ble diffusion preimage, 2024.

Spatially Visual Perception for End-to-End Robotic Learning Primedepth: Efficient monocular depth estimation with a sta- ble diffusion preimage, 2024

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:11:06.264018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:11:06.159789Z digest=sha256:68183195af23574d69933ba94b561116a5bb74cb422c39cdb13013f56296f782

Observation 19b0e112-9ecf-4397-b796-7360df63bafe · outbound

This paper cites Multi-object detection at night for traffic in- vestigations based on improved ssd framework.

Spatially Visual Perception for End-to-End Robotic Learning Multi-object detection at night for traffic in- vestigations based on improved ssd framework

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:11:06.254669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:11:06.163389Z digest=sha256:df58cea884ff114a91157c9d514232b6d33f0962a2f07dc658c48015b1607791

Observation 7c29df97-cbed-4b20-b09a-6c7f8b5f010d · outbound

This paper cites Safe occlusion-aware au- tonomous driving via game-theoretic active perception.

Spatially Visual Perception for End-to-End Robotic Learning Safe occlusion-aware au- tonomous driving via game-theoretic active perception

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:11:06.244517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:11:06.167169Z digest=sha256:e21a18efb7f4c6f9352c4c97cbc039c8995ee7a932e2e901f4bd08340fb97456

Observation 6739c108-dbc8-431f-934b-b839510eb26f · outbound

This paper cites Autofed: Heterogeneity-aware federated multimodal learning for robust autonomous driving, 2023.

Spatially Visual Perception for End-to-End Robotic Learning Autofed: Heterogeneity-aware federated multimodal learning for robust autonomous driving, 2023

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:11:06.235125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:11:06.170444Z digest=sha256:121a31df9625171bc4f158f33dd5feca9ba0d3ba4a7c8e099434753aefc02780

Observation a9359d3f-ef55-4a2e-9514-e40df5bb5fa0 · outbound

This paper cites an unresolved cited work.

Spatially Visual Perception for End-to-End Robotic Learning Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-12T12:11:06.223321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:11:06.173763Z digest=sha256:f0ec6476eee1b006e5996e68adff6f30b84d023fb68ab19edb969cbbdbd8088d

Pith citing papers

Observation f87aae73-bfc0-414e-a53d-ef1deb55b78a · inbound

Spatial RoboGrasp: Generalized Robotic Grasping Control Policy cites this paper.

Spatial RoboGrasp: Generalized Robotic Grasping Control Policy Spatially Visual Perception for End-to-End Robotic Learning

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:53:48.437750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T13:53:41.125211Z digest=sha256:aad264af8c49104727ffa007f0498adc4789de4ec285d67d1d401ca5b6e37cb2