Pith. sign in

REVIEW 3 major objections 6 minor 300 references

A Survey of Simultaneous Localization and Mapping with an Envision in 6G Wireless Networks

T0 review · 3 major / 6 minor · reviewed 2026-08-14 · deepseek-v4-flash

Pith's one-line read This survey claims to provide a full-scale map of SLAM, from Lidar and visual systems to fusion and a 6G radio-based future.

desk verdict A broad but sloppy SLAM survey whose only novelty is an unquantified 6G vision section containing a physics-defying NLOS claim about THz. read the letter →

arxiv 1909.05214 v4 pith:CO2UNMZY submitted 2019-08-24 cs.RO

classification cs.RO
keywords SLAMLidarVisualVisual-inertialodometrySensorfusionDeeplearningin6GwirelessnetworksTerahertzcommunication
verification ladder T0 review T1 audit T2 compute T3 formal

The pith

A machine-rendered reading of the paper's core claim, the machinery that carries it, and where it could break.

The reading

This survey paper tries to give a complete, structured map of simultaneous localization and mapping: the sensor hardware, the open-source algorithms, the deep-learning extensions, and the open challenges in each branch. It organizes SLAM into Lidar SLAM, visual SLAM (including visual-inertial odometry), and Lidar-visual fusion, and it argues that the field has moved from filter-based estimators to graph-based optimization and from single sensors to multi-sensor fusion. It then extends the map forward by arguing that 6G terahertz wireless networks could provide a radio-based SLAM alternative with sub-centimeter accuracy, non-line-of-sight sensing, and remote computation. A sympathetic reader would take the paper's central claim to be that the field's many threads can be usefully held together by this taxonomy, and that the next major shift will come from combining sensor-based SLAM with wireless-network-based sensing.

What carries the argument

The machinery carrying the argument is a taxonomic decomposition plus a projected technology leap. The taxonomy sorts SLAM systems by sensing modality (Lidar, camera, fusion), then by map density (sparse, semi-dense, dense), then by algorithmic family (filter-based versus graph/optimization-based), with deep learning inserted at each level as feature extractors, segmenters, pose estimators, and depth predictors. Fusion is organized into three layers: hardware, data, and task. The projected leap is 6G's terahertz band, described as a radio-frequency spectrum above 100 GHz that could supply data rates near 1 Tbps, sub-centimeter positioning, and reconfigurable intelligent surfaces for non-line-of-sight coverage; that band is the object that would turn wireless links into SLAM sensors.

What would settle it

Run a terahertz 6G testbed through a non-line-of-sight indoor route and measure positioning error and link throughput; if the error stays above the sub-centimeter range (for example, decimeters) or the NLOS link cannot support map-grade measurements, the paper's central 6G premise is contradicted.

Watch

Extended reading notes

Core claim

On its own terms, the paper's discovery is organizational: it claims that the entire SLAM landscape can be sorted into three streams—Lidar, vision, and fusion—and that each stream is best understood through four recurring elements: sensors, open-source systems, deep learning, and open challenges. The survey identifies the historical trajectory from early filter-based systems (EKF, particle filters) to graph-based optimization and multi-threaded pipelines, and it catalogs representative systems such as ORB-SLAM, VINS-Mono, Cartographer, and Loam as anchors of that trajectory. It then claims that Lidar-visual fusion is the balanced route for reliability and versatility, organized at hardware, data, and task layers, and that future SLAM will be semantic, multi-sensor, and increasingly dependent on integrated hardware. Finally, it argues that 6G wireless networks with terahertz communication will let SLAM become radio-based: centimeter or sub-centimeter positioning, maps constructed from radio signals even in non-line-of-sight conditions, and heavy computation offloaded to remote servers.

Load-bearing premise

The survey's forward-looking argument rests on the assumption that 6G networks will actually deliver terahertz data rates near 1 Tbps, sub-centimeter positioning, and radio-based mapping through walls; if those capabilities fail to materialize, the 6G vision loses its foundation, even though the survey portion would still stand.

Editorial extensions

If this is right

  • New researchers can use the survey as an entry path: sensor types, open-source packages, and deep-learning roles are matched to each SLAM family.
  • Experienced researchers can use it as a dictionary to locate systems and open problems, especially in Lidar-visual fusion and semantic SLAM.
  • The field's trajectory points to multi-sensor fusion and integrated hardware as the route from algorithms to products.
  • If 6G terahertz capabilities arrive, SLAM could expand from self-contained sensors to network-based radio sensing with NLOS mapping and remote computation.
  • Event cameras and solid-state Lidar are flagged as the sensor trends that will address high-speed and low-texture failure cases.

Reading between the lines

Editorial extensions of the paper, not claims the author makes directly.

  • Beyond the paper: If 6G radio SLAM matures, the same spectrum used for communication could serve as a cooperative sensing channel, so multiple robots or vehicles could fuse their radio maps; the paper does not develop this multi-agent implication.
  • Beyond the paper: The survey's taxonomy predicts that semantic SLAM and deep learning will converge with radio SLAM, which could be tested by building a benchmark that compares terahertz-based positioning with Lidar and visual baselines in the same non-line-of-sight scenes.
  • Beyond the paper: The paper's open question 'Will end-to-end learning dominate SLAM?' could be sharpened into a measurable test: track whether learned pipelines surpass geometry-based systems on long-duration, large-scale datasets over the next several years.
Share X Bluesky LinkedIn Reddit HN

Editorial analysis

A structured set of objections, weighed in public.

Desk editor's note, referee report, and a circularity audit.

Referee Report

3 major / 6 minor

Summary. The paper is a survey of simultaneous localization and mapping (SLAM) that covers lidar-based, vision-based, and fused lidar-visual systems, including sensor types, open-source implementations, deep learning approaches, and open challenges. It closes with a section envisioning how 6G wireless networks, particularly terahertz communication and reconfigurable intelligent surfaces, could contribute to future SLAM, claiming centimeter-level accuracy even in non-line-of-sight conditions. The abstract describes the paper as a high-quality, full-scale overview useful to newcomers and as a dictionary for experienced researchers.

Significance. If the survey were factually reliable, it would provide a useful entry point into a broad and rapidly evolving literature, and it does compile a large number of relevant systems (Cartographer, ORB-SLAM, VINS, RTAB-Map, and many others) and a substantial bibliography. The organization by sensor modality and the attention to deep-learning-based SLAM are genuine strengths. The paper makes no original algorithmic or quantitative contribution, so there is no parameter-fitting or circularity issue to penalize. However, the survey's value depends on the accuracy of the material it transmits, and the forward-looking claims in Section V are stated as part of the advertised contribution but are left unquantified and, in one respect, contrary to established propagation physics.

major comments (3)
  1. [V] Section V contains load-bearing quantitative claims about THz-enabled SLAM that are made without support. The sentence "As for the difference with VLC, 6G with the THz communications will not affected by the light changes and NLOS" conflicts with established terahertz propagation physics, where high path loss and susceptibility to blockage make NLOS operation a central challenge rather than an inherent advantage. The subsequent assertions that 6G will create "centimeter level accuracy even in NLOS environment" and later "sub-centimeter level" accuracy with 3D maps constructed "without any calibration and prior knowledge" are made with no derivation, measurement, or cited reference. Because the title advertises the 6G envision as part of the paper's contribution, these unsupported claims need either quantitative grounding or an explicit reframing as speculation.
  2. [V.A] The background statements on 5G and 6G in the opening of Section V and in Section V.A are factually inaccurate. "Unlike 100 Gbps of data rates for 5G" misstates 5G peak data rates, since IMT-2020 targets at most 20 Gbps, and "The technology of 6G will need no supports such as multiple-input multiple-output (MIMO) in 5G represented as mmWave communications" is a mischaracterization, as terahertz systems are widely expected to rely on massive MIMO and beamforming to overcome path loss. These errors affect the reliability of the survey's forward-looking comparison and should be corrected with appropriate references.
  3. [III.A] In the paragraph on monocular cameras, the paper states that "visual slam based on monocular camera have a scale with real size of track and map," which is incorrect and directly contradicts the following sentence: "That's say that the real depth can't be got by monocular camera, which called Scale Ambiguity." Monocular SLAM is up-to-scale and does not recover absolute scale. Since Section III is a central part of the survey, this is a load-bearing factual error for readers using the paper as a reference.
minor comments (6)
  1. [Abstract] The English grammar in the abstract and throughout needs editing; for example, "The paper makes an overview" should be "The paper presents an overview," and "the paper can be considered as dictionary" needs an article.
  2. [II.B.1] In the Gmapping bullet, "Rao-Blackwellisation Partical Filter" should read "Rao-Blackwellisation Particle Filter."
  3. [III.B] The sentence "ATAM7 is a visual SLAM toolkit for beginners" appears to be a typo, as no system named ATAM7 is described; the context suggests PTAM or another toolkit.
  4. [III.B.2] In the EVO bullet, "Our algorithm is unaffected by motion blur" should be reworded to "The algorithm is unaffected..." because the paper is a survey rather than an original system description.
  5. [IV.A] In the Camera & Lidar bullet, "Other work can be seen follows as but not limited to" is grammatically broken and should be rewritten.
  6. [References] Several references contain malformed author names or incomplete information, including [47] with "Emanuelea Palazzolo" and [126] with an incomplete author list; the reference list should be checked systematically.

Circularity Check

0 steps flagged · score 0.0 of 10

No circular structure: the paper is a survey, and its 6G section is an attributed vision rather than a derivation from its own inputs.

full rationale

This manuscript is a literature survey, not a derivation chain: it compiles existing SLAM systems, sensors, deep-learning methods, calibration toolboxes, and open challenges, and it makes no quantitative predictions from fitted parameters or from its own assumptions. The 6G discussion in Section V is explicitly framed as an 'envision' and its specific claims (e.g., '6G with the THz communications will not affected by the light changes and NLOS' and '6G with THz will create centimeter level accuracy even in NLOS environment') are supported only by citations to external visionary 6G papers such as [301]–[305]; those claims may be unquantified or physically questionable, but they are not circular because they are not derived from, nor equivalent to, any input defined by this paper. The self-citations that appear ([3], [4], [5], [231], [252], [253]) are used only as ordinary background references for indoor positioning, dynamic-scene SLAM, and LiDAR extrinsic calibration; none is load-bearing for the paper's advertised contribution of providing a 'high quality and full-scale overview.' Since the survey makes no claim that is forced by its own definitions, renames no external result as a new derivation, and contains no fitted-input-called-prediction structure, the appropriate circularity score is 0.

Assumptions & free parameters 0 free parameters · 3 assumptions · 0 invented entities

The paper is a survey, so it introduces no free parameters or invented entities. It relies on the accuracy of its citations and on speculative premises about future 6G networks.

assumptions (3)
  • domain assumption Cited descriptions of SLAM systems are accurate representations of the original papers.
    The survey relies on one-line summaries of dozens of systems without independent verification. See Sections II-B and III-B.
  • domain assumption Future 6G wireless networks will provide terahertz communication, very high data rates, sub-cm positioning, and radio-based mapping capability.
    Section V builds its SLAM vision on these capabilities, citing forward-looking 6G papers rather than demonstrated systems.
  • standard math Standard SLAM mathematics (Kalman filtering, pose graph optimization, bundle adjustment) is assumed from prior textbooks and papers.
    The survey introduces no derivations and references [6] for probabilistic robotics foundations.

how reviews work

0 comments
Cite this review

Pith. "Pith review of A Survey of Simultaneous Localization and Mapping with an Envision in 6G Wireless Networks." pith.science (2026). https://pith.science/paper/CO2UNMZY

@misc{pith2026190905214,
  author       = {Pith},
  title        = {Pith review of: A Survey of Simultaneous Localization and Mapping with an Envision in 6G Wireless Networks},
  year         = {2026},
  howpublished = {\url{https://pith.science/paper/CO2UNMZY}},
  note         = {Machine review of arXiv:1909.05214}
}
read the original abstract

Simultaneous Localization and Mapping (SLAM) achieves the purpose of simultaneous positioning and map construction based on self-perception. The paper makes an overview in SLAM including Lidar SLAM, visual SLAM, and their fusion. For Lidar or visual SLAM, the survey illustrates the basic type and product of sensors, open source system in sort and history, deep learning embedded, the challenge and future. Additionally, visual inertial odometry is supplemented. For Lidar and visual fused SLAM, the paper highlights the multi-sensors calibration, the fusion in hardware, data, task layer. The open question and forward thinking with an envision in 6G wireless networks end the paper. The contributions of this paper can be summarized as follows: the paper provides a high quality and full-scale overview in SLAM. It's very friendly for new researchers to hold the development of SLAM and learn it very obviously. Also, the paper can be considered as a dictionary for experienced researchers to search and find new interesting orientation.

Discussion (0). Continue with ORCID to comment.

Reference graph

Works this paper leans on

300 extracted references · 61 canonical work pages

  1. [1]

    Simultaneous ma p build- ing and localization for an autonomous mobile robot

    John J Leonard and Hugh F Durrant-Whyte. Simultaneous ma p build- ing and localization for an autonomous mobile robot. In Proceedings IROS’91: IEEE/RSJ International W orkshop on Intelligent R obots and Systems’ 91 , pages 1442–1447. Ieee, 1991

  2. [2]

    Estim ating un- certain spatial relationships in robotics

    Randall Smith, Matthew Self, and Peter Cheeseman. Estim ating un- certain spatial relationships in robotics. In Autonomous robot vehicles , pages 167–193. Springer, 1990

  3. [3]

    A robu st indoor positioning method based on bluetooth low energy wit h separate channel information

    Baichuan Huang, Jingbin Liu, Wei Sun, and Fan Y ang. A robu st indoor positioning method based on bluetooth low energy wit h separate channel information. Sensors, 19(16):3487, 2019

  4. [4]

    iparking: An intelligent indoor location-based smartphon e parking service

    Jingbin Liu, Ruizhi Chen, Y uwei Chen, Ling Pei, and Liang Chen. iparking: An intelligent indoor location-based smartphon e parking service. Sensors, 12(11):14612–14629, 2012

  5. [5]

    A hybrid smartphone indoor positioning solutio n for mobile lbs

    Jingbin Liu, Ruizhi Chen, Ling Pei, Robert Guinness, and Heidi Kuusniemi. A hybrid smartphone indoor positioning solutio n for mobile lbs. Sensors, 12(12):17208–17233, 2012

  6. [6]

    Probabilistic robotics

    Sebastian Thrun, Wolfram Burgard, and Dieter Fox. Probabilistic robotics. MIT press, 2005

  7. [7]

    An e valu- ation of 2d slam techniques available in robot operating sys tem

    Joao Machado Santos, David Portugal, and Rui P Rocha. An e valu- ation of 2d slam techniques available in robot operating sys tem. In 2013 IEEE International Symposium on Safety, Security, and Rescue Robotics (SSRR) , pages 1–6. IEEE, 2013

  8. [8]

    Improved techniques for grid mapping with rao-blackwellized partic le filters

    Giorgio Grisetti, Cyrill Stachniss, Wolfram Burgard, e t al. Improved techniques for grid mapping with rao-blackwellized partic le filters. IEEE transactions on Robotics , 23(1):34, 2007

Show all 300 references
  1. [9]

    Fastslam: A factored solution to the simultaneous loc alization and mapping problem

    Michael Montemerlo, Sebastian Thrun, Daphne Koller, Be n Wegbreit, et al. Fastslam: A factored solution to the simultaneous loc alization and mapping problem. Aaai/iaai, 593598, 2002

  2. [10]

    Fastslam 2.0: An improved particle filtering algorith m for simultaneous localization and mapping that provably conve rges

    Michael Montemerlo, Sebastian Thrun, Daphne Koller, B en Wegbreit, et al. Fastslam 2.0: An improved particle filtering algorith m for simultaneous localization and mapping that provably conve rges. In IJCAI, pages 1151–1156, 2003

  3. [11]

    A flexible and scalable slam system with full 3d mot ion estimation

    Stefan Kohlbrecher, Oskar V on Stryk, Johannes Meyer, a nd Uwe Klingauf. A flexible and scalable slam system with full 3d mot ion estimation. In 2011 IEEE International Symposium on Safety, Security, and Rescue Robotics , pages 155–160. IEEE, 2011

  4. [12]

    Efficient sparse pose a djustment for 2d mapping

    Kurt Konolige, Giorgio Grisetti, Rainer K¨ ummerle, Wo lfram Burgard, Benson Limketkai, and Regis Vincent. Efficient sparse pose a djustment for 2d mapping. In 2010 IEEE/RSJ International Conference on Intelligent Robots and Systems , pages 22–29. IEEE, 2010

  5. [13]

    A linear approximation for graph-based simultaneous local ization and mapping

    Luca Carlone, Rosario Aragues, Jos´ e A Castellanos, an d Basilio Bona. A linear approximation for graph-based simultaneous local ization and mapping. Robotics: Science and Systems VII , pages 41–48, 2012

  6. [14]

    A slam algorithm in le ss than 200 lines c-language program

    B Steux and O TinySLAM El Hamzaoui. A slam algorithm in le ss than 200 lines c-language program. Proceedings of the Control Automation Robotics & Vision (ICARCV), Singapore , pages 7–10, 2010

  7. [15]

    Real- time loop closure in 2d lidar slam

    Wolfgang Hess, Damon Kohler, Holger Rapp, and Daniel An dor. Real- time loop closure in 2d lidar slam. In 2016 IEEE International Conference on Robotics and Automation (ICRA) , pages 1271–1278. IEEE, 2016

  8. [16]

    Loam: Lidar odometry and mapp ing in real-time

    Ji Zhang and Sanjiv Singh. Loam: Lidar odometry and mapp ing in real-time. In Robotics: Science and Systems , volume 2, page 9, 2014

  9. [17]

    Lego-loam: Lightweigh t and ground- optimized lidar odometry and mapping on variable terrain

    Tixiao Shan and Brendan Englot. Lego-loam: Lightweigh t and ground- optimized lidar odometry and mapping on variable terrain. I n 2018 IEEE/RSJ International Conference on Intelligent Robots a nd Systems (IROS), pages 4758–4765. IEEE, 2018

  10. [18]

    Imls-slam: scan-to-model ma tching based on 3d data

    Jean-Emmanuel Deschaud. Imls-slam: scan-to-model ma tching based on 3d data. In 2018 IEEE International Conference on Robotics and Automation (ICRA) , pages 2480–2485. IEEE, 2018

  11. [19]

    Pointnetvlad: Deep point cloud based retrieval for large-scale place recognition

    Mikaela Angelina Uy and Gim Hee Lee. Pointnetvlad: Deep point cloud based retrieval for large-scale place recognition. I n Proceedings of the IEEE Conference on Computer Vision and Pattern Recogn ition, pages 4470–4479, 2018

  12. [20]

    V oxelnet: End-to-end learnin g for point cloud based 3d object detection

    Yin Zhou and Oncel Tuzel. V oxelnet: End-to-end learnin g for point cloud based 3d object detection. In Proceedings of the IEEE Confer- ence on Computer Vision and Pattern Recognition , pages 4490–4499, 2018

  13. [21]

    Birdne t: a 3d object detection framework from lidar information

    Jorge Beltr´ an, Carlos Guindel, Francisco Miguel More no, Daniel Cruzado, Fernando Garcia, and Arturo De La Escalera. Birdne t: a 3d object detection framework from lidar information. In 2018 21st International Conference on Intelligent Transportation S ystems (ITSC), pages 3...

  14. [22]

    Lmnet: Real-time multiclass object detection on cpu using 3 d lidar

    Kazuki Minemura, Hengfui Liau, Abraham Monrroy, and Sh inpei Kato. Lmnet: Real-time multiclass object detection on cpu using 3 d lidar. In 2018 3rd Asia-Pacific Conference on Intelligent Robot Syste ms (ACIRS), pages 28–34. IEEE, 2018

  15. [23]

    Pixor: Real-t ime 3d object detection from point clouds

    Bin Y ang, Wenjie Luo, and Raquel Urtasun. Pixor: Real-t ime 3d object detection from point clouds. In Proceedings of the IEEE conference on Computer Vision and Pattern Recognition , pages 7652–7660, 2018

  16. [24]

    Y olo3d: End-to-end real-time 3d orient ed object bounding box detection from lidar point cloud

    Waleed Ali, Sherif Abdelkarim, Mahmoud Zidan, Mohamed Zahran, and Ahmad El Sallab. Y olo3d: End-to-end real-time 3d orient ed object bounding box detection from lidar point cloud. In Proceedings of the European Conference on Computer Vision (ECCV) , pages 0–0, 2018

  17. [25]

    Pointcnn: Convolution on x-transformed points

    Y angyan Li, Rui Bu, Mingchao Sun, Wei Wu, Xinhan Di, and B aoquan Chen. Pointcnn: Convolution on x-transformed points. In Advances in Neural Information Processing Systems , pages 820–830, 2018

  18. [26]

    Mult i-view 3d object detection network for autonomous driving

    Xiaozhi Chen, Huimin Ma, Ji Wan, Bo Li, and Tian Xia. Mult i-view 3d object detection network for autonomous driving. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recogn ition, pages 1907–1915, 2017

  19. [27]

    Pu-gan: A point cloud upsampling adversarial netw ork

    Ruihui Li, Xianzhi Li, Chi-Wing Fu, Daniel Cohen-Or, an d Pheng- Ann Heng. Pu-gan: A point cloud upsampling adversarial netw ork. In Proceedings of the IEEE International Conference on Comput er Vision, pages 7203–7212, 2019

  20. [28]

    Splatnet: Spa rse lattice networks for point cloud processing

    Hang Su, V arun Jampani, Deqing Sun, Subhransu Maji, Eva ngelos Kalogerakis, Ming-Hsuan Y ang, and Jan Kautz. Splatnet: Spa rse lattice networks for point cloud processing. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , pages 2530– 2539, 2018

  21. [29]

    A review of point clou ds segmentation and classification algorithms

    E Grilli, F Menna, and F Remondino. A review of point clou ds segmentation and classification algorithms. The International Archives 11 of Photogrammetry, Remote Sensing and Spatial Information Sciences, 42:339, 2017

  22. [30]

    P ointnet: Deep learning on point sets for 3d classification and segment ation

    Charles R Qi, Hao Su, Kaichun Mo, and Leonidas J Guibas. P ointnet: Deep learning on point sets for 3d classification and segment ation. arXiv preprint arXiv:1612.00593 , 2016

  23. [31]

    Pointn et++: Deep hierarchical feature learning on point sets in a metric spac e

    Charles R Qi, Li Yi, Hao Su, and Leonidas J Guibas. Pointn et++: Deep hierarchical feature learning on point sets in a metric spac e. arXiv preprint arXiv:1706.02413, 2017

  24. [32]

    Deep hough voting for 3d object detection in point clouds

    Charles R Qi, Or Litany, Kaiming He, and Leonidas J Guiba s. Deep hough voting for 3d object detection in point clouds. arXiv preprint arXiv:1904.09664, 2019

  25. [33]

    SegMap: 3d segment mapping usin g data-driven descriptors

    Renaud Dube, Andrei Cramariuc, Daniel Dugas, Juan Niet o, Roland Siegwart, and Cesar Cadena. SegMap: 3d segment mapping usin g data-driven descriptors. In Robotics: Science and Systems (RSS) , 2018

  26. [34]

    Squ eezeseg: Convolutional neural nets with recurrent crf for real-time road-object segmentation from 3d lidar point cloud

    Bichen Wu, Alvin Wan, Xiangyu Y ue, and Kurt Keutzer. Squ eezeseg: Convolutional neural nets with recurrent crf for real-time road-object segmentation from 3d lidar point cloud. ICRA, 2018

  27. [35]

    Squeezesegv2: Improved model structure and unsup ervised domain adaptation for road-object segmentation from a lida r point cloud

    Bichen Wu, Xuanyu Zhou, Sicheng Zhao, Xiangyu Y ue, and K urt Keutzer. Squeezesegv2: Improved model structure and unsup ervised domain adaptation for road-object segmentation from a lida r point cloud. In ICRA, 2019

  28. [36]

    A lidar point cloud generator: f rom a virtual world to autonomous driving

    Xiangyu Y ue, Bichen Wu, Sanjit A Seshia, Kurt Keutzer, a nd Alberto L Sangiovanni-Vincentelli. A lidar point cloud generator: f rom a virtual world to autonomous driving. In ICMR, pages 458–464. ACM, 2018

  29. [37]

    Pointsift: A sift-like network module for 3d point cloud sem antic segmentation

    Mingyang Jiang, Yiran Wu, Tianqi Zhao, Zelin Zhao, and C ewu Lu. Pointsift: A sift-like network module for 3d point cloud sem antic segmentation. arXiv preprint arXiv:1807.00652 , 2018

  30. [38]

    Point wise convo- lutional neural networks

    Binh-Son Hua, Minh-Khoi Tran, and Sai-Kit Y eung. Point wise convo- lutional neural networks. In Computer Vision and Pattern Recognition (CVPR), 2018

  31. [39]

    3d recurrent neural networks with context fusion for point c loud semantic segmentation

    Xiaoqing Y e, Jiamao Li, Hexiao Huang, Liang Du, and Xiao lin Zhang. 3d recurrent neural networks with context fusion for point c loud semantic segmentation. In Proceedings of the European Conference on Computer Vision (ECCV) , pages 403–417, 2018

  32. [40]

    Large-scale poin t cloud semantic segmentation with superpoint graphs

    Loic Landrieu and Martin Simonovsky. Large-scale poin t cloud semantic segmentation with superpoint graphs. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , pages 4558–4567, 2018

  33. [41]

    Segmatch: Segment based place recog nition in 3d point clouds

    Renaud Dub´ e, Daniel Dugas, Elena Stumm, Juan Nieto, Ro land Sieg- wart, and Cesar Cadena. Segmatch: Segment based place recog nition in 3d point clouds. In 2017 IEEE International Conference on Robotics and Automation (ICRA) , pages 5266–5272. IEEE, 2017

  34. [42]

    Escape from cells: D eep kd- networks for the recognition of 3d point cloud models

    Roman Klokov and Victor Lempitsky. Escape from cells: D eep kd- networks for the recognition of 3d point cloud models. In Proceedings of the IEEE International Conference on Computer Vision , pages 863– 872, 2017

  35. [43]

    Deeptemporalseg: Tem porally consistent semantic segmentation of 3d lidar scans

    Ayush Dewan and Wolfram Burgard. Deeptemporalseg: Tem porally consistent semantic segmentation of 3d lidar scans. arXiv preprint arXiv:1906.06962, 2019

  36. [44]

    Lu-net: An efficient network for 3d lidar p oint cloud semantic segmentation based on end-to-end-learned 3 d features and u-net

    Pierre Biasutti, Vincent Lepetit, Jean-Franois Aujol , Mathieu Brdif, and Aurlie Bugeau. Lu-net: An efficient network for 3d lidar p oint cloud semantic segmentation based on end-to-end-learned 3 d features and u-net. 08 2019

  37. [45]

    Pointr cnn: 3d ob- ject proposal generation and detection from point cloud

    Shaoshuai Shi, Xiaogang Wang, and Hongsheng Li. Pointr cnn: 3d ob- ject proposal generation and detection from point cloud. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recogn ition, pages 770–779, 2019

  38. [46]

    L3-net: Towards learning based lidar localization for auto nomous driv- ing

    Lu Weixin, Zhou Y ao, Wan Guowei, Hou Shenhua, and Song Sh iyu. L3-net: Towards learning based lidar localization for auto nomous driv- ing. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR), 2019

  39. [47]

    Suma++: Efficient lidar-based semantic slam

    Chen Xieyuanli, Milioto Andres, and Emanuelea Palazzo lo. Suma++: Efficient lidar-based semantic slam. In 2019 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . IEEE, 2019

  40. [48]

    Imu- assisted 2d slam method for low-texture and dynamic environ ments

    Zhongli Wang, Y an Chen, Y ue Mei, Kuo Y ang, and Baigen Cai . Imu- assisted 2d slam method for low-texture and dynamic environ ments. Applied Sciences , 8(12):2534, 2018

  41. [49]

    Dynamic pose graph slam: Long-term mapping in low dynamic environments

    Aisha Walcott-Bryant, Michael Kaess, Hordur Johannss on, and John J Leonard. Dynamic pose graph slam: Long-term mapping in low dynamic environments. In 2012 IEEE/RSJ International Conference on Intelligent Robots and Systems , pages 1871–1878. IEEE, 2012

  42. [50]

    I llusion and dazzle: Adversarial optical channel exploits against l idars for automotive applications

    Hocheol Shin, Dohyun Kim, Y ujin Kwon, and Y ongdae Kim. I llusion and dazzle: Adversarial optical channel exploits against l idars for automotive applications. In International Conference on Cryptographic Hardware and Embedded Systems , pages 445–467. Springer, 2017

  43. [51]

    Adversarial sensor attack on lidar-based perception in aut onomous driving

    Y ulong Cao, Chaowei Xiao, Benjamin Cyr, Yimeng Zhou, Wo n Park, Sara Rampazzi, Qi Alfred Chen, Kevin Fu, and Z Morley Mao. Adversarial sensor attack on lidar-based perception in aut onomous driving. arXiv preprint arXiv:1907.06826 , 2019

  44. [52]

    Orb-slam: a versatile and accurate monocular slam system

    Raul Mur-Artal, Jose Maria Martinez Montiel, and Juan D Tardos. Orb-slam: a versatile and accurate monocular slam system. IEEE transactions on robotics , 31(5):1147–1163, 2015

  45. [53]

    Vins-mono: A ro bust and versatile monocular visual-inertial state estimator

    Tong Qin, Peiliang Li, and Shaojie Shen. Vins-mono: A ro bust and versatile monocular visual-inertial state estimator. IEEE Transactions on Robotics , 34(4):1004–1020, 2018

  46. [54]

    Parallel tracking and map ping on a camera phone

    Georg Klein and David Murray. Parallel tracking and map ping on a camera phone. In 2009 8th IEEE International Symposium on Mixed and Augmented Reality , pages 83–86. IEEE, 2009

  47. [55]

    K¨ ahler, V

    O. K¨ ahler, V . A. Prisacariu, C. Y . Ren, X. Sun, P . H. S Tor r, and D. W. Murray. V ery High Frame Rate V olumetric Integration of Depth Images on Mobile Device. IEEE Transactions on Visualization and Computer Graphics (Proceedings International Symposium o n Mixed and Aug...

  48. [56]

    Get out of my lab: Large-sca le, real- time visual-inertial localization

    Simon Lynen, Torsten Sattler, Michael Bosse, Joel A Hes ch, Marc Pollefeys, and Roland Siegwart. Get out of my lab: Large-sca le, real- time visual-inertial localization. In Robotics: Science and Systems , volume 1, 2015

  49. [57]

    14 Lectures on Visual SLAM: From Theory to Practice

    Xiang Gao, Tao Zhang, Yi Liu, and Qinrui Y an. 14 Lectures on Visual SLAM: From Theory to Practice . Publishing House of Electronics Industry, 2017

  50. [58]

    Vi sual slam algorithms: A survey from 2010 to 2016

    Takafumi Taketomi, Hideaki Uchiyama, and Sei Ikeda. Vi sual slam algorithms: A survey from 2010 to 2016. IPSJ Transactions on Computer Vision and Applications , 9(1):16, 2017

  51. [59]

    An fft-bas ed technique for translation, rotation, and scale-invariant image regi stration

    B Srinivasa Reddy and Biswanath N Chatterji. An fft-bas ed technique for translation, rotation, and scale-invariant image regi stration. IEEE transactions on image processing , 5(8):1266–1271, 1996

  52. [60]

    Visual odometry ba sed on the fourier-mellin transform for a rover using a monocular grou nd-facing camera

    Tim Kazik and Ali Haydar G¨ okto˘ gan. Visual odometry ba sed on the fourier-mellin transform for a rover using a monocular grou nd-facing camera. In 2011 IEEE International Conference on Mechatronics , pages 469–474. IEEE, 2011

  53. [61]

    Visual odometry based on the fourier transform using a monocular gr ound- facing camera

    Merwan Birem, Richard Kleihorst, and Norddin El-Ghout i. Visual odometry based on the fourier transform using a monocular gr ound- facing camera. Journal of Real-Time Image Processing, 14(3):637–646, 2018

  54. [62]

    A survy of mono cular simultaneous localization and mapping

    Liu Haomin, Zhang Guofeng, and Bao hujun. A survy of mono cular simultaneous localization and mapping. Journal of Computer-Aided Design & Computer Graphics , 28(6):855–868, 2016

  55. [63]

    Event-based vision: A survey

    Guillermo Gallego, Tobi Delbruck, Garrick Orchard, Ch iara Bartolozzi, and Davide Scaramuzza. Event-based vision: A survey. 2019

  56. [64]

    A 128x128 120db 15us latency asynchronous temporal contrast vision s ensor

    Patrick Lichtsteiner, Christoph Posch, and Tobi Delbr uck. A 128x128 120db 15us latency asynchronous temporal contrast vision s ensor. IEEE journal of solid-state circuits , 43(2):566–576, 2008

  57. [65]

    4.1 a 640 × 480 dynamic vision sensor with a 9 µ m pixel and 300meps address-event representation

    Bongki Son, Y unjae Suh, Sungho Kim, Heejae Jung, Jun-Se ok Kim, Changwoo Shin, Keunju Park, Kyoobin Lee, Jinman Park, Jooye on Woo, et al. 4.1 a 640 × 480 dynamic vision sensor with a 9 µ m pixel and 300meps address-event representation. In 2017 IEEE International Solid-State...

  58. [66]

    A microbolometer asynchronous dy namic vision sensor for lwir

    Christoph Posch, Daniel Matolin, Rainer Wohlgenannt, Thomas Maier, and Martin Litzenberger. A microbolometer asynchronous dy namic vision sensor for lwir. IEEE Sensors Journal , 9(6):654–664, 2009

  59. [67]

    A sparc- compatible general purpose address-event processor with 2 0-bit l0ns- resolution asynchronous sensor data interface in 0.18 µ m cmos

    Michael Hofstatter, Peter Sch¨ on, and Christoph Posch . A sparc- compatible general purpose address-event processor with 2 0-bit l0ns- resolution asynchronous sensor data interface in 0.18 µ m cmos. In Proceedings of 2010 IEEE International Symposium on Circui ts and Systems,...

  60. [68]

    A du al-line optical transient sensor with on-chip precision time-stam p generation

    Christoph Posch, Michael Hofstatter, Daniel Matolin, Guy V anstraelen, Peter Schon, Nikolaus Donath, and Martin Litzenberger. A du al-line optical transient sensor with on-chip precision time-stam p generation. In 2007 IEEE International Solid-State Circuits Conference. Digest...

  61. [69]

    A 240 × 180 130 db 3 µ s latency global shutter spatiotemporal vision sensor

    Christian Brandli, Raphael Berner, Minhao Y ang, Shih- Chii Liu, and Tobi Delbruck. A 240 × 180 130 db 3 µ s latency global shutter spatiotemporal vision sensor. IEEE Journal of Solid-State Circuits , 49(10):2333–2341, 2014

  62. [70]

    A qvga 143 db dynamic range frame-free pwm image sensor with lossle ss pixel-level video compression and time-domain cds

    Christoph Posch, Daniel Matolin, and Rainer Wohlgenan nt. A qvga 143 db dynamic range frame-free pwm image sensor with lossle ss pixel-level video compression and time-domain cds. IEEE Journal of Solid-State Circuits, 46(1):259–275, 2010

  63. [71]

    Monoslam: Real-time single camera slam

    Andrew J Davison, Ian D Reid, Nicholas D Molton, and Oliv ier Stasse. Monoslam: Real-time single camera slam. IEEE Transactions on Pattern Analysis & Machine Intelligence , (6):1052–1067, 2007

  64. [72]

    Parallel tracking and map ping for small ar workspaces

    Georg Klein and David Murray. Parallel tracking and map ping for small ar workspaces. In Proceedings of the 2007 6th IEEE and ACM 12 International Symposium on Mixed and Augmented Reality , pages 1–

  65. [73]

    IEEE Computer Society, 2007

  66. [74]

    Improving the agility of k eyframe- based slam

    Georg Klein and David Murray. Improving the agility of k eyframe- based slam. In European Conference on Computer Vision , pages 802–

  67. [75]

    Orb: An efficient alternative to sift or surf

    Ethan Rublee, Vincent Rabaud, Kurt Konolige, and Gary R Bradski. Orb: An efficient alternative to sift or surf. In ICCV, volume 11, page 2. Citeseer, 2011

  68. [76]

    Orb-slam2: An open-s ource slam system for monocular, stereo, and rgb-d cameras

    Raul Mur-Artal and Juan D Tard´ os. Orb-slam2: An open-s ource slam system for monocular, stereo, and rgb-d cameras. IEEE Transactions on Robotics , 33(5):1255–1262, 2017

  69. [77]

    Cubemapslam: A piecewise-pinhole mon ocular fisheye slam system

    Y ahui Wang, Shaojun Cai, Shi-Jie Li, Y un Liu, Y angyan Gu o, Tao Li, and Ming-Ming Cheng. Cubemapslam: A piecewise-pinhole mon ocular fisheye slam system. In Asian Conference on Computer Vision , pages 34–49. Springer, 2018

  70. [78]

    Visual-inertial mo nocular slam with map reuse

    Ra´ ul Mur-Artal and Juan D Tard´ os. Visual-inertial mo nocular slam with map reuse. IEEE Robotics and Automation Letters , 2(2):796– 803, 2017

  71. [79]

    On-manifold preintegration for real-time visual–i nertial odom- etry

    Christian Forster, Luca Carlone, Frank Dellaert, and D avide Scara- muzza. On-manifold preintegration for real-time visual–i nertial odom- etry. IEEE Transactions on Robotics , 33(1):1–21, 2016

  72. [80]

    Schlegel, M

    D. Schlegel, M. Colosi, and G. Grisetti. ProSLAM: Graph SLAM from a Programmer’s Perspective. In 2018 IEEE International Conference on Robotics and Automation (ICRA) , pages 1–9, 2018

  73. [81]

    Efficient non-consecutive feature trac king for robust structure-from-motion

    Guofeng Zhang, Haomin Liu, Zilong Dong, Jiaya Jia, Tien -Tsin Wong, and Hujun Bao. Efficient non-consecutive feature trac king for robust structure-from-motion. IEEE Transactions on Image Processing, 25(12):5957–5970, 2016

  74. [82]

    Ope nvslam: a versatile visual slam framework, 2019

    Shinya Sumikura, Mikiya Shibuya, and Ken Sakurada. Ope nvslam: a versatile visual slam framework, 2019

  75. [83]

    Tagslam: Robust slam with fiducial markers

    Bernd Pfrommer and Kostas Daniilidis. Tagslam: Robust slam with fiducial markers. arXiv preprint arXiv:1910.00679 , 2019

  76. [84]

    Uco slam: Simul- taneous localization and mapping by fusion of keypoints and squared planar markers

    Rafael Munoz-Salinas and Rafael Medina-Carnicer. Uco slam: Simul- taneous localization and mapping by fusion of keypoints and squared planar markers. arXiv preprint arXiv:1902.03729 , 2019

  77. [85]

    Lsd-s lam: Large- scale direct monocular slam

    Jakob Engel, Thomas Sch¨ ops, and Daniel Cremers. Lsd-s lam: Large- scale direct monocular slam. In European conference on computer vision, pages 834–849. Springer, 2014

  78. [86]

    Larg e-scale direct slam with stereo cameras

    Jakob Engel, J¨ org St¨ uckler, and Daniel Cremers. Larg e-scale direct slam with stereo cameras. In 2015 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) , pages 1935–1942. IEEE, 2015

  79. [87]

    Large-s cale direct slam for omnidirectional cameras

    David Caruso, Jakob Engel, and Daniel Cremers. Large-s cale direct slam for omnidirectional cameras. In 2015 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) , pages 141–148. IEEE, 2015

  80. [88]

    Spherical-m odel-based slam on full-view images for indoor environments

    Jianfeng Li, Xiaowei Wang, and Shigang Li. Spherical-m odel-based slam on full-view images for indoor environments. Applied Sciences , 8(11):2268, 2018

  81. [89]

    Svo: Semidirect visual odom etry for monocular and multicamera systems

    Christian Forster, Zichao Zhang, Michael Gassner, Man uel Werl- berger, and Davide Scaramuzza. Svo: Semidirect visual odom etry for monocular and multicamera systems. IEEE Transactions on Robotics , 33(2):249–265, 2016

  82. [90]

    Cnn-svo: Improving the mapping in semi-dire ct visual odometry using single-image depth prediction

    Shing Y an Loo, Ali Jahani Amiri, Syamsiah Mashohor, Sai Hong Tang, and Hong Zhang. Cnn-svo: Improving the mapping in semi-dire ct visual odometry using single-image depth prediction. arXiv preprint arXiv:1810.01011, 2018

  83. [91]

    Direc t sparse odometry

    Jakob Engel, Vladlen Koltun, and Daniel Cremers. Direc t sparse odometry. CoRR, abs/1607.02565, 2016

  84. [92]

    Direc t sparse odom- etry

    Jakob Engel, Vladlen Koltun, and Daniel Cremers. Direc t sparse odom- etry. IEEE transactions on pattern analysis and machine intellig ence, 40(3):611–625, 2017

  85. [93]

    Evo: A geometric approach to event-based 6-dof parallel tracking and mapping in real time

    Henri Rebecq, Timo Horstsch¨ afer, Guillermo Gallego, and Davide Scaramuzza. Evo: A geometric approach to event-based 6-dof parallel tracking and mapping in real time. IEEE Robotics and Automation Letters, 2(2):593–600, 2016

  86. [94]

    Semi-dense 3d reconstruction wi th a stereo event camera

    Yi Zhou, Guillermo Gallego, Henri Rebecq, Laurent Knei p, Hongdong Li, and Davide Scaramuzza. Semi-dense 3d reconstruction wi th a stereo event camera. In Proceedings of the European Conference on Computer Vision (ECCV), pages 235–251, 2018

  87. [95]

    Simulta- neous localization and mapping for event-based vision syst ems

    David Weikersdorfer, Raoul Hoffmann, and J¨ org Conrad t. Simulta- neous localization and mapping for event-based vision syst ems. In International Conference on Computer Vision Systems , pages 133–142. Springer, 2013

  88. [96]

    Event-based 3d slam with a depth-augmented dynami c vision sensor

    David Weikersdorfer, David B Adrian, Daniel Cremers, a nd J¨ org Conradt. Event-based 3d slam with a depth-augmented dynami c vision sensor. In 2014 IEEE International Conference on Robotics and Automation (ICRA) , pages 359–364. IEEE, 2014

  89. [97]

    Inverse depth parametrization for monocular slam

    Javier Civera, Andrew J Davison, and JM Martinez Montie l. Inverse depth parametrization for monocular slam. IEEE transactions on robotics, 24(5):932–945, 2008

  90. [98]

    Dtam: Dense tracking and mapping in real-time

    Richard A Newcombe, Steven J Lovegrove, and Andrew J Dav ison. Dtam: Dense tracking and mapping in real-time. In 2011 international conference on computer vision , pages 2320–2327. IEEE, 2011

  91. [99]

    Multi-level mapping: Real-time dense monocular slam

    W Nicholas Greene, Kyel Ok, Peter Lommel, and Nicholas R oy. Multi-level mapping: Real-time dense monocular slam. In 2016 IEEE International Conference on Robotics and Automation (ICRA ), pages 833–840. IEEE, 2016

  92. [100]

    Kinectfusi on: Real- time dense surface mapping and tracking

    Richard A Newcombe, Shahram Izadi, Otmar Hilliges, Dav id Molyneaux, David Kim, Andrew J Davison, Pushmeet Kohli, Jam ie Shotton, Steve Hodges, and Andrew W Fitzgibbon. Kinectfusi on: Real- time dense surface mapping and tracking. In ISMAR, volume 11, pages 127–136, 2011

  93. [101]

    Kinectfusion: real-time 3d recon- struction and interaction using a moving depth camera

    Shahram Izadi, David Kim, Otmar Hilliges, David Molyn eaux, Richard Newcombe, Pushmeet Kohli, Jamie Shotton, Steve Hodges, Dus tin Freeman, Andrew Davison, et al. Kinectfusion: real-time 3d recon- struction and interaction using a moving depth camera. In Proceedings of the 24t...

  94. [102]

    Real-time visual odometry from dense rgb-d images

    Frank Steinbr¨ ucker, J¨ urgen Sturm, and Daniel Creme rs. Real-time visual odometry from dense rgb-d images. In 2011 IEEE International Conference on Computer Vision W orkshops (ICCV W orkshops) , pages 719–722. IEEE, 2011

  95. [103]

    Ro bust odometry estimation for rgb-d cameras

    Christian Kerl, J¨ urgen Sturm, and Daniel Cremers. Ro bust odometry estimation for rgb-d cameras. In 2013 IEEE International Conference on Robotics and Automation , pages 3748–3754. IEEE, 2013

  96. [104]

    De nse visual slam for rgb-d cameras

    Christian Kerl, J¨ urgen Sturm, and Daniel Cremers. De nse visual slam for rgb-d cameras. In 2013 IEEE/RSJ International Conference on Intelligent Robots and Systems , pages 2100–2106. IEEE, 2013

  97. [105]

    3-d mapping with an rgb-d camera

    Felix Endres, J¨ urgen Hess, J¨ urgen Sturm, Daniel Cremers, and Wolfram Burgard. 3-d mapping with an rgb-d camera. IEEE transactions on robotics, 30(1):177–187, 2013

  98. [106]

    Kintinuous: Spatially ex tended kinectfusion

    Thomas Whelan, Michael Kaess, Maurice Fallon, Hordur Johannsson, John J Leonard, and John McDonald. Kintinuous: Spatially ex tended kinectfusion. 2012

  99. [107]

    Real-time large-scale de nse rgb- d slam with volumetric fusion

    Thomas Whelan, Michael Kaess, Hordur Johannsson, Mau rice Fallon, John J Leonard, and John McDonald. Real-time large-scale de nse rgb- d slam with volumetric fusion. The International Journal of Robotics Research, 34(4-5):598–626, 2015

  100. [108]

    Leonard, and John Mcdonald

    Thomas Whelan, Hordur Johannsson, Michael Kaess, Joh n J. Leonard, and John Mcdonald. Robust real-time visual odometry for den se rgb-d mapping. In IEEE International Conference on Robotics and Automation, 2011

  101. [109]

    Online global lo op closure detection for large-scale multi-session graph-based slam

    Mathieu Labbe and Franc ¸ois Michaud. Online global lo op closure detection for large-scale multi-session graph-based slam . In 2014 IEEE/RSJ International Conference on Intelligent Robots a nd Systems, pages 2661–2666. IEEE, 2014

  102. [110]

    Appearance-based loop closur e detection in real-time for large-scale and long-term operation

    MM Labb´ e and F Michaud. Appearance-based loop closur e detection in real-time for large-scale and long-term operation. IEEE Transactions on Robotics , pages 734–745

  103. [111]

    Memory managem ent for real- time appearance-based loop closure detection

    Mathieu Labb´ e and Franc ¸ois Michaud. Memory managem ent for real- time appearance-based loop closure detection. In 2011 IEEE/RSJ International Conference on Intelligent Robots and System s, pages 1271–1276. IEEE, 2011

  104. [112]

    Rtab-map as an o pen-source lidar and visual simultaneous localization and mapping lib rary for large-scale and long-term online operation

    Mathieu Labb´ e and Franc ¸ois Michaud. Rtab-map as an o pen-source lidar and visual simultaneous localization and mapping lib rary for large-scale and long-term online operation. Journal of Field Robotics , 36(2):416–446, 2019

  105. [113]

    Dyn amicfu- sion: Reconstruction and tracking of non-rigid scenes in re al-time

    Richard A Newcombe, Dieter Fox, and Steven M Seitz. Dyn amicfu- sion: Reconstruction and tracking of non-rigid scenes in re al-time. In Proceedings of the IEEE conference on computer vision and pa ttern recognition, pages 343–352, 2015

  106. [114]

    V olumedeform: Real-time vo lumetric non-rigid reconstruction

    Matthias Innmann, Michael Zollh¨ ofer, Matthias Nieß ner, Christian Theobalt, and Marc Stamminger. V olumedeform: Real-time vo lumetric non-rigid reconstruction. In European Conference on Computer Vision , pages 362–379. Springer, 2016

  107. [115]

    Fusion4d: Real- time performance capture of challenging scenes

    Mingsong Dou, Sameh Khamis, Y ury Degtyarev, Philip Da vidson, Sean Ryan Fanello, Adarsh Kowdle, Sergio Orts Escolano, Chr istoph Rhemann, David Kim, Jonathan Taylor, et al. Fusion4d: Real- time performance capture of challenging scenes. ACM Transactions on Graphics (TOG) , 35...

  108. [116]

    Elasticfusion: Dense slam without a pos e graph

    Thomas Whelan, Stefan Leutenegger, R Salas-Moreno, B en Glocker, and Andrew Davison. Elasticfusion: Dense slam without a pos e graph. Robotics: Science and Systems, 2015

  109. [117]

    Elasticfusion: Real-tim e dense slam and light source estimation

    Thomas Whelan, Renato F Salas-Moreno, Ben Glocker, An drew J Davison, and Stefan Leutenegger. Elasticfusion: Real-tim e dense slam and light source estimation. The International Journal of Robotics Research, 35(14):1697–1716, 2016

  110. [118]

    InfiniTAM v3: A Framework for Large-Scale 3D Reconstruction with Loop Closure

    V A Prisacariu, O K¨ ahler, S Golodetz, M Sapienza, T Cav allari, P H S Torr, and D W Murray. InfiniTAM v3: A Framework for Large-Scale 3D Reconstruction with Loop Closure. arXiv pre-print arXiv:1708.00783v1, 2017

  111. [119]

    Olaf K¨ ahler, Victor Adrian Prisacariu, and David W. M urray. Real-time large-scale dense 3d reconstruction with loop closure. In Computer Vision - ECCV 2016 - 14th European Conference, Amsterdam, Th e Netherlands, October 11-14, 2016, Proceedings, Part VIII , pages 500– 516, 2016

  112. [120]

    Bundlefusion: Real-time globally con sistent 3d re- construction using on-the-fly surface re-integration

    Angela Dai, Matthias Nießner, Michael Zoll¨ ofer, Sha hram Izadi, and Christian Theobalt. Bundlefusion: Real-time globally con sistent 3d re- construction using on-the-fly surface re-integration. ACM Transactions on Graphics 2017 (TOG) , 2017

  113. [121]

    Ko- fusion: Dense visual slam with tightly-coupled kinematic a nd odo- metric tracking

    Charlie Houseago, Michael Bloesch, and Stefan Leuten egger. Ko- fusion: Dense visual slam with tightly-coupled kinematic a nd odo- metric tracking. In 2019 International Conference on Robotics and Automation (ICRA) , pages 4054–4060. IEEE, 2019

  114. [122]

    Soft- slam: Computationally efficient stereo visual slam for auto nomous uavs

    Igor Cviˇ sic, Josip Cesic, Ivan Markovic, and Ivan Pet rovic. Soft- slam: Computationally efficient stereo visual slam for auto nomous uavs. Journal of field robotics , 2017

  115. [123]

    Stereo odometry based on careful feature selection and tracking

    Igor Cviˇ si´ c and Ivan Petrovi´ c. Stereo odometry based on careful feature selection and tracking. In 2015 European Conference on Mobile Robots (ECMR), pages 1–6. IEEE, 2015

  116. [124]

    Robust keyframe-based dense slam with an rgb- d camera

    Haomin Liu, Chen Li, Guojun Chen, Guofeng Zhang, Micha el Kaess, and Hujun Bao. Robust keyframe-based dense slam with an rgb- d camera. arXiv preprint arXiv:1711.05166 , 2017

  117. [125]

    Rgb-d sl am in dynamic environments using points correlations

    Weichen Dai, Y u Zhang, Ping Li, and Zheng Fang. Rgb-d sl am in dynamic environments using points correlations. arXiv preprint arXiv:1811.03217, 2018

  118. [126]

    Schneider, M

    T. Schneider, M. T. Dymczyk, M. Fehr, K. Egger, S. Lynen , I. Gilitschenski, and R. Siegwart. maplab: An open framewor k for research in visual-inertial mapping and localization. IEEE Robotics and Automation Letters , 2018

  119. [127]

    Point-based mult i-view stereo network

    Jing Xu Hao Su Rui Chen, Songfang Han. Point-based mult i-view stereo network. arXiv preprint arXiv:1908.04422 , 2019

  120. [128]

    Mid-fusion: Octree-base d object- level multi-instance dynamic slam

    Binbin Xu, Wenbin Li, Dimos Tzoumanikas, Michael Bloe sch, Andrew Davison, and Stefan Leutenegger. Mid-fusion: Octree-base d object- level multi-instance dynamic slam. arXiv preprint arXiv:1812.07976 , 2018

  121. [129]

    M. Runz, M. Buffier, and L. Agapito. Maskfusion: Real-t ime recog- nition, tracking and reconstruction of multiple moving obj ects. In 2018 IEEE International Symposium on Mixed and Augmented Re ality (ISMAR), pages 10–20, Oct 2018

  122. [130]

    Scale drift-aware large scale monocular slam

    Hauke Strasdat, J Montiel, and Andrew J Davison. Scale drift-aware large scale monocular slam. Robotics: Science and Systems VI , 2(3):7, 2010

  123. [131]

    Keyframe-based visual–inertial odometr y using non- linear optimization

    Stefan Leutenegger, Simon Lynen, Michael Bosse, Rola nd Siegwart, and Paul Furgale. Keyframe-based visual–inertial odometr y using non- linear optimization. The International Journal of Robotics Research , 34(3):314–334, 2015

  124. [132]

    Towa rds consis- tent visual-inertial navigation

    Guoquan Huang, Michael Kaess, and John J Leonard. Towa rds consis- tent visual-inertial navigation. In 2014 IEEE International Conference on Robotics and Automation (ICRA) , pages 4926–4933. IEEE, 2014

  125. [133]

    High-precisio n, consistent ekf-based visual-inertial odometry

    Mingyang Li and Anastasios I Mourikis. High-precisio n, consistent ekf-based visual-inertial odometry. The International Journal of Robotics Research, 32(6):690–711, 2013

  126. [134]

    Carlos Campos, J. M. M. Montiel, and Juan D. Tard´ os. Fa st and robust initialization for visual-inertial slam. 2019 International Conference on Robotics and Automation (ICRA) , pages 1288–1294, 2019

  127. [135]

    An investigation of google tango R© tablet for low cost 3d scanning

    Mark Froehlich, Salman Azhar, and Matthew V anture. An investigation of google tango R© tablet for low cost 3d scanning. In ISARC. Pro- ceedings of the International Symposium on Automation and R obotics in Construction , volume 34. Vilnius Gediminas Technical University, Depa...

  128. [136]

    Real-time high reso lution 3d data on the hololens

    Mathieu Garon, Pierre-Olivier Boulet, Jean-Philipp e Doironz, Luc Beaulieu, and Jean-Franc ¸ois Lalonde. Real-time high reso lution 3d data on the hololens. In 2016 IEEE International Symposium on Mixed and Augmented Reality (ISMAR-Adjunct) , pages 189–191. IEEE, 2016

  129. [137]

    Penncosyvio: A challenging visual inertial odometry bench mark

    Bernd Pfrommer, Nitin Sanket, Kostas Daniilidis, and Jonas Cleveland. Penncosyvio: A challenging visual inertial odometry bench mark. In 2017 IEEE International Conference on Robotics and Automat ion (ICRA), pages 3847–3854. IEEE, 2017

  130. [138]

    A benchmark comparison of monocular visual-inertial odometry algorithms for flyin g robots

    Jeffrey Delmerico and Davide Scaramuzza. A benchmark comparison of monocular visual-inertial odometry algorithms for flyin g robots. In 2018 IEEE International Conference on Robotics and Automat ion (ICRA), pages 2502–2509. IEEE, 2018

  131. [139]

    Vision based navigation for micro helicopters

    Stephan M Weiss. Vision based navigation for micro helicopters . PhD thesis, ETH Zurich, 2012

  132. [140]

    A mu lti-state con- straint kalman filter for vision-aided inertial navigation

    Anastasios I Mourikis and Stergios I Roumeliotis. A mu lti-state con- straint kalman filter for vision-aided inertial navigation . In Proceedings 2007 IEEE International Conference on Robotics and Automat ion, pages 3565–3572. IEEE, 2007

  133. [141]

    Rob ust stereo visual inertial odometry for fast autonomous flight

    Ke Sun, Kartik Mohta, Bernd Pfrommer, Michael Watters on, Sikang Liu, Y ash Mulgaonkar, Camillo J Taylor, and Vijay Kumar. Rob ust stereo visual inertial odometry for fast autonomous flight. IEEE Robotics and Automation Letters , 3(2):965–972, 2018

  134. [142]

    Robust visual inertial odometry using a direct ekf-based ap proach

    Michael Bloesch, Sammy Omari, Marco Hutter, and Rolan d Siegwart. Robust visual inertial odometry using a direct ekf-based ap proach. In 2015 IEEE/RSJ international conference on intelligent rob ots and systems (IROS) , pages 298–304. IEEE, 2015

  135. [143]

    Monocular visual-inertial state estimation for mobile aug mented reality

    Peiliang Li, Tong Qin, Botao Hu, Fengyuan Zhu, and Shao jie Shen. Monocular visual-inertial state estimation for mobile aug mented reality. In 2017 IEEE International Symposium on Mixed and Augmented Reality (ISMAR) , pages 11–21. IEEE, 2017

  136. [144]

    Online temporal calibratio n for monocular visual-inertial systems

    Tong Qin and Shaojie Shen. Online temporal calibratio n for monocular visual-inertial systems. In 2018 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) , pages 3662–3669. IEEE, 2018

  137. [145]

    Robust initialization of mo nocular visual- inertial estimation on aerial robots

    Tong Qin and Shaojie Shen. Robust initialization of mo nocular visual- inertial estimation on aerial robots. In 2017 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) , pages 4225–

  138. [146]

    Monocular visual–iner tial state esti- mation with online initialization and camera–imu extrinsi c calibration

    Zhenfei Y ang and Shaojie Shen. Monocular visual–iner tial state esti- mation with online initialization and camera–imu extrinsi c calibration. IEEE Transactions on Automation Science and Engineering , 14(1):39– 51, 2016

  139. [147]

    Ice-ba: Incremental, consistent and efficient bundle a djustment for visual-inertial slam

    Haomin Liu, Mingyu Chen, Guofeng Zhang, Hujun Bao, and Yingze Bao. Ice-ba: Incremental, consistent and efficient bundle a djustment for visual-inertial slam. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , pages 1974–1982, 2018

  140. [148]

    Structvio: Visual-inertial odometry with structural regu larity of man- made environments

    Danping Zou, Y uanxin Wu, Ling Pei, Haibin Ling, and Wen xian Y u. Structvio: Visual-inertial odometry with structural regu larity of man- made environments. IEEE Transactions on Robotics , 2019

  141. [149]

    Robust keyfr ame-based monocular slam for augmented reality

    Haomin Liu, Guofeng Zhang, and Hujun Bao. Robust keyfr ame-based monocular slam for augmented reality. In 2016 IEEE International Symposium on Mixed and Augmented Reality (ISMAR) , pages 1–10. IEEE, 2016

  142. [150]

    Continuous-time visual-inertial odometry for even t cameras

    Elias Mueggler, Guillermo Gallego, Henri Rebecq, and Davide Scara- muzza. Continuous-time visual-inertial odometry for even t cameras. IEEE Transactions on Robotics , 34(6):1425–1440, 2018

  143. [151]

    Event-based visual inertial odometry

    Alex Zihao Zhu, Nikolay Atanasov, and Kostas Daniilid is. Event-based visual inertial odometry. In 2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , pages 5816–5824. IEEE, 2017

  144. [152]

    Event-based visual-inertial odometr y on a fixed-wing unmanned aerial vehicle

    Kaleb J Nelson. Event-based visual-inertial odometr y on a fixed-wing unmanned aerial vehicle. Technical report, AIR FORCE INSTI TUTE OF TECHNOLOGY WRIGHT-PA TTERSON AFB OH WRIGHT- PA TTERSON, 2019

  145. [153]

    S ensor- failure-resilient multi-imu visual-inertial navigation

    Kevin Eckenhoff, Patrick Geneva, and Guoquan Huang. S ensor- failure-resilient multi-imu visual-inertial navigation. 2019 International Conference on Robotics and Automation (ICRA) , pages 3542–3548, 2019

  146. [154]

    Unsupervised deep visual-inertial odometry wit h online error correction for rgb-d imagery

    E Jared Shamwell, Kyle Lindgren, Sarah Leung, and Will iam D Nothwang. Unsupervised deep visual-inertial odometry wit h online error correction for rgb-d imagery. IEEE transactions on pattern analysis and machine intelligence , 2019

  147. [155]

    Vis ual- inertial odometry for unmanned aerial vehicle using deep le arning

    Hongyun Lee, Matthew McCrink, and James W Gregory. Vis ual- inertial odometry for unmanned aerial vehicle using deep le arning. In AIAA Scitech 2019 F orum , page 1410, 2019

  148. [156]

    Pop- up slam: Semantic monocular plane slam for low-texture envi ronments

    Shichao Y ang, Y u Song, Michael Kaess, and Sebastian Sc herer. Pop- up slam: Semantic monocular plane slam for low-texture envi ronments. In 2016 IEEE/RSJ International Conference on Intelligent Rob ots and Systems (IROS) , pages 1222–1229. IEEE, 2016

  149. [157]

    6-dof object pose from semantic keypoints

    Georgios Pavlakos, Xiaowei Zhou, Aaron Chan, Konstan tinos G Derpa- nis, and Kostas Daniilidis. 6-dof object pose from semantic keypoints. In 2017 IEEE International Conference on Robotics and Automat ion (ICRA), pages 2011–2018. IEEE, 2017. 14

  150. [158]

    Lift: Learned invariant feature transform

    Kwang Moo Yi, Eduard Trulls, Vincent Lepetit, and Pasc al Fua. Lift: Learned invariant feature transform. In European Conference on Computer Vision, pages 467–483. Springer, 2016

  151. [159]

    Toward geometric deep slam

    Daniel DeTone, Tomasz Malisiewicz, and Andrew Rabino vich. Toward geometric deep slam. arXiv preprint arXiv:1707.07410 , 2017

  152. [160]

    Su- perpoint: Self-supervised interest point detection and de scription

    Daniel DeTone, Tomasz Malisiewicz, and Andrew Rabino vich. Su- perpoint: Self-supervised interest point detection and de scription. In Proceedings of the IEEE Conference on Computer Vision and Pa ttern Recognition W orkshops, pages 224–236, 2018

  153. [161]

    Stereo vision -based semantic 3d object and ego-motion tracking for autonomous driving

    Peiliang Li, Qin Tong, and Shaojie Shen. Stereo vision -based semantic 3d object and ego-motion tracking for autonomous driving. I n Euro- pean Conference on Computer Vision , 2018

  154. [162]

    Gcnv2: Efficient correspondence prediction for real-time s lam

    Jiexiong Tang, Ludvig Ericson, John Folkesson, and Pa tric Jensfelt. Gcnv2: Efficient correspondence prediction for real-time s lam. arXiv preprint arXiv:1902.11046, 2019

  155. [163]

    V olumetric i nstance- aware semantic mapping and 3d object discovery

    Margarita Grinvald, Fadri Furrer, Tonci Novkovic, Je n Jen Chung, Cesar Cadena, Roland Siegwart, and Juan Nieto. V olumetric i nstance- aware semantic mapping and 3d object discovery. arXiv preprint arXiv:1903.00268, 2019

  156. [164]

    Salientdso: Bringing attention to direct spars e odometry

    Huai-Jen Liang, Nitin J Sanket, Cornelia Ferm ¨ uller, and Yiannis Aloimonos. Salientdso: Bringing attention to direct spars e odometry. IEEE Transactions on Automation Science and Engineering , 2019

  157. [165]

    Structure aware slam using quadrics and planes

    Mehdi Hosseinzadeh, Y asir Latif, Trung Pham, Niko Sue nderhauf, and Ian Reid. Structure aware slam using quadrics and planes. In Asian Conference on Computer Vision , pages 410–426. Springer, 2018

  158. [166]

    Cubeslam: Monocu lar 3-d object slam

    Shichao Y ang and Sebastian Scherer. Cubeslam: Monocu lar 3-d object slam. IEEE Transactions on Robotics , 2019

  159. [167]

    Monocular object and plane slam in structured environments

    Shichao Y ang and Sebastian Scherer. Monocular object and plane slam in structured environments. IEEE Robotics and Automation Letters , 4(4):3145–3152, 2019

  160. [168]

    Monogrnet: A geome tric reasoning network for 3d object localization

    Zengyi Qin, Jinglu Wang, and Y an Lu. Monogrnet: A geome tric reasoning network for 3d object localization. The Thirty-Third AAAI Conference on Artificial Intelligence (AAAI-19) , 2019

  161. [169]

    Event -based features for robotic vision

    Xavier Lagorce, Sio Hoi Ieng, and Ryad Benosman. Event -based features for robotic vision. In IEEE/RSJ International Conference on Intelligent Robots & Systems , 2013

  162. [170]

    Fast event- based corner detection

    Elias Mueggler, Chiara Bartolozzi, and Davide Scaram uzza. Fast event- based corner detection. In BMVC, 2017

  163. [171]

    Recent adv ances in deep learning for object detection

    Xiongwei Wu, Doyen Sahoo, and Steven CH Hoi. Recent adv ances in deep learning for object detection. arXiv preprint arXiv:1908.03673 , 2019

  164. [172]

    Salas-Moreno, Richard A

    Renato F. Salas-Moreno, Richard A. Newcombe, Hauke St rasdat, Paul H. J. Kelly, and Andrew J. Davison. Slam++: Simultaneous loc alisation and mapping at the level of objects. In Computer Vision & Pattern Recognition, 2013

  165. [173]

    Semi-dense 3d sema ntic mapping from monocular slam

    Xuanpeng Li and Rachid Belaroussi. Semi-dense 3d sema ntic mapping from monocular slam. 2016

  166. [174]

    Semanticfusion: Dense 3d semantic mapping with convol utional neural networks

    John McCormac, Ankur Handa, Andrew Davison, and Stefa n Leuteneg- ger. Semanticfusion: Dense 3d semantic mapping with convol utional neural networks. In 2017 IEEE International Conference on Robotics and automation (ICRA) , pages 4628–4635. IEEE, 2017

  167. [175]

    Pham, Y asir Latif, Michael M ilford, and Ian Reid

    Niko Sunderhauf, Trung T. Pham, Y asir Latif, Michael M ilford, and Ian Reid. Meaningful maps with object-oriented semantic ma pping. In IEEE/RSJ International Conference on Intelligent Robots & Systems, 2017

  168. [176]

    Marrnet: 3d shape reconstruction via 2.5 d s ketches

    Jiajun Wu, Yifan Wang, Tianfan Xue, Xingyuan Sun, Bill Freeman, and Josh Tenenbaum. Marrnet: 3d shape reconstruction via 2.5 d s ketches. In Advances in neural information processing systems , pages 540–550, 2017

  169. [177]

    3dmv: Joint 3d-multi -view pre- diction for 3d semantic scene segmentation

    Angela Dai and Matthias Nießner. 3dmv: Joint 3d-multi -view pre- diction for 3d semantic scene segmentation. In Proceedings of the European Conference on Computer Vision (ECCV) , pages 452–468, 2018

  170. [178]

    Pix3d: Dataset and methods for single-image 3d shape modeli ng

    Xingyuan Sun, Jiajun Wu, Xiuming Zhang, Zhoutong Zhan g, Chengkai Zhang, Tianfan Xue, Joshua B Tenenbaum, and William T Freema n. Pix3d: Dataset and methods for single-image 3d shape modeli ng. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR), 2018

  171. [179]

    Scancomplete: Large-scale scene com pletion and semantic segmentation for 3d scans

    Angela Dai, Daniel Ritchie, Martin Bokeloh, Scott Ree d, J¨ urgen Sturm, and Matthias Nießner. Scancomplete: Large-scale scene com pletion and semantic segmentation for 3d scans. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , pages 4578– 4587, 2018

  172. [180]

    Davison, and Stefan Leutenegger

    John McCormac, Ronald Clark, Michael Bloesch, Andrew J. Davison, and Stefan Leutenegger. Fusion++: V olumetric object-level slam. 2018 International Conference on 3D Vision (3DV) , pages 32–41, 2018

  173. [181]

    Segmap: 3d segment mapping usin g data- driven descriptors

    Renaud Dub´ e, Andrei Cramariuc, Daniel Dugas, Juan Ni eto, Roland Siegwart, and Cesar Cadena. Segmap: 3d segment mapping usin g data- driven descriptors. arXiv preprint arXiv:1804.09557 , 2018

  174. [182]

    3d-sis: 3d se mantic instance segmentation of rgb-d scans

    Ji Hou, Angela Dai, and Matthias Nießner. 3d-sis: 3d se mantic instance segmentation of rgb-d scans. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , pages 4421–4430, 2019

  175. [183]

    Da-rnn: Semantic mapping with data associated recurrent neural networks

    Y u Xiang and Dieter Fox. Da-rnn: Semantic mapping with data associated recurrent neural networks. arXiv preprint arXiv:1703.03098, 2017

  176. [184]

    Densefusion: 6d object pos e estimation by iterative dense fusion

    Chen Wang, Danfei Xu, Y uke Zhu, Roberto Mart´ ın-Mart´ın, Cewu Lu, Li Fei-Fei, and Silvio Savarese. Densefusion: 6d object pos e estimation by iterative dense fusion. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , pages 3343–3352, 2019

  177. [185]

    Ccnet: Criss-cross attention for semant ic segmentation

    Zilong Huang, Xinggang Wang, Lichao Huang, Chang Huan g, Y unchao Wei, and Wenyu Liu. Ccnet: Criss-cross attention for semant ic segmentation. arXiv preprint arXiv:1811.11721 , 2018

  178. [186]

    An event-based classifier fo r dynamic vision sensor and synthetic data

    Evangelos Stromatias, Miguel Soto, Mar´ ıa Teresa Serrano Gotarredona, and Bernab´ e Linares Barranco. An event-based classifier fo r dynamic vision sensor and synthetic data. Frontiers in Neuroscience, 11 (art´ ıculo 360), 2017

  179. [187]

    Event-based ge sture recog- nition with dynamic background suppression using smartpho ne com- putational capabilities

    Jean-Matthieu Maro and Ryad Benosman. Event-based ge sture recog- nition with dynamic background suppression using smartpho ne com- putational capabilities. arXiv preprint arXiv:1811.07802 , 2018

  180. [188]

    Investigation of event-ba sed mem- ory surfaces for high-speed detection, unsupervised featu re extraction, and object recognition

    Saeed Afshar, Tara Julia Hamilton, Jonathan C Tapson, Andr´ e van Schaik, and Gregory Kevin Cohen. Investigation of event-ba sed mem- ory surfaces for high-speed detection, unsupervised featu re extraction, and object recognition. Frontiers in neuroscience, 12:1047, 2018

  181. [189]

    Dynamic vision sensor integrat ion on fpga-based cnn accelerators for high-speed visual classifi cation

    Alejandro Linares-Barranco, Antonio Rios-Navarro, Ricardo Tapiador- Morales, and Tobi Delbruck. Dynamic vision sensor integrat ion on fpga-based cnn accelerators for high-speed visual classifi cation. arXiv preprint arXiv:1905.07419, 2019

  182. [190]

    Cnn- slam: Real-time dense monocular slam with learned depth pre diction

    Keisuke Tateno, Federico Tombari, Iro Laina, and Nass ir Navab. Cnn- slam: Real-time dense monocular slam with learned depth pre diction. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pages 6243–6252, 2017

  183. [191]

    Deepvo: A de ep learning approach for monocular visual odometry

    Vikram Mohanty, Shubh Agrawal, Shaswat Datta, Arna Gh osh, Vishnu Dutt Sharma, and Debashish Chakravarty. Deepvo: A de ep learning approach for monocular visual odometry. arXiv preprint arXiv:1611.06069, 2016

  184. [192]

    Gs3d: An efficient 3d object detection framework for autonom ous driving

    Buyu Li, Wanli Ouyang, Lu Sheng, Xingyu Zeng, and Xiaog ang Wang. Gs3d: An efficient 3d object detection framework for autonom ous driving. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , pages 1019–1028, 2019

  185. [193]

    Un deepvo: Monocular visual odometry through unsupervised deep learn ing

    Ruihao Li, Sen Wang, Zhiqiang Long, and Dongbing Gu. Un deepvo: Monocular visual odometry through unsupervised deep learn ing. In 2018 IEEE International Conference on Robotics and Automat ion (ICRA), pages 7286–7291. IEEE, 2018

  186. [194]

    Learning the depths of moving p eople by watching frozen people

    Zhengqi Li, Tali Dekel, Forrester Cole, Richard Tucke r, Noah Snavely, Ce Liu, and William T Freeman. Learning the depths of moving p eople by watching frozen people. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , pages 4521–4530, 2019

  187. [195]

    Frost, V

    D. Frost, V . Prisacariu, and D. Murray. Recovering sta ble scale in monocular slam using object-supplemented bundle adjustme nt. IEEE Transactions on Robotics , 34(3):736–747, June 2018

  188. [196]

    Bayesian scale est imation for monocular slam based on generic object detection for correc ting scale drift

    Edgar Sucar and Jean Bernard Hayet. Bayesian scale est imation for monocular slam based on generic object detection for correc ting scale drift. 2017

  189. [197]

    Geonet: Unsupervised le arning of dense depth, optical flow and camera pose

    Zhichao Yin and Jianping Shi. Geonet: Unsupervised le arning of dense depth, optical flow and camera pose. In CVPR, 2018

  190. [198]

    Codeslamlearning a compact, optimisa ble representation for dense visual slam

    Michael Bloesch, Jan Czarnowski, Ronald Clark, Stefa n Leutenegger, and Andrew J Davison. Codeslamlearning a compact, optimisa ble representation for dense visual slam. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , pages 2560– 2568, 2018

  191. [199]

    Mono- stixels: monocular depth reconstruction of dynamic street scenes

    Fabian Brickwedde, Steffen Abraham, and Rudolf Meste r. Mono- stixels: monocular depth reconstruction of dynamic street scenes. In 2018 IEEE International Conference on Robotics and Automat ion (ICRA), pages 1–7. IEEE, 2018

  192. [200]

    Saputra, Pedro Por to Buarque de Gusm˜ ao, Andrew Markham, and Agathoniki Trigoni

    Y asin Almalioglu, Muhamad Risqi U. Saputra, Pedro Por to Buarque de Gusm˜ ao, Andrew Markham, and Agathoniki Trigoni. Ganvo: Unsupervised deep monocular visual odometry and depth esti mation with generative adversarial networks. 2019 International Conference on Robotics and A...

  193. [201]

    Gen- slam: Generative modeling for monocular simultaneous loca lization and mapping

    Punarjay Chakravarty, Praveen Narayanan, and Tom Rou ssel. Gen- slam: Generative modeling for monocular simultaneous loca lization and mapping. arXiv preprint arXiv:1902.02086 , 2019. 15

  194. [202]

    Towards robust monocular depth estimation: Mixing dataset s for zero- shot cross-dataset transfer

    Katrin Lasinger, Ren´ e Ranftl, Konrad Schindler, and Vladlen Koltun. Towards robust monocular depth estimation: Mixing dataset s for zero- shot cross-dataset transfer. arXiv:1907.01341, 2019

  195. [203]

    Deepmvs: Learning multi-view stereopsis

    Po-Han Huang, Kevin Matzen, Johannes Kopf, Narendra A huja, and Jia-Bin Huang. Deepmvs: Learning multi-view stereopsis. I n Pro- ceedings of the IEEE Conference on Computer Vision and Patte rn Recognition, pages 2821–2830, 2018

  196. [204]

    Deepv2d: Video to depth with dif ferentiable structure from motion

    Jia Deng Zachary Teed. Deepv2d: Video to depth with dif ferentiable structure from motion. arXiv:1812.04605, 2018

  197. [205]

    A spiking neural network model of depth from defocus for event- based neuromorphic vision

    Germain Haessig, Xavier Berthelon, Sio-Hoi Ieng, and Ryad Benos- man. A spiking neural network model of depth from defocus for event- based neuromorphic vision. Scientific reports , 9(1):3744, 2019

  198. [206]

    A unifying contrast maximization framework for event cameras, with ap plications to motion, depth, and optical flow estimation

    Guillermo Gallego, Henri Rebecq, and Davide Scaramuz za. A unifying contrast maximization framework for event cameras, with ap plications to motion, depth, and optical flow estimation. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , pages 3867...

  199. [207]

    Event- based stereo depth estimation using belief propagation

    Zhen Xie, Shengyong Chen, and Garrick Orchard. Event- based stereo depth estimation using belief propagation. Frontiers in neuroscience , 11:535, 2017

  200. [208]

    Learning vi sual odom- etry with a convolutional network

    Kishore Reddy Konda and Roland Memisevic. Learning vi sual odom- etry with a convolutional network. In VISAPP (1) , pages 486–490, 2015

  201. [209]

    Exploring representation learning with cnns f or frame-to- frame ego-motion estimation

    Gabriele Costante, Michele Mancini, Paolo V aligi, an d Thomas A Ciarfuglia. Exploring representation learning with cnns f or frame-to- frame ego-motion estimation. IEEE robotics and automation letters , 1(1):18–25, 2015

  202. [210]

    Po senet: A convolutional network for real-time 6-dof camera relocali zation

    Alex Kendall, Matthew Grimes, and Roberto Cipolla. Po senet: A convolutional network for real-time 6-dof camera relocali zation. In Proceedings of the IEEE international conference on comput er vision , pages 2938–2946, 2015

  203. [211]

    Vinet: Visual-inertial odometry as a sequence-to -sequence learning problem

    Ronald Clark, Sen Wang, Hongkai Wen, Andrew Markham, a nd Niki Trigoni. Vinet: Visual-inertial odometry as a sequence-to -sequence learning problem. In Thirty-First AAAI Conference on Artificial Intelligence, 2017

  204. [212]

    Deepvo: Towards end-to-end visual odometry with deep recurrent con volutional neural networks

    Sen Wang, Ronald Clark, Hongkai Wen, and Niki Trigoni. Deepvo: Towards end-to-end visual odometry with deep recurrent con volutional neural networks. In 2017 IEEE International Conference on Robotics and Automation (ICRA) , pages 2043–2050. IEEE, 2017

  205. [213]

    Unsupervised learning of depth and ego-motion from video

    Tinghui Zhou, Matthew Brown, Noah Snavely, and David G Lowe. Unsupervised learning of depth and ego-motion from video. I n Proceedings of the IEEE Conference on Computer Vision and Pa ttern Recognition, pages 1851–1858, 2017

  206. [214]

    Sfm-net: Learning o f structure and motion from video

    Sudheendra Vijayanarasimhan, Susanna Ricco, Cordel ia Schmid, Rahul Sukthankar, and Katerina Fragkiadaki. Sfm-net: Learning o f structure and motion from video. arXiv preprint arXiv:1704.07804 , 2017

  207. [215]

    Schnberg er, Marc Polle- feys, and Torsten Sattler

    Konstantinos Nektarios Lianos, Johannes L. Schnberg er, Marc Polle- feys, and Torsten Sattler. Vso: Visual semantic odometry. I n European Conference on Computer Vision (ECCV) , 2018

  208. [216]

    Vidloc: A deep spatio-temporal model for 6-dof video-c lip relocalization

    Ronald Clark, Sen Wang, Andrew Markham, Niki Trigoni, and Hongkai Wen. Vidloc: A deep spatio-temporal model for 6-dof video-c lip relocalization. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , pages 6856–6864, 2017

  209. [217]

    Event-based camera pose tracking using a gener ative event model

    Guillermo Gallego, Christian Forster, Elias Mueggle r, and Davide Scaramuzza. Event-based camera pose tracking using a gener ative event model. arXiv preprint arXiv:1510.01972 , 2015

  210. [218]

    Neuromorphic event-based 3d pose estimation

    David Reverter V aleiras, Garrick Orchard, Sio-Hoi Ie ng, and Ryad B Benosman. Neuromorphic event-based 3d pose estimation. Frontiers in neuroscience, 9:522, 2016

  211. [219]

    Bowman, Nikolay Atanasov, Kostas Daniilidis, and George J

    Sean L. Bowman, Nikolay Atanasov, Kostas Daniilidis, and George J. Pappas. Probabilistic data association for semantic slam. In IEEE International Conference on Robotics & Automation , 2017

  212. [220]

    Lightweight unsuperv ised deep loop closure

    Nate Merrill and Guoquan Huang. Lightweight unsuperv ised deep loop closure. arXiv preprint arXiv:1805.07703 , 2018

  213. [221]

    Long -term visual localization using semantically segmented images

    Erik Stenborg, Carl Toft, and Lars Hammarstrand. Long -term visual localization using semantically segmented images. pages 6 484–6490, 2018

  214. [222]

    X-view: Graph-based semantic multi-view localiza tion

    Abel Gawel, Carlo Del Don, Roland Siegwart, Juan Nieto , and Cesar Cadena. X-view: Graph-based semantic multi-view localiza tion. IEEE Robotics and Automation Letters , 3(3):1687–1694, 2018

  215. [223]

    Multi modal seman- tic slam with probabilistic data association

    Kevin Doherty, Dehann Fourie, and John Leonard. Multi modal seman- tic slam with probabilistic data association. In 2019 IEEE International Conference on Robotics and Automation (ICRA). IEEE , 2019

  216. [224]

    Low-latency localization by active led markers tracking using a dynamic vision sensor

    Andrea Censi, Jonas Strubel, Christian Brandli, Tobi Delbruck, and Davide Scaramuzza. Low-latency localization by active led markers tracking using a dynamic vision sensor. In 2013 IEEE/RSJ Interna- tional Conference on Intelligent Robots and Systems , pages 891–898. IEEE, 2013

  217. [225]

    Robust monocular slam in dynamic environments

    Wei Tan, Haomin Liu, Zilong Dong, Guofeng Zhang, and Hu jun Bao. Robust monocular slam in dynamic environments. In 2013 IEEE International Symposium on Mixed and Augmented Reality (IS MAR), pages 209–218. IEEE, 2013

  218. [226]

    Ds-slam: A semantic visual slam towards dynamic en vi- ronments

    Chao Y u, Zuxin Liu, Xin-Jun Liu, Fugui Xie, Yi Y ang, Qi W ei, and Qiao Fei. Ds-slam: A semantic visual slam towards dynamic en vi- ronments. In 2018 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) , pages 1168–1174. IEEE, 2018

  219. [227]

    Mask r-cnn for object detection and in stance segmen- tation on keras and tensorflow, 2017

    Waleed Abdulla. Mask r-cnn for object detection and in stance segmen- tation on keras and tensorflow, 2017

  220. [228]

    Co-fusion: Real-ti me segmentation, tracking and fusion of multiple objects

    Martin R¨ unz and Lourdes Agapito. Co-fusion: Real-ti me segmentation, tracking and fusion of multiple objects. In 2017 IEEE International Conference on Robotics and Automation (ICRA) , pages 4471–4478. IEEE, 2017

  221. [229]

    Detect-slam: Making object detection and slam mutual ly beneficial

    Fangwei Zhong, Wang Sheng, Ziqi Zhang, China Chen, and Yizhou Wang. Detect-slam: Making object detection and slam mutual ly beneficial. In IEEE Winter Conference on Applications of Computer Vision, 2018

  222. [230]

    Dynaslam: Tracking, mapping, and inpainting in dynamic scenes

    Berta Bescos, Jos´ e M F´ acil, Javier Civera, and Jos´ e Neira. Dynaslam: Tracking, mapping, and inpainting in dynamic scenes. IEEE Robotics and Automation Letters , 3(4):4076–4083, 2018

  223. [231]

    Staticfusion: Background reconstruction for dense rgb-d slam in dynamic environments

    Raluca Scona, Mariano Jaimez, Yvan R Petillot, Mauric e Fallon, and Daniel Cremers. Staticfusion: Background reconstruction for dense rgb-d slam in dynamic environments. In 2018 IEEE International Conference on Robotics and Automation (ICRA) , pages 1–9. IEEE, 2018

  224. [232]

    A computationally efficient semantic slam solution for dynamic scenes

    Zemin Wang, Qian Zhang, Jiansheng Li, Shuming Zhang, a nd Jingbin Liu. A computationally efficient semantic slam solution for dynamic scenes. Remote Sensing , 11(11):1363, 2019

  225. [233]

    Dynamic-slam: Semantic monocular visual localizatio n and mapping based on deep learning in dynamic environment

    Linhui Xiao, Jinge Wang, Xiaosong Qiu, Zheng Rong, and Xudong Zou. Dynamic-slam: Semantic monocular visual localizatio n and mapping based on deep learning in dynamic environment. Robotics and Autonomous Systems , 117:1–16, 2019

  226. [234]

    Robust dense mapping for large-scale dynamic environments

    Ioan Andrei Bˆ arsan, Peidong Liu, Marc Pollefeys, and Andreas Geiger. Robust dense mapping for large-scale dynamic environments . In 2018 IEEE International Conference on Robotics and Automation ( ICRA), pages 7510–7517. IEEE, 2018

  227. [235]

    Simvodis: Sim ultane- ous visual odometry, object detection, and instance segmen tation

    Se-Ho Kim Ue-Hwan Kim and Jong-Hwan Kim. Simvodis: Sim ultane- ous visual odometry, object detection, and instance segmen tation. IEEE Transactions on Pattern Analysis and Machine Intelligence , Under Review, 2019

  228. [236]

    Simultaneous locali zation and mapping in the epoch of semantics: A survey

    Muhammad Sualeh and Gon-Woo Kim. Simultaneous locali zation and mapping in the epoch of semantics: A survey. International Journal of Control, Automation and Systems , 17(3):729–742, 2019

  229. [237]

    Dels-3d: Deep localization and segmentation with a 3d seman tic map

    Peng Wang, Ruigang Y ang, Binbin Cao, Wei Xu, and Y uanqi ng Lin. Dels-3d: Deep localization and segmentation with a 3d seman tic map. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pages 5860–5869, 2018

  230. [238]

    A unifying view of geometry, semantics, and data ass ociation in slam

    Nikolay Atanasov, Sean L Bowman, Kostas Daniilidis, a nd George J Pappas. A unifying view of geometry, semantics, and data ass ociation in slam. In IJCAI, pages 5204–5208, 2018

  231. [239]

    Pl-slam: a stereo slam system through the combination of points and line segme nts

    Ruben Gomez-Ojeda, Francisco-Angel Moreno, David Zu ˜ niga-No¨ el, Davide Scaramuzza, and Javier Gonzalez-Jimenez. Pl-slam: a stereo slam system through the combination of points and line segme nts. IEEE Transactions on Robotics , 2019

  232. [240]

    Structslam: Visual slam with building struc ture lines

    Huizhong Zhou, Danping Zou, Ling Pei, Rendong Ying, Pe ilin Liu, and Wenxian Y u. Structslam: Visual slam with building struc ture lines. IEEE Transactions on V ehicular Technology, 64(4):1364–1375, 2015

  233. [241]

    Slam for dummies : A tutorial approach to simultaneous localization and mapping

    Sren Riisgaard and Morten Rufus Blas. Slam for dummies : A tutorial approach to simultaneous localization and mapping. Techni cal report, 2005

  234. [242]

    Extending kalibr: Calibrating the ext rinsics of multiple imus and of individual axes

    Joern Rehder, Janosch Nikolic, Thomas Schneider, Tim o Hinzmann, and Roland Siegwart. Extending kalibr: Calibrating the ext rinsics of multiple imus and of individual axes. In 2016 IEEE International Conference on Robotics and Automation (ICRA) , pages 4304–4311. IEEE, 2016

  235. [243]

    Multi-camera visual-inertial navigation with online intr insic and ex- trinsic calibration

    Kevin Eckenhoff, Patrick Geneva, Jesse Bloecker, and Guoquan Huang. Multi-camera visual-inertial navigation with online intr insic and ex- trinsic calibration. 2019 International Conference on Robotics and Automation (ICRA) , pages 3158–3164, 2019

  236. [244]

    Tedaldi, A

    A. Tedaldi, A. Pretto, and E. Menegatti. A robust and ea sy to implement method for imu calibration without external equipments. In Proc. of: IEEE International Conference on Robotics and Automation ( ICRA), pages 3042–3049, 2014. 16

  237. [245]

    Pretto and G

    A. Pretto and G. Grisetti. Calibration and performanc e evaluation of low-cost imus. In Proc. of: 20th IMEKO TC4 International Symposium , pages 429–434, 2014

  238. [246]

    High-fidelity sensor modeling and self-calibration in visi on-aided in- ertial navigation

    Mingyang Li, Hongsheng Y u, Xing Zheng, and Anastasios I Mourikis. High-fidelity sensor modeling and self-calibration in visi on-aided in- ertial navigation. In 2014 IEEE International Conference on Robotics and Automation (ICRA) , pages 409–416. IEEE, 2014

  239. [247]

    Selective sensor fusi on for neural visual-inertial odometry

    Changhao Chen, Stefano Rosa, Yishu Miao, Chris Xiaoxu an Lu, Wei Wu, Andrew Markham, and Niki Trigoni. Selective sensor fusi on for neural visual-inertial odometry. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , pages 10542–10551, 2019

  240. [248]

    B ad slam: Bundle adjusted direct rgb-d slam

    Thomas Schops, Torsten Sattler, and Marc Pollefeys. B ad slam: Bundle adjusted direct rgb-d slam. In The IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , June 2019

  241. [249]

    Multi- camera tracking and mapping for unmanned aerial vehicles in unstru ctured environments

    Adam Harmat, Michael Trentini, and Inna Sharf. Multi- camera tracking and mapping for unmanned aerial vehicles in unstru ctured environments. Journal of Intelligent & Robotic Systems , 78(2):291– 317, 2015

  242. [250]

    MultiCol-SLAM - a modul ar real-time multi-camera slam system

    Steffen Urban and Stefan Hinz. MultiCol-SLAM - a modul ar real-time multi-camera slam system. arXiv preprint arXiv:1610.07336 , 2016

  243. [251]

    Iterated extended kalman filter based visu al-inertial odometry using direct photometric feedback

    Michael Bloesch, Michael Burri, Sammy Omari, Marco Hu tter, and Roland Siegwart. Iterated extended kalman filter based visu al-inertial odometry using direct photometric feedback. The International Journal of Robotics Research , 36(10):1053–1072, 2017

  244. [252]

    Tightly coupled 3d lidar inertial odometry and mapping

    Haoyang Y e, Y uying Chen, and Ming Liu. Tightly coupled 3d lidar inertial odometry and mapping. arXiv preprint arXiv:1904.06993 , 2019

  245. [253]

    Extrinsic calibration of 2d laser rangefinders using a n existing cuboid-shaped corridor as the reference

    Deyu Yin, Jingbin Liu, Teng Wu, Keke Liu, Juha Hyypp¨ a, and Ruizhi Chen. Extrinsic calibration of 2d laser rangefinders using a n existing cuboid-shaped corridor as the reference. Sensors, 18(12):4371, 2018

  246. [254]

    Extr insic calibration of 2d laser rangefinders based on a mobile sphere

    Shoubin Chen, Jingbin Liu, Teng Wu, Wenchao Huang, Kek e Liu, Deyu Yin, Xinlian Liang, Juha Hyypp¨ a, and Ruizhi Chen. Extr insic calibration of 2d laser rangefinders based on a mobile sphere . Remote Sensing, 10(8):1176, 2018

  247. [255]

    Automatic laser calibration, mapping, and local- ization for autonomous vehicles

    Jesse Sol Levinson. Automatic laser calibration, mapping, and local- ization for autonomous vehicles . Stanford University, 2011

  248. [256]

    Automatic online calibration of cameras and lasers

    Jesse Levinson and Sebastian Thrun. Automatic online calibration of cameras and lasers. In Robotics: Science and Systems , volume 2, 2013

  249. [257]

    Dhall, K

    A. Dhall, K. Chelani, V . Radhakrishnan, and K. M. Krish na. LiDAR- Camera Calibration using 3D-3D Point correspondences. ArXiv e- prints, May 2017

  250. [258]

    Regnet: Multimodal sensor registration using deep neural n etworks

    Nick Schneider, Florian Piewak, Christoph Stiller, a nd Uwe Franke. Regnet: Multimodal sensor registration using deep neural n etworks. In 2017 IEEE intelligent vehicles symposium (IV) , pages 1803–1810. IEEE, 2017

  251. [259]

    Limo: Lidar-monocular visual odometry

    Johannes Graeter, Alexander Wilczynski, and Martin L auer. Limo: Lidar-monocular visual odometry. 2018

  252. [260]

    Calib- net: self-supervised extrinsic calibration using 3d spati al transformer networks

    Ganesh Iyer, J Krishna Murthy, K Madhava Krishna, et al . Calib- net: self-supervised extrinsic calibration using 3d spati al transformer networks. arXiv preprint arXiv:1803.08181 , 2018

  253. [261]

    3d lidar–camera intrinsic and extrinsic calibration: Iden tifiability and analytical least-squares-based initialization

    Faraz M Mirzaei, Dimitrios G Kottas, and Stergios I Rou meliotis. 3d lidar–camera intrinsic and extrinsic calibration: Iden tifiability and analytical least-squares-based initialization. The International Journal of Robotics Research , 31(4):452–467, 2012

  254. [262]

    Lidar and camera calibration using motions estimated by sensor fusio n odometry

    Ryoichi Ishikawa, Takeshi Oishi, and Katsushi Ikeuch i. Lidar and camera calibration using motions estimated by sensor fusio n odometry. In 2018 IEEE/RSJ International Conference on Intelligent Rob ots and Systems (IROS) , pages 7342–7349. IEEE, 2018

  255. [263]

    Automatic online calibration of cameras and lasers

    Jesse Levinson and Sebastian Thrun. Automatic online calibration of cameras and lasers. In Robotics: Science and Systems , 2013

  256. [264]

    Svin2: An underwater slam system using sonar, visual, inertial, an d depth sensor

    Sharmin Rahman, Alberto Quattrini Li, and Ioannis Rek leitis. Svin2: An underwater slam system using sonar, visual, inertial, an d depth sensor

  257. [265]

    Environment drive n under- water camera-imu calibration for monocular visual-inerti al slam

    Changjun Gu, Y ang Cong, and Gan Sun. Environment drive n under- water camera-imu calibration for monocular visual-inerti al slam. 2019 International Conference on Robotics and Automation (ICRA ), pages 2405–2411, 2019

  258. [266]

    Improving underwater obstacle detection using s emantic image segmentation

    Bilal Arain, Chris McCool, Paul Rigby, Daniel Cagara, and Matthew Dunbabin. Improving underwater obstacle detection using s emantic image segmentation. 2019 International Conference on Robotics and Automation (ICRA) , pages 9271–9277, 2019

  259. [267]

    Wifi-sla m using gaussian process latent variable models

    Brian Ferris, Dieter Fox, and Neil D Lawrence. Wifi-sla m using gaussian process latent variable models. In IJCAI, volume 7, pages 2480–2485, 2007

  260. [268]

    Leveraging mmwave imaging and communications for simu l- taneous localization and mapping

    Mohammed Aladsani, Ahmed Alkhateeb, and Georgios C Tr ichopou- los. Leveraging mmwave imaging and communications for simu l- taneous localization and mapping. In ICASSP 2019-2019 IEEE International Conference on Acoustics, Speech and Signal P rocessing (ICASSP), pages 4539–454...

  261. [269]

    Position locat ioning for millimeter wave systems

    Ojas Kanhere and Theodore S Rappaport. Position locat ioning for millimeter wave systems. In 2018 IEEE Global Communications Conference (GLOBECOM) , pages 206–212. IEEE, 2018

  262. [270]

    Wireless communications and applicat ions above 100 ghz: Opportunities and challenges for 6g and beyond

    Theodore S Rappaport, Y unchou Xing, Ojas Kanhere, Shi hao Ju, Arjuna Madanayake, Soumyajit Mandal, Ahmed Alkhateeb, and Geor- gios C Trichopoulos. Wireless communications and applicat ions above 100 ghz: Opportunities and challenges for 6g and beyond. IEEE Access, 7:78729–78757, 2019

  263. [271]

    Keyframe- based direct thermalinertial odometry

    Shehryar Khattak, Christos Papachristos, and Kostas Alexis. Keyframe- based direct thermalinertial odometry. 2019 International Conference on Robotics and Automation (ICRA) , pages 3563–3569, 2019

  264. [272]

    Image guided depth upsampling using anisotr opic total generalized variation

    David Ferstl, Christian Reinbacher, Rene Ranftl, Mat thias R¨ uther, and Horst Bischof. Image guided depth upsampling using anisotr opic total generalized variation. In Proceedings of the IEEE International Conference on Computer Vision , pages 993–1000, 2013

  265. [273]

    In defen se of classical image processing: Fast depth completion on the cpu

    Jason Ku, Ali Harakeh, and Steven L Waslander. In defen se of classical image processing: Fast depth completion on the cpu. In 2018 15th Conference on Computer and Robot Vision (CRV) , pages 16–22. IEEE, 2018

  266. [274]

    Sparse-to-dense: D epth prediction from sparse depth samples and a single image

    Fangchang Mal and Sertac Karaman. Sparse-to-dense: D epth prediction from sparse depth samples and a single image. In 2018 IEEE International Conference on Robotics and Automation (ICRA ), pages 1–8. IEEE, 2018

  267. [275]

    Sparsity invariant cnns

    Jonas Uhrig, Nick Schneider, Lukas Schneider, Uwe Fra nke, Thomas Brox, and Andreas Geiger. Sparsity invariant cnns. In 2017 Interna- tional Conference on 3D Vision (3DV) , pages 11–20. IEEE, 2017

  268. [276]

    Dfusenet: Deep fusion of rgb and sparse depth informati on for im- age guided dense depth completion

    Shreyas S Shivakumar, Ty Nguyen, Steven W Chen, and Cam illo J Tay- lor. Dfusenet: Deep fusion of rgb and sparse depth informati on for im- age guided dense depth completion. arXiv preprint arXiv:1902.00761 , 2019

  269. [277]

    Estimating depth from rgb and sparse sensing

    Zhao Chen, Vijay Badrinarayanan, Gilad Drozdov, and A ndrew Rabi- novich. Estimating depth from rgb and sparse sensing. In Proceedings of the European Conference on Computer Vision (ECCV) , pages 167– 182, 2018

  270. [278]

    Propagating confidences through cnns for sparse data regres sion

    Abdelrahman Eldesokey, Michael Felsberg, and Fahad S hahbaz Khan. Propagating confidences through cnns for sparse data regres sion. arXiv preprint arXiv:1805.11913, 2018

  271. [279]

    Lic-fusion: Lidar-inertial-camera odometry

    Xingxing Zuo, Patrick Geneva, Woosik Lee, Y ong Liu, an d Guoquan Huang. Lic-fusion: Lidar-inertial-camera odometry. arXiv preprint arXiv:1909.04102, 2019

  272. [280]

    Intersection safety using lidar and stereo vision s ensors

    Olivier Aycard, Qadeer Baig, Siviu Bota, Fawzi Nashas hibi, Sergiu Nedevschi, Cosmin Pantilie, Michel Parent, Paulo Resende, and Trung- Dung Vu. Intersection safety using lidar and stereo vision s ensors. In 2011 IEEE Intelligent V ehicles Symposium (IV) , pages 863–869. IEEE, 2011

  273. [281]

    Multi ple sensor fusion and classification for moving object detection and tr acking

    Ricardo Omar Chavez-Garcia and Olivier Aycard. Multi ple sensor fusion and classification for moving object detection and tr acking. IEEE Transactions on Intelligent Transportation Systems , 17(2):525– 534, 2015

  274. [282]

    A multi-sensor fusion system for movin g object detection and tracking in urban driving environments

    Hyunggi Cho, Y oung-Woo Seo, BVK Vijaya Kumar, and Ragu - nathan Raj Rajkumar. A multi-sensor fusion system for movin g object detection and tracking in urban driving environments. In 2014 IEEE International Conference on Robotics and Automation (ICRA ), pages 1836–1843. IEEE, 2014

  275. [283]

    Int egrating millimeter wave radar with a monocular vision sensor for on- road obstacle detection applications

    Tao Wang, Nanning Zheng, Jingmin Xin, and Zheng Ma. Int egrating millimeter wave radar with a monocular vision sensor for on- road obstacle detection applications. Sensors, 11(9):8992–9008, 2011

  276. [284]

    Robust and precise vehicle localizati on based on multi-sensor fusion in diverse city scenes

    Guowei Wan, Xiaolong Y ang, Renlan Cai, Hao Li, Y ao Zhou , Hao Wang, and Shiyu Song. Robust and precise vehicle localizati on based on multi-sensor fusion in diverse city scenes. In 2018 IEEE International Conference on Robotics and Automation (ICRA ), pages 4670–4677. IEEE, 2018

  277. [285]

    Real-time d epth enhanced monocular odometry

    Ji Zhang, Michael Kaess, and Sanjiv Singh. Real-time d epth enhanced monocular odometry. In 2014 IEEE/RSJ International Conference on Intelligent Robots and Systems , pages 4973–4980. IEEE, 2014

  278. [286]

    Visual-lidar odometry and m apping: Low- drift, robust, and fast

    Ji Zhang and Sanjiv Singh. Visual-lidar odometry and m apping: Low- drift, robust, and fast. In 2015 IEEE International Conference on Robotics and Automation (ICRA) , pages 2174–2181. IEEE, 2015

  279. [287]

    Visual-LiDAR SLAM with loop closure

    Y oshua Nava. Visual-LiDAR SLAM with loop closure . PhD thesis, Masters thesis, KTH Royal Institute of Technology, 2018. 17

  280. [288]

    Slam of robo t based on the fusion of vision and lidar

    Yinglei Xu, Y ongsheng Ou, and Tiantian Xu. Slam of robo t based on the fusion of vision and lidar. In 2018 IEEE International Conference on Cyborg and Bionic Systems (CBS) , pages 121–126. IEEE, 2018

  281. [289]

    Stereo visual inertial lidar simultaneous localization an d mapping

    Weizhao Shao, Srinivasan Vijayarangan, Cong Li, and G eorge Kantor. Stereo visual inertial lidar simultaneous localization an d mapping. arXiv preprint arXiv:1902.10741 , 2019

  282. [290]

    Lidar -aided cam- era feature tracking and visual slam for spacecraft low-orb it navigation and planetary landing

    Franz Andert, Nikolaus Ammann, and Bolko Maass. Lidar -aided cam- era feature tracking and visual slam for spacecraft low-orb it navigation and planetary landing. In Advances in Aerospace Guidance, Navigation and Control, pages 605–623. Springer, 2015

  283. [291]

    Pointf usion: Deep sensor fusion for 3d bounding box estimation

    Danfei Xu, Dragomir Anguelov, and Ashesh Jain. Pointf usion: Deep sensor fusion for 3d bounding box estimation. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , pages 244–253, 2018

  284. [292]

    Roar- net: A robust 3d object detection based on region approximat ion refinement

    Kiwoo Shin, Y oungwook Paul Kwon, and Masayoshi Tomizu ka. Roar- net: A robust 3d object detection based on region approximat ion refinement. arXiv preprint arXiv:1811.03818 , 2018

  285. [293]

    Joint 3d proposal generation and object detecti on from view aggregation

    Jason Ku, Melissa Mozifian, Jungwook Lee, Ali Harakeh, and Steven Waslander. Joint 3d proposal generation and object detecti on from view aggregation. IROS, 2018

  286. [294]

    Fusenet: Incorporating depth into semantic segmentation v ia fusion- based cnn architecture

    Caner Hazirbas, Lingni Ma, Csaba Domokos, and Daniel C remers. Fusenet: Incorporating depth into semantic segmentation v ia fusion- based cnn architecture. In Asian conference on computer vision , pages 213–228. Springer, 2016

  287. [295]

    Fusing birds eye view lidar point cloud and front view camera image for 3d obje ct detection

    Zining Wang, Wei Zhan, and Masayoshi Tomizuka. Fusing birds eye view lidar point cloud and front view camera image for 3d obje ct detection. In 2018 IEEE Intelligent V ehicles Symposium (IV) , pages 1–6. IEEE, 2018

  288. [296]

    Deep continuous fusion for multi-sensor 3d object detection

    Ming Liang, Bin Y ang, Shenlong Wang, and Raquel Urtasu n. Deep continuous fusion for multi-sensor 3d object detection. In Proceedings of the European Conference on Computer Vision (ECCV) , pages 641– 656, 2018

  289. [297]

    3-d lidar+ monocular camera: An inverse-depth-induc ed fusion framework for urban road detection

    Shuo Gu, Tao Lu, Yigong Zhang, Jose M Alvarez, Jian Y ang , and Hui Kong. 3-d lidar+ monocular camera: An inverse-depth-induc ed fusion framework for urban road detection. IEEE Transactions on Intelligent V ehicles, 3(3):351–360, 2018

  290. [298]

    Past, present, and future of simultaneous localization and mapping: Towar d the robust-perception age

    Cesar Cadena, Luca Carlone, Henry Carrillo, Y asir Lat if, Davide Scaramuzza, Jos´ e Neira, Ian Reid, and John J Leonard. Past, present, and future of simultaneous localization and mapping: Towar d the robust-perception age. IEEE Transactions on robotics , 32(6):1309– 1332, 2016

  291. [299]

    Data- efficient decentralized visual slam

    Titus Cieslewski, Siddharth Choudhary, and Davide Sc aramuzza. Data- efficient decentralized visual slam. In 2018 IEEE International Con- ference on Robotics and Automation (ICRA) , pages 2466–2473. IEEE, 2018

  292. [300]

    Differential privacy

    Cynthia Dwork. Differential privacy. Encyclopedia of Cryptography and Security , pages 338–340, 2011

Pith tools

Reviewed August 14, 2026 · model on record in the stance chip above.