AORQ Press
Interdisciplinary Intelligence and Emerging Technologies

From Range Images to Global Descriptors: An Evidence Synthesis of Efficient LiDAR Place Recognition

Read & download PDF
Abstract

LiDAR place recognition retrieves previously mapped locations from current three-dimensional scans and supports loop closure, global localization, and recovery after tracking failure. Its central engineering problem is to compress a large, irregular scan into a descriptor that is discriminative, robust to viewpoint and environmental change, and cheap enough for continuous search. This review examines range-image learning as a response to that problem. It situates RangePlace within the progression from handcrafted polar descriptors and unordered point encoders to sparse convolutional and transformer-based systems. RangePlace is especially informative because it combines a hierarchical range-image transformer, shifted-window attention, depth-wise convolution, and multiscale feature mixing. That design offers a coherent account of how global context and local geometry can be modeled without the quadratic cost of full-resolution attention. The available evidence across KITTI, NCLT, Ford Campus, and nuScenes suggests that the method is competitive, efficient, robust to yaw variation, and capable of cross-domain generalization. However, benchmark recall alone cannot establish deployment readiness. Range projection creates sensor-dependent sampling, dynamic objects can dominate descriptors, and common evaluation protocols may conceal temporal leakage, ambiguous ground truth, or weak geometric verification. A stronger evidence framework should report retrieval accuracy, calibration, latency, memory, cross-sensor transfer, adverse-condition robustness, and end-to-end loop-closure consequences. The review concludes that range images are not merely a convenient conversion of point clouds into pictures. They are a useful computational interface whose value depends on explicit treatment of circular geometry, scale, sensor configuration, and verification.

Keywords
LiDAR place recognitionrange imagesglobal descriptorshierarchical transformersloop closurecross-domain generalizationautonomous localization
References
  1. Li, J., Liu, Q., Wang, B., Liu, H., & Han, Y. (2024). RangePlace: A hierarchical range image transformer for LiDAR-based place recognition. IEEE Transactions on Intelligent Vehicles.
  2. Caesar, H., Bankiti, V., Lang, A. H., Vora, S., Liong, V. E., Xu, Q., ... & Beijbom, O. (2020). nuScenes: A multimodal dataset for autonomous driving. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (pp. 11621-11631).
  3. Carlevaris-Bianco, N., Ushani, A. K., & Eustice, R. M. (2016). University of Michigan North Campus long-term vision and lidar dataset. The International Journal of Robotics Research, 35(9), 1023-1035.
  4. Dosovitskiy, A., Beyer, L., Kolesnikov, A., Weissenborn, D., Zhai, X., Unterthiner, T., ... & Houlsby, N. (2021). An image is worth 16x16 words: Transformers for image recognition at scale. In International Conference on Learning Representations.
  5. Geiger, A., Lenz, P., Stiller, C., & Urtasun, R. (2013). Vision meets robotics: The KITTI dataset. The International Journal of Robotics Research, 32(11), 1231-1237.
  6. Jegou, H., Douze, M., & Schmid, C. (2011). Product quantization for nearest neighbor search. IEEE Transactions on Pattern Analysis and Machine Intelligence, 33(1), 117-128.
  7. Kim, G., & Kim, A. (2018). Scan Context: Egocentric spatial descriptor for place recognition within 3D point cloud map. In 2018 IEEE/RSJ International Conference on Intelligent Robots and Systems (pp. 4802-4809). IEEE.
  8. Knights, J., Vidanapathirana, K., Ramezani, M., Sridharan, S., Fookes, C., & Moghadam, P. (2023). Wild-Places: A large-scale dataset for LiDAR place recognition in unstructured natural environments. In 2023 IEEE International Conference on Robotics and Automation (pp. 11322-11328). IEEE.
  9. Komorowski, J. (2021). MinkLoc3D: Point cloud based large-scale place recognition. In Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision (pp. 1790-1799).
  10. Liu, Z., Lin, Y., Cao, Y., Hu, H., Wei, Y., Zhang, Z., Lin, S., & Guo, B. (2021). Swin Transformer: Hierarchical vision transformer using shifted windows. In Proceedings of the IEEE/CVF International Conference on Computer Vision (pp. 10012-10022).
  11. Ma, J., Zhang, J., Xu, J., Ai, R., Gu, W., & Chen, X. (2022). OverlapTransformer: An efficient and yaw-angle-invariant transformer network for LiDAR-based place recognition. IEEE Robotics and Automation Letters, 7(3), 6958-6965.
  12. Pandey, G., McBride, J. R., & Eustice, R. M. (2011). Ford Campus vision and lidar data set. The International Journal of Robotics Research, 30(13), 1543-1552.
  13. Shan, T., Englot, B., Duarte, F., Ratti, C., & Rus, D. (2021). Robust place recognition using an imaging lidar. In 2021 IEEE International Conference on Robotics and Automation (pp. 5469-5475). IEEE.
  14. Uy, M. A., & Lee, G. H. (2018). PointNetVLAD: Deep point cloud based retrieval for large-scale place recognition. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (pp. 4470-4479).
Publication details
Journal
Interdisciplinary Intelligence and Emerging Technologies
Volume
1 (2026)
Article number
ajg20260010
License
CC BY 4.0