GLAM-SLAM: Real-time Gaussian Large-scale Mapping via Flow Densification and Spatial Decomposition
AuthorsPanagiotis Mermigkas, Argyris Manetas, Petros Maragos
Resources
GLAM-SLAM makes Gaussian-splatting-based robot mapping faster and more scalable for large outdoor environments.
Key results
Reported improvement over the second-best performer.
Average PSNR improvement over GigaSLAM before offline refinement on KITTI Odometry.
Frames per second achieved on KITTI Odometry.
Frames per second achieved on Oxford RobotCar.
Frames per second achieved on Málaga.
Frames processed without interruption on KITTI sequence 08.
What the paper found
GLAM-SLAM, developed by researchers at the Athena Research Center and the National Technical University of Athens, is a decoupled monocular mapping system that combines an ORB-SLAM2 CPU frontend for robust tracking with an asynchronous GPU backend for scalable 3D Gaussian Splatting. Its Flow-Guided Densification Module uses LiteFlowNet3 optical flow, epipolar-consistency filtering, and triangulation to populate sparse regions with geometrically reliable Gaussian anchors, while dynamic spatial decomposition assigns localized MLPs to different outdoor regions, reducing interference across changing illumination and appearance. Evaluated on KITTI Odometry, Oxford RobotCar, and Málaga, GLAM-SLAM reports a 15% reconstruction-quality improvement over the second-best method; against GigaSLAM before offline refinement, its average PSNR improves by 11.6% on KITTI. The system sustains 10 FPS on KITTI, 16 FPS on Oxford RobotCar, and 20 FPS on Málaga, using an NVIDIA RTX 5090 with 32 GB of VRAM. On KITTI, it averages 11.6 GiB of GPU memory and completes sequences of up to 4071 frames, whereas competing systems encounter out-of-memory failures on long sequences. Ablations show that flow densification and localized MLPs are complementary: together they increase representation density and photometric fidelity while maintaining real-time operation.
Original abstract
Existing Gaussian-splatting-based monocular Simultaneous Localization and Mapping (SLAM) systems are either tailored to short sequences, are not real-time, or suffer from prohibitive GPU memory requirements, limiting their applicability in realistic, long-horizon scenarios. To address this, we present GLAM-SLAM, a real-time, decoupled Gaussian-splatting SLAM system designed for large-scale outdoor scenes. We ensure lightweight tracking using a robust, feature-based SLAM frontend, while for mapping, we adopt a structured, sparse anchor grid representation that ensures scalable operation and maintains scene coherence across long-term sequences. To satisfy the dense initialization requirements of 3D Gaussian Splatting (3DGS), we introduce a geometry-based flow-densification anchoring strategy using epipolar constraints. Furthermore, by treating mapping as a multi-scene problem, we propose a scene-partitioning strategy that introduces a strong spatial inductive bias via MLP initializations to generate localized Gaussians. We evaluate our system on the challenging, long-sequence KITTI Odometry, Oxford RobotCar, and M'alaga datasets. Extensive ablations and comparisons demonstrate a 15% improvement in reconstruction quality over the second-best performer, while maintaining real-time performance and the ability to scale to longer sequences. Code is publicly available for the benefit of the community.
Read the original paperMore in Robotics
Browse all 50 papers →JAMB: Joint Action-Motion Diffusion for Bimanual Manipulation
Chuyang Xiao, Peilin Meng, David Held
JAMB helps two robot arms coordinate by jointly imagining their future movements and the changing 3D scene before acting.
Rolling-WAM: World Action Models with Rolling Imagination
Yinghua Zhou, Junjie Ye, Yiqi Zhao, Hao Dong, Celina Shiyu Wang, Ruohai Ge, Tingyi Yang, Basile Van Hoorick, Gaurav Sukhatme, Vitor Guizilini, Yue Wang
Rolling-WAM keeps future robot actions partially imagined and refined over time, making world-model-based manipulation replan 4.5 times faster.
Training-free Behavior Cloning
Maximilian Adang, Timothy Chen, Lars Osterberg, Aiden Swann, Mac Schwager
A fast, training-free robot controller reuses and corrects demonstration trajectories to deliver traceable behavior at real-time speeds.