MOGS: Monocular Object-guided Gaussian Splatting in Large Scenes

Zhang, Shengkai and Liu, Yuhe and He, Jianhua MOGS: Monocular Object-guided Gaussian Splatting in Large Scenes. In: IEEE International Conference on Robotics and Automation (ICRA), 2026-06-01 - 2026-06-05, Vienna, Austria. (In Press)

Abstract

Recent advances in 3D Gaussian Splatting (3DGS) deliver striking photorealism, and extending it to large scenes opens new opportunities for semantic reasoning and prediction in applications such as autonomous driving. Today’s state-of-theart systems for large scenes primarily originate from LiDARbased pipelines that utilize long-range depth sensing. However, they require costly high-channel sensors whose dense point clouds strain memory and computation, limiting scalability, fleet deployment, and optimization speed. We present MOGS, a monocular 3DGS framework that replaces active LiDAR depth with object-anchored, metrized dense depth derived from sparse visual-inertial (VI) structure-from-motion (SfM) cues. Our key idea is to exploit image semantics to hypothesize per-object shape priors, anchor them with sparse but metrically reliable SfM points, and propagate the resulting metric constraints across each object to produce dense depth. To address two key challenges, i.e., insu cient SfM coverage within objects and cross-object geometric inconsistency, MOGS introduces 1) a multi-scale shape consensus module that adaptively merges small segments into coarse objects best supported by SfM and fits them with parametric shape models, and 2) a cross-object depth refinement module that optimizes per-pixel depth under a ombinatorial objective combining geometric consistency, prior anchoring, and edge-aware smoothness. Experiments on public datasets show that, with a low-cost VI sensor suite, MOGS reduces training time by up to 30.4% and memory consumption by 19.8%, while achieving high-quality rendering competitive with costly LiDARbased approaches in large scenes. The source code is publicly available at https://github.com/ClarenceZSK/MOGS/.

Item Metadata

Item Type:	Conference or Workshop Item (Paper)
Additional Information:	Published proceedings: _not provided_
Subjects:	Z Bibliography. Library Science. Information Resources > ZR Rights Retention
Divisions:	Faculty of Science and Health > Computer Science and Electronic Engineering, School of
SWORD Depositor:	Unnamed user with email elements@essex.ac.uk
Depositing User:	Unnamed user with email elements@essex.ac.uk
Date Deposited:	21 Apr 2026 12:00
Last Modified:	21 Apr 2026 12:00
URI:	http://repository.essex.ac.uk/id/eprint/42774

Available files

Accepted Version

Filename: final-MOGS.pdf

Licence: Creative Commons: Attribution 4.0

Download

MOGS: Monocular Object-guided Gaussian Splatting in Large Scenes

Abstract

Item Metadata

Share and export

Available files

Accepted Version

Statistics

Downloads