VIM-GS: Visual-Inertial Monocular Gaussian Splatting via Object-level Guidance in Large Scenes

Shengkai Zhang

Yuhe Liu

Guanjun Wu

Jianhua He

Xinggang Wang

Mozi Chen

Kezhong Liu

VIM-GS is a Gaussian Splatting (GS) framework using monocular images for novel-view synthesis (NVS) in large scenes. GS typically requires accurate depth to initiate Gaussian ellipsoids using RGB-D/stereo cameras. Their limited depth sensing range makes it difficult for GS to work in large scenes. Monocular images, however, lack depth to guide the learning and lead to inferior NVS results. Although large foundation models (LFMs) for monocular depth estimation are available, they suffer from cross-frame inconsistency, inaccuracy for distant scenes, and ambiguity in deceptive texture cues. This paper aims to generate dense, accurate depth images from monocular RGB inputs for high-definite GS rendering. The key idea is to leverage the accurate but sparse depth from visual-inertial Structure-from-Motion (SfM) to refine the dense but coarse depth from LFMs. To bridge the sparse input and dense output, we propose an object-segmented depth propagation algorithm that renders the depth of pixels of structured objects. Then we develop a dynamic depth refinement module to handle the crippled SfM depth of dynamic objects and refine the coarse LFM depth. Experiments using public and customized datasets demonstrate the superior rendering quality of VIM-GS in large scenes.

PDF URL

Featured

Platforms

SplatCapture 3.0.0 Adds Spline Capture

SplatCapture 3.0.0 adds a spline capture mode, Unreal Engine 5.8 support, and square-aspect camera rigs that give Gaussian splatting trainers clean intrinsics.

Michael Rubloff

Jul 17, 2026

Platforms

Marble x Nuke Loads World Labs Marble Splats Straight Into Nuke's 3D Viewport

Marble x Nuke is a free Nuke 17 toolset that turns World Labs Marble text or image prompts into Gaussian splats loaded straight into Nuke's native 3D viewport.

Michael Rubloff

Jul 15, 2026

Platforms

Houdini 22 Ships, Making Native Gaussian Splats Generally Available

Houdini 22 is now generally available, making SideFX's native Gaussian splatting pipeline, PDG training and Copernicus compositing, production ready today.

Michael Rubloff

Jul 15, 2026

360 Gaussian v1.4.5 Adds Per-Clip Extraction Settings

360 Gaussian v1.4.5 updates the 360 video to 3DGS tool with per-clip extraction settings and fixes swapped GPS coordinates for accurate scale.

Michael Rubloff

Jul 15, 2026