Papers
arxiv:2608.02437

InfiniSplat: Implicit Gaussian Decoding for Large-Baseline Monocular View Synthesis

Published on Aug 3
ยท Submitted by
PLUS-WAVE
on Aug 4
Authors:
,
,
,
,
,
,
,
,
,

Abstract

Single-image feed-forward 3D Gaussian Splatting (3DGS) aims to directly generate a renderable 3D scene representation from one input image, avoiding the cost of multi-view capture and per-scene optimization. However, existing methods are often constrained by a pixel-aligned representation, where Gaussians are predicted from fixed image-grid locations. Such pixel-aligned primitives can produce promising nearby-view renderings, but they remain weakly coupled to underlying scene surfaces and struggle to preserve coherent structures under large viewpoint shifts. We present InfiniSplat, a feed-forward single-image 3DGS framework that moves from a pixel-aligned representation toward a surface-aligned representation. InfiniSplat constructs this representation by first using geometry-guided sampling to place 2D supports according to depth-induced local surface structure, and then applying a query-conditioned implicit decoder to predict Gaussian attributes from the image features queried at these supports.By grounding support locations in geometry while decoupling Gaussian prediction from fixed pixel centers, InfiniSplat produces Gaussian layouts that better follow scene surfaces and reduce scattered primitives caused by grid discretization.Across multiple cross-dataset NVS evaluations, InfiniSplat achieves state-of-the-art performance compared with single-image feed-forward baselines, and demonstrates zero-shot generalization from Hypersim indoor synthetic training to complex open-world scenes.Project page: https://zju3dv.github.io/InfiniSplat.

Community

Paper submitter

We are excited to share InfiniSplat, accepted to ACM Transactions on Graphics (SIGGRAPH Asia 2026 Journal Track).

InfiniSplat introduces a surface-aligned representation for feed-forward single-image 3D Gaussian Splatting, moving beyond pixel-aligned Gaussian prediction with geometry-guided sampling and implicit Gaussian decoding. It enables more coherent 3D structures and improved large-baseline novel view synthesis from a single image.

๐Ÿ“„ Paper: https://arxiv.org/abs/2608.02437
๐ŸŒ Project page: https://zju3dv.github.io/InfiniSplat
๐Ÿš€ Code: https://github.com/zju3dv/InfiniSplat
๐Ÿค— Demo: https://huggingface.co/spaces/PLUS-WAVE/InfiniSplat

This is an automated message from the Librarian Bot. I found the following papers similar to this paper.

The following papers were recommended by the Semantic Scholar API

Please give a thumbs up to this comment if you found it helpful!

If you want recommendations for any Paper on Hugging Face checkout this Space

You can directly ask Librarian Bot for paper recommendations by tagging it in a comment: @librarian-bot recommend

Sign up or log in to comment

Get this paper in your agent:

hf papers read 2608.02437
Don't have the latest CLI?
curl -LsSf https://hf.co/cli/install.sh | bash

Models citing this paper 1

Datasets citing this paper 0

No dataset linking this paper

Cite arxiv.org/abs/2608.02437 in a dataset README.md to link it from this page.

Spaces citing this paper 2

Collections including this paper 0

No Collection including this paper

Add this paper to a collection to link it from this page.