SplatsThe evolution of media, in brief
RSS

One exposure in, a whole 3D scene out

Three columns: the raw compressive measurement as an unreadable scatter of speckle, the method's reconstruction of a plate of hot dogs and a vending machine, and the ground truth beside it

Figure: Yang et al., Westlake University AGI Lab · Hardware

Amara Osei

Amara Osei

Aug 15, 2026, 8:20 AM ET-Hardware

Yanming Yang, Chenxi Song and colleagues at Westlake University's AGI Lab have built GS²CI, which reconstructs a 3D gaussian splatting scene from a single snapshot compressive imaging measurement by leaning on the priors inside large vision foundation models.

Why it matters: Snapshot compressive imaging is a camera trick with a real payoff: modulate the incoming light with a set of masks during one exposure, and many temporal frames land encoded in a single 2D readout. It is how you get high-speed capture without a high-speed sensor, and without the data rate that comes with one.

If a scene moves relative to the camera during that exposure, the single measurement also contains multiple viewpoints. Which means one frame, in principle, holds a 3D scene.

The problem: In practice that measurement is a bad citizen. Compression discards fine detail, viewpoint diversity is thin, pose estimation is badly ill-posed, and the optimiser has to solve for the 3D representation and the camera poses at the same time.

Earlier attempts leaned on NeRF and needed tens of thousands of iterations per scene. The supervision signal is ambiguous enough that gaussian splatting optimisation goes unstable — gaussians inflate their opacity to paper over the loss rather than actually fitting the scene.

Zoom in:

  • Initialisation comes from a 3D vision foundation model run on the measurement, followed by SCI-aware gaussian optimisation.
  • After the coarse stage converges, a 2D foundation model supplies pseudo-view supervision at synthesised viewpoints to sharpen local appearance.
  • Opacity-Guided Splitting and Growth Regulation is the stability fix: it uses local opacity statistics to pick split candidates, penalises mean-opacity inflation, and caps how far the representation can grow.
  • Code is published at github.com/Westlake-AGI-Lab/GS2CI.

The tell: The comparison figure is the argument. The raw measurement is an unreadable scatter of noise; the reconstruction beside it is a plate of hot dogs and a vending machine, close enough to ground truth that you have to hunt for the difference.

The authors report best or second-best results on nearly every scene and metric across six scenes, with the clearest gains in perceptual quality rather than raw PSNR.

What's next: The dependence on foundation-model priors is the thing to watch. Performance is partly inherited from whatever those models already know about the world, which is a strength on ordinary scenes and an open question on anything unusual.

But the direction is clear enough: capture hardware gets cheaper and dumber, and the reconstruction absorbs the difficulty.

Go deeper:

  • GS²CI: Robust Gaussian Splatting For Snapshot Compressive Imaging on arXiv
⟵ Back to the brief

More stories

The SuperSplat 3.3.0 editor with a Gaussian splat capture of a street cafe loaded — furled yellow umbrellas over metal tables, parked cars and a tree-lined street behind. The scene manager and transform panels sit at the left, the tool strip along the bottom, and the status bar reads two million splats

SuperSplat rewrote itself on WebGPU and deleted the fallback

Today

Two rows of photoacoustic reconstructions of a branching vascular phantom, shown for SlingBAG, for PAGS, and as the ground-truth digital phantom. The SlingBAG panels carry a mottled noise floor around the vessels; the PAGS panels are cleaner, with the vessel network closer to the crisp white tracery of the phantom

Splatting, but the light is sound and the camera is a transducer

Sep 1, 2026

A schematic of a scene divided into a wireframe grid of cells against black. One cell is outlined in yellow and holds a sharp green cylinder; a blurred blue slab sits behind it and a red slab in front, standing in for the frozen regions flattened into single background and foreground images

Their trick makes VRAM independent of scene size. The test scenes were too small to show it.

Aug 31, 2026

A fairground drop-tower ride rendered twice: on the left from a degraded reconstruction, where the tower and foliage dissolve into white streaks and smears, and on the right after refinement, sharp and photographic against a clear sky

Give it scattered keypoints and it matches a full splat reconstruction

Aug 29, 2026

splats

Short daily briefs on the evolution of media — gaussian splats, volumetric video, dome theaters, headsets, and the research underneath.

Newsroom

  • Latest
  • All stories
  • RSS feed
© 2026 Splats · Terms · Privacy