arXiv · 2610.00017
Spatial Lifting for Dense Prediction
Abstract
We present Spatial Lifting (SL), a novel methodology for dense prediction tasks. SL operates by lifting standard inputs, such as 2D images, into a higher-dimensional space and subsequently processing them using networks designed for that higher dimension, such as a 3D U-Net. Counterintuitively, this dimensionality lifting allows us to achieve good performance on benchmark tasks compared to conventional approaches, while reducing inference costs and \textbf{drastically lowering the number of model parameters}. The SL framework produces intrinsically structured outputs along the lifted dimension. This emergent structure facilitates dense supervision during training and enables single-forward-pass self-consistency-based quality and uncertainty estimation at test time. Spatial Lifting introduces a simple and general modeling strategy that offers a promising path toward more efficient, accurate, and reliable deep networks for dense prediction tasks in vision.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Mingzhi Xu, Tao Zhou, Yong Li, Yizhe Zhang. 2026-07-27. Spatial Lifting for Dense Prediction. https://arxiv.org/abs/2610.00017
Cite the original work for its findings. Save a collection to share your selection of sources.