NSF PAR Search | NSF Public Access Repository

Note: When clicking on a Digital Object Identifier (DOI) number, you will be taken to an external site maintained by the publisher. Some full text articles may not yet be available without a charge during the embargo (administrative interval).
What is a DOI Number?

Some links on this page may take you to non-federal websites. Their policies may differ from this site.

Doppelgangers++: Improved Visual Disambiguation with Geometric 3D Features

Xiangli, Yuanbo; Cai, Ruojin; Chen, Hanyu; Byrne, Jeffrey; Snavely, Noah (June 2025, Computer Vision and Pattern Recognition (CVPR))

Visual aliasing, or doppelgangers, poses severe challenges to 3D reconstruction. We propose Doppelganger++, an enhanced pairwise image classifier that excels in visual disambiguation across diverse and challenging scenes. We seamlessly integrate Doppelganger++ into SfM, successfully disambiguating each scene. (Middle) Compared to prior work (which we refer to as DG-OG), Doppelgangers++ is more robust for everyday scenes, showing improved accuracy and robustness. We show pairs that DG-OG classifies incorrectly and ours gets correct. Our new VisymScenes dataset, featuring complex daily scenes, is particularly challenging for COLMAP and DG-OG, but our method can achieve correct and complete reconstructions.
more » « less
Free, publicly-accessible full text available June 9, 2026
LVSM: A Large View Synthesis Model with Minimal 3D Inductive Bias

Jin, Haian; Jiang, Hanwen; Tan, Hao; Zhang, Kai; Bi, Sai; Zhang, Tianyuan; Luan, Fujun; Snavely, Noah; Xu, Zexiang (April 2025, International Conference on Learning Representations (ICLR))

We propose the Large View Synthesis Model (LVSM), a novel transformer-based approach for scalable and generalizable novel view synthesis from sparse-view inputs. We introduce two architectures: (1) an encoder-decoder LVSM, which encodes input image tokens into a fixed number of 1D latent tokens, functioning as a fully learned scene representation, and decodes novel-view images from them; and (2) a decoder-only LVSM, which directly maps input images to novel-view outputs, completely eliminating intermediate scene representations. Both models bypass the 3D inductive biases used in previous methods—from 3D representations (e.g., NeRF, 3DGS) to network designs (e.g., epipolar projections, plane sweeps)—addressing novel view synthesis with a fully data-driven approach. While the encoder-decoder model offers faster inference due to its independent latent representation, the decoder-only LVSM achieves superior quality, scalability, and zero-shot generalization, outperforming previous state-of-the-art methods by 1.5 to 3.5 dB PSNR. Comprehensive evaluations across multiple datasets demonstrate that both LVSM variants achieve state-of-the-art novel view synthesis quality. Notably, our models surpass all previous methods even with reduced computational resources (1-2 GPUs).
more » « less
Free, publicly-accessible full text available April 24, 2026
ObjectCarver: Semi-automatic segmentation, reconstruction and separation of 3D objects

Hassena, Gemmechu; Moon, Jonathan; Fujii, Ryan; Yuen, Andrew; Snavely, Noah; Marschner, Steve; Hariharan, Bharath (March 2025, IEEE)

Free, publicly-accessible full text available March 25, 2026
Appearance Modeling of Iridescent Feathers with Diverse Nanostructures

https://doi.org/10.1145/3687983

Yu, Yunchen; Weidlich, Andrea; Walter, Bruce; d'Eon, Eugene; Marschner, Steve (December 2024, ACM Transactions on Graphics)

Many animals exhibit structural colors, which are often iridescent, meaning that the perceived colors change with illumination conditions and viewing perspectives. Biological iridescence is usually caused by multilayers or other periodic structures in animal tissues, which selectively reflect light of certain wavelengths and often result in a shiny appearance---which almost always comes with spatially varying highlights, thanks to randomness and irregularities in the structures. Previous models for biological iridescence tend to each target one specific structure, and most models only compute large-area averages, overlooking spatial variation in iridescent appearance. In this work, we build appearance models for biological iridescence using bird feathers as our case study, investigating different types of feathers with a variety of structural coloration mechanisms. We propose an approximate wave simulation method that takes advantage of quasi-regular structures while efficiently modeling the effects of natural structural irregularities. We further propose a method to distill our simulation results into distributions of BRDFs, generated using noise functions, that preserve relevant statistical properties of the simulated BRDFs. This allows us to model the spatially varying, glittery appearance commonly seen on feathers. Our BRDFs are practical and efficient, and we present renderings of multiple types of iridescent feathers with comparisons to photographic images.
more » « less
Free, publicly-accessible full text available December 19, 2025
Reconstructing translucent thin objects from photos

https://doi.org/10.1145/3680528.3687572

Deng, Xi; Wu, Lifan; Walter, Bruce; Ramamoorthi, Ravi; d'Eon, Eugene; Marschner, Steve; Weidlich, Andrea (December 2024, ACM)

Free, publicly-accessible full text available December 3, 2025
A Simple Approach to Differentiable Rendering of SDFs

https://doi.org/10.1145/3680528.3687573

Wang, Zichen; Deng, Xi; Zhang, Ziyi; Jakob, Wenzel; Marschner, Steve (December 2024, ACM)

Free, publicly-accessible full text available December 3, 2025
MegaScenes: Scene-Level View Synthesis at Scale

https://doi.org/10.1007/978-3-031-73397-0_12

Tung, Joseph; Chou, Gene; Cai, Ruojin; Yang, Guandao; Zhang, Kai; Wetzstein, Gordon; Hariharan, Bharath; Snavely, Noah (November 2024, Springer Nature Switzerland)

Free, publicly-accessible full text available November 3, 2025
Neural Caches for Monte Carlo Partial Differential Equation Solvers

https://doi.org/10.1145/3610548.3618141

Li, Zilu; Yang, Guandao; Deng, Xi; De_Sa, Christopher; Hariharan, Bharath; Marschner, Steve (December 2023, ACM)

This paper presents a method that uses neural networks as a caching mechanism to reduce the variance of Monte Carlo Partial Differential Equation solvers, such as the Walk-on-Spheres algorithm [Sawhney and Crane 2020]. While these Monte Carlo PDE solvers have the merits of being unbiased and discretization-free, their high variance often hinders real-time applications. On the other hand, neural networks can approximate the PDE solution, and evaluating these networks at inference time can be very fast. However, neural-network-based solutions may suffer from convergence difficulties and high bias. Our hybrid system aims to combine these two potentially complementary solutions by training a neural field to approximate the PDE solution using supervision from a WoS solver. This neural field is then used as a cache in the WoS solver to reduce variance during inference. We demonstrate that our neural field training procedure is better than the commonly used self-supervised objectives in the literature. We also show that our hybrid solver exhibits lower variance than WoS with the same computational budget: it is significantly better for small compute budgets and provides smaller improvements for larger budgets, reaching the same performance as WoS in the limit.
more » « less
Full Text Available

Search for: All records