PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
February 5, 20260 citationsOpen Access

Learning neural representations for 2d and 3d reconstruction

View Full Paper
ACAnis ChihoubRutgers, The State University of New Jersey

Key Points

  • This research aims to identify the most effective neural network architectures for 2D and 3D reconstruction tasks across various applications.
  • Explored implicit neural representations (INR) for novel view synthesis in NeRF architectures.
  • Evaluated generative methods (VAE and U-Nets) for reconstructing noisy medical images.
  • Assessed 3D foundational models for pose estimation and feature field reconstruction.
  • Addressed challenges in sparsity using geometric priors.
  • Higher quality reconstructions observed with optimized INR choices in NeRF-based architectures.
  • Generative models effectively reconstructed undersampled medical images in various scenarios.
  • Emerging 3D models demonstrated improved accuracy in 3D pose estimation and localization.

Abstract

Two-dimension (2D) and three-dimension (3D) reconstruction are foundational tasks in computer vision that underlie various downstream analyses, with applications ranging from medical imaging and computer graphics to video comprehension. Advances in computational power and deep learning have led to rapid progress in these tasks, motivating the need to properly evaluate which model architectures are best suited for different reconstruction tasks and application domains. In this thesis, we will explore various problems in 2D and 3D reconstruction and the novel methods that can be utilized to solve them. In the first chapter, we examine implicit neural representations (INR) used for novel view synthesis and analyze how the choice of INR influences reconstruction quality in NeRF-based architectures. In the second chapter, we explore 2D and 3D reconstruction methods for medical imaging data, focusing particularly on evaluating how effectively generative methods such as Variational Auto-Encoders (VAE) and U-Nets can reconstruct undersampled or noisy medical images. In the final chapter, we address key challenges in 3D reconstruction: pose estimation, feature field reconstruction, and sparsity. We begin with an assessment of the utility of emerging 3D foundational models, such as VGGT, Map-Anything, and MAST3R, in reconstructing accurate 3D poses. We then evaluate the efficacy of models such as LeRF and Feature Splatting at learning feature fields that can be used to localize objects in a set of egocentric data. Lastly, we finish with a discussion on sparsity and how geometric priors can be used to overcome limited amounts of data. By evaluating these methods on a series of tasks, we seek to evaluate which types of methods perform the best on their respective task and uncover broader trends in visual reconstruction that can be used to inform future work.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Anis Chihoub (2026) studied this question.

synapsesocial.com/papers/69843433f1d9ada3c1fb1ffbhttps://doi.org/10.7282/t3-gr3j-zq74
Ask AI
Helpful
Bookmark
Share
View Full Paper