PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
May 8, 2026IEEE Transactions on Pattern Analysis and Machine Intelligence0 citations

DP-SfM: Dual-Pixel Structure-from-Motion without Scale Ambiguity

View Full Paper
LMLilika MakabeKAKohei AshidaHSHiroaki Santo

Key Points

  • The aim is to resolve scale ambiguity in multi-view 3D reconstruction using dual-pixel sensors without reference objects.
  • Developed a linear method to estimate absolute scale from dual-pixel images and depth maps.
  • Utilized intensity-based optimization to align left and right dual-pixel images.
  • Conducted experiments across various scenes with different cameras and lenses.
  • Successfully resolved scale ambiguity with estimated absolute scale, providing increased accuracy.
  • Demonstrated effectiveness in diverse scenes, confirming the robustness of the proposed method.

Abstract

Multi-view 3D reconstruction, namely, structure-from-motion followed by multi-view stereo, is a fundamental component of 3D computer vision. In general, multi-view 3D reconstruction suffers from an unknown scale ambiguity unless a reference object of known size is present in the scene. In this article, we show that multi-view images captured using a dual-pixel (DP) sensor can automatically resolve the scale ambiguity, without requiring a reference object or prior calibration. Specifically, the defocus blur observed in DP images provides sufficient information to determine the absolute scale when paired with depth maps (up to scale) recovered from multi-view 3D reconstruction. Based on this observation, we develop a simple yet effective linear method to estimate the absolute scale, followed by the intensity-based optimization stage that aligns the left and right DP images by shifting them back toward each other using cross-view blur kernels. Experiments demonstrate the effectiveness of the proposed approach across diverse scenes captured with different cameras and lenses. Code and data are available at https://github.com/lilika-makabe/dp-sfm-tpami.git.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Makabe et al. (2026) studied this question.

synapsesocial.com/papers/69fd7d94bfa21ec5bbf05f69https://doi.org/10.1109/tpami.2026.3690655
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Resolving Scale Ambiguity in Multi-view 3D Reconstruction Using Dual-Pixel Sensors2024 · 3 citations
  2. 2UniDepth: Universal Monocular Metric Depth Estimation2024 · 171 citations
  3. 3UniDepthV2: Universal Monocular Metric Depth Estimation Made Simpler2025 · 38 citations
  4. 4Depth Anything: Unleashing the Power of Large-Scale Unlabeled Data2024 · 1,086 citations
  5. 5Synthetic depth-of-field with a single-camera mobile phone2018 · 187 citations