SOTAVerified

Monocular Depth Estimation

Monocular Depth Estimation is the task of estimating the depth value (distance relative to the camera) of each pixel given a single (monocular) RGB image. This challenging task is a key prerequisite for determining scene understanding for applications such as 3D scene reconstruction, autonomous driving, and AR. State-of-the-art methods usually fall into one of two categories: designing a complex network that is powerful enough to directly regress the depth map, or splitting the input into bins or windows to reduce computational complexity. The most popular benchmarks are the KITTI and NYUv2 datasets. Models are typically evaluated using RMSE or absolute relative error.

Source: Defocus Deblurring Using Dual-Pixel Data

Papers

Showing 76–100 of 876 papers

TitleStatusHype
Distilling Monocular Foundation Model for Fine-grained Depth Completion—0
GeoDepth: From Point-to-Depth to Plane-to-Depth Modeling for Self-Supervised Monocular Depth Estimation—0
Improved Monocular Depth Prediction Using Distance Transform Over Pre-semantic Contours with Self-supervised Neural Networks—0
MetricDepth: Enhancing Monocular Depth Estimation with Deep Metric Learning—0
Learning Monocular Depth from Events via Egomotion Compensation—0
Revisiting Monocular 3D Object Detection from Scene-Level Depth Retargeting to Instance-Level Spatial Refinement—0
Marigold-DC: Zero-Shot Monocular Depth Completion with Guided Diffusion—0
Foundation Models Meet Low-Cost Sensors: Test-Time Adaptation for Rescaling Disparity for Zero-Shot Metric Depth Estimation—0
V-MIND: Building Versatile Monocular Indoor 3D Detector with Diverse 2D Annotations—0
Balancing Shared and Task-Specific Representations: A Hybrid Approach to Depth-Aware Video Panoptic Segmentation—0
GVDepth: Zero-Shot Monocular Depth Estimation for Ground Vehicles based on Probabilistic Cue Fusion—0
LAA-Net: A Physical-prior-knowledge Based Network for Robust Nighttime Depth Estimation—0
Align3R: Aligned Monocular Depth Estimation for Dynamic Videos—0
STATIC : Surface Temporal Affine for TIme Consistency in Video Monocular Depth Estimation—0
FiffDepth: Feed-forward Transformation of Diffusion-Based Generators for Detailed Depth Estimation—0
MonoPP: Metric-Scaled Self-Supervised Monocular Depth Estimation by Planar-Parallax Geometry in Automotive Applications—0
Spatially Visual Perception for End-to-End Robotic Learning—0
PriorDiffusion: Leverage Language Prior in Diffusion Models for Monocular Depth Estimation—0
OceanLens: An Adaptive Backscatter and Edge Correction using Deep Learning Model for Enhanced Underwater ImagingCode1
MGNiceNet: Unified Monocular Geometric Scene UnderstandingCode0
Scalable Autoregressive Monocular Depth Estimation—0
MetricGold: Leveraging Text-To-Image Latent Diffusion Models for Metric Depth EstimationCode0
Mono2Stereo: Monocular Knowledge Transfer for Enhanced Stereo Matching—0
OSMLoc: Single Image-Based Visual Localization in OpenStreetMap with Fused Geometric and Semantic GuidanceCode2
D^3epth: Self-Supervised Depth Estimation with Dynamic Mask in Dynamic ScenesCode0
Show:102550
← PrevPage 4 of 36Next →

No leaderboard results yet.