Papers › FutureDepth: Learning to Predict the Future Improves Video Depth Estimation

FutureDepth: Learning to Predict the Future Improves Video Depth Estimation

19 Mar 2024arXiv:2403.12953archive 2025-07-28

Rajeev Yasarla, Manish Kumar Singh, Hong Cai, Yunxiao Shi, Jisoo Jeong, Yinhao Zhu, Shizhong Han, Risheek Garrepalli, Fatih Porikli

In this paper, we propose a novel video depth estimation approach, FutureDepth, which enables the model to implicitly leverage multi-frame and motion cues to improve depth estimation by making it learn to predict the future at training. More specifically, we propose a future prediction network, F-Net, which takes the features of multiple consecutive frames and is trained to predict multi-frame features one time step ahead iteratively. In this way, F-Net learns the underlying motion and correspondence information, and we incorporate its features into the depth decoding process. Additionally, to enrich the learning of multiframe correspondence cues, we further leverage a reconstruction network, R-Net, which is trained via adaptively masked auto-encoding of multiframe feature volumes. At inference time, both F-Net and R-Net are used to produce queries to work with the depth decoder, as well as a final refinement network. Through extensive experiments on several benchmarks, i.e., NYUDv2, KITTI, DDAD, and Sintel, which cover indoor, driving, and open-domain scenarios, we show that FutureDepth significantly improves upon baseline models, outperforms existing video depth estimation methods, and sets new state-of-the-art (SOTA) accuracy. Furthermore, FutureDepth is more efficient than existing SOTA video depth estimation models and has similar latencies when comparing to monocular models

PaperPDF

In Syntology Open this paper in Syntology's Atlas, the map of the papers in Syntology's graph and their citations.

Code

No code repository is listed for this paper in the archive or in Syntology's graph.

Code Syntology ran Syntology

Not run by Syntology. Nothing on this page verifies that the listed code works.

Tasks

DecoderDepth EstimationFuture predictionMonocular Depth Estimation

Results from the paper archive 2025-07-28

TaskDatasetModelMetricValueRank at snapshotLeaderboardReport
Monocular Depth Estimation KITTI Eigen split FutureDepth Delta < 1.25 0.984 #6 of 79 Archive leaderboard report
Monocular Depth Estimation KITTI Eigen split FutureDepth Delta < 1.25^2 0.998 #6 of 79 Archive leaderboard report
Monocular Depth Estimation KITTI Eigen split FutureDepth Delta < 1.25^3 1.000 #6 of 79 Archive leaderboard report
Monocular Depth Estimation KITTI Eigen split FutureDepth RMSE 1.856 #6 of 79 Archive leaderboard report
Monocular Depth Estimation KITTI Eigen split FutureDepth RMSE log 0.066 #6 of 79 Archive leaderboard report
Monocular Depth Estimation KITTI Eigen split FutureDepth Sq Rel 0.117 #6 of 79 Archive leaderboard report
Monocular Depth Estimation KITTI Eigen split FutureDepth Square relative error (SqRel) 0.117 #6 of 79 Archive leaderboard report
Monocular Depth Estimation KITTI Eigen split FutureDepth absolute relative error 0.041 #6 of 79 Archive leaderboard report
Monocular Depth Estimation NYU-Depth V2 FutureDepth Delta < 1.25 0.981 #18 of 85 Archive leaderboard report
Monocular Depth Estimation NYU-Depth V2 FutureDepth Delta < 1.25^2 0.996 #18 of 85 Archive leaderboard report
Monocular Depth Estimation NYU-Depth V2 FutureDepth Delta < 1.25^3 0.999 #18 of 85 Archive leaderboard report
Monocular Depth Estimation NYU-Depth V2 FutureDepth RMSE 0.233 #18 of 85 Archive leaderboard report
Monocular Depth Estimation NYU-Depth V2 FutureDepth absolute relative error 0.063 #18 of 85 Archive leaderboard report
Monocular Depth Estimation NYU-Depth V2 FutureDepth log 10 0.027 #18 of 85 Archive leaderboard report

Ranks are positions in the archive's leaderboards as they stood at the 2025-07-28 snapshot. Results published since then are not among these rows, so a rank here is not a current standing.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections