Pseudo-depth-based deep neural network model for object detection

S Si-Qi Li W Wei Feng (Materdicine Lab, School of Life Sciences) B Bin Liu X Xin Tong Q Qiang Li

Abstract

Abstract Current machine learning methods only utilize the three-channel color features of optical images for computer visual tasks. However, the optical images only explicitly present information of RGB color and two-dimensional planar shape, where the third-dimensional spatial features are not fully exploited. This limitation restricts the potential improvement in recognition performance. To address this issue, we propose a detection scheme to enhance model’s detection capabilities based on four independent features by combining the pseudo-depth and the RGB features without adding any additional hardware sensors. The monocular depth estimation model is first used as a virtual depth sensor to extract the pseudo-depth features from input optical images. Then the fused Depth-RGB features are fed into the neural network model for object detection training and inference to enhance capability for extracting spatial features. Experiments show that the proposed method has improved the detection metric mAP $$_{50}$$ by 3.8 and 8.0 percentage points on the public M $$^3$$ FD and COCO datasets, respectively. Notably, the scheme can be easily embedded into any machine learning models to definitely improve the detection performance.

Article Details

Volume / Issue Vol. 16, Issue 1
Published March 26, 2026
ISSN 2045-2322
Publisher Nature Portfolio

Journal Info

Scientific Reports

Nature Portfolio

ISSN: 2045-2322 Open Access Life Sciences

Authors (5)

S

Si-Qi Li

W

Wei Feng

Materdicine Lab, School of Life Sciences

B

Bin Liu

X

Xin Tong

Q

Qiang Li