Cited time in webofscience Cited time in scopus

Depth-discriminative Metric Learning for Monocular 3D Object Detection

Title
Depth-discriminative Metric Learning for Monocular 3D Object Detection
Author(s)
Choi, WonhyeokShin, MingyuIm, Sunghoon
Issued Date
2023-12-13
Citation
Conference on Neural Information Processing Systems, pp.1 - 13
Type
Conference Paper
ISSN
1049-5258
Abstract
Monocular 3D object detection poses a significant challenge due to the lack of depth information in RGB images. Many existing methods strive to enhance the object depth estimation performance by allocating additional parameters for object depth estimation, utilizing extra modules or data. In contrast, we introduce a novel metric learning scheme that encourages the model to extract depth-discriminative features regardless of the visual attributes without increasing inference time and model size. Our method employs the distance-preserving function to organize the feature space manifold in relation to ground-truth object depth. The proposed (K, B, ϵ)-quasiisometric loss leverages predetermined pairwise distance restriction as guidance for adjusting the distance among object descriptors without disrupting the non-linearity of the natural feature manifold. Moreover, we introduce an auxiliary head for object-wise depth estimation, which enhances depth quality while maintaining the inference time. The broad applicability of our method is demonstrated through experiments that show improvements in overall performance when integrated into various baselines. The results show that our method consistently improves the performance of various baselines by 25.27% and 4.54% on average across KITTI and Waymo, respectively.
URI
http://hdl.handle.net/20.500.11750/47799
Publisher
Neural Information Processing Systems Foundation (NeurIPS Foundation)
Related Researcher
  • 임성훈 Im, Sunghoon
  • Research Interests Computer Vision; Deep Learning; Robot Vision
Files in This Item:

There are no files associated with this item.

Appears in Collections:
Department of Electrical Engineering and Computer Science Computer Vision Lab. 2. Conference Papers

qrcode

  • twitter
  • facebook
  • mendeley

Items in Repository are protected by copyright, with all rights reserved, unless otherwise indicated.

BROWSE