ZoomNet: Part-Aware Adaptive Zooming Neural Network for 3D Object Detection-Reference-Cited by-同舟云学术

ZoomNet: Part-Aware Adaptive Zooming Neural Network for 3D Object Detection

Published:2020-04-03 Issue:07 Volume:34 Page:12557-12564
ISSN:2374-3468
Container-title:Proceedings of the AAAI Conference on Artificial Intelligence
language:
Short-container-title:AAAI

Author:

Xu Zhenbo,Zhang Wei,Ye Xiaoqing,Tan Xiao,Yang Wei,Wen Shilei,Ding Errui,Meng Ajin,Huang Liusheng

Abstract

3D object detection is an essential task in autonomous driving and robotics. Though great progress has been made, challenges remain in estimating 3D pose for distant and occluded objects. In this paper, we present a novel framework named ZoomNet for stereo imagery-based 3D detection. The pipeline of ZoomNet begins with an ordinary 2D object detection model which is used to obtain pairs of left-right bounding boxes. To further exploit the abundant texture cues in rgb images for more accurate disparity estimation, we introduce a conceptually straight-forward module – adaptive zooming, which simultaneously resizes 2D instance bounding boxes to a unified resolution and adjusts the camera intrinsic parameters accordingly. In this way, we are able to estimate higher-quality disparity maps from the resized box images then construct dense point clouds for both nearby and distant objects. Moreover, we introduce to learn part locations as complementary features to improve the resistance against occlusion and put forward the 3D fitting score to better estimate the 3D detection quality. Extensive experiments on the popular KITTI 3D detection dataset indicate ZoomNet surpasses all previous state-of-the-art methods by large margins (improved by 9.4% on APbv (IoU=0.7) over pseudo-LiDAR). Ablation study also demonstrates that our adaptive zooming strategy brings an improvement of over 10% on AP3d (IoU=0.7). In addition, since the official KITTI benchmark lacks fine-grained annotations like pixel-wise part locations, we also present our KFG dataset by augmenting KITTI with detailed instance-wise annotations including pixel-wise part location, pixel-wise disparity, etc.. Both the KFG dataset and our codes will be publicly available at https://github.com/detectRecog/ZoomNet.

Publisher

Association for the Advancement of Artificial Intelligence (AAAI)

Subject

General Medicine

Cited by 31 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Self-supervised 3D vehicle detection based on monocular images;Signal Processing: Image Communication;2024-09

2. Two‐dimensional adaptive Whittaker–Shannon Sinc‐based zooming;Applied Research;2024-06-25

3. A review of 3D object detection based on autonomous driving;The Visual Computer;2024-06-14

4. A survey on 3D object detection in real time for autonomous driving;Frontiers in Robotics and AI;2024-03-06

5. ER3D: An Efficient Real-time 3D Object Detection Framework for Autonomous Driving;2023 IEEE 29th International Conference on Parallel and Distributed Systems (ICPADS);2023-12-17