Search CORE

37,244 research outputs found

Higher-order neural networks applied to 2D and 3D object recognition

Author: D. Casasent
D.A. Jared
D.E. Rumelhart
F.P. Kuhl
G.L. Giles
J. Sullins
J.R. Quinlan
L. Spirkovska
Lilly Spirkovska
M. Hu
M.B. Reid
M.B. Reid
M.B. Reid
Max B. Reid
R. Haberman
W. Pitts
Y.-N. Hsu
Z. Chen
Publication venue: 'Springer Science and Business Media LLC'
Publication date
Field of study

PIXOR: Real-time 3D Object Detection from Point Clouds

Author: Luo Wenjie
Urtasun Raquel
Yang Bin
Publication venue
Publication date: 01/03/2019
Field of study

We address the problem of real-time 3D object detection from point clouds in the context of autonomous driving. Computation speed is critical as detection is a necessary component for safety. Existing approaches are, however, expensive in computation due to high dimensionality of point clouds. We utilize the 3D data more efficiently by representing the scene from the Bird's Eye View (BEV), and propose PIXOR, a proposal-free, single-stage detector that outputs oriented 3D object estimates decoded from pixel-wise neural network predictions. The input representation, network architecture, and model optimization are especially designed to balance high accuracy and real-time efficiency. We validate PIXOR on two datasets: the KITTI BEV object detection benchmark, and a large-scale 3D vehicle detection benchmark. In both datasets we show that the proposed detector surpasses other state-of-the-art methods notably in terms of Average Precision (AP), while still runs at >28 FPS.Comment: Update of CVPR2018 paper: correct timing, fix typos, add acknowledgemen

arXiv.org e-Print Archive

Crossref

Learning a Hierarchical Latent-Variable Model of 3D Shapes

Author: Giles C. Lee
Liu Shikun
Ororbia II Alexander G.
Publication venue
Publication date: 04/08/2018
Field of study

We propose the Variational Shape Learner (VSL), a generative model that learns the underlying structure of voxelized 3D shapes in an unsupervised fashion. Through the use of skip-connections, our model can successfully learn and infer a latent, hierarchical representation of objects. Furthermore, realistic 3D objects can be easily generated by sampling the VSL's latent probabilistic manifold. We show that our generative model can be trained end-to-end from 2D images to perform single image 3D model retrieval. Experiments show, both quantitatively and qualitatively, the improved generalization of our proposed model over a range of tasks, performing better or comparable to various state-of-the-art alternatives.Comment: Accepted as oral presentation at International Conference on 3D Vision (3DV), 201

arXiv.org e-Print Archive

Crossref