Search CORE

171,143 research outputs found

Detect-and-Track: Efficient Pose Estimation in Videos

Author: Girdhar Rohit
Gkioxari Georgia
Paluri Manohar
Torresani Lorenzo
Tran Du
Publication venue
Publication date: 02/05/2018
Field of study

This paper addresses the problem of estimating and tracking human body keypoints in complex, multi-person video. We propose an extremely lightweight yet highly effective approach that builds upon the latest advancements in human detection and video understanding. Our method operates in two-stages: keypoint estimation in frames or short clips, followed by lightweight tracking to generate keypoint predictions linked over the entire video. For frame-level pose estimation we experiment with Mask R-CNN, as well as our own proposed 3D extension of this model, which leverages temporal information over small clips to generate more robust frame predictions. We conduct extensive ablative experiments on the newly released multi-person video pose estimation benchmark, PoseTrack, to validate various design choices of our model. Our approach achieves an accuracy of 55.2% on the validation and 51.8% on the test set using the Multi-Object Tracking Accuracy (MOTA) metric, and achieves state of the art performance on the ICCV 2017 PoseTrack keypoint tracking challenge.Comment: In CVPR 2018. Ranked first in ICCV 2017 PoseTrack challenge (keypoint tracking in videos). Code: https://github.com/facebookresearch/DetectAndTrack and webpage: https://rohitgirdhar.github.io/DetectAndTrack

arXiv.org e-Print Archive

Crossref

Multi-Domain Pose Network for Multi-Person Pose Estimation and Tracking

Author: Chen Riwei
Guo Hengkai
Lu Yongchen
Luo Guozhong
Tang Tang
Wen Linfu
Publication venue: 'Springer Science and Business Media LLC'
Publication date: 18/10/2018
Field of study

Multi-person human pose estimation and tracking in the wild is important and challenging. For training a powerful model, large-scale training data are crucial. While there are several datasets for human pose estimation, the best practice for training on multi-dataset has not been investigated. In this paper, we present a simple network called Multi-Domain Pose Network (MDPN) to address this problem. By treating the task as multi-domain learning, our methods can learn a better representation for pose prediction. Together with prediction heads fine-tuning and multi-branch combination, it shows significant improvement over baselines and achieves the best performance on PoseTrack ECCV 2018 Challenge without additional datasets other than MPII and COCO.Comment: Extended abstract for the ECCV 2018 PoseTrack Worksho

arXiv.org e-Print Archive

Crossref

Assessment of a photogrammetric approach for urban DSM extraction from tri-stereoscopic satellite imagery

Author: Buyuksalih Gurcan
Goossens Rudi
Tack Frederik
Publication venue: 'Wiley'
Publication date: 01/01/2012
Field of study

Built-up environments are extremely complex for 3D surface modelling purposes. The main distortions that hamper 3D reconstruction from 2D imagery are image dissimilarities, concealed areas, shadows, height discontinuities and discrepancies between smooth terrain and man-made features. A methodology is proposed to improve automatic photogrammetric extraction of an urban surface model from high resolution satellite imagery with the emphasis on strategies to reduce the effects of the cited distortions and to make image matching more robust. Instead of a standard stereoscopic approach, a digital surface model is derived from tri-stereoscopic satellite imagery. This is based on an extensive multi-image matching strategy that fully benefits from the geometric and radiometric information contained in the three images. The bundled triplet consists of an IKONOS along-track pair and an additional near-nadir IKONOS image. For the tri-stereoscopic study a densely built-up area, extending from the centre of Istanbul to the urban fringe, is selected. The accuracy of the model extracted from the IKONOS triplet, as well as the model extracted from only the along-track stereopair, are assessed by comparison with 3D check points and 3D building vector data

Ghent University Academic Bibliography

Archivsystem Ask23

İstanbul Üniversitesi Açık Erişim Sistemi