Search CORE

4,404 research outputs found

SANet: Structure-Aware Network for Visual Tracking

Author: Fan Heng
Ling Haibin
Publication venue
Publication date: 01/05/2017
Field of study

Convolutional neural network (CNN) has drawn increasing interest in visual tracking owing to its powerfulness in feature extraction. Most existing CNN-based trackers treat tracking as a classification problem. However, these trackers are sensitive to similar distractors because their CNN models mainly focus on inter-class classification. To address this problem, we use self-structure information of object to distinguish it from distractors. Specifically, we utilize recurrent neural network (RNN) to model object structure, and incorporate it into CNN to improve its robustness to similar distractors. Considering that convolutional layers in different levels characterize the object from different perspectives, we use multiple RNNs to model object structure in different levels respectively. Extensive experiments on three benchmarks, OTB100, TC-128 and VOT2015, show that the proposed algorithm outperforms other methods. Code is released at http://www.dabi.temple.edu/~hbling/code/SANet/SANet.html.Comment: In CVPR Deep Vision Workshop, 201

arXiv.org e-Print Archive

Crossref

Wavelet Estimation of Time Series Regression with Long Memory Processes

Author: Haibin Wu
Publication venue
Publication date
Field of study

This paper studies the estimation of time series regression when both regressors and disturbances have long memory. In contrast with the frequency domain estimation as in Robinson and Hidalgo (1997), we propose to estimate the same regression model with discrete wavelet transform (DWT) of the original series. Due to the approximate de-correlation property of DWT, the transformed series can be estimated using the traditional least squares techniques. We consider both the ordinary least squares and feasible generalized least squares estimator. Finite sample Monte Carlo simulation study is performed to examine the relative efficiency of the wavelet estimation.Discrete Wavelet Transform

Research Papers in Economics

End-to-end Projector Photometric Compensation

Author: Huang Bingyao
Ling Haibin
Publication venue: 'Institute of Electrical and Electronics Engineers (IEEE)'
Publication date: 08/04/2019
Field of study

Projector photometric compensation aims to modify a projector input image such that it can compensate for disturbance from the appearance of projection surface. In this paper, for the first time, we formulate the compensation problem as an end-to-end learning problem and propose a convolutional neural network, named CompenNet, to implicitly learn the complex compensation function. CompenNet consists of a UNet-like backbone network and an autoencoder subnet. Such architecture encourages rich multi-level interactions between the camera-captured projection surface image and the input image, and thus captures both photometric and environment information of the projection surface. In addition, the visual details and interaction information are carried to deeper layers along the multi-level skip convolution layers. The architecture is of particular importance for the projector compensation task, for which only a small training dataset is allowed in practice. Another contribution we make is a novel evaluation benchmark, which is independent of system setup and thus quantitatively verifiable. Such benchmark is not previously available, to our best knowledge, due to the fact that conventional evaluation requests the hardware system to actually project the final results. Our key idea, motivated from our end-to-end problem formulation, is to use a reasonable surrogate to avoid such projection process so as to be setup-independent. Our method is evaluated carefully on the benchmark, and the results show that our end-to-end learning solution outperforms state-of-the-arts both qualitatively and quantitatively by a significant margin.Comment: To appear in the 2019 IEEE Conference on Computer Vision and Pattern Recognition (CVPR). Source code and dataset are available at https://github.com/BingyaoHuang/compenne

arXiv.org e-Print Archive

Crossref