Search CORE

5 research outputs found

Discrete Multi-modal Hashing with Canonical Views for Robust Mobile Landmark Search

Author: He Xiangnan
Huang Zi
Liu Xiaobai
Song Jingkuan
Zhou Xiaofang
Zhu Lei
Publication venue
Publication date: 13/07/2017
Field of study

Mobile landmark search (MLS) recently receives increasing attention for its great practical values. However, it still remains unsolved due to two important challenges. One is high bandwidth consumption of query transmission, and the other is the huge visual variations of query images sent from mobile devices. In this paper, we propose a novel hashing scheme, named as canonical view based discrete multi-modal hashing (CV-DMH), to handle these problems via a novel three-stage learning procedure. First, a submodular function is designed to measure visual representativeness and redundancy of a view set. With it, canonical views, which capture key visual appearances of landmark with limited redundancy, are efficiently discovered with an iterative mining strategy. Second, multi-modal sparse coding is applied to transform visual features from multiple modalities into an intermediate representation. It can robustly and adaptively characterize visual contents of varied landmark images with certain canonical views. Finally, compact binary codes are learned on intermediate representation within a tailored discrete binary embedding model which preserves visual relations of images measured with canonical views and removes the involved noises. In this part, we develop a new augmented Lagrangian multiplier (ALM) based optimization method to directly solve the discrete binary codes. We can not only explicitly deal with the discrete constraint, but also consider the bit-uncorrelated constraint and balance constraint together. Experiments on real world landmark datasets demonstrate the superior performance of CV-DMH over several state-of-the-art methods

arXiv.org e-Print Archive

University of Queensland eSpace

Real-Time Calibration and Registration Method for Indoor Scene with Joint Depth and Color Camera

Author: Cai X.
Chang Jian
Lei T.
Li J.
Shao X.
Tian Feng
Zhang F.
Publication venue: 'World Scientific Pub Co Pte Lt'
Publication date: 14/03/2018
Field of study

Traditional vision registration technologies require the design of precise markers or rich texture information captured from the video scenes, and the vision-based methods have high computational complexity while the hardware-based registration technologies lack accuracy. Therefore, in this paper, we propose a novel registration method that takes advantages of RGB-D camera to obtain the depth information in real-time, and a binocular system using the Time of Flight (ToF) camera and a commercial color camera is constructed to realize the three-dimensional registration technique. First, we calibrate the binocular system to get their position relationships. The systematic errors are fitted and corrected by the method of B-spline curve. In order to reduce the anomaly and random noise, an elimination algorithm and an improved bilateral filtering algorithm are proposed to optimize the depth map. For the real-time requirement of the system, it is further accelerated by parallel computing with CUDA. Then, the Camshift-based tracking algorithm is applied to capture the real object registered in the video stream. In addition, the position and orientation of the object are tracked according to the correspondence between the color image and the 3D data. Finally, some experiments are implemented and compared using our binocular system. Experimental results are shown to demonstrate the feasibility and effectiveness of our method

Bournemouth University Research Online

The Effects of Multiple Query Evidences on Social Image Retrieval

Author: CHENG Zhiyong
MIAO Haiyan
SHEN Jialie
Publication venue: 'Springer Science and Business Media LLC'
Publication date: 01/07/2016
Field of study

Institutional Knowledge at Singapore Management University

Content-Based Visual Landmark Search via Multimodal Hypergraph Learning

Author: JIN Hai
SHEN Jialie
XIE Liang
ZHENG Ran
ZHU Lei
Publication venue: 'Institute of Electrical and Electronics Engineers (IEEE)'
Publication date: 01/12/2015
Field of study

Formerly IEEE Transactions on Systems, Man, and Cybernetics, Part B: Cybernetics</p

Queen's University Belfast Research Portal

Crossref

Institutional Knowledge at Singapore Management University

Building a Large Scale Test Collection for Effective Benchmarking of Mobile Landmark Search

Author: B.S. Manjunath
D. Lowe
M. Wang
M. Wang
Ritendra Datta
Rongrong Ji
Publication venue: 'Springer Science and Business Media LLC'
Publication date: 01/01/2013
Field of study

Crossref

Institutional Knowledge at Singapore Management University