Search CORE

1,568 research outputs found

Face Centered Image Analysis Using Saliency and Deep Learning Based Techniques

Author: Guo Rui
Publication venue: TRACE: Tennessee Research and Creative Exchange
Publication date: 01/08/2016
Field of study

Image analysis starts with the purpose of configuring vision machines that can perceive like human to intelligently infer general principles and sense the surrounding situations from imagery. This dissertation studies the face centered image analysis as the core problem in high level computer vision research and addresses the problem by tackling three challenging subjects: Are there anything interesting in the image? If there is, what is/are that/they? If there is a person presenting, who is he/she? What kind of expression he/she is performing? Can we know his/her age? Answering these problems results in the saliency-based object detection, deep learning structured objects categorization and recognition, human facial landmark detection and multitask biometrics. To implement object detection, a three-level saliency detection based on the self-similarity technique (SMAP) is firstly proposed in the work. The first level of SMAP accommodates statistical methods to generate proto-background patches, followed by the second level that implements local contrast computation based on image self-similarity characteristics. At last, the spatial color distribution constraint is considered to realize the saliency detection. The outcome of the algorithm is a full resolution image with highlighted saliency objects and well-defined edges. In object recognition, the Adaptive Deconvolution Network (ADN) is implemented to categorize the objects extracted from saliency detection. To improve the system performance, L1/2 norm regularized ADN has been proposed and tested in different applications. The results demonstrate the efficiency and significance of the new structure. To fully understand the facial biometrics related activity contained in the image, the low rank matrix decomposition is introduced to help locate the landmark points on the face images. The natural extension of this work is beneficial in human facial expression recognition and facial feature parsing research. To facilitate the understanding of the detected facial image, the automatic facial image analysis becomes essential. We present a novel deeply learnt tree-structured face representation to uniformly model the human face with different semantic meanings. We show that the proposed feature yields unified representation in multi-task facial biometrics and the multi-task learning framework is applicable to many other computer vision tasks

University of Tennessee, Knoxville: Trace

Project SEMACODE : a scale-invariant object recognition system for content-based queries in image databases

Author: Arlt Björn
Brause Rüdiger W.
Tratar Erwin
Publication venue
Publication date: 01/01/1999
Field of study

For the efficient management of large image databases, the automated characterization of images and the usage of that characterization for searching and ordering tasks is highly desirable. The purpose of the project SEMACODE is to combine the still unsolved problem of content-oriented characterization of images with scale-invariant object recognition and modelbased compression methods. To achieve this goal, existing techniques as well as new concepts related to pattern matching, image encoding, and image compression are examined. The resulting methods are integrated in a common framework with the aid of a content-oriented conception. For the application, an image database at the library of the university of Frankfurt/Main (StUB; about 60000 images), the required operations are developed. The search and query interfaces are defined in close cooperation with the StUB project “Digitized Colonial Picture Library”. This report describes the fundamentals and first results of the image encoding and object recognition algorithms developed within the scope of the project

Hochschulschriftenserver - Universität Frankfurt am Main

Face recognition by cortical multi-scale line and edge representations

Author: du Buf J. M. H.
Rodrigues J. M. F.
Publication venue: Póvoa do Varzim
Publication date: 01/01/2006
Field of study

Empirical studies concerning face recognition suggest that faces may be stored in memory by a few canonical representations. Models of visual perception are based on image representations in cortical area V1 and beyond, which contain many cell layers for feature extraction. Simple, complex and end-stopped cells provide input for line, edge and keypoint detection. Detected events provide a rich, multi-scale object representation, and this representation can be stored in memory in order to identify objects. In this paper, the above context is applied to face recognition. The multi-scale line/edge representation is explored in conjunction with keypoint-based saliency maps for Focus-of-Attention. Recognition rates of up to 96% were achieved by combining frontal and 3/4 views, and recognition was quite robust against partial occlusions

Sapientia

Attention in hierarchical models of object recognition

Author: Amit
Bar
Biederman
Biederman
Bülthoff
Carmi
Cave
Chelazzi
Chelazzi
Connor
Darrel
Deco
Deco
Desimone
Duncan
Duncan
Edelman
Egly
Eriksen
Fei-Fei
Freedman
Frintrop
Fukushima
Gauthier
Grossberg
Hamker
Hamker
Hamker
Hayworth
Heinke
Hershler
Hopfield
Hubel
Hummel
Itti
Itti
Itti
Koch
Logothetis
Lowe
Luck
Marr
Marr
McAdams
McAdams
Mel
Mitchell
Moore
Motter
Mozer
Mozer
Navalpakkam
Navalpakkam
Olshausen
O’Craven
Peters
Poggio
Posner
Raizada
Rees
Rensink
Rensink
Reynolds
Reynolds
Riesenhuber
Roelfsema
Rolls
Rybak
Saenz
Schill
Serre
Spitzer
Tarr
Treisman
Treue
Tsotsos
Tsotsos
Ullman
Ullman
Ullman
VanRullen
Wallis
Walther
Walther
Walther
Wolfe
Wolfe
Publication venue: 'Elsevier BV'
Publication date: 01/01/2007
Field of study

Object recognition and visual attention are tightly linked processes in human perception. Over the last three decades, many models have been suggested to explain these two processes and their interactions, and in some cases these models appear to contradict each other. We suggest a unifying framework for object recognition and attention and review the existing modeling literature in this context. Furthermore, we demonstrate a proof-of-concept implementation for sharing complex features between recognition and attention as a mode of top-down attention to particular objects or object categories

Crossref

Caltech Authors

Multi-scale keypoints in V1 and face detection

Author: du Buf J. M. H.
Rodrigues J. M. F.
Publication venue: Naples
Publication date: 01/01/2005
Field of study

End-stopped cells in cortical area V1, which combine out- puts of complex cells tuned to different orientations, serve to detect line and edge crossings (junctions) and points with a large curvature. In this paper we study the importance of the multi-scale keypoint representa- tion, i.e. retinotopic keypoint maps which are tuned to different spatial frequencies (scale or Level-of-Detail). We show that this representation provides important information for Focus-of-Attention (FoA) and object detection. In particular, we show that hierarchically-structured saliency maps for FoA can be obtained, and that combinations over scales in conjunction with spatial symmetries can lead to face detection through grouping operators that deal with keypoints at the eyes, nose and mouth, especially when non-classical receptive field inhibition is employed. Al- though a face detector can be based on feedforward and feedback loops within area V1, such an operator must be embedded into dorsal and ventral data streams to and from higher areas for obtaining translation-, rotation- and scale-invariant face (object) detection

Sapientia