Search CORE

61 research outputs found

Dynamic, Task-Related and Demand-Driven Scene Representation

Author: AL Yarbus
CA Rothkopf
D Soto
DH Ballard
FH Hamker
J Ferrante
J Triesch
JK Tsotsos
JM Henderson
Julian Eggert
L Itti
M Corbetta
M Hayhoe
M Hayhoe
MA Just
MM Chun
R Sedgewick
RA Rensink
RA Rensink
S Amari
S Frintrop
S Ullman
SP Lloyd
Sven Rebhan
V Navalpakkam
V Navalpakkam
Y Aloimonos
Publication venue: Springer-Verlag
Publication date: 01/01/2010
Field of study

Humans selectively process and store details about the vicinity based on their knowledge about the scene, the world and their current task. In doing so, only those pieces of information are extracted from the visual scene that is required for solving a given task. In this paper, we present a flexible system architecture along with a control mechanism that allows for a task-dependent representation of a visual scene. Contrary to existing approaches, our system is able to acquire information selectively according to the demands of the given task and based on the system’s knowledge. The proposed control mechanism decides which properties need to be extracted and how the independent processing modules should be combined, based on the knowledge stored in the system’s long-term memory. Additionally, it ensures that algorithmic dependencies between processing modules are resolved automatically, utilizing procedural knowledge which is also stored in the long-term memory. By evaluating a proof-of-concept implementation on a real-world table scene, we show that, while solving the given task, the amount of data processed and stored by the system is considerably lower compared to processing regimes used in state-of-the-art systems. Furthermore, our system only acquires and stores the minimal set of information that is relevant for solving the given task

Crossref

Springer - Publisher Connector

PubMed Central

Disambiguating Multi–Modal Scene Representations Using Perceptual Grouping Constraints

Author: A Baumberg
A Sha'ashua
A Verri
C Harris
C Schmid
D Crevier
D Field
D Kraft
D Lowe
D Lowe
D Scharstein
E Baseski
E Brunswik
F Schaffalitzky
Florentin Wörgötter
HH Nagel
J Elder
J Elder
J Elder
J Koenderink
J Mayhew
J Rodrigues
J Rodrigues
J Shi
K Koffka
K Köhler
K Mikolajczyk
L van Gool
L Wolff
M Brown
M Felsber
M Felsberg
M Oram
M Popović
N Kim
N Krüger
N Krüger
N Krüger
N Pugeault
N Pugeault
N Pugeault
N Pugeault
N Pugeault
Nicolas Pugeault
Norbert Krüger
O Faugeras
P Kovesi
P König
P Parent
P Perona
R Chung
R Hartley
R Horaud
R Mohan
S Geman
S Sarkar
S Se
SH Lee
Teresa Serrano-Gotarredona
W Freeman
W Geisler
Y Aloimonos
Y Ohta
Publication venue: Public Library of Science
Publication date: 01/01/2010
Field of study

In its early stages, the visual system suffers from a lot of ambiguity and noise that severely limits the performance of early vision algorithms. This article presents feedback mechanisms between early visual processes, such as perceptual grouping, stereopsis and depth reconstruction, that allow the system to reduce this ambiguity and improve early representation of visual information. In the first part, the article proposes a local perceptual grouping algorithm that — in addition to commonly used geometric information — makes use of a novel multi–modal measure between local edge/line features. The grouping information is then used to: 1) disambiguate stereopsis by enforcing that stereo matches preserve groups; and 2) correct the reconstruction error due to the image pixel sampling using a linear interpolation over the groups. The integration of mutual feedback between early vision processes is shown to reduce considerably ambiguity and noise without the need for global constraints

Public Library of Science (PLOS)

Crossref

Directory of Open Access Journals

PubMed Central

Open Research Exeter

GRO.publications (Univ. Göttingen)

Enlighten

Syddansk Universitets Forskerportal

Surrey Research Insight