Search CORE

76 research outputs found

Approximate information maximization for bandit games

Author: Barbier-Chebbah Alex
Boursier Etienne
Masson Jean-Baptiste
Vestergaard Christian L.
Publication venue
Publication date: 30/10/2023
Field of study

Entropy maximization and free energy minimization are general physical principles for modeling the dynamics of various physical systems. Notable examples include modeling decision-making within the brain using the free-energy principle, optimizing the accuracy-complexity trade-off when accessing hidden variables with the information bottleneck principle (Tishby et al., 2000), and navigation in random environments using information maximization (Vergassola et al., 2007). Built on this principle, we propose a new class of bandit algorithms that maximize an approximation to the information of a key variable within the system. To this end, we develop an approximated analytical physics-based representation of an entropy to forecast the information gain of each action and greedily choose the one with the largest information gain. This method yields strong performances in classical bandit settings. Motivated by its empirical success, we prove its asymptotic optimality for the two-armed bandit problem with Gaussian rewards. Owing to its ability to encompass the system's properties in a global physical functional, this approach can be efficiently adapted to more complex bandit settings, calling for further investigation of information maximization approaches for multi-armed bandit problems

arXiv.org e-Print Archive

Natural search algorithms as a bridge between organisms, evolution, and ecology

Author: Brumley Douglas Richard
Carrara Francesco
Hein Andrew M.
Levin Simon A.
Stocker Roman
Publication venue: 'Proceedings of the National Academy of Sciences'
Publication date: 20/04/2018
Field of study

The ability to navigate is a hallmark of living systems, from single cells to higher animals. Searching for targets, such as food or mates in particular, is one of the fundamental navigational tasks many organisms must execute to survive and reprod uce. Here, we argue that a recent surge of studies of the proximate mechanisms that underlie search behavior offers a new opportunity to integrate the biophysics and neuroscience of sensory systems with ecological and evolutionary processes, closing a feedback loop that promises exciting new avenues of scientific exploration at the frontier of systems biology. Keywords: sensing; navigation; evolutionary strategy; encounter rates; exploration–exploitationGordon and Betty Moore Foundation (Award GBMF3783

DSpace@MIT

Backwards is the way forward: feedback in the cortical hierarchy predicts the expected future

Author: Muckli L.
Petro L.S.
Smith F.W.
Publication venue: 'Cambridge University Press (CUP)'
Publication date: 10/05/2013
Field of study

Clark offers a powerful description of the brain as a prediction machine, which offers progress on two distinct levels. First, on an abstract conceptual level, it provides a unifying framework for perception, action, and cognition (including subdivisions such as attention, expectation, and imagination). Second, hierarchical prediction offers progress on a concrete descriptive level for testing and constraining conceptual elements and mechanisms of predictive coding models (estimation of predictions, prediction errors, and internal models)

Enlighten

University of East Anglia digital repository

Towards Data-centric Graph Machine Learning: Review and Outlook

Author: Bao Zhifeng
Fang Meng
Hu Xia
Liew Alan Wee-Chung
Liu Yixin
Pan Shirui
Zheng Xin
Publication venue
Publication date: 19/09/2023
Field of study

Data-centric AI, with its primary focus on the collection, management, and utilization of data to drive AI models and applications, has attracted increasing attention in recent years. In this article, we conduct an in-depth and comprehensive review, offering a forward-looking outlook on the current efforts in data-centric AI pertaining to graph data-the fundamental data structure for representing and capturing intricate dependencies among massive and diverse real-life entities. We introduce a systematic framework, Data-centric Graph Machine Learning (DC-GML), that encompasses all stages of the graph data lifecycle, including graph data collection, exploration, improvement, exploitation, and maintenance. A thorough taxonomy of each stage is presented to answer three critical graph-centric questions: (1) how to enhance graph data availability and quality; (2) how to learn from graph data with limited-availability and low-quality; (3) how to build graph MLOps systems from the graph data-centric view. Lastly, we pinpoint the future prospects of the DC-GML domain, providing insights to navigate its advancements and applications.Comment: 42 pages, 9 figure

arXiv.org e-Print Archive

Information driven self-organization of complex robotic behaviors

Author: Ay Nihat
Der Ralf
Martius Georg
Publication venue: 'Public Library of Science (PLoS)'
Publication date: 27/03/2013
Field of study

Information theory is a powerful tool to express principles to drive autonomous systems because it is domain invariant and allows for an intuitive interpretation. This paper studies the use of the predictive information (PI), also called excess entropy or effective measure complexity, of the sensorimotor process as a driving force to generate behavior. We study nonlinear and nonstationary systems and introduce the time-local predicting information (TiPI) which allows us to derive exact results together with explicit update rules for the parameters of the controller in the dynamical systems framework. In this way the information principle, formulated at the level of behavior, is translated to the dynamics of the synapses. We underpin our results with a number of case studies with high-dimensional robotic systems. We show the spontaneous cooperativity in a complex physical system with decentralized control. Moreover, a jointly controlled humanoid robot develops a high behavioral variety depending on its physics and the environment it is dynamically embedded into. The behavior can be decomposed into a succession of low-dimensional modes that increasingly explore the behavior space. This is a promising way to avoid the curse of dimensionality which hinders learning systems to scale well.Comment: 29 pages, 12 figure

arXiv.org e-Print Archive

Directory of Open Access Journals

PubMed Central

FigShare

Training an unsupervised GNN model for load balanced clustering of IoT networks

Author: Dănilă Andrei
Publication venue
Publication date: 18/07/2022
Field of study

Pure OAI Repository

Attention is more than prediction precision [Commentary on target article]

Author: Bowman Howard
Filetti Marco
Olivers Christian
Wyble Brad
Publication venue: 'Cambridge University Press (CUP)'
Publication date: 01/01/2013
Field of study

A cornerstone of the target article is that, in a predictive coding framework, attention can be modelled by weighting prediction error with a measure of precision. We argue that this is not a complete explanation, especially in the light of ERP (event-related potentials) data showing large evoked responses for frequently presented target stimuli, which thus are predicted

VU Research Portal

University of Birmingham Research Portal

Kent Academic Repository