Search CORE

17 research outputs found

Performance Evaluation of an Unmanned Airborne Vehicle Multi-Agent System

Author: Abhijit Deshmukh
Zhaotong Lian
Publication venue: 'IntechOpen'
Publication date: 01/01/2009
Field of study

Un ejemplo de aplicación de la tecnica bayesiana y razonamiento basado en casos en el juego del fútbol

Author: Gonzaléz Ariel
Publication venue
Publication date: 23/10/2012
Field of study

En el presente artículo se describe una propuesta de solución al problema del dominio del fútbol, la cual resuelve el problema de las acciones que un jugador de fútbol debería realizar, mediante la integración cooperativa de las Redes Bayesianas ( [7] y [8] ) y el Razonamiento Basado en Casos [1]. Esto incluye las tareas dinámicas de dos equipos, y este artículo se concentra en el fútbol simulado como un ejemplo. Primero, se analizan cuales son los elementos del problema en cuestión, en base a ellos se proponen distintos sensores para obtener información de un jugador y de los objetos de un campo de juego. Por último se presenta un set acciones abstractas que un jugador podría realizar. Se utilizan las redes Bayesianas para caracterizar la selección de una acción donde el método Razonamiento Basado en Casos es usado para determinar cómo llevar a cabo tales acciones (ambos temas son tratados en conjunto pero con una visión diferente en [6]).Eje: Agentes y Sistemas Inteligentes (ASI)Red de Universidades con Carreras en Informática (RedUNCI

Servicio de Difusión de la Creación Intelectual

Planning while Executing: A Constraint-Based Approach

Author: D.S. Weld
Publication venue: 'Springer Science and Business Media LLC'
Publication date
Field of study

Crossref

Risk-sensitive reinforcement learning applied to control under constraints

Author: Fritz Wysotzki
Peter Geibel
Publication venue
Publication date
Field of study

In this paper, we consider Markov Decision Processes (MDPs) with error states. Error states are those states entering which is undesirable or dangerous. We define the risk with respect to a policy as the probability of entering such a state when the policy is pursued. We consider the problem of finding good policies whose risk is smaller than some user-specified threshold, and formalize it as a constrained MDP with two criteria. The first criterion corresponds to the value function originally given. We will show that the risk can be formulated as a second criterion function based on a cumulative return, whose definition is independent of the original value function. We present a model free, heuristic reinforcement learning algorithm that aims at finding good deterministic policies. It is based on weighting the original value function and the risk. The weight parameter is adapted in order to find a feasible solution for the constrained problem that has a good performance with respect to the value function. The algorithm was successfully applied to the control of a feed tank with stochastic inflows that lies upstream of a distillation column. This control task was originally formulated as an optimal control problem with chance constraints, and it was solved under certain assumptions on the model to obtain an optimal solution. The power of our learning algorithm is that it can be used even when some of these restrictive assumptions are relaxed. 1

CiteSeerX