7,441 research outputs found
Hierarchical Decomposition of Nonlinear Dynamics and Control for System Identification and Policy Distillation
The control of nonlinear dynamical systems remains a major challenge for
autonomous agents. Current trends in reinforcement learning (RL) focus on
complex representations of dynamics and policies, which have yielded impressive
results in solving a variety of hard control tasks. However, this new
sophistication and extremely over-parameterized models have come with the cost
of an overall reduction in our ability to interpret the resulting policies. In
this paper, we take inspiration from the control community and apply the
principles of hybrid switching systems in order to break down complex dynamics
into simpler components. We exploit the rich representational power of
probabilistic graphical models and derive an expectation-maximization (EM)
algorithm for learning a sequence model to capture the temporal structure of
the data and automatically decompose nonlinear dynamics into stochastic
switching linear dynamical systems. Moreover, we show how this framework of
switching models enables extracting hierarchies of Markovian and
auto-regressive locally linear controllers from nonlinear experts in an
imitation learning scenario.Comment: 2nd Annual Conference on Learning for Dynamics and Contro
Minimally Constrained Stable Switched Systems and Application to Co-simulation
We propose an algorithm to restrict the switching signals of a constrained
switched system in order to guarantee its stability, while at the same time
attempting to keep the largest possible set of allowed switching signals. Our
work is motivated by applications to (co-)simulation, where numerical stability
is a hard constraint, but should be attained by restricting as little as
possible the allowed behaviours of the simulators. We apply our results to
certify the stability of an adaptive co-simulation orchestration algorithm,
which selects the optimal switching signal at run-time, as a function of
(varying) performance and accuracy requirements.Comment: Technical report complementing the following conference publication:
Gomes, Cl\'audio, Beno\^it Legat, Rapha\"el Jungers, and Hans Vangheluwe.
"Minimally Constrained Stable Switched Systems and Application to
Co-Simulation." In IEEE Conference on Decision and Control. Miami Beach, FL,
USA, 201
- …