Hierarchical Decomposition of Nonlinear Dynamics and Control for System Identification and Policy Distillation

Abdulsamad, Hany; Peters, Jan

Computer Science > Machine Learning

arXiv:2005.01432 (cs)

[Submitted on 4 May 2020 (v1), last revised 12 May 2020 (this version, v2)]

Title:Hierarchical Decomposition of Nonlinear Dynamics and Control for System Identification and Policy Distillation

Authors:Hany Abdulsamad, Jan Peters

View PDF

Abstract:The control of nonlinear dynamical systems remains a major challenge for autonomous agents. Current trends in reinforcement learning (RL) focus on complex representations of dynamics and policies, which have yielded impressive results in solving a variety of hard control tasks. However, this new sophistication and extremely over-parameterized models have come with the cost of an overall reduction in our ability to interpret the resulting policies. In this paper, we take inspiration from the control community and apply the principles of hybrid switching systems in order to break down complex dynamics into simpler components. We exploit the rich representational power of probabilistic graphical models and derive an expectation-maximization (EM) algorithm for learning a sequence model to capture the temporal structure of the data and automatically decompose nonlinear dynamics into stochastic switching linear dynamical systems. Moreover, we show how this framework of switching models enables extracting hierarchies of Markovian and auto-regressive locally linear controllers from nonlinear experts in an imitation learning scenario.

Comments:	2nd Annual Conference on Learning for Dynamics and Control
Subjects:	Machine Learning (cs.LG); Machine Learning (stat.ML)
Cite as:	arXiv:2005.01432 [cs.LG]
	(or arXiv:2005.01432v2 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2005.01432

Submission history

From: Hany Abdulsamad [view email]
[v1] Mon, 4 May 2020 12:40:59 UTC (3,060 KB)
[v2] Tue, 12 May 2020 14:54:33 UTC (3,060 KB)

Computer Science > Machine Learning

Title:Hierarchical Decomposition of Nonlinear Dynamics and Control for System Identification and Policy Distillation

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Hierarchical Decomposition of Nonlinear Dynamics and Control for System Identification and Policy Distillation

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators