Hybrid system identification using switching density networks

Burke, Michael G.; Hristov, Yordan; Ramamoorthy, Subramanian

doi:10.48550/arxiv.1907.04360

Cited by 2 publications

(2 citation statements)

References 25 publications

Supporting

Mentioning

Contrasting

Order By: Relevance

“…Switching systems have also served as a powerful tool in various imitation learning approaches. Calinon et al ( 2010) combine traditional HMMs with Gaussian mixture regression to represent trajectory distributions, while Daniel et al (2016) use a hidden semi-Markov model to learn hierarchical policies and Burke et al (2019) introduced switching density networks for system identification and behavioral cloning. Finally, excellent work on hierarchical decomposition of policies in a fully Bayesian framework is introduced by Šošić et al (2017), albeit under known transition dynamics.…”

Section: Related Workmentioning

confidence: 99%

Hierarchical Decomposition of Nonlinear Dynamics and Control for System Identification and Policy Distillation

Abdulsamad,

Peters

2020

Preprint

View full text Add to dashboard Cite

The control of nonlinear dynamical systems remains a major challenge for autonomous agents. Current trends in reinforcement learning (RL) focus on complex representations of dynamics and policies, which have yielded impressive results in solving a variety of hard control tasks. However, this new sophistication and extremely over-parameterized models have come with the cost of an overall reduction in our ability to interpret the resulting policies. In this paper, we take inspiration from the control community and apply the principles of hybrid switching systems in order to break down complex dynamics into simpler components. We exploit the rich representational power of probabilistic graphical models and derive an expectation-maximization (EM) algorithm for learning a sequence model to capture the temporal structure of the data and automatically decompose nonlinear dynamics into stochastic switching linear dynamical systems. Moreover, we show how this framework of switching models enables extracting hierarchies of Markovian and auto-regressive locally linear controllers from nonlinear experts in an imitation learning scenario.

show abstract

Section: Related Workmentioning

confidence: 99%

Hierarchical Decomposition of Nonlinear Dynamics and Control for System Identification and Policy Distillation

Abdulsamad,

Peters

2020

Preprint

View full text Add to dashboard Cite

show abstract

“…The extraction of the policy implemented by a controller (being human or automatic) is often referred to as behavior cloning within the machine learning community, and we inherit here the same terminology. Specifically, the scope of behavior cloning is to learn a policy by imitation, i.e., the action to be performed in a given system state by extrapolating experience from a set of observationaction sequences [40]. With reference to system identification, the target is to imitate the behavior performed by an instructor (e.g., an automated control system or advanced logics on supervisory systems) by observing it operating in closed-loop over the controlled system (e.g., a device, a machine or a process).…”

Section: Behavior Cloning In Industrymentioning

confidence: 99%