Learning Policies for Continuous Control via Transition Models

Huebotter, Justus; Thill, Serge; van Gerven, Marcel; Lanillos, Pablo

Computer Science > Robotics

arXiv:2209.08033 (cs)

[Submitted on 16 Sep 2022]

Title:Learning Policies for Continuous Control via Transition Models

Authors:Justus Huebotter, Serge Thill, Marcel van Gerven, Pablo Lanillos

View PDF

Abstract:It is doubtful that animals have perfect inverse models of their limbs (e.g., what muscle contraction must be applied to every joint to reach a particular location in space). However, in robot control, moving an arm's end-effector to a target position or along a target trajectory requires accurate forward and inverse models. Here we show that by learning the transition (forward) model from interaction, we can use it to drive the learning of an amortized policy. Hence, we revisit policy optimization in relation to the deep active inference framework and describe a modular neural network architecture that simultaneously learns the system dynamics from prediction errors and the stochastic policy that generates suitable continuous control commands to reach a desired reference position. We evaluated the model by comparing it against the baseline of a linear quadratic regulator, and conclude with additional steps to take toward human-like motor control.

Subjects:	Robotics (cs.RO); Artificial Intelligence (cs.AI); Machine Learning (cs.LG); Systems and Control (eess.SY)
Cite as:	arXiv:2209.08033 [cs.RO]
	(or arXiv:2209.08033v1 [cs.RO] for this version)
	https://doi.org/10.48550/arXiv.2209.08033

Submission history

From: Justus Huebotter [view email]
[v1] Fri, 16 Sep 2022 16:23:48 UTC (709 KB)

Computer Science > Robotics

Title:Learning Policies for Continuous Control via Transition Models

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Robotics

Title:Learning Policies for Continuous Control via Transition Models

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators