Inverse Optimal Control from Incomplete Trajectory Observations

Jin, Wanxin; Kulić, Dana; Mou, Shaoshuai; Hirche, Sandra

Computer Science > Robotics

arXiv:1803.07696v3 (cs)

[Submitted on 21 Mar 2018 (v1), revised 31 Aug 2020 (this version, v3), latest version 22 Jan 2021 (v4)]

Title:Inverse Optimal Control from Incomplete Trajectory Observations

Authors:Wanxin Jin, Dana Kulić, Shaoshuai Mou, Sandra Hirche

View PDF

Abstract:This article develops a methodology that enables learning an objective function of an optimal control system from incomplete trajectory observations. The objective function is assumed to be a weighted sum of features (or basis functions) with unknown weights, and the observed data is a segment of a trajectory of system states and inputs. The proposed technique introduces the concept of the recovery matrix to establish the relationship between any available segment of the trajectory and the weights of given candidate features. The rank of the recovery matrix indicates whether a subset of relevant features can be found among the candidate features and the corresponding weights can be learned from the segment data. The recovery matrix can be obtained iteratively and its rank non-decreasing property shows that additional observations may contribute to the objective learning. Based on the recovery matrix, a method for using incomplete trajectory observations to learn the weights of selected features is established, and an incremental inverse optimal control algorithm is developed by automatically finding the minimal required observation. The effectiveness of the proposed method is demonstrated on a linear quadratic regulator system and a simulated robot manipulator.

Comments:	Codes: this https URL
Subjects:	Robotics (cs.RO); Systems and Control (eess.SY)
Cite as:	arXiv:1803.07696 [cs.RO]
	(or arXiv:1803.07696v3 [cs.RO] for this version)
	https://doi.org/10.48550/arXiv.1803.07696

Submission history

From: Wanxin Jin [view email]
[v1] Wed, 21 Mar 2018 00:04:19 UTC (619 KB)
[v2] Thu, 23 May 2019 18:24:35 UTC (1,240 KB)
[v3] Mon, 31 Aug 2020 12:58:23 UTC (837 KB)
[v4] Fri, 22 Jan 2021 04:10:18 UTC (870 KB)

Computer Science > Robotics

Title:Inverse Optimal Control from Incomplete Trajectory Observations

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Robotics

Title:Inverse Optimal Control from Incomplete Trajectory Observations

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators