State-Based Confidence Bounds for Data-Driven Stochastic Reachability Using Hilbert Space Embeddings

Thorpe, Adam J.; Ortiz, Kendric R.; Oishi, Meeko M. K.

Mathematics > Optimization and Control

arXiv:2010.08036 (math)

[Submitted on 15 Oct 2020 (v1), last revised 7 Dec 2021 (this version, v2)]

Title:State-Based Confidence Bounds for Data-Driven Stochastic Reachability Using Hilbert Space Embeddings

Authors:Adam J. Thorpe, Kendric R. Ortiz, Meeko M. K. Oishi

View PDF

Abstract:In this paper, we compute finite sample bounds for data-driven approximations of the solution to stochastic reachability problems. Our approach uses a nonparametric technique known as kernel distribution embeddings, and provides probabilistic assurances of safety for stochastic systems in a model-free manner. By implicitly embedding the stochastic kernel of a Markov control process in a reproducing kernel Hilbert space, we can approximate the safety probabilities for stochastic systems with arbitrary stochastic disturbances as simple matrix operations and inner products. We present finite sample bounds for point-based approximations of the safety probabilities through construction of probabilistic confidence bounds that are state- and input-dependent. One advantage of this approach is that the bounds are responsive to non-uniformly sampled data, meaning that tighter bounds are feasible in regions of the state- and input-space with more observations. We numerically evaluate the approach, and demonstrate its efficacy on a neural network-controlled pendulum system.

Subjects:	Optimization and Control (math.OC); Systems and Control (eess.SY)
Cite as:	arXiv:2010.08036 [math.OC]
	(or arXiv:2010.08036v2 [math.OC] for this version)
	https://doi.org/10.48550/arXiv.2010.08036

Submission history

From: Adam Thorpe [view email]
[v1] Thu, 15 Oct 2020 21:47:13 UTC (882 KB)
[v2] Tue, 7 Dec 2021 21:45:43 UTC (1,793 KB)

Mathematics > Optimization and Control

Title:State-Based Confidence Bounds for Data-Driven Stochastic Reachability Using Hilbert Space Embeddings

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Mathematics > Optimization and Control

Title:State-Based Confidence Bounds for Data-Driven Stochastic Reachability Using Hilbert Space Embeddings

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators