<p>A fundamental challenge in learning an unknown dynamical system is to reduce model uncertainty by making measurements while maintaining safety. In this work, we formulate a mathematical definition of what it means to safely learn a dynamical system by sequentially deciding where to initialize the next trajectory. In our framework, the state of the system is required to stay within a safety region for a horizon of <i>T</i> time steps under the action of all dynamical systems that (i) belong to a given initial uncertainty set, and (ii) are consistent with the information gathered so far. For our first set of results, we consider the setting of safely learning a linear dynamical system involving <i>n</i> states. For the case <InlineEquation ID="IEq1"> <EquationSource Format="TEX">\(T=1\)</EquationSource> <EquationSource Format="MATHML"><math> <mrow> <mi>T</mi> <mo>=</mo> <mn>1</mn> </mrow> </math></EquationSource> </InlineEquation>, we present a linear programming-based algorithm that either safely recovers the true dynamics from at most <i>n</i> trajectories, or certifies that safe learning is impossible. For <InlineEquation ID="IEq2"> <EquationSource Format="TEX">\(T=2\)</EquationSource> <EquationSource Format="MATHML"><math> <mrow> <mi>T</mi> <mo>=</mo> <mn>2</mn> </mrow> </math></EquationSource> </InlineEquation>, we give a semidefinite representation of the set of safe initial conditions and show that <InlineEquation ID="IEq3"> <EquationSource Format="TEX">\(\lceil n/2 \rceil \)</EquationSource> <EquationSource Format="MATHML"><math> <mrow> <mo>⌈</mo> <mi>n</mi> <mo stretchy="false">/</mo> <mn>2</mn> <mo>⌉</mo> </mrow> </math></EquationSource> </InlineEquation> trajectories generically suffice for safe learning. For <InlineEquation ID="IEq4"> <EquationSource Format="TEX">\(T = \infty \)</EquationSource> <EquationSource Format="MATHML"><math> <mrow> <mi>T</mi> <mo>=</mo> <mi>∞</mi> </mrow> </math></EquationSource> </InlineEquation>, we provide semidefinite representable inner approximations of the set of safe initial conditions and show that one trajectory generically suffices for safe learning. Finally, we extend a number of our results to the cases where the initial uncertainty set contains sparse, low-rank, or permutation matrices, or when the dynamical system involves a control input. Our second set of results concerns the problem of safely learning a general class of nonlinear dynamical systems. For the case <InlineEquation ID="IEq5"> <EquationSource Format="TEX">\(T=1\)</EquationSource> <EquationSource Format="MATHML"><math> <mrow> <mi>T</mi> <mo>=</mo> <mn>1</mn> </mrow> </math></EquationSource> </InlineEquation>, we give a second-order cone programming based representation of the set of safe initial conditions. For <InlineEquation ID="IEq6"> <EquationSource Format="TEX">\(T=\infty \)</EquationSource> <EquationSource Format="MATHML"><math> <mrow> <mi>T</mi> <mo>=</mo> <mi>∞</mi> </mrow> </math></EquationSource> </InlineEquation>, we provide semidefinite representable inner approximations to the set of safe initial conditions. We show how one can safely collect trajectories and fit a polynomial model of the nonlinear dynamics that is consistent with the initial uncertainty set and best agrees with the observations. We also present extensions of some of our results to the cases where the measurements are noisy or the dynamical system involves disturbances.</p>

错误:搜索内容不能为空,请输入英文关键词
错误:关键词超出字数限制,请精简
高级检索

Safely Learning Dynamical Systems

  • Amir Ali Ahmadi,
  • Abraar Chaudhry,
  • Vikas Sindhwani,
  • Stephen Tu

摘要

A fundamental challenge in learning an unknown dynamical system is to reduce model uncertainty by making measurements while maintaining safety. In this work, we formulate a mathematical definition of what it means to safely learn a dynamical system by sequentially deciding where to initialize the next trajectory. In our framework, the state of the system is required to stay within a safety region for a horizon of T time steps under the action of all dynamical systems that (i) belong to a given initial uncertainty set, and (ii) are consistent with the information gathered so far. For our first set of results, we consider the setting of safely learning a linear dynamical system involving n states. For the case \(T=1\) T = 1 , we present a linear programming-based algorithm that either safely recovers the true dynamics from at most n trajectories, or certifies that safe learning is impossible. For \(T=2\) T = 2 , we give a semidefinite representation of the set of safe initial conditions and show that \(\lceil n/2 \rceil \) n / 2 trajectories generically suffice for safe learning. For \(T = \infty \) T = , we provide semidefinite representable inner approximations of the set of safe initial conditions and show that one trajectory generically suffices for safe learning. Finally, we extend a number of our results to the cases where the initial uncertainty set contains sparse, low-rank, or permutation matrices, or when the dynamical system involves a control input. Our second set of results concerns the problem of safely learning a general class of nonlinear dynamical systems. For the case \(T=1\) T = 1 , we give a second-order cone programming based representation of the set of safe initial conditions. For \(T=\infty \) T = , we provide semidefinite representable inner approximations to the set of safe initial conditions. We show how one can safely collect trajectories and fit a polynomial model of the nonlinear dynamics that is consistent with the initial uncertainty set and best agrees with the observations. We also present extensions of some of our results to the cases where the measurements are noisy or the dynamical system involves disturbances.