-
作者:Leung, Chi Ho; Hota, Ashish R.; Pare, Philip E.
作者单位:Purdue University System; Purdue University; Indian Institute of Technology System (IIT System); Indian Institute of Technology (IIT) - Kharagpur
摘要:In this work, we first show that the problem of parameter identification is often ill-conditioned and lacks the persistence of excitation required for the convergence of online learning schemes. To tackle these challenges, we introduce the notion of optimal and greedy excitation sets, which contain data points with sufficient richness to aid in the identification task. We then present the greedy excitation set-based recursive least squares algorithm to alleviate the problem of the lack of pers...
-
作者:Li, Xin; Shao, Mingjie; Zhang, Ji-Feng; Zhao, Yanlong
作者单位:Chinese Academy of Sciences; Nanjing Institute of Geology & Paleontology, CAS; Academy of Mathematics & System Sciences, CAS; Chinese Academy of Sciences; University of Chinese Academy of Sciences, CAS; Zhongyuan University of Technology
摘要:This article studies the control-oriented recursive identification of finite impulse response systems with binary-valued observations. Inspired by the maximum likelihood method, a novel recursive algorithm is proposed using the statistical property of system noises and observations. Unlike existing research, the gradient of the proposed algorithm is derived from the local likelihood function, which has not been previously considered. The core advantage of the algorithm is the adaptation of the...
-
作者:Luo, Huiping; Wang, JinRong
作者单位:Guizhou University; Guizhou University
摘要:The article investigates the controllability of piecewise systems with noncommutative coefficient matrices and state-dependent delays (PSCMSD). First, an array of functions is constructed and appropriate formulations of the solution to PSCMSD are developed. Then, two necessary and sufficient conditions, specifically the Gram criterion and the rank criterion, are established to evaluate the controllability of linear PSCMSD. Moreover, a minimally normalized control function capable of steering t...
-
作者:Wu, Yiting; Zhang, Junyu
作者单位:Sun Yat Sen University; Guangdong University of Foreign Studies
摘要:This study considers the discounted criterion of nonzero-sum decentralized stochastic games with prospect players. The model is nonstationary. Each player independently controls his/her own Markov chain. The subjective behavior of players is described by prospect theory. Compared with the average criterion under prospect theory studied first in 2018, we concern the time value of utility. Because prospect theory distorts the probability, an optimality equation that plays a significant role in p...
-
作者:Fang, Xuling; Li, Xun; Meng, Qingxin; Tang, Maoning
作者单位:Huzhou Normal University; Hong Kong Polytechnic University
摘要:In this paper, we study the H-2/H-infinity control problem for continuous-time mean-field stochastic systems on an infinite horizon. Our primary objective is to establish a comprehensive framework that ensures both robust performance and optimal control for such systems. First, we derive an infinite-horizon mean-field stochastic bounded real lemma, demonstrating the equivalence between H-infinity robustness and the solution to two generalized algebraic Riccati equations. We then show that the ...
-
作者:Reis, Lucas N. R.; Carvalho, Lilian K.; Moreira, Marcos V.
作者单位:Universidade Federal do Rio de Janeiro
摘要:Several strategies have been proposed in the literature to ensure data security in cyber-physical systems abstracted as networked discrete-event systems, such as enforcing opacity by using obfuscation policies. In these cases, if it is considered the existence of an intended receiver, the useful information can be hidden from him/her. This problem becomes worse when the confidentiality of the data must be ensured, i.e., when the intended receiver must always know the system current state from ...
-
作者:Xie, Kaiyun; Xiong, Junlin
作者单位:Hefei University of Technology; Chinese Academy of Sciences; University of Science & Technology of China, CAS
摘要:This article aims to design sensor transmission strategies to ensure estimation quality under denial-of-service (DoS) attacks. The interaction between sensors and attackers is formulated as a two-person zero-sum stochastic game. To approximate Nash transmission strategies and ensure estimation quality, a receding horizon (RH) approach is presented to design is an element of-optimal transmission strategies. However, this requires sufficient computing resources. Furthermore, is an element of-opt...
-
作者:Zeng, Xiong; Bako, Laurent; Ozay, Necmiye
作者单位:University of Michigan System; University of Michigan; Inria; Universite de Lille; Centre National de la Recherche Scientifique (CNRS); Centrale Lille; CNRS - Institute for Information Sciences & Technologies (INS2I)
摘要:In this article, we study the noise sensitivity of the semidefinite program (SDP) proposed for direct data-driven infinite-horizon linear quadratic regulator (LQR) problem for discrete-time linear time-invariant systems. While this SDP is shown to find the true LQR controller in the noise-free setting, we show that it leads to a trivial solution with zero gain matrices when data are corrupted by noise, even when the noise is arbitrarily small. We then study a variant of the SDP that includes a...
-
作者:Zhang, Yuchen; Chen, Bo; Wang, Zheming; Zhang, Wen-An; Yu, Li; Guo, Lei
作者单位:Zhejiang University of Technology; Zhejiang University of Technology; Beihang University
摘要:Fusion estimation is widely applied in multisensor systems to provide accurate state information, which is crucial for designing efficient control and decision-making strategies. Despite their ability to accommodate unknown noise statistics, set-membership approaches to fusion estimation are still in an early stage of development, especially zonotopic fusion estimation, which is highly advantageous given its computational efficiency and superior representational granularity. This article is co...
-
作者:Kolarijani, Mohamad Amin Sharifi; Esfahani, Peyman Mohajerin
作者单位:Delft University of Technology; University of Toronto
摘要:Recent control algorithms for Markov decision processes (MDPs) have been designed using an implicit analogy with well-established optimization algorithms. In this article, we adopt the quasi-Newton method (QNM) from convex optimization to introduce a novel control algorithm coined as quasi-policy iteration (QPI). In particular, QPI is based on a novel approximation of the Hessian matrix in the policy iteration algorithm, which exploits two linear structural constraints specific to MDPs and all...