English
Related papers

Related papers: Relative Entropy (RE) Based LTI System Modeling Eq…

200 papers

Balanced and efficient information flow is essential for optimizing language generation models. In this work, we propose Entropy-UID, a new token selection method that balances entropy and Uniform Information Density (UID) principles for…

Computation and Language · Computer Science 2025-02-21 Xinpeng Shou

In the context of control and estimation under information constraints, restoration entropy measures the minimal required data rate above which the state of a system can be estimated so that the estimation quality does not degrade over time…

Optimization and Control · Mathematics 2020-09-22 C. Kawan , A. Matveev , A. Pogromsky

In this paper, an adaptive observer is proposed for multi-input multi-output (MIMO) discrete-time linear time-invariant (LTI) systems. Unlike existing MIMO adaptive observer designs, the proposed approach is applicable to LTI systems in…

Systems and Control · Electrical Eng. & Systems 2024-01-31 Anchita Dey , Shubhendu Bhasin

Continuous-time reinforcement learning (CTRL) provides a natural framework for sequential decision-making in dynamic environments where interactions evolve continuously over time. While CTRL has shown growing empirical success, its ability…

Machine Learning · Computer Science 2025-12-04 Runze Zhao , Yue Yu , Ruhan Wang , Chunfeng Huang , Dongruo Zhou

We present a framework for learning of modeling uncertainties in Linear Time Invariant (LTI) systems. We propose a methodology to extend the dynamics of an LTI (without uncertainty) with an uncertainty model, based on measured data, to…

Systems and Control · Electrical Eng. & Systems 2023-11-01 Farhad Ghanipoor , Carlos Murguia , Peyman Mohajerin Esfahani , Nathan van de Wouw

An effective way to scale up test-time compute of large language models is to sample multiple responses and then select the best one, as in Grok Heavy and Gemini Deep Think. Existing selection methods often rely on external reward models,…

Machine Learning · Computer Science 2026-05-04 Wenshuo Zhao , Qi Zhu , Xingshan Zeng , Fei Mi , Lifeng Shang , Yi R. , Fung

Linear time-invariant (LTI) systems appear frequently in natural sciences and engineering contexts. Many LTI systems are described by ordinary differential equations (ODEs). For example, biological gene regulation, analog filter circuits,…

Systems and Control · Electrical Eng. & Systems 2019-12-18 Parker S. Ruth , Herbert M. Sauro

Recent exploration methods have proven to be a recipe for improving sample-efficiency in deep reinforcement learning (RL). However, efficient exploration in high-dimensional observation spaces still remains a challenge. This paper presents…

Machine Learning · Computer Science 2021-06-22 Younggyo Seo , Lili Chen , Jinwoo Shin , Honglak Lee , Pieter Abbeel , Kimin Lee

We consider the problem of estimating a linear time-invariant (LTI) dynamical system from a single trajectory via streaming algorithms, which is encountered in several applications including reinforcement learning (RL) and time-series…

Machine Learning · Computer Science 2021-12-03 Prateek Jain , Suhas S Kowshik , Dheeraj Nagaraj , Praneeth Netrapalli

In this paper, we study the cooperative robust output regulation problem for linear uncertain multi-agent systems with both communication delay and input delay by the distributed internal model approach. The problem includes the…

Optimization and Control · Mathematics 2015-08-19 Maobin Lu , Jie Huang

Best relay selection (BRS) is crucial in enhancing the performance of cooperative networks. In contrast to most previous works, where the guidelines for BRS are limited to Gaussian noise, in this article, we propose a novel relay selection…

Signal Processing · Electrical Eng. & Systems 2019-05-30 Md Sahabul Alam , Georges Kaddoum , Basile L. Agba

We first develop systematic and comprehensive interval observer designs for linear time-invariant (LTI) systems, under standard assumptions of observability and interval bounds on the initial condition and uncertainties. Traditionally, such…

Systems and Control · Electrical Eng. & Systems 2025-06-09 Thach Ngoc Dinh , Gia Quoc Bao Tran

In the present work, an embedded PI controller is designed for speed regulation of DC servomotor over a wireless network. The embedded controller integrates PI controller with a proposed time-delay estimator and an adaptive digital Smith…

Systems and Control · Electrical Eng. & Systems 2019-12-03 Santosh Mohan Rajkumar , Sayan Chakraborty , Rajeeb Dey , Dipankar Deb

Transfer entropy (TE) is an information theoretic measure that reveals the directional flow of information between processes, providing valuable insights for a wide range of real-world applications. This work proposes Transfer Entropy…

Information Theory · Computer Science 2025-07-22 Omer Luxembourg , Dor Tsur , Haim Permuter

In this paper, we provide two new stable online algorithms for the problem of prediction in reinforcement learning, \emph{i.e.}, estimating the value function of a model-free Markov reward process using the linear function approximation…

Machine Learning · Computer Science 2018-06-19 Ajin George Joseph , Shalabh Bhatnagar

Embedded systems power many modern applications and must often meet strict reliability, real-time, thermal, and power requirements. Task replication can improve reliability by duplicating a task's execution to handle transient and permanent…

Machine Learning · Computer Science 2025-03-18 Roozbeh Siyadatzadeh , Mohsen Ansari , Muhammad Shafique , Alireza Ejlali

Machine learning models must continuously self-adjust themselves for novel data distribution in the open world. As the predominant principle, entropy minimization (EM) has been proven to be a simple yet effective cornerstone in existing…

Machine Learning · Statistics 2024-10-16 Qingyang Zhang , Yatao Bian , Xinke Kong , Peilin Zhao , Changqing Zhang

6-Degree of Freedom (6DoF) motion estimation with a combination of visual and inertial sensors is a growing area with numerous real-world applications. However, precise calibration of the time offset between these two sensor types is a…

Robotics · Computer Science 2025-01-06 Yunfei Fan , Tianyu Zhao , Linan Guo , Chen Chen , Xin Wang , Fengyi Zhou

Online optimisation studies the convergence of optimisation methods as the data embedded in the problem changes. Based on this idea, we propose a primal dual online method for nonlinear time-discrete inverse problems. We analyse the method…

Optimization and Control · Mathematics 2025-03-18 Neil Dizon , Jyrki Jauhiainen , Tuomo Valkonen

This paper presents a low-dimensional observer design for stable, single-input single-output, continuous-time linear time-invariant (LTI) systems. Leveraging the model reduction by moment matching technique, we approximate the system with a…

Systems and Control · Electrical Eng. & Systems 2025-08-04 M. F. Shakib , M. Khalil , R. Postoyan