English
Related papers

Related papers: Conformal Lyapunov Optimization: Optimal Resource …

200 papers

Energy-based learning algorithms, such as predictive coding (PC), have garnered significant attention in the machine learning community due to their theoretical properties, such as local operations and biologically plausible mechanisms for…

Machine Learning · Computer Science 2024-10-08 Ankur Mali , Tommaso Salvatori , Alexander Ororbia

We present a distributionally robust optimization (DRO) approach for the transmission expansion planning problem, considering both long- and short-term uncertainties on the system demand and non-dispatchable renewable generation. On the…

Optimization and Control · Mathematics 2020-03-17 Alexandre Velloso , David Pozo , Alexandre Street

Recent advances in large language models (LLMs) have shown strong reasoning capabilities through large-scale pretraining and post-training reinforcement learning, demonstrated by DeepSeek-R1. However, current post-training methods, such as…

Artificial Intelligence · Computer Science 2025-12-04 Boyang Gu , Hongjian Zhou , Bradley Max Segal , Jinge Wu , Zeyu Cao , Hantao Zhong , Lei Clifton , Fenglin Liu , David A. Clifton

This technical note proposes the decentralized-partial-consensus optimization with inequality constraints, and a continuous-time algorithm based on multiple interconnected recurrent neural networks (RNNs) is derived to solve the obtained…

Optimization and Control · Mathematics 2021-03-23 Zicong Xia , Yang Liu , Jianlong Qiu , Qihua Ruan , Jinde Cao

AI/ML-based tools are at the forefront of resource management solutions for communication networks. Deep learning, in particular, is highly effective in facilitating fast and high-performing decision-making whenever representative training…

Networking and Internet Architecture · Computer Science 2025-04-07 George Iosifidis , Naram Mhaisen , Douglas J. Leith

We propose a generalization of modern representation learning objectives by reframing them as recursive divergence alignment processes over localized conditional distributions While recent frameworks like Information Contrastive Learning…

Machine Learning · Computer Science 2025-05-02 Anthony D Martin

Time-distributed Optimization (TDO) is an approach for reducing the computational burden of Model Predictive Control (MPC). When using TDO, optimization iterations are distributed over time by maintaining a running solution estimate and…

Optimization and Control · Mathematics 2021-02-25 Dominic Liao-McPherson , Terrence Skibik , Jordan Leung , Ilya Kolmanovsky , Marco M. Nicotra

Unconstrained Online Linear Optimization (OLO) is a practical problem setting to study the training of machine learning models. Existing works proposed a number of potential-based algorithms, but in general the design of these potential…

Machine Learning · Computer Science 2022-06-16 Zhiyu Zhang , Ashok Cutkosky , Ioannis Paschalidis

Reinforcement learning (RL) has emerged as an effective approach for enhancing the reasoning capabilities of large language models (LLMs), especially in scenarios where supervised fine-tuning (SFT) falls short due to limited…

Machine Learning · Computer Science 2026-04-15 Jian Xiong , Jingbo Zhou , Jingyong Ye , Qiang Huang , Dejing Dou

In safe reinforcement learning (SRL) problems, an agent explores the environment to maximize an expected total reward and meanwhile avoids violation of certain constraints on a number of expected total costs. In general, such SRL problems…

Machine Learning · Computer Science 2021-06-01 Tengyu Xu , Yingbin Liang , Guanghui Lan

Approaches to continual learning aim to successfully learn a set of related tasks that arrive in an online manner. Recently, several frameworks have been developed which enable deep learning to be deployed in this learning scenario. A key…

Machine Learning · Statistics 2020-06-17 Tameem Adel , Han Zhao , Richard E. Turner

Recent advancements in Reinforcement Learning (RL), particularly Group Relative Policy Optimization (GRPO), have significantly enhanced the reasoning capabilities of Large Language Models. However, applying these problem-centric…

Computation and Language · Computer Science 2026-05-26 Yihong Tang , Kehai Chen , Liang Yue , Benyou Wang , Min Zhang

Modern control systems must operate in increasingly complex environments subject to safety constraints and input limits, and are often implemented in a hierarchical fashion with different controllers running at multiple time scales. Yet…

Systems and Control · Electrical Eng. & Systems 2022-04-04 Noel Csomay-Shanklin , Andrew J. Taylor , Ugo Rosolia , Aaron D. Ames

The recursive logit (RL) model has become a widely used framework for route choice modeling, but it suffers from a key limitation: it assigns nonzero probabilities to all paths in the network, including those that are unrealistic, such as…

Econometrics · Economics 2025-09-03 Hung Tran , Tien Mai , Minh Ha Hoang

We consider online convex optimization (OCO) with multi-slot feedback delay, where an agent makes a sequence of online decisions to minimize the accumulation of time-varying convex loss functions, subject to short-term and long-term…

Information Theory · Computer Science 2021-08-17 Juncheng Wang , Ben Liang , Min Dong , Gary Boudreau , Hatem Abou-zeid

We introduce LAGO, a LocAl-Global Optimization algorithm that combines gradient-enhanced Bayesian Optimization (BO) with gradient-based trust region local refinement through an adaptive competition mechanism. At each iteration, global and…

Machine Learning · Computer Science 2026-03-04 Eliott Van Dieren , Tommaso Vanzan , Fabio Nobile

Designing robust algorithms for the optimal power flow (OPF) problem is critical for the control of large-scale power systems under uncertainty. The chance-constrained OPF (CCOPF) problem provides a natural formulation of the trade-off…

Optimization and Control · Mathematics 2025-01-23 Eli Brock , Haixiang Zhang , Javad Lavaei , Somayeh Sojoudi

This paper introduces a chordal decomposition approach for scalable analysis of linear networked systems, including stability, $\mathcal{H}_2$ and $\mathcal{H}_{\infty}$ performance. Our main strategy is to exploit any sparsity within these…

Optimization and Control · Mathematics 2018-03-19 Yang Zheng , Maryam Kamgarpour , Aivar Sootla , Antonis Papachristodoulou

Conventionally, the resource allocation is formulated as an optimization problem and solved online with instantaneous scenario information. Since most resource allocation problems are not convex, the optimal solutions are very difficult to…

Machine Learning · Computer Science 2017-12-20 Jun-Bo Wang , Junyuan Wang , Yongpeng Wu , Jin-Yuan Wang , Huiling Zhu , Min Lin , Jiangzhou Wang

Conditional validity and length efficiency are two crucial aspects of conformal prediction (CP). Conditional validity ensures accurate uncertainty quantification for data subpopulations, while proper length efficiency ensures that the…

Machine Learning · Statistics 2024-12-12 Shayan Kiyani , George Pappas , Hamed Hassani