English
Related papers

Related papers: Impermanent loss and Loss-vs-Rebalancing II

200 papers

Different representations to describe noise processes and finding connections or equivalence between them have been part of active research for decades, in particular for linear time-invariant case. In this paper the linear…

Systems and Control · Computer Science 2016-10-31 Pepijn Bastiaan Cox , Roland Tóth

With Reinforcement Learning (RL) for inventory management (IM) being a nascent field of research, approaches tend to be limited to simple, linear environments with implementations that are minor modifications of off-the-shelf RL algorithms.…

Machine Learning · Computer Science 2023-04-19 Madhav Khirwar , Karthik S. Gurumoorthy , Ankit Ajit Jain , Shantala Manchenahally

Balancing resource efficiency and fairness is critical in networked systems that support modern learning applications. We introduce the Fair Minimum Labeling (FML) problem: the task of designing a minimum-cost temporal edge activation plan…

Social and Information Networks · Computer Science 2025-10-22 Lutz Oettershagen , Othon Michail

While mixture of linear regressions (MLR) is a well-studied topic, prior works usually do not analyze such models for prediction error. In fact, {\em prediction} and {\em loss} are not well-defined in the context of mixtures. In this paper,…

Machine Learning · Statistics 2022-05-27 Avishek Ghosh , Arya Mazumdar , Soumyabrata Pal , Rajat Sen

As AI agents increasingly act on behalf of human stakeholders in economic settings, understanding their behavior in complex market environments becomes critical. This article examines how Large Language Models coordinate on markets that are…

General Economics · Economics 2026-03-11 Alexander Erlei , Lukas Meub

Do workers always work more for more? We investigate how intertemporal and uncompensated labor supply decisions change across observational and experimental windows, within the same workers. Combining a real-effort emoji-counting experiment…

General Economics · Economics 2026-05-29 Mattia Adamo , Michele Cantarella

We consider the effect of parametric uncertainty on properties of Linear Time Invariant systems. Traditional approaches to this problem determine the worst-case gains of the system over the uncertainty set. Whilst such approaches are…

Optimization and Control · Mathematics 2015-05-21 Giorgio Valmorbida , Dhruva Raman , James Anderson

This paper studies distributed Time-Varying Resource Allocation (TVRA) where the local cost functions, global equality constraints, and Local Feasibility Constraints (LFCs) vary with time. Algorithms that mimic the structure of…

Systems and Control · Electrical Eng. & Systems 2025-10-06 Yiqiao Xu , Tengyang Gong , Zhengtao Ding , Alessandra Parisio

Firms increasingly rely on dynamic pricing to respond to evolving customer demand, yet in many applications they observe only the revenue generated by a single posted price in each period. At the same time, market conditions may shift…

Machine Learning · Computer Science 2026-05-21 Xiangyu Yang , Feng Xu , Jian-Qiang Hu , Jiaqiao Hu

Inverse reinforcement learning (IRL) aims to estimate the reward function of optimizing agents by observing their response (estimates or actions). This paper considers IRL when noisy estimates of the gradient of a reward function generated…

Machine Learning · Computer Science 2021-01-19 Vikram Krishnamurthy , George Yin

We propose a multi-agent distributed reinforcement learning algorithm that balances between potentially conflicting short-term reward and sparse, delayed long-term reward, and learns with partial information in a dynamic environment. We…

Machine Learning · Computer Science 2022-04-06 Jing Tan , Ramin Khalili , Holger Karl

Reward engineering, the manual specification of reward functions to induce desired agent behavior, remains a fundamental challenge in multi-agent reinforcement learning. This difficulty is amplified by credit assignment ambiguity,…

Artificial Intelligence · Computer Science 2026-01-14 Haoran Su , Yandong Sun , Congjia Yu

Reinforcement learning with verifiable rewards (RLVR) scales the reasoning ability of large language models (LLMs) but remains bottlenecked by limited labeled samples for continued data scaling. Reinforcement learning with intrinsic rewards…

Machine Learning · Computer Science 2025-10-13 Chuyi Tan , Peiwen Yuan , Xinglin Wang , Yiwei Li , Shaoxiong Feng , Yueqi Zhang , Jiayi Shi , Ji Zhang , Boyuan Pan , Yao Hu , Kan Li

This paper considers the problem of system identification for linear time varying systems. We propose a new system realization approach that uses an "information-state" as the state vector, where the "information-state" is composed of a…

Systems and Control · Electrical Eng. & Systems 2024-04-08 Mohamed Naveed Gul Mohamed , Raman Goyal , Suman Chakravorty , Ran Wang

We study a fair resource scheduling problem, where a set of interval jobs are to be allocated to heterogeneous machines controlled by agents. Each job is associated with release time, deadline, and processing time such that it can be…

Computer Science and Game Theory · Computer Science 2022-01-03 Bo Li , Minming Li , Ruilong Zhang

Safety-aligned LLMs suffer from two failure modes: jailbreak (answering harmful inputs) and over-refusal (declining benign queries). Existing vector steering methods adjust the magnitude of answer vectors, but this creates a fundamental…

Machine Learning · Computer Science 2026-05-05 Haonan Zhang , Dongxia Wang , Yi Liu , Kexin Chen , Wenhai Wang

Reward design plays a pivotal role in aligning large language models (LLMs) with human values, serving as the bridge between feedback signals and model optimization. This survey provides a structured organization of reward modeling and…

Computation and Language · Computer Science 2025-09-03 Miaomiao Ji , Yanqiu Wu , Zhibin Wu , Shoujin Wang , Jian Yang , Mark Dras , Usman Naseem

Reinforcement learning with verifiable rewards (RLVR) has significantly improved reasoning in large language models (LLMs), yet the token-level mechanisms underlying these improvements remain unclear. We present a systematic empirical study…

Computation and Language · Computer Science 2026-03-25 Haoming Meng , Kexin Huang , Shaohang Wei , Chiyu Ma , Shuo Yang , Xue Wang , Guoyin Wang , Bolin Ding , Jingren Zhou

The gain-loss asymmetry, observed in the inverse statistics of stock indices is present for logarithmic return levels that are over $2\%$, and it is the result of the non-Pearson type auto-correlations in the index. These non-Pearson type…

Statistical Finance · Quantitative Finance 2016-08-24 Bulcsú Sándor , Ingve Simonsen , Bálint Zsolt Nagy , Zoltán Néda

This study investigates the short-term asymptotic behavior of the implied volatility surface (IVS), with a particular focus on the at-the-money (ATM) skew and curvature, which are key determinants of the IVS shape and whose are widely…

Pricing of Securities · Quantitative Finance 2025-06-24 Liexin Cheng , Xue Cheng