English
Related papers

Related papers: Multi-Attribute Utility Preference Robust Optimiza…

200 papers

In this paper, the problem of uplink (UL) and downlink (DL) resource optimization, mode selection and power allocation is studied for wireless cellular networks under the assumption of in-band full duplex (IBFD) base stations,…

Information Theory · Computer Science 2017-06-20 Mohammed S. Elbamby , Mehdi Bennis , Walid Saad , Mérouane Debbah , Matti Latva-aho

We study multi-objective reinforcement learning with nonlinear preferences over trajectories. That is, we maximize the expected value of a nonlinear function over accumulated rewards (expected scalarized return or ESR) in a multi-objective…

Machine Learning · Computer Science 2025-02-19 Nianli Peng , Muhang Tian , Brandon Fain

We consider the classical multi-asset Merton investment problem under drift uncertainty, i.e. the asset price dynamics are given by geometric Brownian motions with constant but unknown drift coefficients. The investor assumes a prior drift…

Portfolio Management · Quantitative Finance 2024-02-22 Nicole Bäuerle , Antje Mahayni

We study active preference learning as a framework for intuitively specifying the behaviour of autonomous robots. In active preference learning, a user chooses the preferred behaviour from a set of alternatives, from which the robot learns…

Robotics · Computer Science 2020-09-30 Nils Wilde , Dana Kulic , Stephen L. Smith

A major challenge in robotics is to design robust policies which enable complex and agile behaviors in the real world. On one end of the spectrum, we have model-free reinforcement learning (MFRL), which is incredibly flexible and general…

Robotics · Computer Science 2024-10-01 Jacob Sacks , Rwik Rana , Kevin Huang , Alex Spitzer , Guanya Shi , Byron Boots

The existence of optimal strategy in robust utility maximization is addressed when the utility function is finite on the entire real line. A delicate problem in this case is to find a "good definition" of admissible strategies, so that an…

Portfolio Management · Quantitative Finance 2012-10-16 Keita Owari

Aligning Large Language Models (LLMs) to human preferences in content, style, and presentation is challenging, in part because preferences are varied, context-dependent, and sometimes inherently ambiguous. While successful, Reinforcement…

Machine Learning · Computer Science 2024-10-29 Sam Houliston , Alizée Pace , Alexander Immer , Gunnar Rätsch

This paper proposes a new robust optimization (RO) formulation namely the RO under objective functional uncertainty (ObRO). The ObRO adopts a min-max structure where the inner problem finds the worst-case objective function in a continuous…

Optimization and Control · Mathematics 2026-05-19 Yue Song , Yuxi Lu , Gang Li , Kairui Feng , Qi Liu

For aligning large language models (LLMs), prior work has leveraged reinforcement learning via human feedback (RLHF) or variations of direct preference optimization (DPO). While DPO offers a simpler framework based on maximum likelihood…

Artificial Intelligence · Computer Science 2025-05-27 Anirudhan Badrinath , Prabhat Agarwal , Jiajing Xu

We study a general model on reusable resource allocation under model uncertainty. A heterogeneous population of customers arrive at the decision maker's (DM's) platform sequentially. Upon observing a customer's type, the DM selects an…

Optimization and Control · Mathematics 2022-12-07 Xilin Zhang , Wang Chi Cheung

A differentially private selection algorithm outputs from a finite set the item that approximately maximizes a data-dependent quality function. The most widely adopted mechanisms tackling this task are the pioneering exponential mechanism…

Cryptography and Security · Computer Science 2022-08-05 Gonzalo Munilla Garrido , Florian Matthes

It is typically understood that the training of modern neural networks is a process of fitting the probability distribution of desired output. However, recent paradoxical observations in a number of language generation tasks let one wonder…

Machine Learning · Computer Science 2023-05-31 Huang Bojun , Fei Yuan

Machine learning has recently been widely adopted to address the managerial decision making problems, in which the decision maker needs to be able to interpret the contributions of individual attributes in an explicit form. However, there…

Machine Learning · Computer Science 2019-10-28 Mengzhuo Guo , Qingpeng Zhang , Xiuwu Liao , Frank Youhua Chen , Daniel Dajun Zeng

We consider decision-making problems under decision-dependent uncertainty (DDU), where the distribution of uncertain parameters depends on the decision variables and is only observable through a finite offline dataset. To address this…

Optimization and Control · Mathematics 2025-08-12 Chengrui Qu , Huiwen Jia , Pengcheng You

Distributionally Robust Optimisation (DRO) protects risk-averse decision-makers by considering the worst-case risk within an ambiguity set of distributions based on the empirical distribution or a model. To further guard against finite,…

Machine Learning · Statistics 2025-05-07 Charita Dellaporta , Patrick O'Hara , Theodoros Damoulas

Dybvig (1988a,b) solves in a complete market setting the problem of finding a payoff that is cheapest possible in reaching a given target distribution ("cost-efficient payoff"). In the presence of ambiguity, the distribution of a payoff is,…

Portfolio Management · Quantitative Finance 2023-08-11 Carole Bernard , Gero Junike , Thibaut Lux , Steven Vanduffel

In this paper, we propose a modified polyhedral method to elicit a decision maker's (DM's) nonlinear univariate utility function, which does not rely on explicit information about the shape structure, Lipschitz modulus, and the inflection…

Optimization and Control · Mathematics 2025-04-01 Sainan Zhang , Shaoyan Guo , Melvyn Sim , Huifu Xu

We develop efficient algorithms to construct utility maximizing mechanisms in the presence of risk averse players (buyers and sellers) in Bayesian settings. We model risk aversion by a concave utility function, and players play…

Computer Science and Game Theory · Computer Science 2012-06-28 Anand Bhalgat , Tanmoy Chakraborty , Sanjeev Khanna

Mathematical reasoning presents a significant challenge for Large Language Models (LLMs) as it requires ensuring the correctness of each reasoning step. Researchers have been strengthening the mathematical reasoning abilities of LLMs…

Machine Learning · Computer Science 2025-06-23 Yunze Lin

This paper deals with optimal policy learning (OPL) with observational data, i.e. data-driven optimal decision-making, in multi-action (or multi-arm) settings, where a finite set of decision options is available. It is organized in three…

Machine Learning · Statistics 2024-04-01 Giovanni Cerulli
‹ Prev 1 4 5 6 7 8 10 Next ›