English
Related papers

Related papers: Asymmetric Release Planning-Compromising Satisfact…

200 papers

In the field of large language models (LLMs), aligning models with the diverse preferences of users is a critical challenge. Direct Preference Optimization (DPO) has played a key role in this area. It works by using pairs of preferences…

Computation and Language · Computer Science 2024-05-29 Yueqin Yin , Zhendong Wang , Yi Gu , Hai Huang , Weizhu Chen , Mingyuan Zhou

Efficiently solving multi-objective optimization problems for simulation optimization of important scientific and engineering applications such as materials design is becoming an increasingly important research topic. This is due largely to…

Artificial Intelligence · Computer Science 2023-06-27 Eric Hans Lee , Bolong Cheng , Michael McCourt

This paper aims to design a set of transmitting waveforms in cognitive colocated Multi-Input Multi-Output (MIMO) radar systems considering the simultaneous minimization of spatial- and the range- Integrated Sidelobe Level Ratio (ISLR). The…

Signal Processing · Electrical Eng. & Systems 2021-07-07 Ehsan Raei , Mohammad Alaee-Kerahrood , M. R. Bhavani Shankar

The Constraint Satisfaction Problem (CSP) framework offers a simple and sound basis for representing and solving simple decision problems, without uncertainty. This paper is devoted to an extension of the CSP framework enabling us to deal…

Artificial Intelligence · Computer Science 2013-02-21 Helene Fargier , Jerome Lang , Roger Martin-Clouaire , Thomas Schiex

Direct Preference Optimization (DPO) has emerged as a more computationally efficient alternative to Reinforcement Learning from Human Feedback (RLHF) with Proximal Policy Optimization (PPO), eliminating the need for reward models and online…

Computation and Language · Computer Science 2024-10-28 Xin Mao , Feng-Lin Li , Huimin Xu , Wei Zhang , Wang Chen , Anh Tuan Luu

Asynchronous Partial Overlay (APO) is a search algorithm that uses cooperative mediation to solve Distributed Constraint Satisfaction Problems (DisCSPs). The algorithm partitions the search into different subproblems of the DisCSP. The…

Artificial Intelligence · Computer Science 2014-01-16 Tal Grinshpoun , Amnon Meisels

In this paper we consider two problems regarding the scheduling of available personnel in order to perform a given quantity of work, which can be arbitrarily decomposed into a sequence of activities. We are interested in schedules which…

Data Structures and Algorithms · Computer Science 2013-01-22 Mugurel Ionut Andreica , Romulus Andreica , Angela Andreica

Distributionally robust optimization (DRO) has shown lot of promise in providing robustness in learning as well as sample based optimization problems. We endeavor to provide DRO solutions for a class of sum of fractionals, non-convex…

Machine Learning · Computer Science 2022-06-01 Avinandan Bose , Arunesh Sinha , Tien Mai

We present the Integrated Size and Price Optimization Problem (ISPO) for a fashion discounter with many branches. Based on a two-stage stochastic programming model with recourse, we develop an exact algorithm and a production-compliant…

Optimization and Control · Mathematics 2014-02-03 Miriam Kießling , Sascha Kurz , Jörg Rambau

To ensure a successful bid while maximizing of profits, generation companies (GENCOs) need a self-scheduling strategy that can cope with a variety of scenarios. So distributionally robust opti-mization (DRO) is a good choice because that it…

Optimization and Control · Mathematics 2021-05-05 Linfeng Yang , Ying Yang , Guo Chen , Zhaoyang Dong

Large-scale text-to-image foundation models have achieved remarkable visual realism, yet generating human images with correct anatomical structures remains challenging. Existing approaches enforce anatomical constraints through…

Computer Vision and Pattern Recognition · Computer Science 2026-05-26 Bao Li , Yuliang Xiu , Zhen Liu

Human preference alignment is critical in building powerful and reliable large language models (LLMs). However, current methods either ignore the multi-dimensionality of human preferences (e.g. helpfulness and harmlessness) or struggle with…

Machine Learning · Computer Science 2024-10-14 Xingzhou Lou , Junge Zhang , Jian Xie , Lifeng Liu , Dong Yan , Kaiqi Huang

Direct Alignment Algorithms (DAAs), such as Direct Preference Optimisation (DPO) and Identity Preference Optimisation (IPO), have emerged as alternatives to online Reinforcement Learning from Human Feedback (RLHF) algorithms such as…

Computation and Language · Computer Science 2024-10-21 Zhengyan Shi , Sander Land , Acyr Locatelli , Matthieu Geist , Max Bartolo

The Joint Replenishment Problem (JRP) is a fundamental optimization problem in supply-chain management, concerned with optimizing the flow of goods from a supplier to retailers. Over time, in response to demands at the retailers, the…

Data Structures and Algorithms · Computer Science 2015-12-04 Marcin Bienkowski , Jaroslaw Byrka , Marek Chrobak , Neil Dobbs , Tomasz Nowicki , Maxim Sviridenko , Grzegorz Swirszcz , Neal E. Young

Finding the optimal solution is often the primary goal in combinatorial optimization (CO). However, real-world applications frequently require diverse solutions rather than a single optimum, particularly in two key scenarios. The first…

Machine Learning · Statistics 2025-08-15 Yuma Ichikawa , Hiroaki Iwashita

Decision-making pipelines are generally characterized by tradeoffs among various risk functions. It is often desirable to manage such tradeoffs in a data-adaptive manner. As we demonstrate, if this is done naively, state-of-the art…

In a unified system of passive radar and communication systems of joint transmitter platform, information intended for a communication receiver may be eavesdropped by a passive radar receiver (RR), thereby undermining the security of…

Information Theory · Computer Science 2018-08-30 Batu K. Chalise , Moeness G. Amin

We develop a framework to optimize the tradeoff between diversity, multiplexing, and delay in MIMO systems to minimize end-to-end distortion. We first focus on the diversity-multiplexing tradeoff in MIMO systems, and develop analytical…

Information Theory · Computer Science 2008-04-09 Tim Holliday , Andrea J. Goldsmith , H. Vincent Poor

To address efficiency and design challenges in choice-based matching platforms, we introduce a two-sided assortment optimization framework under general choice preferences. The goal in this problem is to maximize the expected number of…

Optimization and Control · Mathematics 2026-05-08 Omar El Housni , Ulysse Hennebelle , Alfredo Torrico

We propose Soft Preference Optimization (SPO), a method for aligning generative models, such as Large Language Models (LLMs), with human preferences, without the need for a reward model. SPO optimizes model outputs directly over a…

Machine Learning · Computer Science 2024-10-07 Arsalan Sharifnassab , Saber Salehkaleybar , Sina Ghiassian , Surya Kanoria , Dale Schuurmans