中文
相关论文

相关论文: Recovering Markov Models from Closed-Loop Data

200 篇论文

Recommender systems shape individual choices through feedback loops in which user behavior and algorithmic recommendations coevolve over time. The systemic effects of these loops remain poorly understood, in part due to unrealistic…

We present an algorithm that can efficiently compute a broad class of inferences for discrete-time imprecise Markov chains, a generalised type of Markov chains that allows one to take into account partially specified probabilities and other…

概率论 · 数学 2019-07-02 Natan T'Joens , Thomas Krak , Jasper De Bock , Gert de Cooman

In this paper we present a method for reformulating the Recommender Systems problem in an Information Retrieval one. In our tests we have a dataset of users who give ratings for some movies; we hide some values from the dataset, and we try…

信息检索 · 计算机科学 2011-06-03 Alberto Costa , Fabio Roda

Autonomous agents are often tasked with operating in an area where feedback is unavailable. Inspired by such applications, this paper develops a novel switched systems-based control method for uncertain nonlinear systems with temporary loss…

系统与控制 · 计算机科学 2018-03-16 Hsi-Yuan Chen , Zachary I. Bell , Patryk Deptula , Warren E. Dixon

The innovations algorithm is a classical recursive forecasting algorithm used in time series analysis. We develop the innovations algorithm for a class of nonnegative regularly varying time series models constructed via transformed-linear…

统计理论 · 数学 2023-09-20 Nehali Mhatre , Daniel Cooley

Recommender systems have been widely adopted by electronic commerce and entertainment industries for individualized prediction and recommendation, which benefit consumers and improve business intelligence. In this article, we propose an…

机器学习 · 统计学 2017-11-07 Xuan Bi , Annie Qu , Xiaotong Shen

We consider online reinforcement learning in episodic Markov decision process (MDP) with unknown transition function and stochastic rewards drawn from some fixed but unknown distribution. The learner aims to learn the optimal policy and…

机器学习 · 计算机科学 2024-03-12 Vincent Leon , S. Rasoul Etesami

Conversational recommender systems (CRS) increasingly rely on user simulators for automated evaluation of sales agents. A key requirement for such simulators is the ability to model human decision-making. However, most existing simulation…

信息检索 · 计算机科学 2026-05-08 Yuan-Chi Li , Li-Chi Chen , Sung-Yi Wu , Yu-Che Tsai , Shou-De Lin

This paper considers a half-duplex scenario where an interferer behaves according to a parametric model but the values of the model parameters are unknown. We explore the necessary number of sensing steps to gather sufficient knowledge…

信息论 · 计算机科学 2024-10-11 Vincent Corlay , Jean-Christophe Sibel , Nicolas Gresset

State-space models (SSM) with Markov switching offer a powerful framework for detecting multiple regimes in time series, analyzing mutual dependence and dynamics within regimes, and asserting transitions between regimes. These models…

统计方法学 · 统计学 2021-06-14 David Degras , Chee-Ming Ting , Hernando Ombao

We study reinforcement learning from human feedback in general Markov decision processes, where agents learn from trajectory-level preference comparisons. A central challenge in this setting is to design algorithms that select informative…

机器学习 · 计算机科学 2025-12-05 Andreas Schlaginhaufen , Reda Ouhamma , Maryam Kamgarpour

Conversational information access is an emerging research area. Currently, human evaluation is used for end-to-end system evaluation, which is both very time and resource intensive at scale, and thus becomes a bottleneck of progress. As an…

信息检索 · 计算机科学 2020-06-17 Shuo Zhang , Krisztian Balog

This paper studies the performative prediction problem which optimizes a stochastic loss function with data distribution that depends on the decision variable. We consider a setting where the agent(s) provides samples adapted to the…

最优化与控制 · 数学 2021-10-05 Qiang Li , Hoi-To Wai

Our work revisits the design of mechanisms via the learning-augmented framework. In this model, the algorithm is enhanced with imperfect (machine-learned) information concerning the input, usually referred to as prediction. The goal is to…

计算机科学与博弈论 · 计算机科学 2024-10-29 George Christodoulou , Alkmini Sgouritsa , Ioannis Vlachos

Dynamic and evolving operational and economic environments present significant challenges for decision-making. We explore a simulation optimization problem characterized by non-stationary input distributions with regime-switching dynamics…

最优化与控制 · 数学 2025-08-19 Jianglin Xia , Haowei Wang , Songhao Wang , Szu Hui Ng

User simulation is increasingly vital to develop and evaluate recommender systems (RSs). While Large Language Models (LLMs) offer promising avenues to simulate user behavior, they often struggle with the absence of specific task alignment…

人机交互 · 计算机科学 2026-04-20 Tianjun Wei , Huizhong Guo , Yingpeng Du , Zhu Sun , Huang Chen , Dongxia Wang , Jie Zhang

Recommender systems are embracing conversational technologies to obtain user preferences dynamically, and to overcome inherent limitations of their static models. A successful Conversational Recommender System (CRS) requires proper handling…

信息检索 · 计算机科学 2020-02-24 Wenqiang Lei , Xiangnan He , Yisong Miao , Qingyun Wu , Richang Hong , Min-Yen Kan , Tat-Seng Chua

Markov-switching models are a powerful tool for modelling time series data that are driven by underlying latent states. As such, they are widely used in behavioural ecology, where discrete states can serve as proxies for behavioural modes…

统计方法学 · 统计学 2025-08-26 Jan-Ole Koslik

Pedestrian trajectory prediction remains a challenge for autonomous systems, particularly due to the intricate dynamics of social interactions. Accurate forecasting requires a comprehensive understanding not only of each pedestrian's…

计算机视觉与模式识别 · 计算机科学 2024-12-09 Haleh Damirchi , Ali Etemad , Michael Greenspan

Using supervised machine learning approaches to recognize human activities from on-body wearable accelerometers generally requires a large amount of labelled data. When ground truth information is not available, too expensive, time…

机器学习 · 统计学 2013-12-30 Dorra Trabelsi , Samer Mohammed , Faicel Chamroukhi , Latifa Oukhellou , Yacine Amirat