English
Related papers

Related papers: An instance-based learning approach for evaluating…

200 papers

Inverse Reinforcement Learning (IRL) describes the problem of learning an unknown reward function of a Markov Decision Process (MDP) from observed behavior of an agent. Since the agent's behavior originates in its policy and MDP policies…

Artificial Intelligence · Computer Science 2016-04-14 Michael Herman , Tobias Gindele , Jörg Wagner , Felix Schmitt , Wolfram Burgard

Motivated by the observation that overexposure to unwanted marketing activities leads to customer dissatisfaction, we consider a setting where a platform offers a sequence of messages to its users and is penalized when users abandon the…

Machine Learning · Computer Science 2019-03-21 Junyu Cao , Wei Sun

This paper presents an inverse reinforcement learning~(IRL) framework for Bayesian stopping time problems. By observing the actions of a Bayesian decision maker, we provide a necessary and sufficient condition to identify if these actions…

Machine Learning · Computer Science 2023-03-29 Kunal Pattanayak , Vikram Krishnamurthy

Commuter comfort in cab rides affects driver rating as well as the reputation of ride-hailing firms like Uber/Lyft. Existing research has revealed that commuter comfort not only varies at a personalized level but also is perceived…

Human-Computer Interaction · Computer Science 2022-03-23 Rohit Verma , Sugandh Pargal , Debasree Das , Tanusree Parbat , Sai Shankar Kambalapalli , Bivas Mitra , Sandip Chakraborty

We address the problem of inverse reinforcement learning in Markov decision processes where the agent is risk-sensitive. In particular, we model risk-sensitivity in a reinforcement learning framework by making use of models of human…

Machine Learning · Computer Science 2017-11-23 Lillian J. Ratliff , Eric Mazumdar

We study the use of inverse reinforcement learning (IRL) as a tool for the recognition of agents' behavior on the basis of observation of their sequential decision behavior interacting with the environment. We model the problem faced by the…

Machine Learning · Computer Science 2013-03-22 Qifeng Qiao , Peter A. Beling

Temporal credit assignment is crucial for learning and skill development in natural and artificial intelligence. While computational methods like the TD approach in reinforcement learning have been proposed, it's unclear if they accurately…

Artificial Intelligence · Computer Science 2023-07-18 Thuy Ngoc Nguyen , Chase McDonald , Cleotilde Gonzalez

Motivated by applications of the Erlang-B blocking model and the extended $M/M/k/k+N$ model that allows for some queueing, beyond communication networks to sizing and pricing in production, messaging, and app-based parking systems, we study…

Systems and Control · Electrical Eng. & Systems 2025-11-25 Saghar Adler , Mehrdad Moharrami , Vijay Subramanian

Today, GPS-equipped mobile devices are ubiquitous, and they generate Location-Based Service (LBS) data, which has become a critical resource for understanding human mobility. However, inherent limitations in LBS datasets, primarily…

Computational Engineering, Finance, and Science · Computer Science 2024-11-26 Xinhua Wu , Yanchao Wang , Ekin Ugurel , Cynthia Chen , Shuai Huang , Qi R. Wang

We consider a model describing the waiting time of a server alternating between two service points. This model is described by a Lindley-type equation. We are interested in the time-dependent behaviour of this system and derive explicit…

Probability · Mathematics 2014-04-23 Maria Vlasiou , Bert Zwart

For many reinforcement learning (RL) applications, specifying a reward is difficult. This paper considers an RL setting where the agent obtains information about the reward only by querying an expert that can, for example, evaluate…

Machine Learning · Computer Science 2022-02-01 David Lindner , Matteo Turchetta , Sebastian Tschiatschek , Kamil Ciosek , Andreas Krause

The road safety of traffic is greatly affected by the driving performance of online ride-hailing, which has become an increasingly popular travel option for many people. Little attention has been paid to the fact that the use of cell phone…

Computers and Society · Computer Science 2023-11-28 Xiangnan Song , Xianghong Li , Kai Yin , Huimin Qi , Xufei Fang

Online experiments are the gold standard for evaluating impact on user experience and accelerating innovation in software. However, since experiments are typically limited in duration, observed treatment effects are not always permanently…

Human-Computer Interaction · Computer Science 2021-02-26 Soheil Sadeghi , Somit Gupta , Stefan Gramatovici , Jiannan Lu , Hao Ai , Ruhan Zhang

In human-robot interaction (HRI) systems, such as autonomous vehicles, understanding and representing human behavior are important. Human behavior is naturally rich and diverse. Cost/reward learning, as an efficient way to learn and…

Robotics · Computer Science 2020-08-24 Liting Sun , Zheng Wu , Hengbo Ma , Masayoshi Tomizuka

Strategic customer behavior is strongly influenced by the level of information that is provided to customers. Hence, to optimize the design of queueing systems, many studies consider various versions of the same service model and compare…

Optimization and Control · Mathematics 2019-06-14 Yiannis Dimitrakopoulos , Antonis Economou , Stefanos Leonardos

Ride sharing services combine trips of multiple users in the same vehicle and may provide more sustainable transport than private cars. As mobility demand varies during the day, the travel times experienced by passengers may substantially…

Physics and Society · Physics 2023-06-26 Charlotte Lotze , Philip Marszal , Malte Schröder , Marc Timme

We study a vacation-type queueing model, and a single-server multi-queue polling model, with the special feature of retrials. Just before the server arrives at a station there is some deterministic glue period. Customers (both new arrivals…

Probability · Mathematics 2016-02-03 Murtuza Ali Abidini , Onno Boxma , Jacques Resing

The study examines heterogeneity in travel behaviour among ride-hailing services (RHS) users by including attitudes, in order to reinforce conventional user-segmentation approaches. Simultaneously, prioritization of ride-hailing specific…

Applications · Statistics 2022-08-17 Eeshan Bhaduri , Shagufta Pal , Arkopal Kishore Goswami

In urban settings, bus transit stands as a significant mode of public transportation, yet faces hurdles in delivering accurate and reliable arrival times. This discrepancy often culminates in delays and a decline in ridership, particularly…

Machine Learning · Computer Science 2024-03-05 Narges Rashvand , Sanaz Sadat Hosseini , Mona Azarbayjani , Hamed Tabkhi

A fundamental question in any peer-to-peer ride-sharing system is how to, both effectively and efficiently, meet the request of passengers to balance the supply and demand in real time. On the passenger side, traditional approaches focus on…

Machine Learning · Computer Science 2022-11-08 Yanqiu Wu , Qingyang Li , Zhiwei Qin
‹ Prev 1 3 4 5 6 7 10 Next ›