中文
相关论文

相关论文: Smart Containers With Bidding Capacity: A Policy G…

200 篇论文

Policy gradient is a generic and flexible reinforcement learning approach that generally enjoys simplicity in analysis, implementation, and deployment. In the last few decades, this approach has been extensively advanced for fully…

机器学习 · 计算机科学 2020-05-26 Kamyar Azizzadenesheli , Yisong Yue , Animashree Anandkumar

Robust control policy learning for autonomous driving requires training environments to be both physically realistic and computationally scalable, properties that existing simulators provide only in isolation. We introduce Sim2Sim2Sim, a…

机器人学 · 计算机科学 2026-05-05 Xunjiang Gu , Kashyap Chitta , Mahsa Golchoubian , Vladimir Suplin , Igor Gilitschenski

With a novel search algorithm or assortment planning or assortment optimization algorithm that takes into account a Bayesian approach to information updating and two-stage assortment optimization techniques, the current research provides a…

理论经济学 · 经济学 2023-07-19 Dipankar Das

Shared micromobility systems, such as electric scooters and bikes, have gained widespread popularity as sustainable alternatives to traditional transportation modes. However, these systems face persistent challenges due to spatio-temporal…

多智能体系统 · 计算机科学 2025-10-07 Heng Tan , Hua Yan , Lucas Yang , Yu Yang

Resource allocation in distributed and networked systems such as the Cloud is becoming increasingly flexible, allowing these systems to dynamically adjust toward the workloads they serve, in a demand-aware manner. Online balanced…

数据结构与算法 · 计算机科学 2024-10-24 Harald Räcke , Stefan Schmid , Ruslan Zabrodin

We consider a freight platform that serves as an intermediary between shippers and carriers in a truckload transportation network. The platform's objective is to design a policy that determines prices for shippers and payments to carriers,…

最优化与控制 · 数学 2025-03-07 Ruoran Chen , Sungwoo Kim , He Wang , Xuan Wang

A smart grid connects wind or solar or storage farms, fossil fuel plants, industrialor commercial loads, or load serving entities, modeled as stochastic dynamical systems. In each time period, they consume or supply electrical energy, with…

最优化与控制 · 数学 2016-06-30 Rahul Singh , P. R. Kumar , Le Xie

Bin covering is a dual version of classic bin packing. Thus, the goal is to cover as many bins as possible, where covering a bin means packing items of total size at least one in the bin. For online bin covering, competitive analysis fails…

数据结构与算法 · 计算机科学 2014-02-28 Marie G. Christ , Lene M. Favrholdt , Kim S. Larsen

We formulate offloading of computational tasks from a dynamic group of mobile agents (e.g., cars) as decentralized decision making among autonomous agents. We design an interaction mechanism that incentivizes such agents to align private…

多智能体系统 · 计算机科学 2022-08-11 Jing Tan , Ramin Khalili , Holger Karl , Artur Hecker

We propose a distributed bidding-aided Matern carrier sense multiple access (CSMA) policy for device-to-device (D2D) content distribution. The network is composed of D2D receivers and potential D2D transmitters, i.e., transmitters are…

信息论 · 计算机科学 2017-10-17 Derya Malak , Mazin Al-Shalash , Jeffrey G. Andrews

Many cluster management systems (CMSs) have been proposed to share a single cluster with multiple distributed computing systems. However, none of the existing approaches can handle distributed machine learning (ML) workloads given the…

分布式、并行与集群计算 · 计算机科学 2017-06-19 Peng Sun , Yonggang Wen , Ta Nguyen Binh Duong , Shengen Yan

As the industry of autonomous driving grows, so does the potential interaction of groups of autonomous cars. Combined with the advancement of Artificial Intelligence and simulation, such groups can be simulated, and safety-critical models…

机器学习 · 计算机科学 2024-02-22 Omar Tanner

As service robots become more and more capable of performing useful tasks for us, there is a growing need to teach robots how we expect them to carry out these tasks. However, different users typically have their own preferences, for…

机器人学 · 计算机科学 2015-12-22 Nichola Abdo , Cyrill Stachniss , Luciano Spinello , Wolfram Burgard

Policy-gradient methods have received increased attention recently as a mechanism for learning to act in partially observable environments. They have shown promise for problems admitting memoryless policies but have been less successful…

机器学习 · 计算机科学 2025-12-08 Douglas Aberdeen , Jonathan Baxter

The idea of reusing or transferring information from previously learned tasks (source tasks) for the learning of new tasks (target tasks) has the potential to significantly improve the sample efficiency of a reinforcement learning agent. In…

人工智能 · 计算机科学 2022-09-28 Thommen George Karimpanal , Roland Bouffanais

In targeted online advertising, advertisers look for maximizing campaign performance under delivery constraint within budget schedule. Most of the advertisers typically prefer to impose the delivery constraint to spend budget smoothly over…

人工智能 · 计算机科学 2015-06-22 Jian Xu , Kuang-chih Lee , Wentong Li , Hang Qi , Quan Lu

Sharing economy is a transformative socio-economic phenomenon built around the idea of sharing underused resources and services, e.g. transportation and housing, thereby reducing costs and extracting value. Anticipating continued reduction…

系统与控制 · 计算机科学 2018-07-24 Pratyush Chakraborty , Enrique Baeyens , Kameshwar Poolla , Pramod P. Khargonekar , Pravin Varaiya

SmartFlow is a multi-layered framework that integrates Reinforcement Learning and Agentic AI to address the dynamic rebalancing problem in urban bike-sharing services. Its architecture separates strategic, tactical, and communication…

In this paper, spectrum access in cognitive radio networks is modeled as a repeated auction game subject to monitoring and entry costs. For secondary users, sensing costs are incurred as the result of primary users' activity. Furthermore,…

信息论 · 计算机科学 2009-10-14 Zhu Han , Rong Zheng , Vincent H. Poor

The capacity regions of semideterministic multiuser channels, such as the semideterministic relay channel and the multiple access channel with partially cribbing encoders, have been characterized using the idea of partial-decode-forward.…

信息论 · 计算机科学 2016-11-17 Ritesh Kolte , Ayfer Özgür , Haim Permuter
‹ 上一页 1 8 9 10 下一页 ›