中文
相关论文

相关论文: Improving Efficiency in Near-State and State-Optim…

200 篇论文

We study here the problem of determining the majority type in an arbitrary connected network, each vertex of which has initially two possible types. The vertices may have a few additional possible states and can interact in pairs only if…

分布式、并行与集群计算 · 计算机科学 2014-05-01 George B. Mertzios , Sotiris E. Nikoletseas , Christoforos L. Raptopoulos , Paul G. Spirakis

We consider a discrete-time system comprising a first-come-first-served queue, a non-preemptive server, and a stationary non-work-conserving scheduler. New tasks enter the queue according to a Bernoulli process with a pre-specified arrival…

应用统计 · 统计学 2020-08-05 Michael Lin , Nuno C. Martins , Richard J. La

Among the many variants of RL, an important class of problems is where the state and action spaces are continuous -- autonomous robots, autonomous vehicles, optimal control are all examples of such problems that can lend themselves…

人工智能 · 计算机科学 2024-10-16 Rajesh Mangannavar , Gopalakrishnan Srinivasaraghavan

Control of continuous time dynamics with multiplicative noise is a classic topic in stochastic optimal control. This work addresses the problem of designing infinite horizon optimal controls with stability guarantees for \textit{a single…

最优化与控制 · 数学 2020-10-02 Kaivalya Bakshi , Evangelos A. Theodorou , Piyush Grover

Consider a discrete-time linear time-invariant descriptor system $Ex(k+1)=Ax(k)$ for $k \in \mathbb Z_{+}$. In this paper, we tackle for the first time the problem of stabilizing such systems by computing a nearby regular index one stable…

最优化与控制 · 数学 2019-10-11 Nicolas Gillis , Michael Karow , Punit Sharma

This paper addresses the problem of consensus tracking with fixed-time convergence, for leader-follower multi-agent systems with double-integrator dynamics, where only a subset of followers has access to the state of the leader. The control…

系统与控制 · 电气工程与系统科学 2026-02-19 Miguel A. Trujillo , Rodrigo Aldana-López , David Gomez Gutierrez , Michael Defoort , Javier Ruiz Leon , Hector M. Becerra

Reinforcement learning (RL) has recently proven itself as a powerful instrument for solving complex problems and even surpassed human performance in several challenging applications. This signifies that RL algorithms can be used in the…

机器学习 · 计算机科学 2023-03-07 Ahmet Semih Tasbas , Safa Onur Sahin , Nazim Kemal Ure

Automated decision-making tools increasingly assess individuals to determine if they qualify for high-stakes opportunities. A recent line of research investigates how strategic agents may respond to such scoring tools to receive favorable…

机器学习 · 计算机科学 2021-10-28 Keegan Harris , Hoda Heidari , Zhiwei Steven Wu

Stochastic resetting, the procedure of stopping and re-initializing random processes, has recently emerged as a powerful tool for accelerating processes ranging from queuing systems to molecular simulations. However, its usefulness is…

统计力学 · 物理学 2025-03-18 Tommer D. Keidar , Ofir Blumer , Barak Hirshberg , Shlomi Reuveni

In this article, bipartite ranking, a statistical learning problem involved in many applications and widely studied in the passive context, is approached in a much more general \textit{active setting} than the discrete one previously…

机器学习 · 统计学 2026-03-02 James Cheshire , Stephan Clémençon

In this paper, we mainly investigate an integrated system operating under a software defined network (SDN) protocol. SDN is a new networking paradigm in which network intelligence is centrally administered and data is communicated via…

最优化与控制 · 数学 2018-12-04 Cheng Tan , Wing Shing Wong , Huanshui Zhang

This paper examines the objective of optimally harvesting a single species in a stochastic environment. This problem has previously been analyzed in Alvarez (2000) using dynamic programming techniques and, due to the natural payoff…

最优化与控制 · 数学 2016-08-02 Richard H. Stockbridge , Chao Zhu

We propose a new sequential decision-making setting, combining key aspects of two established online learning problems with bandit feedback. The optimal action to play at any given moment is contingent on an underlying changing state which…

机器学习 · 计算机科学 2023-11-07 Alexander Galozy , Slawomir Nowaczyk , Mattias Ohlsson

Motivated by growing evidence of agents' mistakes in strategically simple environments, we propose a solution concept -- robust equilibrium -- that requires only an asymptotically optimal behavior. We use it to study large random matching…

理论经济学 · 经济学 2023-09-26 Georgy Artemov , Yeon-Koo Che , YingHua He

In this paper, containment control of multi-agent systems with measurement noises is studied under directed networks. When the leaders are stationary, a stochastic approximation type protocol is employed to solve the containment control of…

系统与控制 · 计算机科学 2014-05-20 Yuanshi Zheng , Tao Li , Long Wang

We consider the classic problem of establishing a statistical ranking of a set of n items given a set of inconsistent and incomplete pairwise comparisons between such items. Instantiations of this problem occur in numerous applications in…

机器学习 · 计算机科学 2015-04-07 Mihai Cucuringu

We study an optimal-control problem of polling systems with large switchover times, when a holding cost is incurred on the queues. In particular, we consider a stochastic network with a single server that switches between several buffers…

概率论 · 数学 2020-09-01 Yue Hu , Jing Dong , Ohad Perry

We consider an agent interacting with an environment in a single stream of actions, observations, and rewards, with no reset. This process is not assumed to be a Markov Decision Process (MDP). Rather, the agent has several representations…

机器学习 · 计算机科学 2013-03-19 Odalric-Ambrym Maillard , Phuong Nguyen , Ronald Ortner , Daniil Ryabko

In this paper we consider prioritized maximal scheduling in multi-hop wireless networks, where the scheduler chooses a maximal independent set greedily according to a sequence specified by certain priorities. We show that if the probability…

信息论 · 计算机科学 2009-01-20 Qiao Li , Rohit Negi

This paper aims to provide a methodology for generating autonomous and non-autonomous systems with a fixed-time stable equilibrium point where an Upper Bound of the Settling Time (UBST) is set a priori as a parameter of the system. In…

‹ 上一页 1 8 9 10 下一页 ›