中文
相关论文

相关论文: Lock-step simulation is child's play

200 篇论文

To enable automated software testing, the ability to automatically navigate to a state of interest and to explore all, or at least sufficient number of, instances of such a state is fundamental. When testing a computer game the problem has…

Modern computational neuroscience strives to develop complex network models to explain dynamics and function of brains in health and disease. This process goes hand in hand with advancements in the theory of neuronal networks and increasing…

Model-free reinforcement learning (RL) requires a large number of trials to learn a good policy, especially in environments with sparse rewards. We explore a method to improve the sample efficiency when we have access to demonstrations. Our…

机器学习 · 计算机科学 2022-04-22 Cinjon Resnick , Roberta Raileanu , Sanyam Kapoor , Alexander Peysakhovich , Kyunghyun Cho , Joan Bruna

In repeated interactions between individuals, we do not expect that exactly the same situation will occur from one time to another. Contrary to what is common in models of repeated games in the literature, most real situations may differ a…

种群与进化 · 定量生物学 2007-05-23 Anders Eriksson , Kristian Lindgren

Reinforcement learning (RL) algorithms, due to their reliance on external systems to learn from, require digital environments (e.g., simulators) with very simple interfaces, which in turn constrain significantly the implementation of such…

编程语言 · 计算机科学 2025-04-29 Massimo Fioravanti , Samuele Pasini , Giovanni Agosta

Formal analyses of incentives for compliance with network protocols often appeal to game-theoretic models and concepts. Applications of game-theoretic analysis to network security have generally been limited to highly stylized models, where…

计算机科学与博弈论 · 计算机科学 2013-06-04 Michael P. Wellman , Tae Hyung Kim , Quang Duong

Learning to program could possibly be analogous to acquiring expertise in abstract mathematics, which may be boring or dull for a majority of students. Thus, among the countless options to approach learning coding [1-14], acquiring concepts…

计算机与社会 · 计算机科学 2019-03-14 Kailash Chandra , Shyamal Suhana Chandra

We introduce a formal notion of masking fault-tolerance between probabilistic transition systems based on a variant of probabilistic bisimulation (named masking simulation). We also provide the corresponding probabilistic game…

计算机科学中的逻辑 · 计算机科学 2022-07-06 Pablo F. Castro , Pedro D'Argenio , Luciano Putruele , Ramiro Demasi

Simulation has become the evaluation method of choice for many areas of distributing computing research. However, most existing simulation packages have several limitations on the size and complexity of the system being modeled. Fine…

分布式、并行与集群计算 · 计算机科学 2011-07-01 Dobre Ciprian , Cristea Valentin , Iosif C. Legrand

Safe and economic operation of networked systems is often challenging. Optimization-based schemes are frequently considered, since they achieve near-optimality while ensuring safety via the explicit consideration of constraints. In…

最优化与控制 · 数学 2024-01-30 Alexander Engelmann , Maisa B. Bandeira , Timm Faulwasser

There are many benefits in providing formal specifications for our software. However, teaching students to do this is not always easy as courses on formal methods are often experienced as dry by students. This paper presents a game called…

Many popular puzzle and matching games have been analyzed through the lens of computational complexity. Prominent examples include Sudoku, Candy Crush, and Flood-It. A common theme among these widely played games is that their generalized…

计算复杂性 · 计算机科学 2026-03-03 Linus Klocker , Simon D. Fink

Lockstep processing is a recognized technique for helping to secure functional-safety relevant processing against, for instance, single upset errors that might cause faulty execution of code. Lockstepping processors does however bind…

硬件体系结构 · 计算机科学 2021-07-20 Hans Dermot Doran , Timo Lang

In this paper, we study a class of bilevel programming problem where the inner objective function is strongly convex. More specifically, under some mile assumptions on the partial derivatives of both inner and outer objective functions, we…

最优化与控制 · 数学 2018-02-08 Saeed Ghadimi , Mengdi Wang

In multi-agent environments in which coordination is desirable, the history of play often causes lock-in at sub-optimal outcomes. Notoriously, technologies with a significant environmental footprint or high social cost persist despite the…

计算机科学与博弈论 · 计算机科学 2020-07-28 Stefanos Leonardos , Iosif Sakos , Costas Courcoubetis , Georgios Piliouras

There is a broad consensus that the inability to form long-term plans is one of the key limitations of current foundational models and agents. However, the existing planning benchmarks remain woefully inadequate to truly measure their…

人工智能 · 计算机科学 2026-04-07 Michael Katz , Harsha Kokel , Sarath Sreedharan

Strategic interactions often take place in an environment rife with uncertainty. As a result, the equilibrium of a game is intimately related to the information available to its players. The \emph{signaling problem} abstracts the task faced…

计算机科学与博弈论 · 计算机科学 2014-10-14 Yu Cheng , Ho Yee Cheung , Shaddin Dughmi , Shanghua Teng

Deep Reinforcement Learning (DRL) has demonstrated impressive results in domains such as games and robotics, where task formulations are well-defined. However, few DRL benchmarks are grounded in complex, real-world environments, where…

机器学习 · 计算机科学 2025-05-13 Henrique Donâncio , Laurent Vercouter , Harald Roclawski

Simulation is used extensively in autonomous systems, particularly in robotic manipulation. By far, the most common approach is to train a controller in simulation, and then use it as an initial starting point for the real system. We…

机器学习 · 统计学 2021-10-06 Shirli Di Castro Shashua , Dotan Di Castro , Shie Mannor

We introduce a formal notion of masking fault-tolerance between probabilistic transition systems using stochastic games. These games are inspired in bisimulation games, but they also take into account the possible faulty behavior of…

计算机科学中的逻辑 · 计算机科学 2023-09-15 Pablo F. Castro , Pedro R. D'Argenio , Ramiro Demasi , Luciano Putruele