English
Related papers

Related papers: Uniform Substitution At One Fell Swoop

200 papers

This paper proposes a payoff perturbation technique for the Mirror Descent (MD) algorithm in games where the gradient of the payoff functions is monotone in the strategy profile space, potentially containing additive noise. The optimistic…

Computer Science and Game Theory · Computer Science 2024-06-25 Kenshi Abe , Kaito Ariu , Mitsuki Sakamoto , Atsushi Iwasaki

Lane changes are complex safety and throughput critical driver actions. Most lane changing models deal with lane-changing maneuvers solely from the merging driver's standpoint and thus ignore driver interaction. To overcome this…

Physics and Society · Physics 2020-08-11 Kyungwon Kang , Hesham A Rakha

We prove the dynamic programming principe for uniformly nondegenerate stochastic differential games in the framework of time-homogeneous diffusion processes considered up to the first exit time from a domain. The zeroth-order "coefficient"…

Optimization and Control · Mathematics 2012-07-17 N. V. Krylov

Post-alignment of large language models (LLMs) is critical in improving their utility, safety, and alignment with human intentions. Direct preference optimisation (DPO) has become one of the most widely used algorithms for achieving this…

Machine Learning · Computer Science 2025-01-06 Rasul Tutnov , Antoine Grosnit , Haitham Bou-Ammar

Significant progress has been recently achieved in developing efficient solutions for simple stochastic games (SSGs), focusing on reachability objectives. While reductions from stochastic parity games (SPGs) to SSGs have been presented in…

Computer Science and Game Theory · Computer Science 2025-06-09 Raphaël Berthon , Joost-Pieter Katoen , Zihan Zhou

A safe Pareto improvement (SPI) [41] is a modification of a game that leaves all players better off with certainty. SPIs are typically proven under qualitative assumptions about the way different games are played. For example, we assume…

Computer Science and Game Theory · Computer Science 2025-11-24 Nathaniel Sauerberg , Caspar Oesterheld

Prioritized experience replay (PER) samples important transitions, rather than uniformly, to improve the performance of a deep reinforcement learning agent. We claim that such prioritization has to be balanced with sample diversity for…

Machine Learning · Computer Science 2020-11-30 Sanghwa Lee , Jaeyoung Lee , Ichiro Hasuo

We study a stopping game of preemption type between two players who both act under uncertain competition. In this framework we introduce, and study the effect of, (i) asymmetry of payoffs, allowing e.g. for different investment costs, and…

Probability · Mathematics 2024-11-08 Erik Ekström , Yuqiong Wang

We develop a flexible stochastic approximation framework for analyzing the long-run behavior of learning in games (both continuous and finite). The proposed analysis template incorporates a wide array of popular learning algorithms,…

Computer Science and Game Theory · Computer Science 2023-07-04 Panayotis Mertikopoulos , Ya-Ping Hsieh , Volkan Cevher

One of the most striking quantum phenomena is superposition, where one particle simultaneously inhabits different states. Most methods to verify coherent superposition are indirect, in that they require the distinct states to be recombined.…

Game-theoretic characterizations of process equivalences traditionally form a central topic in concurrency; for example, most equivalences on the classical linear-time / branching-time spectrum come with such characterizations. Recent work…

Logic in Computer Science · Computer Science 2025-01-28 Jonas Forster , Lutz Schröder , Paul Wild

We derive a primal discontinuous Galerkin (DG) formulation for heterogeneous and anisotropic diffusion, obtained by exact algebraic elimination of the skeletal unknown in a compact hybridized interior penalty (H-IP) method. The resulting…

Numerical Analysis · Mathematics 2026-05-20 Gregory Etangsale , Vincent Fontaine , Anis Younes

This paper reviews the fully complete hypergames model of system $F$, presented a decade ago in the author's thesis. Instantiating type variables is modelled by allowing ``games as moves''. The uniformity of a quantified type variable…

Logic · Mathematics 2008-01-18 Dominic Hughes

This paper investigates a class of general linear-quadratic mean field games with common noise, where the diffusion terms of the system contain the state variables, control variables, and the average state terms. We solve the problem using…

Optimization and Control · Mathematics 2025-08-29 Yu Si , Jingtao Shi

Uniform interpolation is the property that, for any formula and set of atoms, there exists the strongest consequence omitting those atoms. It plays a central role in knowledge representation and reasoning tasks such as knowledge update and…

Logic in Computer Science · Computer Science 2026-03-31 Kexu Wang , Liangda Fang

In this paper we present a matricial result that generalizes Hironaka's game and Perron transforms simultaneously. We also show how one can deduce the various forms in which the algorithm of Perron appears in proofs of local uniformization…

Commutative Algebra · Mathematics 2019-07-05 Michael de Moraes , Josnei Novacoski

In the present paper, based on the previous work (Part I), we present a game semantics for the intensional variant of intuitionistic type theory that refutes the principle of uniqueness of identity proofs and validates the univalence axiom,…

Logic in Computer Science · Computer Science 2016-04-06 Norihiro Yamada

We study the numerical approximation of a time-dependent variational mean field game system with local couplings and either periodic or Neumann boundary conditions. Following a variational approach, we employ a finite difference…

Numerical Analysis · Mathematics 2026-01-06 Heidi Wolles Ljósheim , Dante Kalise , John W. Pearson , Francisco J. Silva

The cornerstone underpinning deep learning is the guarantee that gradient descent on an objective converges to local minima. Unfortunately, this guarantee fails in settings, such as generative adversarial nets, where there are multiple…

Machine Learning · Computer Science 2018-06-07 David Balduzzi , Sebastien Racaniere , James Martens , Jakob Foerster , Karl Tuyls , Thore Graepel

Game theory provides essential analysis in many applications of strategic interactions. However, the question of how to construct a game model and what is its fidelity is seldom addressed. In this work, we consider learning in a class of…

Computer Science and Game Theory · Computer Science 2021-07-30 Yunian Pan , Quanyan Zhu
‹ Prev 1 8 9 10 Next ›