English
Related papers

Related papers: Foundations and Fundamental Properties of a Two-Pa…

200 papers

The memory model for RISC-V, a newly developed open source ISA, has not been finalized yet and thus, offers an opportunity to evaluate existing memory models. We believe RISC-V should not adopt the memory models of POWER or ARM, because…

Programming Languages · Computer Science 2018-09-20 Sizhuo Zhang , Muralidaran Vijayaraghavan , Arvind

Traditional memory writing operations proceed one bit at a time, where e.g. an individual magnetic domain is force-flipped by a localized external field. One way to increase material storage capacity would be to write several bits at a time…

Soft Condensed Matter · Physics 2022-05-10 Théo Jules , Laura Michel , Adèle Douin , Frédéric Lechenault

The exponential moving average (EMA) is a commonly used statistic for providing stable estimates of stochastic quantities in deep learning optimization. Recently, EMA has seen considerable use in generative models, where it is computed with…

Machine Learning · Computer Science 2023-10-24 Jonathan Patsenker , Henry Li , Yuval Kluger

Supervised operator learning centers on the use of training data, in the form of input-output pairs, to estimate maps between infinite-dimensional spaces. It is emerging as a powerful tool to complement traditional scientific computing,…

Machine Learning · Computer Science 2024-08-14 Nicholas H. Nelsen , Andrew M. Stuart

Neural network based data-driven operator learning schemes have shown tremendous potential in computational mechanics. DeepONet is one such neural network architecture which has gained widespread appreciation owing to its excellent…

Machine Learning · Statistics 2022-06-14 Shailesh Garg , Souvik Chakraborty

Motivated by recent experimental studies in microbiology, we suggest a modification of the classic ballistic deposition model of surface growth, where the memory of a deposition at a site induces more depositions at that site or its…

Cellular Automata and Lattice Gases · Physics 2022-02-24 Ahmed Roman , Ruomin Zhu , Ilya Nemenman

Deep neural networks have excelled on a wide range of problems, from vision to language and game playing. Neural networks very gradually incorporate information into weights as they process data, requiring very low learning rates. If the…

We characterize how memorization is represented in transformer models and show that it can be disentangled in the weights of both language models (LMs) and vision transformers (ViTs) using a decomposition based on the loss landscape…

Computation and Language · Computer Science 2025-11-03 Jack Merullo , Srihita Vatsavaya , Lucius Bushnaq , Owen Lewis

Large language model agents increasingly depend on memory to sustain long horizon interaction, but existing frameworks remain limited. Most expose only a few basic primitives such as encode, retrieve, and delete, while higher order…

Computation and Language · Computer Science 2025-10-24 Yi Wang , Lihai Yang , Boyu Chen , Gongyi Zou , Kerun Xu , Bo Tang , Feiyu Xiong , Siheng Chen , Zhiyu Li

We consider the model equation arising in the theory of viscoelasticity $$\partial_{tt} u-h_t(0)\Delta u -\int_{0}^\infty h_t'(s)\Delta u(t-s)d s+ f(u) = g.$$ Here, the main feature is that the memory kernel $h_t(\cdot)$ depends on time,…

Dynamical Systems · Mathematics 2016-03-24 Monica Conti , Valeria Danese , Claudio Giorgi , Vittorino Pata

We study episodic reinforcement learning in non-stationary linear (a.k.a. low-rank) Markov Decision Processes (MDPs), i.e, both the reward and transition kernel are linear with respect to a given feature map and are allowed to evolve either…

Machine Learning · Computer Science 2021-12-28 Ahmed Touati , Pascal Vincent

This article introduces operator on operator regression in quantum probability. Here in the regression model, the response and the independent variables are certain operator valued observables, and they are linearly associated with unknown…

Methodology · Statistics 2024-08-02 Suprio Bhar , Subhra Sankar Dhar , Soumalya Joardar

Hierarchical Vision-Language-Action (VLA) models have rapidly become a dominant paradigm for robotic manipulation. It typically comprising a Vision-Language backbone for perception and understanding, together with a generative policy for…

Robotics · Computer Science 2026-05-19 Zaijing Li , Bing Hu , Rui Shao , Gongwei Chen , Dongmei Jiang , Pengwei Xie , Jianye Hao , Liqiang Nie

Reinforcement learning (RL) with continuous time and state/action spaces is often data-intensive and brittle under nuisance variability and shift, motivating methods that exploit value-preserving structures to stabilize and improve…

Machine Learning · Computer Science 2026-05-08 Zuyuan Zhang , Fei Xu Yu , Tian Lan

The Multiplicative Weights Exponential Mechanism (MWEM) is a fundamental iterative framework for private data analysis, with broad applications such as answering $m$ linear queries, or privately solving systems of $m$ linear constraints.…

Machine Learning · Computer Science 2026-02-04 Themistoklis Haris , Steve Choi , Mutiraj Laksanawisit

Let $\alpha,\beta$ be orientation-preserving diffeomorphism (shifts) of $\mathbb{R}_+=(0,\infty)$ onto itself with the only fixed points $0$ and $\infty$ and $U_\alpha,U_\beta$ be the isometric shift operators on $L^p(\mathbb{R}_+)$ given…

Functional Analysis · Mathematics 2015-01-16 Alexei Yu. Karlovich , Yuri I. Karlovich , Amarino B. Lebre

Data-driven modeling techniques have been explored in the spatial-temporal modeling of complex dynamical systems for many engineering applications. However, a systematic approach is still lacking to leverage the information from different…

Machine Learning · Computer Science 2024-10-15 Chuanqi Chen , Jin-Long Wu

A continuous time model for multiagent systems governed by reinforcement learning with scale-free memory is developed. The agents are assumed to act independently of one another in optimizing their choice of possible actions via…

Physics and Society · Physics 2015-05-14 Ihor Lubashevsky , Shigeru Kanemoto

Foundation models are transitioning from offline predictors to deployed systems expected to operate over long time horizons. In real deployments, objectives are not fixed: domains drift, user preferences evolve, and new tasks appear after…

Computer Vision and Pattern Recognition · Computer Science 2026-03-10 Tencent HY Team

Accelerators with power-law memory are proposed in the framework of the discrete time approach. To describe discrete accelerators we use the capital stock adjustment principle, which has been suggested by Matthews.The suggested discrete…

Economics · Quantitative Finance 2017-07-25 Valentina V. Tarasova , Vasily E. Tarasov