中文
相关论文

相关论文: Towards Scalable and Stable Parallelization of Non…

200 篇论文

The Newton-Raphson (NR) method is widely used for solving power flow (PF) equations due to its quadratic convergence. However, its performance deteriorates under poor initialization or extreme operating scenarios, e.g., high levels of…

系统与控制 · 电气工程与系统科学 2025-11-26 Zeynab Kaseb , Matthias Moller , Lindsay Spoor , Jerry J. Guo , Yu Xiang , Peter Palensky , Pedro P. Vergara

Sequential fine-tuning of transformers is useful when new data arrive sequentially, especially with shifting distributions. Unlike batch learning, sequential learning demands that training be stabilized despite a small amount of data by…

机器学习 · 计算机科学 2025-09-16 Haoming Jing , Oren Wright , José M. F. Moura , Yorie Nakahira

In this paper, we study the problem of estimating the state of a dynamic state-space system where the output is subject to quantization. We compare some classical approaches and a new development in the literature to obtain the filtering…

系统与控制 · 电气工程与系统科学 2021-12-16 Angel L. Cedeño , Ricardo Albornoz , Boris I. Godoy , Rodrigo Carvajal , Juan C. Agüero

The Newmark/Newton-Raphson (NNR) method is widely employed for solving nonlinear dynamic systems. However, the current NNR method exhibits limited applicability in complex nonlinear dynamic systems, as the acquisition of the Jacobian matrix…

计算工程、金融与科学 · 计算机科学 2025-06-17 Yifan Jiang , Yuhong Jin , Lei Hou , Yi Chen , Andong Cong

This research introduces an extended application of neural networks for solving nonlinear partial differential equations (PDEs). A neural network, combined with a pseudo-arclength continuation, is proposed to construct bifurcation diagrams…

数值分析 · 数学 2025-07-24 Muhammad Luthfi Shahab , Hadi Susanto

Recurrent Neural Networks (RNNs) can encode rich dynamics which makes them suitable for modeling dynamic systems. To train an RNN for multi-step prediction of dynamic systems, it is crucial to efficiently address the state initialization…

神经与进化计算 · 计算机科学 2018-06-05 Nima Mohajerin , Steven L. Waslander

As deep neural networks (DNNs) become deeper, the training time increases. In this perspective, multi-GPU parallel computing has become a key tool in accelerating the training of DNNs. In this paper, we introduce a novel methodology to…

数值分析 · 数学 2024-07-08 Chang-Ock Lee , Youngkyu Lee , Jongho Park

Mathematical models for flow and reactive transport in porous media often involve non-linear, degenerate parabolic equations. Their solutions have low regularity, and therefore lower order schemes are used for the numerical approximation.…

数值分析 · 数学 2021-05-24 Jakub W. Both , Kundan Kumar , Jan M. Nordbotten , Iuliu Sorin Pop , Florin A. Radu

We use Physics-Informed Neural Networks (PINNs) to solve the discrete-time nonlinear observer state estimation problem. Integrated within a single-step exact observer linearization framework, the proposed PINN approach aims at learning a…

In this work, we propose a parallel-in-time solver for linear and nonlinear ordinary differential equations. The approach is based on an efficient multilevel solver of the Schur complement related to a multilevel time partition. For linear…

数值分析 · 数学 2017-09-20 Santiago Badia , Marc Olm

Efficiency, as a critical practical challenge for LLM-driven agentic and reasoning systems, is increasingly constrained by the inherent latency of autoregressive (AR) decoding. Speculative decoding mitigates this cost through a draft-verify…

机器学习 · 计算机科学 2025-12-18 Zicong Cheng , Guo-Wei Yang , Jia Li , Zhijie Deng , Meng-Hao Guo , Shi-Min Hu

We present iterative solvers to approximate the solution of numerical schemes for stochastic Stefan problems. After briefly talking about the convergence results, we tackle the question of efficient strategies for solving the nonlinear…

数值分析 · 数学 2025-08-12 Muhammad Awais Khan , Jérôme Droniou , Kim-Ngan Le , Iuliu Sorin Pop

While autoregressive (AR) LLM-based ASR systems achieve strong accuracy, their sequential decoding limits parallelism and incurs high latency. We propose NLE, a non-autoregressive (NAR) approach that formulates speech recognition as…

音频与语音处理 · 电气工程与系统科学 2026-03-10 Avihu Dekel , Samuel Thomas , Takashi Fukada , George Saon

Building black-box models for dynamical systems from data is a challenging problem in machine learning, especially when asymptotic stability guarantees are required. In this paper, we introduce a novel stability-ensuring and…

机器学习 · 计算机科学 2026-05-15 Sergio Vanegas , Lasse Lensu , Fredy Ruiz

Recently, Deep Neural Networks (DNNs) have recorded great success in handling medical and other complex classification tasks. However, as the sizes of a DNN model and the available dataset increase, the training process becomes more complex…

分布式、并行与集群计算 · 计算机科学 2022-02-08 Samson B. Akintoye , Liangxiu Han , Xin Zhang , Haoming Chen , Daoqiang Zhang

Linear Attention (LA) offers a promising paradigm for scaling large language models (LLMs) to long sequences by avoiding the quadratic complexity of self-attention. Recent LA models such as Mamba2 and GDN interpret linear recurrences as…

机器学习 · 计算机科学 2026-05-08 Yulong Huang , Xiang Liu , Hongxiang Huang , Xiaopeng Lin , Zunchang Liu , Xiaowen Chu , Zeke Xie , Bojun Cheng

Distributed sensor networks often include a multitude of sensors, each measuring parts of a process state space or observing the operations of a system. Communication of measurements between the sensor nodes and estimator(s) cannot…

系统与控制 · 电气工程与系统科学 2023-05-02 Sanjay Chandrasekaran , Vishnu Varadan , Siva Vignesh Krishnan , Florian Dörfler , Mohammad H. Mamduhi

Rapid development in numerical modelling of materials and the complexity of new models increases quickly together with their computational demands. Despite the growing performance of modern computers and clusters, calibration of such models…

神经与进化计算 · 计算机科学 2016-03-08 Tomáš Mareš , Eliška Janouchová , Anna Kučerová

As deep learning becomes more expensive, both in terms of time and compute, inefficiencies in machine learning (ML) training prevent practical usage of state-of-the-art models for most users. The newest model architectures are simply too…

分布式、并行与集群计算 · 计算机科学 2021-07-15 Kabir Nagrecha

The removal of multiplicative Gamma noise is a critical research area in the application of synthetic aperture radar (SAR) imaging, where neural networks serve as a potent tool. However, real-world data often diverges from theoretical…

计算机视觉与模式识别 · 计算机科学 2024-12-24 Yi Ran , Zhichang Guo , Jia Li , Yao Li , Martin Burger , Boying Wu