English
Related papers

Related papers: Maximal Update Parametrization and Zero-Shot Hyper…

200 papers

Parameter-efficient fine-tuning (PEFT) has become a common method for fine-tuning large language models, where a base model can serve multiple users through PEFT module switching. To enhance user experience, base models require periodic…

Computation and Language · Computer Science 2025-06-10 Naibin Gu , Peng Fu , Xiyu Liu , Ke Ma , Zheng Lin , Weiping Wang

Multi-objective optimization (MOO) problems are prevalent in machine learning. These problems have a set of optimal solutions, called the Pareto front, where each point on the front represents a different trade-off between possibly…

Machine Learning · Computer Science 2021-04-27 Aviv Navon , Aviv Shamsian , Gal Chechik , Ethan Fetaya

Fourier Neural Operator (FNO) is a popular operator learning framework. It not only achieves the state-of-the-art performance in many tasks, but also is efficient in training and prediction. However, collecting training data for the FNO can…

Machine Learning · Computer Science 2024-04-01 Shibo Li , Xin Yu , Wei Xing , Mike Kirby , Akil Narayan , Shandian Zhe

Neural operators have emerged as a powerful tool for learning the mapping between infinite-dimensional parameter and solution spaces of partial differential equations (PDEs). In this work, we focus on multiscale PDEs that have important…

Machine Learning · Computer Science 2024-06-11 Xinliang Liu , Bo Xu , Shuhao Cao , Lei Zhang

Bayesian optimization (BO) is a popular methodology to tune the hyperparameters of expensive black-box functions. Traditionally, BO focuses on a single task at a time and is not designed to leverage information from related functions, such…

Machine Learning · Statistics 2021-04-20 David Salinas , Huibin Shen , Valerio Perrone

Recently, several optimization methods have been successfully applied to the hyperparameter optimization of deep neural networks (DNNs). The methods work by modeling the joint distribution of hyperparameter values and corresponding error.…

Machine Learning · Computer Science 2016-08-02 Ilija Ilievski , Jiashi Feng

We introduce a memory- and compute-efficient method for low-communication distributed training. Existing methods reduce communication by performing multiple local updates between infrequent global synchronizations. We demonstrate that their…

Machine Learning · Computer Science 2025-09-29 Anastasiia Filippova , Angelos Katharopoulos , David Grangier , Ronan Collobert

This paper proposes a physics-informed neural operator (PINO) framework for solving inverse scattering problems, enabling rapid and accurate reconstructions under diverse measurement conditions. In the proposed approach, the dielectric…

Computational Physics · Physics 2026-03-27 Q. C. Dong , Zi-Xuan Su , Qing Huo Liu , Wen Chen , Zhizhang , Chen

Global urbanization has underscored the significance of urban microclimates for human comfort, health, and building/urban energy efficiency. They profoundly influence building design and urban planning as major environmental impacts.…

Machine Learning · Computer Science 2023-10-03 Wenhui Peng , Shaoxiang Qin , Senwen Yang , Jianchun Wang , Xue Liu , Liangzhu Leon Wang

Hyper-parameters optimization (HPO) is vital for machine learning models. Besides model accuracy, other tuning intentions such as model training time and energy consumption are also worthy of attention from data analytic service providers.…

Machine Learning · Computer Science 2023-04-21 Hui Dou , Shanshan Zhu , Yiwen Zhang , Pengfei Chen , Zibin Zheng

To obtain fast solutions for governing physical equations in solid mechanics, we introduce a method that integrates the core ideas of the finite element method with physics-informed neural networks and concept of neural operators. This…

This paper deals with linear equalization in massive multi-user multiple-input multiple-output (MU-MIMO) wireless systems. We first provide simple conditions on the antenna configuration for which the well-known linear minimum mean-square…

Information Theory · Computer Science 2018-11-12 Ramina Ghods , Charles Jeon , Gulnar Mirza , Arian Maleki , Christoph Studer

Operator learning is a variant of machine learning that is designed to approximate maps between function spaces from data. The Fourier Neural Operator (FNO) is one of the main model architectures used for operator learning. The FNO combines…

Numerical Analysis · Mathematics 2025-09-29 Samuel Lanthaler , Andrew M. Stuart , Margaret Trautner

The Monte Carlo-type Neural Operator (MCNO) introduces a lightweight architecture for learning solution operators for parametric PDEs by directly approximating the kernel integral using a Monte Carlo approach. Unlike Fourier Neural…

Machine Learning · Computer Science 2025-11-25 Salah Eddine Choutri , Prajwal Chauhan , Othmane Mazhar , Saif Eddin Jabari

Neural operators are becoming the default tools to learn solutions to governing partial differential equations (PDEs) in weather and ocean forecasting applications. Despite early promising achievements, significant challenges remain,…

Machine Learning · Computer Science 2025-10-14 Vahidreza Jahanmard , Ali Ramezani-Kebrya , Robinson Hordoir

Derivative-free optimization (DFO) problems are optimization problems where derivative information is unavailable or extremely difficult to obtain. Model-based DFO solvers have been applied extensively in scientific computing. Powell's…

Optimization and Control · Mathematics 2025-04-07 Pengcheng Xie , Stefan M. Wild

Neural operators such as the Fourier Neural Operator (FNO) have been shown to provide resolution-independent deep learning models that can learn mappings between function spaces. For example, an initial condition can be mapped to the…

Machine Learning · Computer Science 2024-07-02 Aditya Kashi , Arka Daw , Muralikrishnan Gopalakrishnan Meena , Hao Lu

Few-shot node classification on hypergraphs requires models that generalize from scarce labels while capturing high-order structures. Existing hypergraph neural networks (HNNs) effectively encode such structures but often suffer from…

Machine Learning · Computer Science 2025-10-27 Chaewoon Bae , Doyun Choi , Jaehyun Lee , Jaemin Yoo

Solutions to many partial differential equations (PDEs) display coexisting smooth global transport and localized sharp features within a single trajectory: shock fronts, thin interfaces, and concentrated high-frequency content sit on top of…

Machine Learning · Computer Science 2026-05-14 Yingzhe Ma , Xiao Yang , Yuxin Xie , Zihan Xiong , Jinliang Liu

We introduce Neural Organ Transplantation (NOT), a modular adaptation framework that enables trained transformer layers to function as reusable transferable checkpoints for domain adaptation. Unlike conventional fine-tuning approaches that…

Machine Learning · Computer Science 2026-01-21 Ahmad Al-Zuraiqi
‹ Prev 1 8 9 10 Next ›