English
Related papers

Related papers: Mondrian: Transformer Operators via Domain Decompo…

200 papers

While transformers have greatly boosted performance in semantic segmentation, domain adaptive transformers are not yet well explored. We identify that the domain gap can cause discrepancies in self-attention. Due to this gap, the…

Computer Vision and Pattern Recognition · Computer Science 2022-12-22 Kaihong Wang , Donghyun Kim , Rogerio Feris , Kate Saenko , Margrit Betke

Monocular depth estimation (MDE) has attracted intense study due to its low cost and critical functions for robotic tasks such as localization, mapping and obstacle detection. Supervised approaches have led to great success with the advance…

Computer Vision and Pattern Recognition · Computer Science 2022-08-02 Shao-Yuan Lo , Wei Wang , Jim Thomas , Jingjing Zheng , Vishal M. Patel , Cheng-Hao Kuo

The solution of a PDE over varying initial/boundary conditions on multiple domains is needed in a wide variety of applications, but it is computationally expensive if the solution is computed de novo whenever the initial/boundary conditions…

Machine Learning · Computer Science 2024-02-13 Minglang Yin , Nicolas Charon , Ryan Brody , Lu Lu , Natalia Trayanova , Mauro Maggioni

Latest development of neural models has connected the encoder and decoder through a self-attention mechanism. In particular, Transformer, which is solely based on self-attention, has led to breakthroughs in Natural Language Processing (NLP)…

Computation and Language · Computer Science 2019-11-07 Xindian Ma , Peng Zhang , Shuai Zhang , Nan Duan , Yuexian Hou , Dawei Song , Ming Zhou

Neural operators offer a powerful data-driven framework for learning mappings between function spaces, in which the transformer-based neural operator architecture faces a fundamental scalability-accuracy trade-off: softmax attention…

Machine Learning · Computer Science 2025-10-21 Ming Zhong , Zhenya Yan

Transfer learning (TL) enables the transfer of knowledge gained in learning to perform one task (source) to a related but different task (target), hence addressing the expense of data acquisition and labeling, potential computational power…

Machine Learning · Computer Science 2022-12-20 Somdatta Goswami , Katiana Kontolati , Michael D. Shields , George Em Karniadakis

Supervised learning in function spaces is an emerging area of machine learning research with applications to the prediction of complex physical systems such as fluid flows, solid mechanics, and climate modeling. By directly learning maps…

Machine Learning · Computer Science 2022-06-09 Jacob H. Seidman , Georgios Kissas , Paris Perdikaris , George J. Pappas

We consider differential operators on a supermanifold of dimension $1|1$. We define non-degenerate operators as those with an invertible top coefficient in the expansion in the "superderivative" $D$ (which is the square root of the shift…

Differential Geometry · Mathematics 2019-01-08 Simon Li , Ekaterina Shemyakova , Theodore Voronov

Deep neural networks (DNNs) have achieved remarkable success in numerous domains, and their application to PDE-related problems has been rapidly advancing. This paper provides an estimate for the generalization error of learning Lipschitz…

Machine Learning · Computer Science 2023-10-04 Ke Chen , Chunmei Wang , Haizhao Yang

Transformer-based deep learning models have achieved state-of-the-art performance across numerous language and vision tasks. While the self-attention mechanism, a core component of transformers, has proven capable of handling complex data…

Machine Learning · Computer Science 2025-08-05 Laziz Abdullaev , Tan M. Nguyen

We present Polynomial Attention Drop-in Replacement (PADRe), a novel and unifying framework designed to replace the conventional self-attention mechanism in transformer models. Notably, several recent alternative attention mechanisms,…

Computer Vision and Pattern Recognition · Computer Science 2024-07-17 Pierre-David Letourneau , Manish Kumar Singh , Hsin-Pai Cheng , Shizhong Han , Yunxiao Shi , Dalton Jones , Matthew Harper Langston , Hong Cai , Fatih Porikli

Nonlinear phenomena can be analyzed via linear techniques using operator-theoretic approaches. Data-driven method called the extended dynamic mode decomposition (EDMD) and its variants, which approximate the Koopman operator associated with…

Machine Learning · Computer Science 2022-05-18 Hiroaki Terao , Sho Shirasaka , Hideyuki Suzuki

We propose new domain decomposition methods for systems of partial differential equations in two and three dimensions. The algorithms are derived with the help of the Smith factorization of the operator. This could also be validated by…

Numerical Analysis · Mathematics 2009-09-04 Victorita Dolean , Frédéric Nataf , Gerd Rapin

Operators with fractional perturbations are crucial components for robust preconditioning of interface-coupled multiphysics systems. However, in case the perturbation is strong, standard approaches can fail to provide scalable approximation…

Numerical Analysis · Mathematics 2022-12-01 Miroslav Kuchta

The success of vision transformers-especially for generative modeling-is limited by the quadratic cost and weak spatial inductive bias of self-attention. We propose PDE-SSM, a spatial state-space block that replaces attention with a…

Machine Learning · Computer Science 2026-03-17 Eshed Gal , Moshe Eliasof , Siddharth Rout , Eldad Haber

Neural operators improve conventional neural networks by expanding their capabilities of functional mappings between different function spaces to solve partial differential equations (PDEs). One of the most notable methods is the Fourier…

Machine Learning · Computer Science 2024-07-29 Xuanle Zhao , Yue Sun , Tielin Zhang , Bo Xu

We propose the first method to show theoretical limitations for one-layer softmax transformers with arbitrarily many precision bits (even infinite). We establish those limitations for three tasks that require advanced reasoning. The first…

Unsupervised Domain Adaptation (UDA) aims to utilize labeled data from a source domain to solve tasks in an unlabeled target domain, often hindered by significant domain gaps. Traditional CNN-based methods struggle to fully capture complex…

Computer Vision and Pattern Recognition · Computer Science 2024-12-06 A. Enes Doruk , Erhan Oztop , Hasan F. Ates

Domain shift is a formidable issue in Machine Learning that causes a model to suffer from performance degradation when tested on unseen domains. Federated Domain Generalization (FedDG) attempts to train a global model using collaborative…

Computer Vision and Pattern Recognition · Computer Science 2026-02-17 Khiem Le , Long Ho , Cuong Do , Danh Le-Phuoc , Kok-Seng Wong

This work establishes a rigorous bridge between infinite-dimensional delay dynamics and finite-dimensional Koopman learning, with explicit and interpretable error guarantees. While Koopman analysis is well-developed for ordinary…

Systems and Control · Electrical Eng. & Systems 2026-04-06 Santosh Mohan Rajkumar , Dibyasri Barman , Kumar Vikram Singh , Debdipta Goswami
‹ Prev 1 4 5 6 7 8 10 Next ›