English
Related papers

Related papers: Machine learning modularity

200 papers

Mixed-precision quantization improves the budget--accuracy trade-off for large language models (LLMs) by allocating more bits to sensitive modules. However, automating this allocation at LLM scale faces a unique combination of constraints:…

Machine Learning · Computer Science 2026-05-19 Zhangyang Yao , Haiyan Zhao , Haoyu Wang , Tianbo Huang , Lihua Zhang , Xu Han

Motivated by cryptographic applications, we investigate two machine learning approaches to modular multiplication: namely circular regression and a sequence-to-sequence transformer model. The limited success of both methods demonstrated in…

Machine Learning · Computer Science 2024-03-01 Kristin Lauter , Cathy Yuanchen Li , Krystal Maughan , Rachel Newton , Megha Srivastava

We present a deep machine learning (ML) approach to constraining cosmological parameters with multi-wavelength observations of galaxy clusters. The ML approach has two components: an encoder that builds a compressed representation of each…

Instrumentation and Methods for Astrophysics · Physics 2022-02-16 Michelle Ntampaka , Alexey Vikhlinin

Machine learning explorations can make significant inroads into solving difficult problems in pure mathematics. One advantage of this approach is that mathematical datasets do not suffer from noise, but a challenge is the amount of data…

Machine Learning · Computer Science 2026-05-08 Max Petschack , Alexandr Garbali , Jan de Gier

Many machine learning techniques incorporate identity-preserving transformations into their models to generalize their performance to previously unseen data. These transformations are typically selected from a set of functions that are…

Machine Learning · Computer Science 2023-03-30 Marissa Connor , Kion Fallah , Christopher Rozell

This is the first paper in a series where we study arithmetic applications of the multiple elliptic Gamma functions originated from mathematical physics. The main purpose of this paper is the introduction of a framework for applications of…

Number Theory · Mathematics 2026-01-27 Pierre L. L. Morain

The work we describe here is a part of a research program of developing foundations of declarative solving of search problems. We consider the model expansion task as the task representing the essence of search problems where we are given…

Logic in Computer Science · Computer Science 2011-09-06 Shahab Tasharrofi , Xiongnan , Wu , Eugenia Ternovska

We propose a simple method to identify a continuous Lie algebra symmetry in a dataset through regression by an artificial neural network. Our proposal takes advantage of the $ \mathcal{O}(\epsilon^2)$ scaling of the output variable under…

High Energy Physics - Phenomenology · Physics 2022-06-01 Sean Craven , Djuna Croon , Daniel Cutting , Rachel Houtz

As an efficient alternative to conventional full finetuning, parameter-efficient finetuning (PEFT) is becoming the prevailing method to adapt pretrained language models. In PEFT, a lightweight module is learned on each dataset while the…

Computation and Language · Computer Science 2023-12-12 Jinghan Zhang , Shiqi Chen , Junteng Liu , Junxian He

Given a maze populated with different objects, one may task a robot with a sequential goal completion task, e.g. 1) pick up a key then 2) unlock the door then 3) unlock the treasure chest. A typical machine learning (ML) solution would…

Neural and Evolutionary Computing · Computer Science 2024-05-01 Nathan McDonald

Recent work showed that ML-based attacks on Learning with Errors (LWE), a hard problem used in post-quantum cryptography, outperform classical algebraic attacks in certain settings. Although promising, ML attacks struggle to scale to more…

Machine Learning · Computer Science 2025-08-26 Eshika Saxena , Alberto Alfarano , François Charton , Zeyuan Allen-Zhu , Emily Wenger , Kristin Lauter

Autoformalization is the task of automatically translating mathematical content written in natural language to a formal language expression. The growing language interpretation capabilities of Large Language Models (LLMs), including in…

Computation and Language · Computer Science 2025-06-16 Lan Zhang , Xin Quan , Andre Freitas

Quantum Machine Learning (QML) has emerged as a promising framework for exploring how quantum dynamics may enhance data processing tasks. Here we investigate Quantum Extreme Learning Machines (QELMs), a quantum analogue of classical Extreme…

Quantum Physics · Physics 2026-04-27 A. De Lorenzis , M. P. Casado , N. Lo Gullo , T. Lux , F. Plastina , A. Riera

Large language models (LLMs) have revolutionized algorithm development, yet their application in symbolic regression, where algorithms automatically discover symbolic expressions from data, remains limited. In this paper, we propose a…

Neural and Evolutionary Computing · Computer Science 2026-04-01 Hengzhe Zhang , Qi Chen , Bing Xue , Wolfgang Banzhaf , Mengjie Zhang

This is the second paper in a series where we study arithmetic applications of the multiple elliptic Gamma functions originated in mathematical physics. In the first article in this series we defined geometric families of these functions…

Number Theory · Mathematics 2026-02-09 Pierre L. L. Morain

In the vicinity of a phase transition ergodicity can be broken. Here, different initial many-body configurations evolve towards one of several fixed points, which are macroscopically distinguishable through an order parameter. This…

Quantum Physics · Physics 2025-09-24 Mario Boneberg , Simon Kochsiek , Gabriele Perfetto , Igor Lesanovsky

Recent advancements in Large Language Models (LLMs) have set themselves apart with their exceptional performance in complex language modelling tasks. However, these models are also known for their significant computational and storage…

Computation and Language · Computer Science 2025-08-12 Peng Lu , Ivan Kobyzev , Mehdi Rezagholizadeh , Boxing Chen , Philippe Langlais

Methodologies for training machine learning potentials (MLPs) to quantum-mechanical simulation data have recently seen tremendous progress. Experimental data has a very different character than simulated data, and most MLP training…

Decomposing complex tasks into a sequence of simpler subtasks can improve learning efficiency for an autonomous agent. Reinforcement learning (RL) can be used to optimize agent policies to complete subtasks, but requires well-defined…

Machine Learning · Computer Science 2026-05-26 Nicholas Potteiger , Ankita Samaddar , Taylor T. Johnson , Xenofon Koutsoukos

Recent techniques that integrate \emph{solver layers} into Deep Neural Networks (DNNs) have shown promise in bridging a long-standing gap between inductive learning and symbolic reasoning techniques. In this paper we present a set of…

Machine Learning · Computer Science 2023-01-30 Matt Fredrikson , Kaiji Lu , Saranya Vijayakumar , Somesh Jha , Vijay Ganesh , Zifan Wang