English
Related papers

Related papers: Neural Prime Sieves: Density-Driven Generalization…

200 papers

We introduce squared families, which are families of probability densities obtained by squaring a linear transformation of a statistic. Squared families are singular, however their singularity can easily be handled so that they form regular…

Machine Learning · Statistics 2025-03-28 Russell Tsuchida , Jiawei Liu , Cheng Soon Ong , Dino Sejdinovic

Modeling of strong gravitational lenses is a necessity for further applications in astrophysics and cosmology. Especially with the large number of detections in current and upcoming surveys such as the Rubin Legacy Survey of Space and Time…

Instrumentation and Methods for Astrophysics · Physics 2023-03-30 S. Schuldt , R. Cañameras , Y. Shu , S. H. Suyu , S. Taubenberger , T. Meinhardt , L. Leal-Taixé

Bayesian neural networks promise calibrated uncertainty but require $O(mn)$ parameters for standard mean-field Gaussian posteriors. We argue this cost is often unnecessary, particularly when weight matrices exhibit fast singular value…

Machine Learning · Statistics 2026-05-05 Mame Diarra Toure , David A. Stephens

Language model families exhibit striking disparity in their capacity to benefit from reinforcement learning: under identical training, models like Qwen achieve substantial gains, while others like Llama yield limited improvements.…

Computation and Language · Computer Science 2026-01-13 Shaoning Sun , Mingzhu Cai , Huang He , Bingjin Chen , Siqi Bao , Yujiu Yang , Hua Wu , Haifeng Wang

We give an estimation of the existence density for the $2d$ different primes by using a new and simple algorithm for getting the $2d$ different primes. The algorithm is a kind of the sieve method, but the remainders are the central numbers…

Number Theory · Mathematics 2014-02-27 Minoru Fujimoto , Kunihiko Uehara

Precise recall control is critical in large-scale spatial conflation and entity-matching tasks, where missing even a few true matches can break downstream analytics, while excessive manual review inflates cost. Classical confidence-interval…

Machine Learning · Computer Science 2025-10-03 John N. Daras

This paper presents a novel approach at the intersection of machine learning and number theory, focusing on the classification of prime and non-prime numbers. At the core of our research is the development of a highly sparse encoding…

Number Theory · Mathematics 2026-04-01 Serin Lee , S. Kim

We study an LCM-based analogue of Rowland's GCD-based prime-generating recurrence, introduced by the author in 2008. The multiplicative increments of this sequence are conjectured always to be $1$ or prime, but a complete proof requires a…

Number Theory · Mathematics 2026-04-22 Benoit Cloitre

Consider a system \Psi of non-constant affine-linear forms \psi_1,...,\psi_t: Z^d -> Z, no two of which are linearly dependent. Let N be a large integer, and let K be a convex subset of [-N,N]^d. A famous and difficult open conjecture of…

Number Theory · Mathematics 2008-04-22 Ben Green , Terence Tao

We adopt an empirical approach to the characterization of the distribution of twin primes within the set of primes, rather than in the set of all natural numbers. The occurrences of twin primes in any finite sequence of primes are like…

Number Theory · Mathematics 2007-05-23 P. F. Kelly , Terry Pilling

We introduce a pruning algorithm that provably sparsifies the parameters of a trained model in a way that approximately preserves the model's predictive accuracy. Our algorithm uses a small batch of input points to construct a data-informed…

Machine Learning · Computer Science 2021-03-16 Cenk Baykal , Lucas Liebenwein , Igor Gilitschenski , Dan Feldman , Daniela Rus

The quantitative distribution of twin primes remains a central open problem in number theory. This paper develops a heuristic model grounded in the principles of sieve theory, with the goal of constructing an analytical approximation for…

General Mathematics · Mathematics 2025-07-14 Yuhang Shi

Exactly solvable neural network models with asymmetric weights are rare, and exact solutions are available only in some mean-field approaches. In this article we find exact analytical solutions of an asymmetric spin-glass-like model of…

Neurons and Cognition · Quantitative Biology 2017-02-16 Diego Fasoli , Anna Cattani , Stefano Panzeri

Even though probabilistic treatments of neural networks have a long history, they have not found widespread use in practice. Sampling approaches are often too slow already for simple networks. The size of the inputs and the depth of typical…

Computer Vision and Pattern Recognition · Computer Science 2018-05-30 Jochen Gast , Stefan Roth

Deep retrieval models are widely used for learning entity representations and recommendations. Federated learning provides a privacy-preserving way to train these models without requiring centralization of user data. However, federated deep…

Machine Learning · Computer Science 2021-11-03 Lin Ning , Karan Singhal , Ellie X. Zhou , Sushant Prakash

We adopt a physically motivated empirical approach to the characterisation of the distributions of twin and triplet primes within the set of primes, rather than in the set of all natural numbers. Remarkably, the occurrences of twins or…

High Energy Physics - Theory · Physics 2007-05-23 P. F. Kelly , Terry Pilling

Randomly initialized neural networks induce a prior over functions, but the predictor used in practice is produced only after training. We ask how much of this initial bias survives the training pipeline. To make the question measurable, we…

Machine Learning · Computer Science 2026-05-29 Mohua Das , Pierfrancesco Beneventano , Shibshankar Dey , Gareth H. McKinkey , Tomaso Poggio

Flexible models for probability distributions are an essential ingredient in many machine learning tasks. We develop and investigate a new class of probability distributions, which we call a Squared Neural Family (SNEFY), formed by squaring…

Machine Learning · Computer Science 2023-10-27 Russell Tsuchida , Cheng Soon Ong , Dino Sejdinovic

We study a new class of preferential attachment trees with \emph{self-reinforcement}. At each time, each vertex is assigned a weight equal to the cumulative sum over past times of an affine function of its degree. A new vertex attaches…

Probability · Mathematics 2026-05-21 Shankar Bhamidi , Remco van der Hofstad , Frank den Hollander , Rounak Ray

In a sponsored search engine, generative retrieval models are recently proposed to mine relevant advertisement keywords for users' input queries. Generative retrieval models generate outputs token by token on a path of the target library…

Information Retrieval · Computer Science 2020-10-22 Weizhen Qi , Yeyun Gong , Yu Yan , Jian Jiao , Bo Shao , Ruofei Zhang , Houqiang Li , Nan Duan , Ming Zhou
‹ Prev 1 2 3 10 Next ›