English
Related papers

Related papers: A Two-Parameter Weibull Framework for Diagnosing T…

200 papers

Penalized transformation models (PTMs) are a semiparametric location-scale regression family that estimate a response's conditional distribution directly from the data, and model the location and scale through structured additive…

Methodology · Statistics 2025-09-22 Johannes Brachem , Paul F. V. Wiemann , Thomas Kneib

Graph learning architectures based on the k-dimensional Weisfeiler-Leman (k-WL) hierarchy offer a theoretically well-understood expressive power. However, such architectures often fail to deliver solid predictive performance on real-world…

Machine Learning · Computer Science 2024-11-11 Luis Müller , Daniel Kusuma , Blai Bonet , Christopher Morris

Variational quantum circuits (VQCs) are a leading approach to quantum machine learning on near-term devices, yet it remains unclear which circuit architecture yields the best accuracy-parameter trade-off on classical tabular data. We…

Quantum Physics · Physics 2026-04-28 Chi-Sheng Chen , En-Jui Kuo

State-of-the-art results on neural machine translation often use attentional sequence-to-sequence models with some form of convolution or recursion. Vaswani et al. (2017) propose a new architecture that avoids recurrence and convolution…

Artificial Intelligence · Computer Science 2017-11-08 Karim Ahmed , Nitish Shirish Keskar , Richard Socher

We extend the theory of concentration inequalities to simple random tensors with heavy-tailed coefficients. Specifically, we consider the class of sub-Weibull distributions $\mathcal{S}_\alpha$ for $\alpha \in [1, 2]$. We establish…

Mathematical Finance · Quantitative Finance 2026-03-11 Yunfan Zhao

Rating prediction is a core problem in recommender systems to quantify user's preferences towards items, however, rating imbalance naturally roots in real-world user ratings that cause biased predictions and lead to poor performance on tail…

Information Retrieval · Computer Science 2022-08-18 Yuexin Wu , Xiaolei Huang

Geometric Attention (GA) specifies an attention layer by four independent inputs: a finite carrier (what indices are addressable), an evidence-kernel rule (how masked proto-scores and a link induce nonnegative weights), a probe family…

Machine Learning · Computer Science 2026-01-21 Luis Rosario Freytes

We propose a new class of claim severity distributions with six parameters, that has the standard two-parameter distributions, the log-normal, the log-Gamma, the Weibull, the Gamma and the Pareto, as special cases. This distribution is much…

Methodology · Statistics 2018-05-29 Erik Bølviken , Ingrid Hobæk Haff

Transformers trained on modular arithmetic exhibit sharp transitions between memorization, generalization, and collapse. We show that weight decay acts as a scalar empirical control parameter for these regimes, and introduce two cheap…

Machine Learning · Computer Science 2026-05-21 Lucky Verma

In engineering systems, it is usually assumed that lifetimes of components are independent and identically distributed (iid). But, the failure of a component results in a higher load on the remaining components and hence causes the…

Statistics Theory · Mathematics 2019-12-18 M. Doostparast , M. Hashempour , E. Velayati Moghaddam 1

Stress-strength models are widely used to assess the reliability of systems under uncertain conditions. While most studies assume independence between stress and strength variables, such an assumption may be unrealistic in many practical…

Methodology · Statistics 2026-04-15 Fatih Kızılaslan

We propose a new class of discrete generalized linear models based on the class of Poisson-Tweedie factorial dispersion models with variance of the form $\mu + \phi\mu^p$, where $\mu$ is the mean, $\phi$ and $p$ are the dispersion and…

In this article, we consider statistical inference based on dependent competing risks data from Marshall-Olkin bivariate Weibull distribution. The maximum likelihood estimates of the unknown model parameters have been computed by using the…

Methodology · Statistics 2023-04-20 Subhankar Dutta , Suchandan Kayal

Despite their power, Transformers face challenges with long sequences due to the quadratic complexity of self-attention. To address this limitation, methods like $k$-Nearest-Neighbor ($k$NN) attention have been introduced [Roy, Saffar,…

Machine Learning · Computer Science 2024-11-11 Themistoklis Haris

We study the extreme value distribution of stochastic processes modeled by superstatistics. Classical extreme value theory asserts that (under mild asymptotic independence assumptions) only three possible limit distributions are possible,…

Statistical Mechanics · Physics 2015-06-22 Pau Rabassa , Christian Beck

Ensembles of neural network weight matrices are studied through the training process for the MNIST classification problem, testing the efficacy of matrix models for representing their distributions, under assumptions of Gaussianity and…

Machine Learning · Computer Science 2025-10-08 Edward Hirst , Sanjaye Ramgoolam

We introduce a general, flexible, parametric survival modelling framework which encompasses key shapes of hazard function (constant, increasing, decreasing, up-then-down, down-then-up), various common survival distributions (log-logistic,…

Methodology · Statistics 2019-01-11 Kevin Burke , M. C. Jones , Angela Noufaily

Variational methods are attractive for computing Bayesian inference for highly parametrized models and large datasets where exact inference is impractical. They approximate a target distribution - either the posterior or an augmented…

Computation · Statistics 2019-11-21 Michael Stanley Smith , Ruben Loaiza-Maya , David J. Nott

We present a novel approach to selective model quantization that transcends the limitations of architecture-specific and size-dependent compression methods for Large Language Models (LLMs) using Entropy-Weighted Quantization (EWQ). By…

Machine Learning · Computer Science 2025-03-10 Alireza Behtash , Marijan Fofonjka , Ethan Baird , Tyler Mauer , Hossein Moghimifam , David Stout , Joel Dennison

In this paper, we analyze the relative errors that crop up in the various reliability measures due to the tacit assumption that the components are independently working associated with a $n$-component series system or a parallel system…

Statistics Theory · Mathematics 2025-03-28 Subarna Bhattacharjee , Aninda Kumar Nanda , Subhashree Patra