English
Related papers

Related papers: A Two-Parameter Weibull Framework for Diagnosing T…

200 papers

We present a formal operator-theoretic framework for analyzing Transformer-based language models using free probability theory. By modeling token embeddings and attention mechanisms as self-adjoint operators in a tracial \( W^*…

Machine Learning · Computer Science 2025-08-19 Swagatam Das

The scarcity of labelled data is specifically an urgent challenge in the field of quantum machine learning (QML). Two transfer fusion frameworks are proposed in this paper to predict the labels of a target domain data by aligning its…

Computer Vision and Pattern Recognition · Computer Science 2024-11-05 Xi He , Feiyu Du , Xiaohan Yu , Yang Zhao , Tao Lei

Transfer learning from ImageNet is the go-to approach when applying deep learning to medical images. The approach is either to fine-tune a pre-trained model or use it as a feature extractor. Most modern architecture contain batch…

Computer Vision and Pattern Recognition · Computer Science 2021-02-11 Fahdi Kanavati , Masayuki Tsuneki

Stable distributions provide a flexible framework for modeling heavy-tailed and skewed data, with the stability index $\alpha$ quantifying tail heaviness. We propose a new semiparametric estimator for $\alpha$ that leverages the two-sum…

Methodology · Statistics 2025-08-19 Cornelis J. Potgieter , Jacques van Appel , Sudharshan Samaratunga

Attention-based neural networks have achieved state-of-the-art results on a wide range of tasks. Most such models use deterministic attention while stochastic attention is less explored due to the optimization difficulties or complicated…

Machine Learning · Computer Science 2021-06-10 Shujian Zhang , Xinjie Fan , Bo Chen , Mingyuan Zhou

Although existing variational graph autoencoders (VGAEs) have been widely used for modeling and generating graph-structured data, most of them are still not flexible enough to approximate the sparse and skewed latent node representations,…

Machine Learning · Computer Science 2024-10-15 Chaojie Wang , Xinyang Liu , Dongsheng Wang , Hao Zhang , Bo Chen , Mingyuan Zhou

A new unimodal distribution family indexed by the mode and three other parameters is derived from a mixture of a Gumbel distribution for the maximum and a Gumbel distribution for the minimum. Properties of the proposed distribution are…

Methodology · Statistics 2024-07-02 Qingyang Liu , Xianzheng Huang , Haiming Zhou

Let $X_{1}=(W_{1},Y_{1}),\ldots,X_{n}=(W_{n},Y_{n})$ be $n$ pairs of independent random variables. We assume that, for each $i\in\{1,\ldots,n\}$, the conditional distribution of $Y_{i}$ given $W_{i}$ belongs to a one-parameter exponential…

Statistics Theory · Mathematics 2022-03-15 Juntong Chen

The Weibull--like distributions form a large class of probability distributions that belong to the domain of attraction for the maxima of the Gumbel law. Besides the Weibull distribution, it includes important distributions as the Gamma…

Statistics Theory · Mathematics 2013-08-27 Armengol Gasull , José A. López-Salcedo , Frederic Utzet

By recognizing that the main difficulty of the modeling of daily precipitation amounts is the selection of an appropriate probability distribution, this study aims to establish a model selection framework to identify the appropriate…

Applications · Statistics 2020-09-01 Hsien-Wei Chen

This paper explores the possibility of establishing an analytic form of the distribution of the order parameter fluctuations in a two-dimensional critical spin wave model, or width fluctuations of a two dimensional Edwards-Wilkinson…

Statistical Mechanics · Physics 2022-04-13 Steven T. Bramwell

Based on suitable left-truncated or censored data, two flexible classes of $M$-estimations of Weibull tail coefficient are proposed with two additional parameters bounding the impact of extreme contamination. Asymptotic normality with…

Statistics Theory · Mathematics 2018-10-18 Chengping Gong , Chengxiu Ling

The proposed paper discusses the problem of discrimination between close hypotheses about distributions belonging to the Gumbel maximum domain of attraction. The distinctive feature of the proposed work is using only k higher order…

Statistics Theory · Mathematics 2016-06-29 Igor Rodionov

We develop an econometric framework integrating heavy-tailed Student's $t$ distributions with behavioral probability weighting while preserving infinite divisibility. Using 432{,}752 observations across 86 assets (2004--2024), we…

Mathematical Finance · Quantitative Finance 2025-11-21 Akash Deep , Svetlozar T. Rachev , Frank J. Fabozzi

We investigate whether the Wigner semi-circle and Marcenko-Pastur distributions, often used for deep neural network theoretical analysis, match empirically observed spectral densities. We find that even allowing for outliers, the observed…

Machine Learning · Statistics 2021-11-04 Diego Granziol

Probability density function estimation with weighted samples is the main foundation of all adaptive importance sampling algorithms. Classically, a target distribution is approximated either by a non-parametric model or within a parametric…

Machine Learning · Computer Science 2023-10-16 Julien Demange-Chryst , François Bachoc , Jérôme Morio , Timothé Krauth

Optical wireless communication (OWC) is highly vulnerable to the atmospheric turbulence and pointing error. Performance analysis of the OWC system under the combined channel effects of pointing errors and atmospheric turbulence is desirable…

Signal Processing · Electrical Eng. & Systems 2020-07-16 Kartik Wardhan , S. M. Zafaruddin

Graph neural network architectures aligned with the $k$-dimensional Weisfeiler--Leman ($k$-WL) hierarchy offer theoretically well-understood expressive power. However, these architectures often fail to deliver state-of-the-art predictive…

Machine Learning · Computer Science 2024-06-06 Luis Müller , Christopher Morris

We present a data-driven framework to model the stochastic evolution of volume-price distribution from the New York Stock Exchange (NYSE) equities. The empirical distributions are sampled every 10 minutes over 976 trading days, and fitted…

Neural and Evolutionary Computing · Computer Science 2026-05-08 Anup Budhathoki , Leonardo Rydin Gorjão , Pedro G. Lind , Shailendra Bhandari

Recently, multiple architectures has been proposed to improve the efficiency of the Transformer Language Models through changing the design of the self-attention block to have a linear-cost inference (LCI). A notable approach in this realm…

Computation and Language · Computer Science 2024-04-04 Sehyun Choi
‹ Prev 1 4 5 6 7 8 10 Next ›