English
Related papers

Related papers: Prediction of long memory processes on same-realis…

200 papers

Deep Gaussian Processes learn probabilistic data representations for supervised learning by cascading multiple Gaussian Processes. While this model family promises flexible predictive distributions, exact inference is not tractable.…

Machine Learning · Statistics 2020-10-23 Jakob Lindinger , David Reeb , Christoph Lippert , Barbara Rakitsch

The field of machine have seen rising applications of equivariance criterion. However, there is no systematic way to justify its usage, including why it works, whether there is an optimal solution and if so, what form it carries. In this…

Statistics Theory · Mathematics 2025-09-23 Daowei Wang , Mian Wu , Haojin Zhou

This paper deals with the consistency of the least squares estimator of a convex regression function when the predictor is multidimensional. We characterize and discuss the computation of such an estimator via the solution of certain…

Statistics Theory · Mathematics 2015-03-13 Emilio Seijo , Bodhisattva Sen

Successive Halving is a popular algorithm for hyperparameter optimization which allocates exponentially more resources to promising candidates. However, the algorithm typically relies on intermediate performance values to make resource…

Machine Learning · Computer Science 2025-08-21 Jihao Andreas Lin , Nicolas Mayoraz , Steffen Rendle , Dima Kuzmin , Emil Praun , Berivan Isik

Finitarily Markovian processes are those processes $\{X_n\}_{n=-\infty}^{\infty}$ for which there is a finite $K$ ($K = K(\{X_n\}_{n=-\infty}^0$) such that the conditional distribution of $X_1$ given the entire past is equal to the…

Probability · Mathematics 2015-05-13 Gusztav Morvai , Benjamin Weiss

Iteration of randomly chosen quadratic maps defines a Markov process: X_{n+1}=\epsilon_{n+1}X_n(1-X_n), where \epsilon_n are i.i.d. with values in the parameter space [0,4] of quadratic maps F_{\theta}(x)=\theta x(1-x). Its study is of…

Probability · Mathematics 2007-05-23 Rabi Bhattacharya , Mukul Majumdar

Empirical risk minimization is a standard principle for choosing algorithms in learning theory. In this paper we study the properties of empirical risk minimization for time series. The analysis is carried out in a general framework that…

Machine Learning · Statistics 2021-08-12 Christian Brownlees , Jordi Llorens-Terrazas

In this work, we consider the deterministic optimization using random projections as a statistical estimation problem, where the squared distance between the predictions from the estimator and the true solution is the error metric. In…

Optimization and Control · Mathematics 2020-06-16 Srivatsan Sridhar , Mert Pilanci , Ayfer Özgür

A continuous-time regression model with a jointly strictly sub-Gaussian random noise is considered in the paper. Upper exponential bounds for probabilities of large deviations of the least squares estimator for the regression parameter are…

Probability · Mathematics 2018-06-12 Alexander V. Ivanov , Igor V. Orlovskyi

We consider the multilinear polynomial-form process \[X(n)=\sum_{1\le i_1<\ldots<i_k<\infty}a_{i_1}\ldots a_{i_k}\epsilon_{n-i_1}\ldots\epsilon_{n-i_k},\] obtained by applying a multilinear polynomial-form filter to i.i.d.\ sequence…

Probability · Mathematics 2013-04-19 Murad S. Taqqu , Shuyang Bai

A local linear kernel estimator of the regression function x\mapsto g(x):=E[Y_i|X_i=x], x\in R^d, of a stationary (d+1)-dimensional spatial process {(Y_i,X_i),i\in Z^N} observed over a rectangular domain of the form I_n:={i=(i_1,...,i_N)\in…

Statistics Theory · Mathematics 2007-06-13 Marc Hallin , Zudi Lu , Lanh T. Tran

Employing recent results of Robinson (2005) we consider the asymptotic properties of conditional-sum-of-squares (CSS) estimates of parametric models for stationary time series with long memory. CSS estimation has been considered as a rival…

Statistics Theory · Mathematics 2007-06-13 P. M. Robinson

Given a stationary first-order autoregressive process X_t (with lag-one correlation rho satisfying |rho|<1), we examine the Central Limit Theorem for (1/n)*ln |X_1...X_n| and compute variances to high precision. Given a nonstationary…

Dynamical Systems · Mathematics 2007-12-29 Steven R. Finch

In this paper we propose the first non-parametric Bayesian model using Gaussian Processes to make inference on Poisson Point Processes without resorting to gridding the domain or to introducing latent thinning points. Unlike competing…

Machine Learning · Statistics 2015-06-30 Yves-Laurent Kom Samo , Stephen Roberts

Previous analysis on forecasting theory either assume knowing the true parameters or assume the stationarity of the series. Not much are known on the forecasting theory for nonstationary process with estimated parameters. This paper…

Statistics Theory · Mathematics 2007-06-13 Jin-Lung Lin , Ching-Zong Wei

We develop a Bayesian approach to learning from sequential data by using Gaussian processes (GPs) with so-called signature kernels as covariance functions. This allows to make sequences of different length comparable and to rely on strong…

Machine Learning · Statistics 2020-07-07 Csaba Toth , Harald Oberhauser

We present a purely deep neural network-based approach for estimating long memory parameters of time series models that incorporate the phenomenon of long-range dependence. Parameters, such as the Hurst exponent, are critical in…

In this paper, we study finite-sample properties of the least squares estimator in first order autoregressive processes. By leveraging a result from decoupling theory, we derive upper bounds on the probability that the estimate deviates by…

Statistics Theory · Mathematics 2020-05-26 Rodrigo A. González , Cristian R. Rojas

We study the problem of list-decodable Gaussian mean estimation and the related problem of learning mixtures of separated spherical Gaussians. We develop a set of techniques that yield new efficient algorithms with significantly improved…

Data Structures and Algorithms · Computer Science 2017-11-21 Ilias Diakonikolas , Daniel M. Kane , Alistair Stewart

Causal Transformers are trained to predict the next token for a given context. While it is widely accepted that self-attention is crucial for encoding the causal structure of sequences, the precise underlying mechanism behind this…

Machine Learning · Statistics 2025-03-04 Michael E. Sander , Gabriel Peyré