English
Related papers

Related papers: Modeling Random Directions in 2D Simplex Data

200 papers

Prevalence mapping in low resource settings is an increasingly important endeavor to guide policy making and to spatially and temporally characterize the burden of disease. We will focus our discussion on consideration of the complex design…

Methodology · Statistics 2016-08-15 Jon Wakefield , Daniel Simpson , Jessica Godwin

Environmental and climate processes are often distributed over large space-time domains. Their complexity and the amount of available data make modelling and analysis a challenging task. Statistical modelling of environment and climate data…

Methodology · Statistics 2019-10-02 Behnaz Pirzamanbein

We introduce a multivariate hidden Markov model to jointly cluster time-series observations with different support, i.e. circular and linear. Relying on the general projected normal distribution, our approach allows for bimodal and/or…

Applications · Statistics 2015-01-27 Gianluca Mastrantonio , Antonello Maruotti , Giovanna Jona Lasinio

This paper proposes a new Bayesian machine learning model that can be applied to large datasets arising in macroeconomics. Our framework sums over many simple two-component location mixtures. The transition between components is determined…

Econometrics · Economics 2023-12-05 Florian Huber

When dealing with datasets containing a billion instances or with simulations that require a supercomputer to execute, computational resources become part of the equation. We can improve the efficiency of learning and inference by…

Machine Learning · Computer Science 2014-03-06 Max Welling

Irregularly sampled time series data arise naturally in many application domains including biology, ecology, climate science, astronomy, and health. Such data represent fundamental challenges to many classical models from machine learning…

Machine Learning · Computer Science 2021-01-07 Satya Narayan Shukla , Benjamin M. Marlin

Hierarchical learning models, such as mixture models and Bayesian networks, are widely employed for unsupervised learning tasks, such as clustering analysis. They consist of observable and hidden variables, which represent the given data…

Machine Learning · Statistics 2018-01-08 Keisuke Yamazaki

Fine population distribution both in space and in time is crucial for epidemic management, disaster prevention,urban planning and more. Human mobility data have a great potential for mapping population distribution at a high level of…

Applications · Statistics 2020-06-25 Xiang Liu , Philo Pöllmann

In many settings, a data curator links records from two files to produce datasets that are shared with secondary analysts. Analysts use the linked files to estimate models of interest, such as regressions. Such two-stage approaches do not…

Methodology · Statistics 2025-11-18 Xueyan Hu , Jerome P. Reiter

Bimodal truncated count distributions are frequently observed in aggregate survey data and in user ratings when respondents are mixed in their opinion. They also arise in censored count data, where the highest category might create an…

Methodology · Statistics 2014-01-24 Pragya Sur , Galit Shmueli , Smarajit Bose , Paromita Dubey

In public health applications, spatial data collected are often recorded at different spatial scales and over different correlated variables. Spatial change of support is a key inferential problem in these applications and have become…

Methodology · Statistics 2024-03-28 Shijie Zhou , Jonathan R. Bradley

Fairness in data-driven decision-making studies scenarios where individuals from certain population segments may be unfairly treated when being considered for loan or job applications, access to public resources, or other types of services.…

Databases · Computer Science 2022-10-19 Sina Shaham , Gabriel Ghinita , Cyrus Shahabi

The challenges posed by high-dimensional data and use of the simplex constraint are two major concerns in the empirical application of the synthetic control method (SCM) in econometric studies. To address both issues simultaneously, we…

Methodology · Statistics 2025-12-02 Yihong Xu , Quan Zhou

We use random matrix theory to study the spectrum of random geometric graphs, a fundamental model of spatial networks. Considering ensembles of random geometric graphs we look at short range correlations in the level spacings of the…

Physics and Society · Physics 2017-06-08 Carl P. Dettmann , Orestis Georgiou , Georgie Knight

Learning models of complex spatial density functions, representing the steady-state density of mobile nodes moving on a two-dimensional terrain, can assist in network design and optimization problems, e.g., by accelerating the computation…

Networking and Internet Architecture · Computer Science 2024-11-19 Wanxin Gao , Ioanis Nikolaidis , Janelle Harms

Divergence is not only an important mathematical concept in information theory, but also applied to machine learning problems such as low-dimensional embedding, manifold learning, clustering, classification, and anomaly detection. We…

Computation · Statistics 2016-11-22 Kun Yang , Hao Su , Wing Hung Wong

In low-resource settings, prevalence mapping relies on empirical prevalence data from a finite, often spatially sparse, set of surveys of communities within the region of interest, possibly supplemented by remotely sensed images that can…

Applications · Statistics 2015-05-27 Peter J. Diggle , Emanuele Giorgi

Influenced mixed moving average fields are a versatile modeling class for spatio-temporal data. However, their predictive distribution is not generally known. Under this modeling assumption, we define a novel spatio-temporal embedding and a…

Machine Learning · Statistics 2024-08-05 Imma Valentina Curato , Orkun Furat , Lorenzo Proietti , Bennet Stroeh

In some areas of knowledge there are data representing directions restricted to a specific range of values. Consequently, it is useful to have models for describing variables defined in subsets of the k-dimensional unit sphere. This need…

Methodology · Statistics 2025-07-17 Joel Montesinos-Vazquez , Gabriel Núñez-Antonio

We consider Bayesian analysis of a class of multiple changepoint models. While there are a variety of efficient ways to analyse these models if the parameters associated with each segment are independent, there are few general approaches…

Computation · Statistics 2009-10-19 Paul Fearnhead , Zhen Liu