Related papers: Age-stratified clustering of multiple long-term co…
In electronic health records (EHRs), clustering patients and distinguishing disease subtypes are key tasks to elucidate pathophysiology and aid clinical decision-making. However, clustering in healthcare informatics is still based on…
There are many cluster analysis methods that can produce quite different clusterings on the same dataset. Cluster validation is about the evaluation of the quality of a clustering; "relative cluster validation" is about using such criteria…
This paper derives a formula for computing the conditional probability of a set of candidates, where a candidate is a set of disorders that explain a given set of positive findings. Such candidate sets are produced by a recent method for…
In data containing heterogeneous subpopulations, classification performance benefits from incorporating the knowledge of cluster structure in the classifier. Previous methods for such combined clustering and classification either 1) are…
We present results based on YJKs photometry of star clusters located in the outermost, eastern region of the Small Magellanic Cloud (SMC). We analysed a total of 51 catalogued clusters whose colour--magnitude diagrams (CMDs), having been…
The last decades have not only been characterized by an explosive growth of data, but also an increasing appreciation of data as a valuable resource. Their value comes with the ability to extract meaningful patterns that are of economic,…
The age distribution function of star clusters in the Large Magellanic Cloud (LMC) is known to present a feature called the cluster age gap, a period of time from ~ 4 to 11 Gyr ago with a remarkable small number of clusters identified. In…
Recent studies have started to cast doubt on the assumption that most stars are formed in clusters. Observational studies of field stars and star cluster systems in nearby galaxies can lead to better constraints on the fraction of stars…
Tree-structured models are a powerful alternative to parametric regression models if non-linear effects and interactions are present in the data. Yet, classical tree-structured models might not be appropriate if data comes in clusters of…
Globular cluster ages provide both an important test of models of globular cluster formation and a powerful method to constrain the assembly history of galaxies. Unfortunately, measuring the ages of unresolved old stellar populations has…
The process of manually searching for relevant instances in, and extracting information from, clinical databases underpin a multitude of clinical tasks. Such tasks include disease diagnosis, clinical trial recruitment, and continuing…
We analyzed HST/WFPC2 colour-magnitude diagrams (CMDs) of 15 populous Large Magellanic Cloud (LMC) stellar clusters with ages between ~ 0.3 Gyr and ~ 3 Gyr. These (V, B-V) CMDs are photometrically homogeneous and typically reach V \~ 22.…
Most star clusters at an intermediate age (1-2 Gyr) in the Large and Small Magellanic Clouds show a puzzling feature in their color-magnitude diagrams (CMD) that is not in agreement with a simple stellar population. The main sequence…
Aggregated health data such as claims data from health insurances become more and more available for research purposes. Estimates of excess mortality from prevalence and incidence of a chronic condition have only been possible for ages 50…
Peer-grouping is used in many sectors for organisational learning, policy implementation, and benchmarking. Clustering provides a statistical, data-driven method for constructing meaningful peer groups, but peer groups must be compatible…
Finite mixture models that allow for a broad range of potentially non-elliptical cluster distributions is an emerging methodological field. Such methods allow for the shape of the clusters to match the natural heterogeneity of the data,…
Selecting subsets of features that differentiate between two conditions is a key task in a broad range of scientific domains. In many applications, the features of interest form clusters with similar effects on the data at hand. To recover…
A novel approach rooted on the notion of consensus clustering, a strategy developed for community detection in complex networks, is proposed to cope with the heterogeneity that characterizes connectivity matrices in health and disease. The…
Comparing clusterings is central to evaluating unsupervised models, yet the many existing similarity measures can produce widely divergent, sometimes contradictory, evaluations. Clustering similarity measures are typically organized into…
We derive photometric, structural and dynamical evolution-related parameters of 11 nearby open clusters with ages in the range 70 Myr to 7 Gyr and masses in the range $\approx400$ \ms to $\approx5 300$ \ms. We search for relations of…