Related papers: Analyzing LC-MS/MS data by spectral count and ion …
Correlation functions and correlation lengths are frequently used to describe phase transitions in quantum systems, but they require an explicit choice of observables. The recently introduced information lattice instead provides an…
We present a cross-spectra based approach for the analysis of CMB data at large angular scales to constrain the reionization optical depth $\tau$, the tensor to scalar ratio $r$ and the amplitude of the primordial scalar perturbations…
Benchmark datasets are used to profile and compare algorithms across a variety of tasks, ranging from image classification to segmentation, and also play a large role in image pretraining algorithms. Emphasis is placed on results with…
Astrophysical molecular spectroscopy is an important means of searching for new physics through probing the variation of the proton-to-electron mass ratio, $\mu$. New molecular probes could provide tighter constraints on the variation of…
Information coefficient (IC) is a widely used metric for measuring investment managers' skills in selecting stocks. However, its adequacy and effectiveness for evaluating stock selection models has not been clearly understood, as IC from a…
This paper considers the problem of multi-sample nonparametric comparison of counting processes with panel count data, which arise naturally when recurrent events are considered. Such data frequently occur in medical follow-up studies and…
The study of immune cellular composition has been of great scientific interest in immunology because of the generation of multiple large-scale data. From the statistical point of view, such immune cellular data should be treated as…
Statistical moments of the intensity distributions are used as molecular descriptors. They are used as a basis for defining similarity distances between two model spectra. Parameters which carry the information derived from the comparison…
Comparative meta-analyses of groups of subjects by integrating multiple observational studies rely on estimated propensity scores (PSs) to mitigate covariate imbalances. However, PS estimation grapples with the theoretical and practical…
This paper studies an integrated sensing and communication (ISAC) system within a centralized cell-free massive MIMO (multiple-input multiple-output) network for target detection. ISAC transmit access points serve the user equipments in the…
Materials informatics is increasingly used to support modelling, analysis and design across the length scales of materials science, from atomistic simulations to microstructural characterisation and continuum descriptions. Despite rapid…
The Connectivity Map (CMap) is a large publicly available database of cellular transcriptomic responses to chemical and genetic perturbations built using a standardized acquisition protocol known as the L1000 technique. Databases such as…
Metabonomics, the measure of the fingerprint of biochemical perturbations caused by disease, drugs or toxins, recently has become a major focus of research in various areas especially indications of drug toxicity. Two types of technology…
Oral cancer is a significant global health burden, and early detection remains a critical clinical need. Electrical impedance spectroscopy (EIS) offers a promising non-invasive approach for real-time tissue characterization, but…
Two methods of data analysis are compared: spreadsheet software and a statistics software suite. Their use is compared analyzing data collected in three selected experiments taken from an introductory physics laboratory, which include a…
We propose a one-to-many matching estimator of the average treatment effect based on propensity scores estimated by isotonic regression. This approach is predicated on the assumption of monotonicity in the propensity score function, a…
This paper proposes a class of origin-smooth approximators of indicators underlying the sum-of-negative-part statistic for testing multiple inequalities. The need for simulation or bootstrap to obtain test critical values is thereby…
Enrichment of predictive models with new biomolecular markers is an important task in high-dimensional omic applications. Increasingly, clinical studies include several sets of such omics markers available for each patient, measuring…
We study statistics dependence of the probability distributions and the means of measured moments of conserved quantities, respectively. The required statistics of all interested moments and their products are estimated based on a simple…
In real-world applications, as data availability increases, obtaining labeled data for machine learning (ML) projects remains challenging due to the high costs and intensive efforts required for data annotation. Many ML projects,…