English
Related papers

Related papers: Bayesian Multi Plate High Throughput Screening of …

200 papers

Identifying the interaction targets of bioactive compounds is a foundational element for deciphering their pharmacological effects. Target prediction algorithms equip researchers with an effective tool to rapidly scope and explore potential…

Machine Learning · Computer Science 2025-01-03 Haojie Wang , Zhe Zhang , Haotian Gao , Xiangying Zhang , Jingyuan Li , Zhihang Chen , Xinchong Chen , Yifei Qi , Yan Li , Renxiao Wang

Merging datafiles containing information on overlapping sets of entities is a challenging task in the absence of unique identifiers, and is further complicated when some entities are duplicated in the datafiles. Most approaches to this…

Methodology · Statistics 2021-10-11 Serge Aleshin-Guendel , Mauricio Sadinle

Due to the rapid development of high-throughput experimental techniques and fast-dropping prices, many transcriptomic datasets have been generated and accumulated in the public domain. Meta-analysis combining multiple transcriptomic studies…

Applications · Statistics 2019-05-17 Zhiguang Huo , Chi Song , George Tseng

This thesis responds to the challenges of using a large number, such as thousands, of features in regression and classification problems. There are two situations where such high dimensional features arise. One is when high dimensional…

Machine Learning · Statistics 2007-09-20 Longhai Li

Due to the significant resemblance in visual appearance, pill misuse is prevalent and has become a critical issue, responsible for one-third of all deaths worldwide. Pill identification, thus, is a crucial concern needed to be investigated…

Computer Vision and Pattern Recognition · Computer Science 2023-03-20 Anh Duy Nguyen , Huy Hieu Pham , Huynh Thanh Trung , Quoc Viet Hung Nguyen , Thao Nguyen Truong , Phi Le Nguyen

Approximate Bayesian computation is an established and popular method for likelihood-free inference with applications in many disciplines. The effectiveness of the method depends critically on the availability of well performing summary…

Machine Learning · Statistics 2018-05-23 Prashant Singh , Andreas Hellander

The spectral density matrix is a fundamental object of interest in time series analysis, and it encodes both contemporary and dynamic linear relationships between component processes of the multivariate system. In this paper we develop…

Statistics Theory · Mathematics 2025-02-04 Jinyuan Chang , Qing Jiang , Tucker S. McElroy , Xiaofeng Shao

Biocatalysis is a promising approach to sustainably synthesize pharmaceuticals, complex natural products, and commodity chemicals at scale. However, the adoption of biocatalysis is limited by our ability to select enzymes that will catalyze…

Biomolecules · Quantitative Biology 2022-04-06 Samuel Goldman , Ria Das , Kevin K. Yang , Connor W. Coley

Very often for the same scientific question, there may exist different techniques or experiments that measure the same numerical quantity. Historically, various methods have been developed to exploit the information within each type of data…

Methodology · Statistics 2021-09-22 Yiwen Liu , Xiaoxiao Sun , Wenxuan Zhong , Bing Li

Core-level X-ray photoelectron spectroscopy (XPS) is a useful measurement technique for investigating the electronic states of a strongly correlated electron system. Usually, to extract physical information of a target object from a…

Strongly Correlated Electrons · Physics 2019-03-27 Yoh-ichi Mototake , Masaichiro Mizumaki , Ichiro Akai , Masato Okada

A novel approach for calibrating quantum-chemical properties determined as part of a high-throughput virtual screen to experimental analogs is presented. Information on the molecular graph is extracted through the use of extended…

Chemical Physics · Physics 2015-10-05 Edward O. Pyzer-Knapp , Gregor N. Simm , Alan Aspuru-Guzik

We study the application of a Bayesian method to extract relevant information from data for the case of a signal consisting of two or more decaying particles and its background. The method takes advantage of the dependence that exists in…

High Energy Physics - Phenomenology · Physics 2023-06-06 Ezequiel Alvarez

Multilevel compositional data are data that are repeatedly measured or clustered within groups and are non-negative and sum to a constant value. These data arise in various settings, such as intensive, longitudinal studies using ecological…

Methodology · Statistics 2025-02-21 Flora Le , Tyman E. Stanford , Dorothea Dumuid , Joshua F. Wiley

In a multiple testing framework, we propose a method that identifies the interval with the highest estimated false discovery rate of P-values and rejects the corresponding null hypotheses. Unlike the Benjamini-Hochberg method, which does…

Statistics Theory · Mathematics 2018-08-03 Shiyun Chen , Andrew Ying , Ery Arias-Castro

A hypergraph is a generalization of a graph, in which a hyperedge can connect multiple vertices, modeling complex relationships involving multiple vertices simultaneously. Hypergraph pattern matching, which is to find all isomorphic…

Databases · Computer Science 2025-12-23 Siwoo Song , Wonseok Shin , Kunsoo Park , Giuseppe F. Italiano , Zhengyi Yang , Wenjie Zhang

We consider comparisons of statistical learning algorithms using multiple data sets, via leave-one-in cross-study validation: each of the algorithms is trained on one data set; the resulting model is then validated on each remaining data…

Applications · Statistics 2015-06-02 Lorenzo Trippa , Levi Waldron , Curtis Huttenhower , Giovanni Parmigiani

Scene parsing has attracted a lot of attention in computer vision. While parametric models have proven effective for this task, they cannot easily incorporate new training data. By contrast, nonparametric approaches, which bypass any…

Computer Vision and Pattern Recognition · Computer Science 2016-03-16 Mohammad Najafi , Sarah Taghavi Namin , Mathieu Salzmann , Lars Petersson

Motivated by problems in data clustering, we establish general conditions under which families of nonparametric mixture models are identifiable, by introducing a novel framework involving clustering overfitted \emph{parametric} (i.e.…

Statistics Theory · Mathematics 2020-02-19 Bryon Aragam , Chen Dan , Eric P. Xing , Pradeep Ravikumar

We study the problem of performing automated experiment design for drug screening through Bayesian inference and optimisation. In particular, we compare and contrast the behaviour of linear-Gaussian models and Gaussian processes, when used…

Machine Learning · Computer Science 2021-05-11 Hannes Eriksson , Christos Dimitrakakis , Lars Carlsson

Count outcomes in longitudinal studies are frequent in clinical and engineering studies. In frequentist and Bayesian statistical analysis, methods such as Mixed linear models allow the variability or correlation within individuals to be…

Methodology · Statistics 2024-07-15 Alejandra Estefanía Patiño Hoyos , Johnatan Cardona Jiménez