English
Related papers

Related papers: A Conversation with Richard A. Olshen

200 papers

Understanding human mobility patterns is important in applications as diverse as urban planning, public health, and political organizing. One rich source of data on human mobility is taxi ride data. Using the city of Chicago as a case…

Social and Information Networks · Computer Science 2023-06-22 Harish Chauhan , Nikunj Gupta , Zoe Haskell-Craig

One of the cornerstones in combating the HIV pandemic is being able to assess the current state and evolution of local HIV epidemics. This remains a complex problem, as many HIV infected individuals remain unaware of their infection status,…

Populations and Evolution · Quantitative Biology 2019-10-24 Pieter Libin , Nassim Versbraegen , Ana B. Abecasis , Perpetua Gomes , Tom Lenaerts , Ann Nowé

As language models become more general purpose, increased attention needs to be paid to detecting out-of-distribution (OOD) instances, i.e., those not belonging to any of the distributions seen during training. Existing methods for…

Machine Learning · Computer Science 2024-07-19 Aryan Gulati , Xingjian Dong , Carlos Hurtado , Sarath Shekkizhar , Swabha Swayamdipta , Antonio Ortega

Motivated by the massive deployment of power-hungry data centers for service provisioning, we examine the problem of routing in optical networks with the aim of minimizing traffic-driven power consumption. To tackle this issue, routing must…

Networking and Internet Architecture · Computer Science 2016-05-06 Panayotis Mertikopoulos , Aris L. Moustakas , Anna Tzanakaki

Jayaram Sethuraman was born in the town of Hubli in Bombay Province (now Karnataka State) on October 3, 1937. His early years were spent in Hubli and in 1950 his family moved to Madras (now renamed Chennai). He graduated from Madras…

Methodology · Statistics 2008-09-01 Myles Hollander

In binary and ordinal regression one can distinguish between a location component and a scaling component. While the former determines the location within the range of the response categories, the scaling indicates variance heterogeneity.…

Methodology · Statistics 2019-10-31 Gerhard Tutz , Moritz Berger

Rooted phylogenetic networks provide an explicit representation of the evolutionary history of a set $X$ of sampled species. In contrast to phylogenetic trees which show only speciation events, networks can also accommodate reticulate…

Combinatorics · Mathematics 2021-01-01 Peter L. Erdos , Charles Semple , Mike Steel

Breast cancer is among the most deadly diseases, distressing mostly women worldwide. Although traditional methods for detection have presented themselves as valid for the task, they still commonly present low accuracies and demand…

Machine Learning · Computer Science 2021-01-15 Leandro Aparecido Passos , Danilo Samuel Jodas , Luiz C. F. Ribeiro , Thierry Pinheiro , João P. Papa

The aim of this study is to look at predicting whether a person will complete a drug and alcohol rehabilitation program and the number of times a person attends. The study is based on demographic data obtained from Substance Abuse and…

Machine Learning · Computer Science 2024-04-25 Karen Roberts-Licklider , Theodore Trafalis

This book is a collection of papers dedicated to the memory of Yehuda Vardi. Yehuda was the chair of the Department of Statistics of Rutgers University when he passed away unexpectedly on January 13, 2005. On October 21--22, 2005, some 150…

Statistics Theory · Mathematics 2007-08-22 Regina Liu , William Strawderman , Cun-Hui Zhang

Thin spanning trees lie at the intersection of graph theory, approximation algorithms, and combinatorial optimization. They are central to the long-standing \emph{thin tree conjecture}, which asks whether every $k$-edge-connected graph…

Data Structures and Algorithms · Computer Science 2025-10-15 Mohit Daga

Graph alignment - identifying node correspondences between two graphs - is a fundamental problem with applications in network analysis, biology, and privacy research. While substantial progress has been made in aligning correlated…

Information Theory · Computer Science 2026-03-16 Jakob Maier , Laurent Massoulié

One-class Classification (OCC) is an area of machine learning which addresses prediction based on unbalanced datasets. Basically, OCC algorithms achieve training by means of a single class sample, with potentially some additional…

Machine Learning · Statistics 2020-03-27 Sarah Itani , Fabian Lecron , Philippe Fortemps

An Objective Structured Practical Examination (OSPE) is an effective and robust, but resource-intensive, means of evaluating anatomical knowledge. Since most OSPEs employ short answer or fill-in-the-blank style questions, the format…

Machine Learning · Computer Science 2021-06-02 Jason Bernard , Ranil Sonnadara , Anthony N. Saraco , Josh P. Mitchell , Alex B. Bak , Ilana Bayer , Bruce C. Wainman

In this paper we retrace the recent history of statistics by analyzing all the papers published in five prestigious statistical journals since 1970, namely: Annals of Statistics, Biometrika, Journal of the American Statistical Association,…

Applications · Statistics 2017-09-13 Laura Anderlucci , Angela Montanari , Cinzia Viroli

Detecting test-time distribution shift has emerged as a key capability for safely deployed machine learning models, with the question being tackled under various guises in recent years. In this paper, we aim to provide a consolidated view…

Computer Vision and Pattern Recognition · Computer Science 2024-09-02 Hongjun Wang , Sagar Vaze , Kai Han

This paper considers the problem of clustering a partially observed unweighted graph---i.e., one where for some node pairs we know there is an edge between them, for some others we know there is no edge, and for the remaining we do not know…

Machine Learning · Computer Science 2014-07-25 Yudong Chen , Ali Jalali , Sujay Sanghavi , Huan Xu

In open-set semi-supervised learning (OSSL), we consider unlabeled datasets that may contain unknown classes. Existing OSSL methods often use the softmax confidence for classifying data as in-distribution (ID) or out-of-distribution (OOD).…

Machine Learning · Computer Science 2026-01-26 Erik Wallin , Lennart Svensson , Fredrik Kahl , Lars Hammarstrand

Inferential summaries of tree estimates are useful in the setting of evolutionary biology, where phylogenetic trees have been built from DNA data since the 1960's. In bioinformatics, psychometrics and data mining, hierarchical clustering…

Applications · Statistics 2010-06-08 John Chakerian , Susan Holmes

Tailoring treatments to individual needs is a central goal in fields such as medicine. A key step toward this goal is estimating Heterogeneous Treatment Effects (HTE) - the way treatments impact different subgroups. While crucial, HTE…

Machine Learning · Statistics 2025-07-30 Tomer Meir , Uri Shalit , Malka Gorfine