中文
相关论文

相关论文: Model-free Envelope Dimension Selection

200 篇论文

Multi-dimensional classification (MDC) can be employed in a range of applications where one needs to predict multiple class variables for each given instance. Many existing MDC methods suffer from at least one of inaccuracy, scalability,…

机器学习 · 计算机科学 2023-11-28 Vu-Linh Nguyen , Yang Yang , Cassio de Campos

Establishing a low-dimensional representation of the data leads to efficient data learning strategies. In many cases, the reduced dimension needs to be explicitly stated and estimated from the data. We explore the estimation of dimension in…

统计方法学 · 统计学 2022-02-10 Wei Q. Deng , Radu V. Craiu

Biphilic microdome arrays are ubiquitous in nature, but their synthetic counterparts have been scarce. To bridge that gap, we leverage condensed droplet polymerization (CDP) to enable their template-free synthesis. During CDP, monomer…

软凝聚态物质 · 物理学 2025-10-02 Haobo Xu , Haonian Shu , Rong Yang

When model predictions inform downstream decision making, a natural question is under what conditions can the decision-makers simply respond to the predictions as if they were the true outcomes. Calibration suffices to guarantee that simple…

机器学习 · 计算机科学 2025-04-23 Jingwu Tang , Jiayun Wu , Zhiwei Steven Wu , Jiahao Zhang

Bayesian inverse problems use observed data to update a prior probability distribution for an unknown state or parameter of a scientific system to a posterior distribution conditioned on the data. In many applications, the unknown parameter…

数值分析 · 数学 2026-05-12 Josie König , Elizabeth Qian , Melina A. Freitag

Likelihood-free inference provides a rigorous approach to preform Bayesian analysis using forward simulations only. The main advantage of likelihood-free methods is its ability to account for complex physical processes and observational…

宇宙学与河外天体物理 · 物理学 2022-02-09 Sut-Ieng Tam , Keiichi Umetsu , Adam Amara

We develop an envelope model for joint mean and covariance regression in the large $p$, small $n$ setting. In contrast to existing envelope methods, which improve mean estimates by incorporating estimates of the covariance structure, we…

统计方法学 · 统计学 2020-10-02 Alexander Franks

Practical large-scale recommender systems usually contain thousands of feature fields from users, items, contextual information, and their interactions. Most of them empirically allocate a unified dimension to all feature fields, which is…

信息检索 · 计算机科学 2020-10-23 Xiangyu Zhao , Haochen Liu , Hui Liu , Jiliang Tang , Weiwei Guo , Jun Shi , Sida Wang , Huiji Gao , Bo Long

The increasing availability of full-field displacement data from imaging techniques in experimental mechanics is determining a gradual shift in the paradigm of material model calibration and discovery, from using several simple-geometry…

计算工程、金融与科学 · 计算机科学 2025-07-01 Saeid Ghouli , Moritz Flaschel , Siddhant Kumar , Laura De Lorenzis

Model-free knockoffs is a recently proposed technique for identifying covariates that is likely to have an effect on a response variable. The method is an efficient method to control the false discovery rate in hypothesis tests for separate…

统计方法学 · 统计学 2019-03-29 Lars Holden , Kristoffer Hellton

We present novel reductions from sample compression schemes in multiclass classification, regression, and adversarially robust learning settings to binary sample compression schemes. Assuming we have a compression scheme for binary classes…

机器学习 · 计算机科学 2025-04-09 Idan Attias , Steve Hanneke , Arvind Ramaswami

Diffusion Probabilistic Field (DPF) models the distribution of continuous functions defined over metric spaces. While DPF shows great potential for unifying data generation of various modalities including images, videos, and 3D geometry, it…

计算机视觉与模式识别 · 计算机科学 2023-05-25 Kangfu Mei , Mo Zhou , Vishal M. Patel

Feature selection aims to identify the most pattern-discriminative feature subset. In prior literature, filter (e.g., backward elimination) and embedded (e.g., Lasso) methods have hyperparameters (e.g., top-K, score thresholding) and tie to…

机器学习 · 计算机科学 2024-03-07 Wangyang Ying , Dongjie Wang , Haifeng Chen , Yanjie Fu

Graphical models have found widespread applications in many areas of modern statistics and machine learning. Iterative Proportional Fitting (IPF) and its variants have become the default method for undirected graphical model estimation, and…

统计方法学 · 统计学 2024-08-22 Kshitij Khare , Syed Rahman , Bala Rajaratnam , Jiayuan Zhou

In this paper we propose new approaches to estimating large dimensional monotone index models. This class of models has been popular in the applied and theoretical econometrics literatures as it includes discrete choice, nonparametric…

计量经济学 · 经济学 2023-02-22 Shakeeb Khan , Xiaoying Lan , Elie Tamer , Qingsong Yao

This paper proposes a new method for estimating high-dimensional binary choice models. We consider a semiparametric model that places no distributional assumptions on the error term, allows for heteroskedastic errors, and permits endogenous…

计量经济学 · 经济学 2025-07-15 Fu Ouyang , Thomas Tao Yang

External biasing forces are often applied to enhance sampling in regions of phase space which would otherwise be rarely observed. While the typical goal of these experiments is to calculate the potential of mean force (PMF) along the…

统计力学 · 物理学 2010-10-27 David D. L. Minh

Dynamic feature selection (DFS) is a machine learning framework in which features are acquired sequentially for individual samples under budget constraints. The exponential growth in the number of possible feature acquisition paths forces a…

机器学习 · 计算机科学 2026-05-13 Javier Fumanal-Idocin , Raquel Fernandez-Peralta , Javier Andreu-Perez

We introduce a general framework for large-scale model-based derivative-free optimization based on iterative minimization within random subspaces. We present a probabilistic worst-case complexity analysis for our method, where in particular…

最优化与控制 · 数学 2021-02-25 Coralia Cartis , Lindon Roberts

In the context of supervised parametric models, we introduce the concept of e-values. An e-value is a scalar quantity that represents the proximity of the sampling distribution of parameter estimates in a model trained on a subset of…

机器学习 · 统计学 2022-07-19 Subhabrata Majumdar , Snigdhansu Chatterjee