中文
相关论文

相关论文: An improved method for model selection based on In…

200 篇论文

Knowledge reduction of dynamic covering information systems involves with the time in practical situations. In this paper, we provide incremental approaches to computing the type-1 and type-2 characteristic matrices of dynamic coverings…

信息论 · 计算机科学 2023-11-30 Mingjie Cai

In machine learning or scientific computing, model performance is measured with an objective function. But why choose one objective over another? Information theory gives one answer: To maximize the information in the model, select the…

机器学习 · 计算机科学 2024-06-05 Timothy O. Hodson , Thomas M. Over , Tyler J. Smith , Lucy M. Marshall

In this paper, we provide a detailed overview of the models used for information retrieval in the first and second stages of the typical processing chain. We discuss the current state-of-the-art models, including methods based on terms,…

信息检索 · 计算机科学 2024-02-16 Kailash A. Hambarde , Hugo Proenca

Standard selection criteria for forecasting models focus on information that is calculated for each series independently, disregarding the general tendencies and performances of the candidate models. In this paper, we propose a new way to…

统计方法学 · 统计学 2021-04-21 Fotios Petropoulos , Evangelos Spiliotis , Anastasios Panagiotelis

In this paper, we present and prove some consistency results about the performance of classification models using a subset of features. In addition, we propose to use beam search to perform feature selection, which can be viewed as a…

机器学习 · 计算机科学 2022-03-10 Nicolas Fraiman , Zichao Li

For many scientific questions, understanding the underlying mechanism is the goal. To help investigators better understand the underlying mechanism, variable selection is a crucial step that permits the identification of the most associated…

统计方法学 · 统计学 2025-10-06 Shuangshuang Xu , Marco A. R. Ferreira , Allison N. Tegge

Selecting the number of topics in LDA models is considered to be a difficult task, for which alternative approaches have been proposed. The performance of the recently developed singular Bayesian information criterion (sBIC) is evaluated…

计算与语言 · 计算机科学 2023-02-17 Victor Bystrov , Viktoriia Naboka , Anna Staszewska-Bystrova , Peter Winker

Vine copulas allow to build flexible dependence models for an arbitrary number of variables using only bivariate building blocks. The number of parameters in a vine copula model increases quadratically with the dimension, which poses new…

统计方法学 · 统计学 2018-11-20 Thomas Nagler , Christian Bumann , Claudia Czado

Principal component analysis (PCA) is the most commonly used statistical procedure for dimension reduction. An important issue for applying PCA is to determine the rank, which is the number of dominant eigenvalues of the covariance matrix.…

统计方法学 · 统计学 2020-08-06 Hung Hung , Su-Yun Huang , Ching-Kang Ing

Large databases are often organized by hand-labeled metadata, or criteria, which are expensive to collect. We can use unsupervised learning to model database variation, but these models are often high dimensional, complex to parameterize,…

计算机视觉与模式识别 · 计算机科学 2017-06-14 James Tompkin , Kwang In Kim , Hanspeter Pfister , Christian Theobalt

Classification and characterization of variable phenomena and transient phenomena are critical for astrophysics and cosmology. These objects are commonly studied using photometric time series or spectroscopic data. Given that many ongoing…

天体物理仪器与方法 · 物理学 2020-01-08 Javiera Astudillo , Pavlos Protopapas , Karim Pichara , Pablo Huijse

Large amount of unstructured designed information is difficult to deal with. Obtaining specific information is a hard mission and takes a lot of time. Information Retrieval System (IR) is a way to solve this kind of problem. IR is a good…

信息检索 · 计算机科学 2018-04-03 Maher Abdullah , Mohammed GH. I. Al Zamil

Using predictive adaptive arithmetic coding and the Minimum Description Length principle, we derive an efficient tool for model selection problems : the RIC information criterion. We then present an extension of these coding techniques to…

统计方法学 · 统计学 2007-05-23 Guilhem Coq , Olivier Alata , Marc Arnaudon , Christian Olivier

This work presents a systematic study of objective evaluations of abstaining classifications using Information-Theoretic Measures (ITMs). First, we define objective measures for which they do not depend on any free parameter. This…

计算机视觉与模式识别 · 计算机科学 2012-08-16 Bao-Gang Hu , Ran He , XiaoTong Yuan

We present a two-stage approach for learning dictionaries for object classification tasks based on the principle of information maximization. The proposed method seeks a dictionary that is compact, discriminative, and generative. In the…

计算机视觉与模式识别 · 计算机科学 2015-03-20 Qiang Qiu , Vishal M. Patel , Rama Chellappa

Model selection is a pivotal process in the quantitative sciences, where researchers must navigate between numerous candidate models of varying complexity. Traditional information criteria, such as the corrected Akaike Information Criterion…

定量方法 · 定量生物学 2025-12-16 Jakob Vanhoefer , Antonia Körner , Domagoj Doresic , Jan Hasenauer , Dilan Pathirana

Consider the setting where there are B>1 candidate statistical models, and one is interested in model selection. Two common approaches to solve this problem are to select a single model or to combine the candidate models through model…

统计方法学 · 统计学 2021-04-22 Qingying Zong , Jonathan R. Bradley

Model selection is of fundamental importance to high dimensional modeling featured in many contemporary applications. Classical principles of model selection include the Kullback-Leibler divergence principle and the Bayesian principle,…

统计理论 · 数学 2016-05-12 Jinchi Lv , Jun S. Liu

We address the issue of model selection in beta regressions with varying dispersion. The model consists of two submodels, namely: for the mean and for the dispersion. Our focus is on the selection of the covariates for each submodel. Our…

统计计算 · 统计学 2017-02-08 Fábio M. Bayer , Francisco Cribari-Neto

The information bottleneck (IB) method seeks a compressed representation of data that preserves information relevant to a target variable for prediction while discarding irrelevant information from the original data. In its classical…

信息论 · 计算机科学 2026-02-23 Akira Kamatsuka , Takahiro Yoshida