中文
相关论文

相关论文: Zipf's Law in Importance of Genes for Cancer Class…

200 篇论文

The power law is useful in describing count phenomena such as network degrees and word frequencies. With a single parameter, it captures the main feature that the frequencies are linear on the log-log scale. Nevertheless, there have been…

应用统计 · 统计学 2024-07-24 Clement Lee , Emma Eastoe , Aiden Farrell

Machine Learning methods have of late made significant efforts to solving multidisciplinary problems in the field of cancer classification using microarray gene expression data. Feature subset selection methods can play an important role in…

计算工程、金融与科学 · 计算机科学 2013-03-04 G. Prat , Ll. Belanche

Statistical methods for analyzing large-scale biomolecular data are commonplace in computational biology. A notable example is phenotype prediction from gene expression data, for instance, detecting human cancers, differentiating subtypes…

基因组学 · 定量生物学 2014-11-24 Bahman Afsari , Ulisses M. Braga-Neto , Donald Geman

Bayesian modelling and statistical text analysis rely on informed probability priors to encourage good solutions. This paper empirically analyses whether text in medical discharge reports follow Zipf's law, a commonly assumed statistical…

According to Zipf's meaning-frequency law, words that are more frequent tend to have more meanings. Here it is shown that a linear dependency between the frequency of a form and its number of meanings is found in a family of models of…

计算与语言 · 计算机科学 2016-10-14 Ramon Ferrer-i-Cancho

The frequency distributions of DNA k-mers are shaped by fundamental biological processes and offer a window into genome structure and evolution. Inspired by analogies to natural language, prior studies have attempted to model genomic k-mer…

Zipf's law is a paradigm describing the importance of different elements in communication systems, especially in linguistics. Despite the complexity of the hierarchical structure of language, music has in some sense an even more complex…

物理与社会 · 物理学 2023-11-20 Marc Serra-Peralta , Joan Serrà , Álvaro Corral

Various approaches to gene selection for cancer classification based on microarray data can be found in the literature and they may be grouped into two categories: univariate methods and multivariate methods. Univariate methods look at each…

定量方法 · 定量生物学 2015-06-18 Min Xu , Rudy Setiono

It has been shown recently that a specific class of path-dependent stochastic processes, which reduce their sample space as they unfold, lead to exact scaling laws in frequency and rank distributions. Such Sample Space Reducing processes…

物理与社会 · 物理学 2017-10-02 Bernat Corominas-Murtra , Rudolf Hanel , Stefan Thurner

Zipf's law describes the empirical size distribution of the components of many systems in natural and social sciences and humanities. We show, by solving a statistical model, that Zipf's law co-occurs with the maximization of the diversity…

The analysis of the leukemia data from Whitehead/MIT group is a discriminant analysis (also called a supervised learning). Among thousands of genes whose expression levels are measured, not all are needed for discriminant analysis: a gene…

生物物理 · 物理学 2007-05-23 Wentian Li , Yaning Yang

Cancer detection is one of the key research topics in the medical field. Accurate detection of different cancer types is valuable in providing better treatment facilities and risk minimization for patients. This paper deals with the…

定量方法 · 定量生物学 2022-05-31 Yasamin Kowsari , Sanaz Nakhodchi , Davoud Gholamiangonabadi

Identification of essential genes is one of the ultimate goals of drug designs. Here we introduce an {\it in silico} method to select essential genes through the microarray assay. We construct a graph of genes, called the gene transcription…

统计力学 · 物理学 2007-05-23 K. Rho , H. Jeong , B. Kahng

History-dependent processes are ubiquitous in natural and social systems. Many such stochastic processes, especially those that are associated with complex systems, become more constrained as they unfold, meaning that their sample-space, or…

物理与社会 · 物理学 2015-04-16 Bernat Corominas-Murtra , Rudolf Hanel , Stefan Thurner

A common practice in microarray analysis is to transform the microarray raw data (light intensity) by a logarithmic transformation, and the justification for this transformation is to make the distribution more symmetric and Gaussian-like.…

定量方法 · 定量生物学 2016-11-17 Wentian Li , Young Ju Suh , Jingshan Zhang

Recent works have highlighted optimization difficulties faced by gradient descent in training the first and last layers of transformer-based language models, which are overcome by optimizers such as Adam. These works suggest that the…

机器学习 · 计算机科学 2025-05-27 Frederik Kunstner , Francis Bach

The Cancer Genome Atlas (TCGA) provides researchers with clinicopathological data and genomic characterizations of various carcinomas. These data sets include expression microarrays for genes and microRNAs -- short, non-coding strands of…

定量方法 · 定量生物学 2013-07-05 Siddharth G. Reddy , Weimin Xiao , Preethi H. Gunaratne

An important body of quantitative linguistics is constituted by a series of statistical laws about language usage. Despite the importance of these linguistic laws, some of them are poorly formulated, and, more importantly, there is no…

物理与社会 · 物理学 2020-11-09 Alvaro Corral , Isabel Serra

Heaps' or Herdan's law is a linguistic law describing the relationship between the vocabulary/dictionary size (type) and word counts (token) to be a power-law function. Its existence in genomes with certain definition of DNA words is…

基因组学 · 定量生物学 2024-07-02 Wentian Li , Yannis Almirantis , Astero Provata

miRNA and gene expression profiles have been proved useful for classifying cancer samples. Efficient classifiers have been recently sought and developed. A number of attempts to classify cancer samples using miRNA/gene expression profiles…

计算工程、金融与科学 · 计算机科学 2014-01-21 Rania Ibrahim , Noha A. Yousri , Mohamed A. Ismail , Nagwa M. El-Makky