相关论文: Distribution of Korean Family Names
Large scale databases are available that contain homologous gene families constructed from hundreds of complete genome sequences from across the three domains of Life. Here we discuss approches of increasing complexity aimed at extracting…
A common sample descriptor in human genomics studies is that of 'genetic ancestry group', with terms such as 'European genetic ancestry' or 'East Asian genetic ancestry' frequently used in publications to describe the genetics of groups of…
South Korea has become one of the most important economies in Asia. The largest Korean multinational firms are affiliated with influential family-owned business groups known as the chaebol. Despite the surging academic popularity of the…
Qian, Luscombe and Gerstein [J. Molecular Biol. 313 (2001) 673--681] introduced a model of the diversification of protein folds in a genome that we may formulate as follows. Consider a multitype Yule process starting with one individual in…
This paper investigates the rank distribution, cumulative probability, and probability density of price returns for the stocks traded in the KSE and the KOSDAQ market. This research demonstrates that the rank distribution is consistent…
Fractional counting of citations can improve on ranking of multi-disciplinary research units (such as universities) by normalizing the differences among fields of science in terms of differences in citation behavior. Furthermore,…
The study of human mobility is both of fundamental importance and of great potential value. For example, it can be leveraged to facilitate efficient city planning and improve prevention strategies when faced with epidemics. The newfound…
We investigated the temporally evolving network structures of the Japanese and Korean stock markets through the minimum spanning trees composed of listed stocks. We tested the validity of conventional grouping by industrial categories, and…
This article brings forward an estimation of the proportion of homonyms in large scale groups based on the distribution of first names and last names in a subset of these groups. The estimation is based on the generalization of the…
Background: A wide range of diseases show some degree of clustering in families; family history is therefore an important aspect for clinicians when making risk predictions. Familial aggregation is often quantified in terms of a familial…
This study deals with a fairly simply formulated problem -- how to estimate the number of people bearing the same full name in a large population. Estimation of name popularity can leverage personal name matching in databases and be of…
Familial Searching is the process of searching in a DNA database for relatives of a certain individual. It is well known that in order to evaluate the genetic evidence in favour of a certain given form of relatedness between two…
The number of extant individuals within a lineage, as exemplified by counts of species numbers across genera in a higher taxonomic category, is known to be a highly skewed distribution. Because the sublineages (such as genera in a clade)…
We present a method to analyse the scientific contributions between research groups. Given multiple research groups, we construct their journal/proceeding graphs and then compute the similarity/gap between them using network analysis. This…
In this article, we investigate the properties of phoneme N-grams across half of the world's languages. We investigate if the sizes of three different N-gram distributions of the world's language families obey a power law. Further, the…
Families form the basis of society, and anthropologists have characterised various family systems. This study developed a multi-level evolutionary model of pre-industrial agricultural societies to simulate the evolution of family systems…
We analyze the social mechanisms that shape the popularity rise and fall of the names given to newborn babies. During the initial stage, popularity increases by imitation. As the people with the same name grow in number, however, its usage…
A transmission interval for an infectious disease is important to understand epidemic processes in complex networks. The transmission interval is defined as a time interval between one person's infection and their infection to another…
We report investigations on the statistical characteristics of the baby names given between 1910 and 2010 in the United States of America. For each year, the 100 most frequent names in the USA are sorted out. For these names, the…
Regarding "Surname distribution in population genetics and in statistical physics" by Rossi, we add comments on the following two points: One is a difference in his renormalization-group formulation from the traditional one in statistical…