中文
相关论文

相关论文: Analysis in HUGIN of Data Conflict

200 篇论文

Mutual information is commonly used as a measure of similarity between competing labelings of a given set of objects, for example to quantify performance in classification and community detection tasks. As argued recently, however, the…

社会与信息网络 · 计算机科学 2025-07-17 Maximilian Jerdee , Alec Kirkley , M. E. J. Newman

Classifiers are often tested on relatively small data sets, which should lead to uncertain performance metrics. Nevertheless, these metrics are usually taken at face value. We present an approach to quantify the uncertainty of…

机器学习 · 统计学 2021-03-05 Niklas Tötsch , Daniel Hoffmann

We are able to unify various disparate claims and results in the literature, that stand in the way of a unified description and understanding of human conflict. First, we provide a reconciliation of the numerically different exponent values…

物理与社会 · 物理学 2019-11-06 Michael Spagat , Stijn van Weezel , Minzhang Zheng , Neil F. Johnson

Unmeasured confounding is a major challenge for identifying causal relationships from non-experimental data. Here, we propose a method that can accommodate unmeasured discrete confounding. Extending recent identifiability results in deep…

机器学习 · 计算机科学 2024-08-13 Patrick Burauel , Frederick Eberhardt , Michel Besserve

Human disease diagnosis is a complicated process and requires high level of expertise. Any attempt of developing a web-based expert system dealing with human disease diagnosis has to overcome various difficulties. This paper describes a…

人工智能 · 计算机科学 2010-06-24 Mir Anamul Hasan , Khaja Md. Sher-E-Alam , Ahsan Raja Chowdhury

In many areas of science multiple sets of data are collected pertaining to the same system. Examples are food products which are characterized by different sets of variables, bio-processes which are on-line sampled with different…

A growing body of literature attempts to learn about contagion using observational (i.e. non-experimental) data collected from a single social network. While the conclusions of these studies may be correct, the methods rely on assumptions…

应用统计 · 统计学 2017-06-30 Elizabeth L. Ogburn

Due to the complexity of the human body, most diseases present a high inter-personal variability in the way they manifest, i.e. in their phenotype, which has important clinical repercussions - as for instance the difficulty in defining…

物理与社会 · 物理学 2018-06-06 Massimiliano Zanin , Juan Manuel Tuñas , Ernestina Menasalvas

We review possible measures of complexity which might in particular be applicable to situations where the complexity seems to arise spontaneously. We point out that not all of them correspond to the intuitive (or "naive") notion, and that…

数据分析、统计与概率 · 物理学 2012-08-20 Peter Grassberger

In a multi-source environment, each source has its own credibility. If there is no external knowledge about credibility then we can use the information provided by the sources to assess their credibility. In this paper, we propose a way to…

信号处理 · 电气工程与系统科学 2018-03-14 Pan Wei , John E. Ball , Derek T. Anderson , Archit Harsh , Christopher Archibald

In this discussion we consider why it is important to estimate causal effect parameters well even they are not identified, propose a partially identified approach for causal inference in the presence of colliders, point out an…

统计方法学 · 统计学 2017-11-01 Edward H. Kennedy , Sivaraman Balakrishnan

With the growing interest in social applications of Natural Language Processing and Computational Argumentation, a natural question is how controversial a given concept is. Prior works relied on Wikipedia's metadata and on content analysis…

Traditionally, statistical and causal inference on human subjects rely on the assumption that individuals are independently affected by treatments or exposures. However, recently there has been increasing interest in settings, such as…

统计方法学 · 统计学 2020-02-25 Elizabeth L. Ogburn , Ilya Shpitser , Youjin Lee

Estimating how a treatment affects different individuals, known as heterogeneous treatment effect estimation, is an important problem in empirical sciences. In the last few years, there has been a considerable interest in adapting machine…

机器学习 · 计算机科学 2024-10-18 Christopher Tran , Keith Burghardt , Kristina Lerman , Elena Zheleva

Comparing two population means of network data is of paramount importance in a wide range of scientific applications. Many existing network inference solutions focus on global testing of entire networks, without comparing individual network…

统计方法学 · 统计学 2019-10-10 Yin Xia , Lexin Li

Knowledge about existence, strength, and dominant direction of causal influences is of paramount importance for understanding complex systems. With limited amounts of realistic data, however, current methods for investigating causal links…

数据分析、统计与概率 · 物理学 2020-10-20 Erik Laminski , Klaus R. Pawelzik

Measurement involves the determination of quantitative estimates of physical quantities from experiment, along with estimates of their associated uncertainties. Herewith an experimental system model is the key to extracting information from…

应用统计 · 统计学 2008-09-01 Vladimir B. Bokov

Observational data is increasingly used as a means for making individual-level causal predictions and intervention recommendations. The foremost challenge of causal inference from observational data is hidden confounding, whose presence…

机器学习 · 统计学 2018-10-30 Nathan Kallus , Aahlad Manas Puli , Uri Shalit

Large-scale data are often characterized by some degree of inhomogeneity as data are either recorded in different time regimes or taken from multiple sources. We look at regression models and the effect of randomly changing coefficients,…

统计方法学 · 统计学 2016-08-11 Nicolai Meinshausen , Peter Bühlmann

Today, data analysts largely rely on intuition to determine whether missing or withheld rows of a dataset significantly affect their analyses. We propose a framework that can produce automatic contingency analysis, i.e., the range of values…

数据库 · 计算机科学 2020-04-09 Xi Liang , Zechao Shang , Aaron J. Elmore , Sanjay Krishnan , Michael J. Franklin