English
Related papers

Related papers: Counterfactual Likelihood Tests for Indirect Influ…

200 papers

The capacity to address counterfactual "what if" inquiries is crucial for understanding and making use of causal influences. Traditional counterfactual inference, under Pearls' counterfactual framework, typically depends on having access to…

Machine Learning · Computer Science 2024-02-29 Shaoan Xie , Biwei Huang , Bin Gu , Tongliang Liu , Kun Zhang

Large Language Models have demonstrated remarkable capabilities across diverse tasks, yet they frequently generate hallucinations outputs that are fluent but factually incorrect or unsupported. We propose Counterfactual Probing, a novel…

Computation and Language · Computer Science 2025-08-05 Yijun Feng

Learning-based approaches, such as reinforcement and imitation learning are gaining popularity in decision-making for autonomous driving. However, learned policies often fail to generalize and cannot handle novel situations well. Asking and…

Machine Learning · Computer Science 2020-11-13 Patrick Hart , Alois Knoll

Likelihood-to-evidence ratio estimation is usually cast as either a binary (NRE-A) or a multiclass (NRE-B) classification task. In contrast to the binary classification framework, the current formulation of the multiclass version has an…

Machine Learning · Statistics 2024-07-08 Benjamin Kurt Miller , Christoph Weniger , Patrick Forré

Correlation between channel state and source symbol is under investigation for a joint source-channel coding problem. We investigate simultaneously the lossless transmission of information and the empirical coordination of channel inputs…

Information Theory · Computer Science 2016-11-17 Maël Le Treust

Since 2016, the amount of academic research with the keyword "misinformation" has more than doubled [2]. This research often focuses on article headlines shown in artificial testing environments, yet misinformation largely spreads through…

Human-Computer Interaction · Computer Science 2020-12-16 Emily Saltz , Claire Leibowicz , Claire Wardle

Despite the increasing effectiveness of language models, their reasoning capabilities remain underdeveloped. In particular, causal reasoning through counterfactual question answering is lacking. This work aims to bridge this gap. We first…

Computation and Language · Computer Science 2025-03-18 Alihan Hüyük , Xinnuo Xu , Jacqueline Maasch , Aditya V. Nori , Javier González

The TRUST democratic discourse analysis pipeline exposes its large language model (LLM) components to peer model identity through multiple structural channels -- a design feature whose bias implications have not previously been empirically…

Computers and Society · Computer Science 2026-04-28 Juergen Dietrich

Many empirical networks are intrinsically pluralistic, with interactions occurring within groups of arbitrary agents. Then the agent in the network can be influenced by types of neighbors, common examples include similarity, opposition, and…

Physics and Society · Physics 2020-05-12 Shuo Liu , Xiwang Guan , Shuangling Luo , Haoxiang Xia

Recent interpretability work has identified model-internal handles on post-trained behavior, including refusal directions, assistant/persona axes, and sparse chat-tuning features. These results localize where behaviors can be read out or…

Machine Learning · Computer Science 2026-05-11 Yifan Zhou

Counterfactual thinking describes a psychological phenomenon that people re-infer the possible results with different solutions about things that have already happened. It helps people to gain more experience from mistakes and thus to…

Machine Learning · Computer Science 2019-08-19 Yue Wang , Yao Wan , Chenwei Zhang , Lixin Cui , Lu Bai , Philip S. Yu

Robustness and counterfactual bias are usually evaluated on a test dataset. However, are these evaluations robust? If the test dataset is perturbed slightly, will the evaluation results keep the same? In this paper, we propose a "double…

Computation and Language · Computer Science 2021-04-13 Chong Zhang , Jieyu Zhao , Huan Zhang , Kai-Wei Chang , Cho-Jui Hsieh

Deliberative processes are often discussed as increasing or decreasing polarization. This approach misses a different, and arguably more diagnostic, dimension of opinion change: whether deliberation reshuffles who agrees with whom, or…

Social and Information Networks · Computer Science 2026-01-21 Mohak Goyal , Lodewijk Gelauff , Naman Gupta , Ashish Goel , Kamesh Munagala

Informally, a 'spurious correlation' is the dependence of a model on some aspect of the input data that an analyst thinks shouldn't matter. In machine learning, these have a know-it-when-you-see-it character; e.g., changing the gender of a…

Machine Learning · Computer Science 2021-11-04 Victor Veitch , Alexander D'Amour , Steve Yadlowsky , Jacob Eisenstein

Can stated preferences inform counterfactual analyses of actual choice? This research proposes a novel approach to researchers who have access to both stated choices in hypothetical scenarios and actual choices, matched or unmatched. The…

Econometrics · Economics 2025-11-18 Romuald Meango , Marc Henry , Ismael Mourifie

While various approaches have recently been studied for bias identification, little is known about how implicit language that does not explicitly convey a viewpoint affects bias amplification in large language models. To examine the…

Computation and Language · Computer Science 2024-08-19 Abeer Aldayel , Areej Alokaili , Rehab Alahmadi

Estimating counterfactual distributions under interventions is central to treatment risk assessment and counterfactual generation tasks. Existing approaches model the counterfactual distribution as a standalone generative target, without…

Machine Learning · Statistics 2026-05-11 Hugh Dance , Johnny Xi , Peter Orbanz , Benjamin Bloem-Reddy

Familiar statistical tests and estimates are obtained by the direct observation of cases of interest: a clinical trial of a new drug, for instance, will compare the drug's effects on a relevant set of patients and controls. Sometimes,…

Methodology · Statistics 2010-12-09 Bradley Efron

Experimental user studies evaluating the effectiveness of different subtypes of post-hoc explanations for black-box models are largely nonexistent. Therefore, the aim of this study was to investigate and evaluate how different types of…

Human-Computer Interaction · Computer Science 2026-04-14 Tabea E. Röber , Paul Festor , Rob Goedhart , S. İlker Birbil , Aldo Faisal

Regression and Bayesian accounts of in-context learning (ICL) explain how demonstrations can induce predictors, while mechanistic analyses often identify compact activation directions that steer prompted behavior. However, it remains…

Machine Learning · Computer Science 2026-05-20 Wei Tang , Xinyan Jiang , Fakhri Karray , Lijie Hu
‹ Prev 1 3 4 5 6 7 10 Next ›