中文
相关论文

相关论文: Measuring religious morality using very limited po…

200 篇论文

Statistical modeling is often used to measure the strength of evidence for or against hypotheses on given data. We have previously proposed an information-dynamic framework in support of a properly calibrated measurement scale for…

统计理论 · 数学 2023-07-19 V. J Vieland , S-C. Seok

In Recommender System (RS), explanations help users understand why items are recommended and can enhance a system's transparency, persuasiveness, engagement, and trust, which are known as explanation goals. However, evaluating the…

信息检索 · 计算机科学 2025-12-17 André Levi Zanon , Marcelo Garcia Manzato , Leonardo Rocha

Surveys are commonly used to facilitate research in epidemiology, health, and the social and behavioral sciences. Often, these surveys are not simple random samples, and respondents are given weights reflecting their probability of…

统计方法学 · 统计学 2024-08-20 Adway S. Wadekar , Jerome P. Reiter

Statistical significance testing is used in natural language processing (NLP) to determine whether the results of a study or experiment are likely to be due to chance or if they reflect a genuine relationship. A key step in significance…

计算与语言 · 计算机科学 2024-01-01 Palash Goyal , Qian Hu , Rahul Gupta

One of the key issues in decision problems is the selection and use of the appropriate response scale. In this paper verbal expressions are converted into numerical scales for a subjective problem instance. The main motivation for our…

最优化与控制 · 数学 2025-07-16 Zsombor Szádoczki , Sándor Bozóki , László Sipos , Zsófia Galambosi

The question of selecting the "best" amongst different choices is a common problem in statistics. In drug development, our motivating setting, the question becomes, for example: what is the dose that gives me a pre-specified risk of…

统计理论 · 数学 2018-03-15 Pavel Mozgunov , Thomas Jaki

As Large Language Models (LLMs) increasingly appear in social science research (e.g., economics and marketing), it becomes crucial to assess how well these models replicate human behavior. In this work, using hypothesis testing, we present…

计算机与社会 · 计算机科学 2025-06-19 Harbin Hong , Sebastian Caldas , Liu Leqi

Though it has been recognized that recommending serendipitous (i.e., surprising and relevant) items can be helpful for increasing users' satisfaction and behavioral intention, how to measure serendipity in the offline environment is still…

人机交互 · 计算机科学 2020-04-23 Li Chen , Ningxia Wang , Yonghua Yang , Keping Yang , Quan Yuan

This paper is a quantitative analysis of the data collected globally by the World Value Survey. The data is used to study the trajectories of change in individuals' religious beliefs, values, and behaviors in societies. Utilizing random…

机器学习 · 计算机科学 2023-10-18 Elaheh Jafarigol , William Keely , Tess Hartog , Tom Welborn , Peyman Hekmatpour , Theodore B. Trafalis

Nearly all statistical analyses that inform policy-making are based on imperfect data. As examples, the data may suffer from measurement errors, missing values, sample selection bias, or record linkage errors. Analysts have to decide how to…

统计方法学 · 统计学 2025-10-24 Adway S. Wadekar , Jerome P. Reiter

In multi-center clinical trials, due to various reasons, the individual-level data are strictly restricted to be assessed publicly. Instead, the summarized information is widely available from published results. With the advance of…

统计方法学 · 统计学 2021-01-05 Jing Qin , Yukun Liu , Pengfei Li

In the last decade, the use of simple rating and comparison surveys has proliferated on social and digital media platforms to fuel recommendations. These simple surveys and their extrapolation with machine learning algorithms shed light on…

社会与信息网络 · 计算机科学 2019-01-29 Nandana Sengupta , Nati Srebro , James Evans

Belief functions are a powerful and popular framework for the mathematical characterisation of uncertainty, in particular in situations in which lack of data renders learning a probability distribution for the problem impractical. The first…

统计理论 · 数学 2026-05-11 Fabio Cuzzolin

We introduce an Item Response Theory (IRT)-based framework to detect and quantify socioeconomic bias in large language models (LLMs) without relying on subjective human judgments. Unlike traditional methods, IRT accounts for item…

人工智能 · 计算机科学 2025-03-18 Jasmin Wachter , Michael Radloff , Maja Smolej , Katharina Kinder-Kurlanda

This work presents a systematic study of objective evaluations of abstaining classifications using Information-Theoretic Measures (ITMs). First, we define objective measures for which they do not depend on any free parameter. This…

计算机视觉与模式识别 · 计算机科学 2012-08-16 Bao-Gang Hu , Ran He , XiaoTong Yuan

We seek to democratise public-opinion research by providing practitioners with a general methodology to make representative inference from cheap, high-frequency, highly unrepresentative samples. We focus specifically on samples which are…

统计方法学 · 统计学 2023-09-13 Roberto Cerina , Raymond Duch

We present EvalMORAAL, a transparent chain-of-thought (CoT) framework that uses two scoring methods (log-probabilities and direct ratings) plus a model-as-judge peer review to evaluate moral alignment in 20 large language models. We assess…

计算与语言 · 计算机科学 2026-05-21 Hadi Mohammadi , Anastasia Giachanou , Robert A. Bagheri

Since Rendle and Krichene argued that commonly used sampling-based evaluation metrics are "inconsistent" with respect to the global metrics (even in expectation), there have been a few studies on the sampling-based recommender system…

信息检索 · 计算机科学 2023-10-12 Dong Li , Ruoming Jin , Zhenming Liu , Bin Ren , Jing Gao , Zhi Liu

We study the task of retrieving relevant experiments given a query experiment. By experiment, we mean a collection of measurements from a set of `covariates' and the associated `outcomes'. While similar experiments can be retrieved by…

机器学习 · 统计学 2014-02-20 Sohan Seth , John Shawe-Taylor , Samuel Kaski

How can we assess the reliability of a dataset without access to ground truth? We introduce the problem of reliability scoring for datasets collected from potentially strategic sources. The true data are unobserved, but we see outcomes of…

机器学习 · 计算机科学 2025-10-21 Yiling Chen , Shi Feng , Paul Kattuman , Fang-Yi Yu