English
Related papers

Related papers: The notion of validity in experimental crowd dynam…

200 papers

Recent protocols and metrics for training and evaluating autonomous robot navigation through crowds are inconsistent due to diversified definitions of "social behavior". This makes it difficult, if not impossible, to effectively compare…

Robotics · Computer Science 2022-11-29 Junxian Wang , Wesley P. Chan , Pamela Carreno-Medrano , Akansel Cosgun , Elizabeth Croft

Focusing on a specific crowd dynamics situation, including real life experiments and measurements, our paper targets a twofold aim: (1) we present a Bayesian probabilistic method to estimate the value and the uncertainty (in the form of a…

Data Analysis, Statistics and Probability · Physics 2018-04-12 Alessandro Corbetta , Adrian Muntean , Federico Toschi , Kiamars Vafayi

Scientific feasibility assessment asks whether a claim is consistent with established knowledge and whether experimental evidence could support or refute it. We frame feasibility assessment as a diagnostic reasoning task in which, given a…

Computation and Language · Computer Science 2026-04-22 Seyedali Mohammadi , Manas Gaur , Francis Ferraro

The rapid evolution of artificial intelligence has led to expectations of transformative impact on science, yet current systems remain fundamentally limited in enabling genuine scientific discovery. This perspective contends that progress…

Artificial Intelligence · Computer Science 2025-12-16 Karthik Duraisamy

Science is a fundamental human activity and we trust its results because it has several error-correcting mechanisms. Its is subject to experimental tests that are replicated by independent parts. Given the huge amount of information…

Physics and Society · Physics 2011-03-25 Andre C. R. Martins

Controlled experiments are a core research method in software engineering (SE) for validating causal claims. However, recruiting a sample of participants that represents the intended target population is often difficult or expensive, which…

Software Engineering · Computer Science 2026-04-27 Julian Frattini , Richard Torkar , Robert Feldt , Carlo A. Furia

Assessing the validity of user simulators when used for the evaluation of information retrieval systems remains an open question, constraining their effective use and the reliability of simulation-based results. To address this issue, we…

Information Retrieval · Computer Science 2026-01-19 Andreas Konstantin Kruff , Nolwenn Bernard , Philipp Schaer

Large language models (LLMs) are rapidly being integrated into psychological research as research tools, evaluation targets, human simulators, and cognitive models. However, recent evidence reveals severe measurement unreliability:…

Human-Computer Interaction · Computer Science 2025-07-08 Zhicheng Lin

Modeling and simulation approaches that express crowd movement with mathematical models are widely and actively studied to understand crowd movement and resolve crowd accidents. Existing literature on crowd modeling focuses on only the…

Multiagent Systems · Computer Science 2023-02-27 Ryo Nishida , Masaki Onishi , Koichi Hashimoto

We here discuss the outcome of an hypothetic experiments of populations dynamics, where a set of independent realizations is made available. The importance of ensemble average is clarified with reference to the registered time evolution of…

Statistical Mechanics · Physics 2008-10-31 Timoteo Carletti , Duccio Fanelli

The notion of experiment precision quantifies the variance of user ratings in a subjective experiment. Although there exist measures that assess subjective experiment precision, there are no systematic analyses of these measures available…

Multimedia · Computer Science 2022-08-05 Jakub Nawała , Tobias Hoßfeld , Lucjan Janowski , Michael Seufert

Reproducibility is central to the credibility of scientific findings, yet complete replication studies are costly and infrequent. However, many biological experiments contain internal replication, which is defined as repetition across…

Applications · Statistics 2025-06-05 Stanley E. Lazic

The NLP community typically relies on performance of a model on a held-out test set to assess generalization. Performance drops observed in datasets outside of official test sets are generally attributed to "out-of-distribution" effects.…

Computation and Language · Computer Science 2024-04-03 Aparna Elangovan , Jiayuan He , Yuan Li , Karin Verspoor

Although the methodology of Design Science Research (DSR) is playing an increasingly important role with the emergence of the "sciences of the artificial", the validity of the resulting artifacts is occasionally questioned. This paper…

Other Computer Science · Computer Science 2025-04-15 Sylvana Kroop

Many published research results are false, and controversy continues over the roles of replication and publication policy in improving the reliability of research. Addressing these problems is frustrated by the lack of a formal framework…

Other Statistics · Statistics 2015-08-27 Richard McElreath , Paul E. Smaldino

This paper provides an overview and critical analysis on the modeling and applications of the dynamics of human crowds, where social interactions can have an important influence on the behavioral dynamics of the crowd viewed as a living,…

Physics and Society · Physics 2017-09-21 Nicola Bellomo , Livio Gibelli , Nisrine Outada

Nowadays, crowd sensing becomes increasingly more popular due to the ubiquitous usage of mobile devices. However, the quality of such human-generated sensory data varies significantly among different users. To better utilize sensory data,…

Cryptography and Security · Computer Science 2018-10-12 Yaliang Li , Houping Xiao , Zhan Qin , Chenglin Miao , Lu Su , Jing Gao , Kui Ren , Bolin Ding

How should researchers analyze randomized experiments in which the main outcome is latent and measured in multiple ways but each measure contains some degree of error? We first identify a critical study-specific noncomparability problem in…

Econometrics · Economics 2026-01-13 Jiawei Fu , Donald P. Green

Generative models at times produce "invalid" outputs, such as images with generation artifacts and unnatural sounds. Validity-constrained distribution learning attempts to address this problem by requiring that the learned distribution have…

Machine Learning · Computer Science 2024-10-22 Nick Rittler , Kamalika Chaudhuri

Security especially in the fields of IoT, industrial automation and critical infrastructure is paramount nowadays and a hot research topic. In order to ensure confidence in research results they need to be reproducible. In the past we…

Hardware Architecture · Computer Science 2024-07-10 Dmytro Petryk , Ievgen Kabin , Peter Langendörfer , Zoya Dyka