English
Related papers

Related papers: Towards Accountable AI: Hybrid Human-Machine Analy…

200 papers

Effective human-AI collaboration hinges on the ability to dynamically integrate the complementary strengths of human experts and AI models across diverse decision contexts. Context-aware weighted combination of human and AI outputs is a…

Human-Computer Interaction · Computer Science 2025-11-07 Renlong Jie

Machine learning workflow development is anecdotally regarded to be an iterative process of trial-and-error with humans-in-the-loop. However, we are not aware of quantitative evidence corroborating this popular belief. A quantitative…

Machine Learning · Computer Science 2018-05-21 Doris Xin , Litian Ma , Shuchen Song , Aditya Parameswaran

This paper investigates the user experience of visualizations of a machine learning (ML) system that recognizes objects in images. This is important since even good systems can fail in unexpected ways as misclassifications on photo-sharing…

Human-Computer Interaction · Computer Science 2020-08-06 Hendrik Heuer , Andreas Breiter

Production machine learning (ML) systems fail silently -- not with crashes, but through wrong decisions. While observability is recognized as critical for ML operations, there is a lack empirical evidence of what practitioners actually…

Software Engineering · Computer Science 2025-10-29 Joran Leest , Ilias Gerostathopoulos , Patricia Lago , Claudia Raibulet

In recent years, the use of sophisticated statistical models that influence decisions in domains of high societal relevance is on the rise. Although these models can often bring substantial improvements in the accuracy and efficiency of…

Machine Learning · Computer Science 2021-04-13 Alfredo Carrillo , Luis F. Cantú , Alejandro Noriega

Reliable and robust evaluation methods are a necessary first step towards developing machine learning models that are themselves robust and reliable. Unfortunately, current evaluation protocols typically used to assess classifiers fail to…

Machine Learning · Computer Science 2025-05-26 Michael W. Spratling

The development and operation of Liquid-Argon Time-Projection Chambers for neutrino physics has created a need for new approaches to pattern recognition in order to fully exploit the imaging capabilities offered by this technology. Whereas…

High Energy Physics - Experiment · Physics 2023-02-17 MicroBooNE collaboration , R. Acciarri , C. Adams , R. An , J. Anthony , J. Asaadi , M. Auger , L. Bagby , S. Balasubramanian , B. Baller , C. Barnes , G. Barr , M. Bass , F. Bay , M. Bishai , A. Blake , T. Bolton , L. Camilleri , D. Caratelli , B. Carls , R. Castillo Fernandez , F. Cavanna , H. Chen , E. Church , D. Cianci , E. Cohen , G. H. Collin , J. M. Conrad , M. Convery , J. I. Crespo-Anadon , M. Del Tutto , D. Devitt , S. Dytman , B. Eberly , A. Ereditato , L. Escudero Sanchez , J. Esquivel , A. A. Fadeeva , B. T. Fleming , W. Foreman , A. P. Furmanski , D. Garcia-Gomez , G. T. Garvey , V. Genty , D. Goeldi , S. Gollapinni , N. Graf , E. Gramellini , H. Greenlee , R. Grosso , R. Guenette , A. Hackenburg , P. Hamilton , O. Hen , V Hewes , C. Hill , J. Ho , G. Horton-Smith , A. Hourlier , E. -C. Huang , C. James , J. Jan de Vries , C. -M. Jen , L. Jiang , R. A. Johnson , J. Joshi , H. Jostlein , D. Kaleko , G. Karagiorgi , W. Ketchum , B. Kirby , M. Kirby , T. Kobilarcik , I. Kreslo , A. Laube , Y. Li , A. Lister , B. R. Littlejohn , S. Lockwitz , D. Lorca , W. C. Louis , M. Luethi , B. Lundberg , X. Luo , A. Marchionni , C. Mariani , J. Marshall , D. A. Martinez Caicedo , V. Meddage , T. Miceli , G. B. Mills , J. Moon , M. Mooney , C. D. Moore , J. Mousseau , R. Murrells , D. Naples , P. Nienaber , J. Nowak , O. Palamara , V. Paolone , V. Papavassiliou , S. F. Pate , Z. Pavlovic , E. Piasetzky , D. Porzio , G. Pulliam , X. Qian , J. L. Raaf , A. Rafique , L. Rochester , C. Rudolf von Rohr , B. Russell , D. W. Schmitz , A. Schukraft , W. Seligman , M. H. Shaevitz , J. Sinclair , A. Smith , E. L. Snider , M. Soderberg , S. Soldner-Rembold , S. R. Soleti , P. Spentzouris , J. Spitz , J. St. John , T. Strauss , A. M. Szelc , N. Tagg , K. Terao , M. Thomson , M. Toups , Y. -T. Tsai , S. Tufanli , T. Usher , W. Van De Pontseele , R. G. Van de Water , B. Viren , M. Weber , D. A. Wickremasinghe , S. Wolbers , T. Wongjirad , K. Woodruff , T. Yang , L. Yates , G. P. Zeller , J. Zennamo , C. Zhang

Sequential multi-agent systems built with large language models (LLMs) can automate complex software tasks, but they are hard to trust because errors quietly pass from one stage to the next. We study a traceable and accountable pipeline,…

Artificial Intelligence · Computer Science 2025-10-10 Amine Barrak

In this position paper, we argue that human baselines in foundation model evaluations must be more rigorous and more transparent to enable meaningful comparisons of human vs. AI performance, and we provide recommendations and a reporting…

Mechanistic interpretability is often motivated for alignment auditing, where a model's verbal explanations can be absent, incomplete, or misleading. Yet many evaluations do not control whether black-box prompting alone can recover the…

Machine Learning · Computer Science 2026-04-14 Ziqian Zhong , Aashiq Muhamed , Mona T. Diab , Virginia Smith , Aditi Raghunathan

Software organizations are increasingly incorporating machine learning (ML) into their product offerings, driving a need for new data management tools. Many of these tools facilitate the initial development of ML applications, but…

Software Engineering · Computer Science 2022-07-19 Shreya Shankar , Aditya Parameswaran

While the most visible part of the safety verification process of automated vehicles concerns the planning and control system, it is often overlooked that safety of the latter crucially depends on the fault-tolerance of the preceding…

Robotics · Computer Science 2021-11-25 Cornelius Buerkle , Florian Geissler , Michael Paulitsch , Kay-Ulrich Scholl

Developing and fielding complex systems requires proof that they are reliably correct with respect to their design and operating requirements. Especially for autonomous systems which exhibit unanticipated emergent behavior, fully…

Software Engineering · Computer Science 2024-02-28 Matthew Litton , Doron Drusinsky , James Bret Michael

Responsible Artificial Intelligence (AI) - the practice of developing, evaluating, and maintaining accurate AI systems that also exhibit essential properties such as robustness and explainability - represents a multifaceted challenge that…

Machine Learning · Computer Science 2022-01-19 Ryan Soklaski , Justin Goodwin , Olivia Brown , Michael Yee , Jason Matterer

Machine learning (ML) provides us with numerous opportunities, allowing ML systems to adapt to new situations and contexts. At the same time, this adaptability raises uncertainties concerning the run-time product quality or dependability,…

Software Engineering · Computer Science 2022-10-18 Lalli Myllyaho , Mikko Raatikainen , Tomi Männistö , Jukka K. Nurminen , Tommi Mikkonen

Despite significant improvements in robot capabilities, they are likely to fail in human-robot collaborative tasks due to high unpredictability in human environments and varying human expectations. In this work, we explore the role of…

Robotics · Computer Science 2023-09-20 Parag Khanna , Elmira Yadollahi , Mårten Björkman , Iolanda Leite , Christian Smith

Machine learning models are often implemented in cohort with humans in the pipeline, with the model having an option to defer to a domain expert in cases where it has low confidence in its inference. Our goal is to design mechanisms for…

Machine Learning · Computer Science 2021-12-14 Vijay Keswani , Matthew Lease , Krishnaram Kenthapadi

Human-AI collaboration has the potential to transform various domains by leveraging the complementary strengths of human experts and Artificial Intelligence (AI) systems. However, unobserved confounding can undermine the effectiveness of…

Human-Computer Interaction · Computer Science 2025-02-27 Ruijiang Gao , Mingzhang Yin

Many safety failures in machine learning arise when models are used to assign predictions to people (often in settings like lending, hiring, or content moderation) without accounting for how individuals can change their inputs. In this…

Machine Learning · Computer Science 2025-07-04 Seung Hyun Cheon , Meredith Stewart , Bogdan Kulynych , Tsui-Wei Weng , Berk Ustun

Automated decision systems increasingly rely on human oversight to ensure accuracy in uncertain cases. This paper presents a practical framework for optimizing such human-in-the-loop classification systems using a double-threshold policy.…

Human-Computer Interaction · Computer Science 2026-01-13 Goran Muric , Steven Minton