中文
相关论文

相关论文: Wisdom of the Silicon Crowd: LLM Ensemble Predicti…

200 篇论文

Despite their performance, large language models (LLMs) can inadvertently perpetuate biases found in the data they are trained on. By analyzing LLM responses to bias-eliciting headlines, we find that these models often mirror human biases.…

计算与语言 · 计算机科学 2025-05-20 Axel Abels , Tom Lenaerts

The puzzling idea that the combination of independent estimates of the magnitude of a quantity results in a very accurate prediction, which is superior to any or, at least, to most of the individual estimates is known as the wisdom of…

应用统计 · 统计学 2021-05-26 Sandro M. Reia , José F. Fontanari

Large language models (LLMs) match and sometimes exceeding human performance in many domains. This study explores the potential of LLMs to augment human judgement in a forecasting task. We evaluate the effect on human forecasters of two LLM…

计算机与社会 · 计算机科学 2024-08-23 Philipp Schoenegger , Peter S. Park , Ezra Karger , Sean Trott , Philip E. Tetlock

Structured deliberation has been found to improve the performance of human forecasters. This study investigates whether a similar intervention, i.e. allowing LLMs to review each other's forecasts before updating, can improve accuracy in…

人工智能 · 计算机科学 2025-12-30 Paul Schneider , Amalie Schramm

This study investigates the forecasting accuracy of human experts versus Large Language Models (LLMs) in the retail sector, particularly during standard and promotional sales periods. Utilizing a controlled experimental setup with 123 human…

机器学习 · 计算机科学 2024-05-20 MAhdi Abolghasemi , Odkhishig Ganbold , Kristian Rotaru

Human groups are able to converge on more accurate beliefs through deliberation, even in the presence of polarization and partisan bias -- a phenomenon known as the "wisdom of partisan crowds." Generated agents powered by Large Language…

Wisdom of the crowd, the collective intelligence derived from responses of multiple human or machine individuals to the same questions, can be more accurate than each individual, and improve social decision-making and prediction accuracy.…

机器学习 · 统计学 2021-10-29 Lingfei Wang , Tom Michoel

The interest in the wisdom of crowds stems mainly from the possibility of combining independent forecasts from experts in the hope that many expert minds are better than a few. Hence the relevant subject of study nowadays is the Vox…

数据分析、统计与概率 · 物理学 2023-01-20 Nilton S. Siqueira Neto , José F. Fontanari

The emergence of large language models (LLMs) has sparked much interest in creating LLM-based digital populations that can be applied to many applications such as social simulation, crowdsourcing, marketing, and recommendation systems. A…

多智能体系统 · 计算机科学 2026-01-15 Ryan Feng Lin , Keyu Tian , Hanming Zheng , Congjing Zhang , Li Zeng , Shuai Huang

Large language models (LLMs) have demonstrated remarkable capabilities across diverse tasks, but their ability to forecast future events remains understudied. A year ago, large language models struggle to come close to the accuracy of a…

机器学习 · 计算机科学 2025-08-06 Janna Lu

Large Language Models (LLMs) offer a promising alternative to traditional survey methods, potentially enhancing efficiency and reducing costs. In this study, we use LLMs to create virtual populations that answer survey questions, enabling…

人机交互 · 计算机科学 2025-03-24 Enzo Sinacola , Arnault Pachot , Thierry Petit

Large language models (LLMs) achieve strong average performance yet remain unreliable at the instance level, with frequent hallucinations, brittle failures, and poorly calibrated confidence. We study reliability through the lens of…

人工智能 · 计算机科学 2026-01-13 Pranav Kallem

Large Language Models (LLMs) have shown capabilities close to human performance in various analytical tasks, leading researchers to use them for time and labor-intensive analyses. However, their capability to handle highly specialized and…

计算与语言 · 计算机科学 2024-10-08 Alexander S. Choi , Syeda Sabrina Akter , JP Singh , Antonios Anastasopoulos

The wisdom of crowds is the idea that the combination of independent estimates of the magnitude of some quantity yields a remarkably accurate prediction, which is always more accurate than the average individual estimate. In addition, it is…

信息论 · 计算机科学 2020-12-29 Davi A. Nobre , José F. Fontanari

The ability to discern subtle emotional cues is fundamental to human social intelligence. As artificial intelligence (AI) becomes increasingly common, AI's ability to recognize and respond to human emotions is crucial for effective human-AI…

人工智能 · 计算机科学 2025-08-13 Mustafa Akben , Vinayaka Gude , Haya Ajjan

Forecasting future events is important for policy and decision making. In this work, we study whether language models (LMs) can forecast at the level of competitive human forecasters. Towards this goal, we develop a retrieval-augmented LM…

机器学习 · 计算机科学 2024-02-29 Danny Halawi , Fred Zhang , Chen Yueh-Han , Jacob Steinhardt

Advances in deep learning systems have allowed large models to match or surpass human accuracy on a number of skills such as image classification, basic programming, and standardized test taking. As the performance of the most capable…

机器学习 · 计算机科学 2024-06-10 Sarah Pratt , Seth Blumberg , Pietro Kreitlon Carolino , Meredith Ringel Morris

Crowd predictions have demonstrated powerful performance in predicting future events. We aim to understand crowd prediction efficacy in ascertaining the veracity of human emotional expressions. We discover that collective discernment can…

人机交互 · 计算机科学 2018-08-17 Zhenyue Qin , Tom Gedeon , Sabrina Caldwell

"Synthetic samples" based on large language models (LLMs) have been argued to serve as efficient alternatives to surveys of humans, assuming that their training data includes information on human attitudes and behavior. However,…

计算机与社会 · 计算机科学 2025-04-21 Leah von der Heyde , Anna-Carolina Haensch , Alexander Wenz , Bolei Ma

Large language models (LLMs) are increasingly used to provide instructions to many agents who interact with one another. Such shared reliance couples agents who appear to act independently: they may in fact be guided by a common model. This…

计算机科学与博弈论 · 计算机科学 2026-05-08 Jonathan Shaki , Eden Hartman , Sarit Kraus , Yonatan Aumann
‹ 上一页 1 2 3 10 下一页 ›