English
Related papers

Related papers: We Should Evaluate Real-World Impact

200 papers

Although most people support climate action, widespread underestimation of others' support stalls individual and systemic changes. In this preregistered experiment, we test whether large language models (LLMs) can reliably predict these…

Computers and Society · Computer Science 2026-01-29 Nattavudh Powdthavee , Sandra J. Geiger

Many networks in natural and human-made systems exhibit scale-free properties and are small worlds. Now we show that people's understanding of complex systems in their cognitive maps also follow a scale-free topology. People focus on a few…

Other Quantitative Biology · Quantitative Biology 2007-05-23 Uygar Ozesmi

Model interpretations are often used in practice to extract real world insights from machine learning models. These interpretations have a wide range of applications; they can be presented as business recommendations or used to evaluate…

Machine Learning · Computer Science 2020-11-20 Brian Liu , Madeleine Udell

While Large Language Models (LLMs) are fundamentally next-token prediction systems, their practical applications extend far beyond this basic function. From natural language processing and text generation to conversational assistants and…

Computation and Language · Computer Science 2025-03-10 Vishakha Agrawal , Archie Chaudhury , Shreya Agrawal

NLP has a significant role in advancing healthcare and has been found to be key in extracting structured information from radiology reports. Understanding recent developments in NLP application to radiology is of significance but recent…

Much recent work seeks to evaluate values and opinions in large language models (LLMs) using multiple-choice surveys and questionnaires. Most of this work is motivated by concerns around real-world LLM applications. For example,…

Computation and Language · Computer Science 2024-06-06 Paul Röttger , Valentin Hofmann , Valentina Pyatkin , Musashi Hinck , Hannah Rose Kirk , Hinrich Schütze , Dirk Hovy

In Natural Language Processing (NLP) classification tasks such as topic categorisation and sentiment analysis, model generalizability is generally measured with standard metrics such as Accuracy, F-Measure, or AUC-ROC. The diversity of…

Computation and Language · Computer Science 2024-01-09 Peter Vickers , Loïc Barrault , Emilio Monti , Nikolaos Aletras

The impact of large language models (LLMs) on critical thinking has provoked growing attention, yet this impact on actual performance may not be uniformly negative or positive. Particularly, the role of time -- the temporal context under…

Human-Computer Interaction · Computer Science 2026-03-11 Jiayin Zhi , Harsh Kumar , Mina Lee

Recently, a growing number of experts in artificial intelligence (AI) and medicine have be-gun to suggest that the use of AI systems, particularly machine learning (ML) systems, is likely to humanise the practice of medicine by…

Computers and Society · Computer Science 2025-04-11 Joshua Hatherley

Both scientific progress and individual researcher careers depend on the quality of peer review, which in turn depends on paper-reviewer matching. Surprisingly, this problem has been mostly approached as an automated recommendation problem…

Computation and Language · Computer Science 2022-05-04 Terne Sasha Thorn Jakobsen , Anna Rogers

Is explainability a false promise? This debate has emerged from the insufficient evidence that explanations help people in situations they are introduced for. More human-centered, application-grounded evaluations of explanations are needed…

Computation and Language · Computer Science 2024-11-06 Fateme Hashemi Chaleshtori , Atreya Ghosal , Alexander Gill , Purbid Bambroo , Ana Marasović

The proliferation of open-source Large Language Models (LLMs) from various institutions has highlighted the urgent need for comprehensive evaluation methods. However, current evaluation platforms, such as the widely recognized HuggingFace…

Computation and Language · Computer Science 2024-11-01 Fanghua Ye , Mingming Yang , Jianhui Pang , Longyue Wang , Derek F. Wong , Emine Yilmaz , Shuming Shi , Zhaopeng Tu

Altmetrics are tools for measuring the impact of research beyond scientific communities. In general, they measure online mentions of scholarly outputs, such as on online social networks, blogs, and news sites. Some stakeholders in higher…

Digital Libraries · Computer Science 2021-12-20 Grischa Fraumann

Although researchers and practitioners are pushing the boundaries and enhancing the capacities of NLP tools and methods, works on African languages are lagging. A lot of focus on well resourced languages such as English, Japanese, German,…

Computation and Language · Computer Science 2020-04-03 Ignatius Ezeani , Paul Rayson , Ikechukwu Onyenwe , Chinedu Uchechukwu , Mark Hepple

Computer-assisted reading and analysis of text has various applications in the humanities and social sciences. The increasing size of many electronic text archives has the advantage of a more complete analysis but the disadvantage of taking…

Databases · Computer Science 2007-05-23 Steven Keith , Owen Kaser , Daniel Lemire

Recent studies show that Natural Language Processing (NLP) technologies propagate societal biases about demographic groups associated with attributes such as gender, race, and nationality. To create interventions and mitigate these biases…

Computation and Language · Computer Science 2022-10-17 Sunipa Dev , Emily Sheng , Jieyu Zhao , Aubrie Amstutz , Jiao Sun , Yu Hou , Mattie Sanseverino , Jiin Kim , Akihiro Nishi , Nanyun Peng , Kai-Wei Chang

Artificial Intelligence (AI) finds widespread application across various domains, but it sparks concerns about fairness in its deployment. The prevailing discourse in classification often emphasizes outcome-based metrics comparing sensitive…

Machine Learning · Computer Science 2024-12-18 Sofie Goethals , Marco Favier , Toon Calders

Recommender systems have generated tremendous value for both users and businesses, drawing significant attention from academia and industry alike. However, due to practical constraints, academic research remains largely confined to offline…

Information Retrieval · Computer Science 2025-09-09 Kuan Zou , Aixin Sun

This paper examines the thin-slicing approach - the ability to make accurate judgments based on minimal information - in the context of scientific presentations. Drawing on research from nonverbal communication and personality psychology,…

Computation and Language · Computer Science 2025-04-16 Ralf Schmälzle , Sue Lim , Yuetong Du , Gary Bente