English
Related papers

Related papers: Trends in Frontier AI Model Count: A Forecast to 2…

200 papers

Context: With artificial intelligence (AI) being well established within the daily lives of research communities, we turn our gaze toward formal methods (FM). FM aim to provide sound and verifiable reasoning about problems in computer…

Logic in Computer Science · Computer Science 2025-08-29 Sebastian Stock , Jannik Dunkelau , Atif Mashkoor

Do leading LLM developers possess a proprietary ``secret sauce'', or is LLM performance driven by scaling up compute? Using training and benchmark data for 809 models released between 2022 and 2025, we estimate scaling-law regressions with…

Artificial Intelligence · Computer Science 2026-05-05 Matthias Mertens , Natalia Fischl-Lanzoni , Neil Thompson

METR's time horizon metric has grown exponentially since 2019, along with compute. However, it is unclear whether compute scaling will persist at current rates through 2030, raising the question of how possible compute slowdowns might…

Computers and Society · Computer Science 2025-11-26 Parker Whitfill , Ben Snodin , Joel Becker

The Model Context Protocol (MCP), introduced by Anthropic in November 2024 and now governed by the Linux Foundation's Agentic AI Foundation, has rapidly become the de facto standard for connecting large language model (LLM)-based agents to…

Cryptography and Security · Computer Science 2026-04-08 Nirajan Acharya , Gaurav Kumar Gupta

Over time, the distribution of medical image data drifts due to factors such as shifts in patient demographics, acquisition devices, and disease manifestations. While human radiologists can adjust their expertise to accommodate such…

Foundation models, particularly large language models, are increasingly integrated into agent architectures for industrial tasks such as decision support, process monitoring, and engineering automation. Yet evidence on their purposes,…

The rapid proliferation of AI models has underscored the importance of thorough documentation, as it enables users to understand, trust, and effectively utilize these models in various applications. Although developers are encouraged to…

Software Engineering · Computer Science 2024-02-09 Weixin Liang , Nazneen Rajani , Xinyu Yang , Ezinwanne Ozoani , Eric Wu , Yiqun Chen , Daniel Scott Smith , James Zou

This technical report presents methods developed by the UK AI Security Institute for assessing whether advanced AI systems reliably follow intended goals. Specifically, we evaluate whether frontier models sabotage safety research when…

Artificial Intelligence · Computer Science 2026-04-02 Alexandra Souly , Robert Kirk , Jacob Merizian , Abby D'Cruz , Xander Davies

Power is becoming an increasingly important concern for large supercomputing centers. Due to cost concerns, data centers are becoming increasingly limited in their ability to enhance their power infrastructure to support increased compute…

Applications · Statistics 2015-05-13 Curtis Storlie , Joe Sexton , Scott Pakin , Michael Lang , Brian Reich , William Rust

At face value, this essay is about understanding a fairly esoteric governance tool called compute thresholds. However, in order to grapple with whether these thresholds will achieve anything, we must first understand how they came to be. To…

Artificial Intelligence · Computer Science 2024-07-31 Sara Hooker

Fault detection in automotive engine systems is one of the most promising research areas. Several works have been done in the field of model-based fault diagnosis. Many researchers have discovered more advanced statistical methods and…

Machine Learning · Computer Science 2024-08-26 Priyanka Kumar

Which jobs can AI learn to do? We examine this for every occupation in the US economy. Existing indices measure the overlap between AI capabilities and occupational tasks rather than which tasks AI systems can learn to perform, and as a…

General Economics · Economics 2026-05-05 Philip Moreira Tomei , Bouke Klein Teeselink

Rapidly increasing AI capabilities have substantial real-world consequences, ranging from AI safety concerns to labor market consequences. The Model Evaluation & Threat Research (METR) report argues that AI capabilities have exhibited…

Artificial Intelligence · Computer Science 2026-02-09 Haosen Ge , Hamsa Bastani , Osbert Bastani

For an AI solution to evolve from a trained machine learning model into a production-ready AI system, many more things need to be considered than just the performance of the machine learning model. A production-ready AI system needs to be…

Artificial Intelligence · Computer Science 2023-03-24 Petra Heck , Gerard Schouten

In June 2024, the EU AI Act came into force. The AI Act includes obligations for the provider of an AI system. Article 10 of the AI Act includes a new obligation for providers to evaluate whether their training, validation and testing…

Computers and Society · Computer Science 2025-02-28 Marvin van Bekkum

Foundation models (FMs) provide societal benefits but also amplify risks. Governments, companies, and researchers have proposed regulatory frameworks, acceptable use policies, and safety benchmarks in response. However, existing public…

Computers and Society · Computer Science 2024-08-07 Yi Zeng , Yu Yang , Andy Zhou , Jeffrey Ziwei Tan , Yuheng Tu , Yifan Mai , Kevin Klyman , Minzhou Pan , Ruoxi Jia , Dawn Song , Percy Liang , Bo Li

Fog and edge computing require adaptive control schemes that can handle partial observability, severe latency requirements, and dynamically changing workloads. Recent research on Agentic AI (AAI) increasingly integrates reasoning systems…

Distributed, Parallel, and Cluster Computing · Computer Science 2026-01-29 Saeed Akbar , Muhammad Waqas , Rahmat Ullah

As foundation models grow in both popularity and capability, researchers have uncovered a variety of ways that the models can pose a risk to the model's owner, user, or others. Despite the efforts of measuring these risks via benchmarks and…

Cryptography and Security · Computer Science 2025-06-04 David Piorkowski , Michael Hind , John Richards , Jacquelyn Martino

As AI workloads increase in scope, generalization capability becomes challenging for small task-specific models and their demand for large amounts of labeled training samples increases. On the contrary, Foundation Models (FMs) are trained…

Artificial Intelligence · Computer Science 2024-04-19 Aristeidis Tsaris , Philipe Ambrozio Dias , Abhishek Potnis , Junqi Yin , Feiyi Wang , Dalton Lunga

The proliferation of Machine Learning (ML) models and their open-source implementations has transformed Artificial Intelligence research and applications. Platforms like Hugging Face (HF) enable this evolving ecosystem, yet a large-scale…

Software Engineering · Computer Science 2025-11-11 Joel Castaño , Rafael Cabañas , Antonio Salmerón , David Lo , Silverio Martínez-Fernández