中文
相关论文

相关论文: An Empirical Study of Safetensors' Usage Trends an…

200 篇论文

Many have observed that the development and deployment of generative machine learning (ML) and artificial intelligence (AI) models follow a distinctive pattern in which pre-trained models are adapted and fine-tuned for specific downstream…

社会与信息网络 · 计算机科学 2025-08-12 Benjamin Laufer , Hamidah Oderinwale , Jon Kleinberg

Model stores offer third-party ML models and datasets for easy project integration, minimizing coding efforts. One might hope to find detailed specifications of these models and datasets in the documentation, leveraging documentation…

软件工程 · 计算机科学 2024-06-19 Ernesto Lang Oreamuno , Rohan Faiyaz Khan , Abdul Ali Bangash , Catherine Stinson , Bram Adams

Software engineering (SE) activities have been revolutionized by the advent of pre-trained models (PTMs), defined as large machine learning (ML) models that can be fine-tuned to perform specific SE tasks. However, users with limited…

软件工程 · 计算机科学 2024-05-24 Claudio Di Sipio , Riccardo Rubei , Juri Di Rocco , Davide Di Ruscio , Phuong T. Nguyen

According to Gartner, more than 70% of organizations will have integrated AI models into their workflows by the end of 2025. In order to reduce cost and foster innovation, it is often the case that pre-trained models are fetched from model…

密码学与安全 · 计算机科学 2026-01-09 Mohamed Nabeel , Oleksii Starov

The proliferation of Machine Learning (ML) models and their open-source implementations has transformed Artificial Intelligence research and applications. Platforms like Hugging Face (HF) enable this evolving ecosystem, yet a large-scale…

软件工程 · 计算机科学 2025-11-11 Joel Castaño , Rafael Cabañas , Antonio Salmerón , David Lo , Silverio Martínez-Fernández

The rise of capabilities expressed by large language models has been quickly followed by the integration of the same complex systems into application level logic. Algorithms, programs, systems, and companies are built around structured…

软件工程 · 计算机科学 2024-02-28 Kaiser Pister , Dhruba Jyoti Paul , Patrick Brophy , Ishan Joshi

Machine Learning (ML) already has been integrated into all kinds of systems, helping developers to solve problems with even higher accuracy than human beings. However, when integrating ML models into a system, developers may accidentally…

密码学与安全 · 计算机科学 2019-08-07 Mingtian Tan , Zhe Zhou

Recent progress in natural language processing has been driven by advances in both model architecture and model pretraining. Transformer architectures have facilitated building higher-capacity models and pretraining has made it possible to…

As machine learning (ML) becomes an integral part of high-autonomy systems, it is critical to ensure the trustworthiness of learning-enabled software systems (LESS). Yet, the nondeterministic and run-time-defined semantics of ML complicate…

软件工程 · 计算机科学 2025-12-10 Nan Jia , Anita Raja , Raffi Khatchadourian

The proliferation of open Pre-trained Language Models (PTLMs) on model registry platforms like Hugging Face (HF) presents both opportunities and challenges for companies building products around them. Similar to traditional software…

软件工程 · 计算机科学 2025-02-20 Adekunle Ajibode , Abdul Ali Bangash , Filipe Roseiro Cogo , Bram Adams , Ahmed E. Hassan

To enhance the performance of large language models (LLMs) in various domain-specific applications, sensitive data such as healthcare, law, and finance are being used to privately customize or fine-tune these models. Such privately adapted…

密码学与安全 · 计算机科学 2025-12-09 Huifeng Zhu , Shijie Li , Qinfeng Li , Yier Jin

Machine learning tools are becoming increasingly powerful and widely used. Unfortunately membership attacks, which seek to uncover information from data sets used in machine learning, have the potential to limit data sharing. In this paper…

计算机视觉与模式识别 · 计算机科学 2021-08-03 Dennis Conway , Loic Simon , Alexis Lechervy , Frederic Jurie

Mechanistic interpretability research requires reliable tools for analyzing transformer internals across diverse architectures. Current approaches face a fundamental tradeoff: custom implementations like TransformerLens ensure consistent…

机器学习 · 计算机科学 2025-12-16 Clément Dumas

Pre-trained language models (PTLMs) have transformed natural language processing (NLP), enabling major advances in tasks such as text generation and translation. Similar to software package management, PTLMs are developed using code and…

软件工程 · 计算机科学 2026-01-27 Adekunle Ajibode , Abdul Ali Bangash , Oussama Ben Sghaier , Bram Adams , Ahmed E. Hassan

As innovation in deep learning continues, many engineers are incorporating Pre-Trained Models (PTMs) as components in computer systems. Some PTMs are foundation models, and others are fine-tuned variations adapted to different needs. When…

软件工程 · 计算机科学 2025-08-20 Wenxin Jiang , Mingyu Kim , Chingwo Cheung , Heesoo Kim , George K. Thiruvathukal , James C. Davis

In recent years, Machine Learning (ML) models have achieved remarkable success in various domains. However, these models also tend to demonstrate unsafe behaviors, precluding their deployment in safety-critical systems. To cope with this…

计算机科学中的逻辑 · 计算机科学 2025-02-17 Andoni Rodriguez , Guy Amir , Davide Corsi , Cesar Sanchez , Guy Katz

Face recognition datasets are often collected by crawling Internet and without individuals' consents, raising ethical and privacy concerns. Generating synthetic datasets for training face recognition models has emerged as a promising…

计算机视觉与模式识别 · 计算机科学 2025-03-04 Hatef Otroshi Shahreza , Sébastien Marcel

The rise of machine learning (ML) systems has exacerbated their carbon footprint due to increased capabilities and model sizes. However, there is scarce knowledge on how the carbon footprint of ML models is actually measured, reported, and…

机器学习 · 计算机科学 2023-12-01 Joel Castaño , Silverio Martínez-Fernández , Xavier Franch , Justus Bogner

Synthetic data generation is gaining increasing popularity in different computer vision applications. Existing state-of-the-art face recognition models are trained using large-scale face datasets, which are crawled from the Internet and…

计算机视觉与模式识别 · 计算机科学 2024-11-01 Hatef Otroshi Shahreza , Sébastien Marcel

Large Language Models (LLMs) have been shown to be susceptible to jailbreak attacks, or adversarial attacks used to illicit high risk behavior from a model. Jailbreaks have been exploited by cybercriminals and blackhat actors to cause…

计算与语言 · 计算机科学 2025-01-07 Joao Fonseca , Andrew Bell , Julia Stoyanovich