English
Related papers

Related papers: Optimal Pricing for Data-Augmented AutoML Marketpl…

200 papers

Data-driven machine learning (ML) has witnessed great successes across a variety of application domains. Since ML model training are crucially relied on a large amount of data, there is a growing demand for high quality data to be collected…

Databases · Computer Science 2020-03-31 Jinfei Liu

As Machine Learning (ML) models are becoming increasingly complex, one of the central challenges is their deployment at scale, such that companies and organizations can create value through Artificial Intelligence (AI). An emerging paradigm…

Machine Learning · Computer Science 2021-12-07 Lam Duc Nguyen , Shashi Raj Pandey , Soret Beatriz , Arne Broering , Petar Popovski

The realization that AI-driven decision-making is indispensable in today's fast-paced and ultra-competitive marketplace has raised interest in industrial machine learning (ML) applications significantly. The current demand for analytics…

Machine Learning · Computer Science 2025-06-03 Marc Schmitt

AutoML services provide a way for non-expert users to benefit from high-quality ML models without worrying about model design and deployment, in exchange for a charge per hour ($21.252 for VertexAI). However, existing AutoML services are…

Databases · Computer Science 2023-05-18 Zezhou Huang , Pranav Subramaniam , Raul Castro Fernandez , Eugene Wu

Data analytics using machine learning (ML) has become ubiquitous in science, business intelligence, journalism and many other domains. While a lot of work focuses on reducing the training cost, inference runtime and storage cost of ML…

Databases · Computer Science 2018-05-30 Lingjiao Chen , Paraschos Koutris , Arun Kumar

Web-scale ranking systems at Meta serving billions of users is complex. Improving ranking models is essential but engineering heavy. Automated Machine Learning (AutoML) can release engineers from labor intensive work of tuning ranking…

The difficulty in acquiring a sufficient amount of training data is a major bottleneck for machine learning (ML) based data analytics. Recently, commoditizing ML models has been proposed as an economical and moderate solution to ML-oriented…

Machine Learning · Computer Science 2023-04-04 Shuyuan Zheng , Yang Cao , Masatoshi Yoshikawa , Huizhong Li , Qiang Yan

Automated machine learning (AutoML) systems aim to enable training machine learning (ML) models for non-ML experts. A shortcoming of these systems is that when they fail to produce a model with high accuracy, the user has no path to improve…

Machine Learning · Computer Science 2021-02-23 Behnaz Arzani , Kevin Hsieh , Haoxian Chen

Big data has been emerging as a new approach in utilizing large datasets to optimize complex system operations. Big data is fueled with Internet-of-Things (IoT) services that generate immense sensory data from numerous sensors and devices.…

Computer Science and Game Theory · Computer Science 2016-08-16 Dusit Niyato , Mohammad Abu Alsheikh , Ping Wang , Dong In Kim , Zhu Han

Data augmentation is arguably the most important regularization technique commonly used to improve generalization performance of machine learning models. It primarily involves the application of appropriate data transformation operations to…

Machine Learning · Computer Science 2025-03-07 Alhassan Mumuni , Fuseini Mumuni

In this work, we aim to design a data marketplace; a robust real-time matching mechanism to efficiently buy and sell training data for Machine Learning tasks. While the monetization of data and pre-trained models is an essential focus of…

Computer Science and Game Theory · Computer Science 2019-05-14 Anish Agarwal , Munther Dahleh , Tuhin Sarkar

Federated Learning (FL), as a mainstream privacy-preserving machine learning paradigm, offers promising solutions for privacy-critical domains such as healthcare and finance. Although extensive efforts have been dedicated from both academia…

Machine Learning · Computer Science 2024-11-19 Zhenyu Wen , Wanglei Feng , Di Wu , Haozhen Hu , Chang Xu , Bin Qian , Zhen Hong , Cong Wang , Shouling Ji

The rise of the machine learning (ML) model economy has intertwined markets for training datasets and pre-trained models. However, most pricing approaches still separate data and model transactions or rely on broker-centric pipelines that…

Machine Learning · Computer Science 2026-05-12 Hongrun Ren , Yun Xiong , Lei You , Yingying Wang , Haixu Xiong , Yangyong Zhu

We study revenue-optimal pricing in data markets with rational, budget-constrained buyers. Such a market offers multiple datasets for sale, and buyers aim to improve the accuracy of their prediction tasks by acquiring data bundles. The…

Computer Science and Game Theory · Computer Science 2026-04-28 Bhaskar Ray Chaudhury , Jugal Garg , Eklavya Sharma , Jiaxin Song

Training data is the backbone of large language models (LLMs), yet today's data markets often operate under exploitative pricing -- sourcing data from marginalized groups with little pay or recognition. This paper introduces a theoretical…

Computer Science and Game Theory · Computer Science 2025-11-20 Luyang Zhang , Cathy Jiao , Beibei Li , Chenyan Xiong

As big data becomes ubiquitous across domains, and more and more stakeholders aspire to make the most of their data, demand for machine learning tools has spurred researchers to explore the possibilities of automated machine learning…

The $\textit{data market design}$ problem is a problem in economic theory to find a set of signaling schemes (statistical experiments) to maximize expected revenue to the information seller, where each experiment reveals some of the…

Computer Science and Game Theory · Computer Science 2023-11-01 Sai Srivatsa Ravindranath , Yanchen Jiang , David C. Parkes

Data only generates value for a few organizations with expertise and resources to make data shareable, discoverable, and easy to integrate. Sharing data that is easy to discover and integrate is hard because data owners lack information…

Databases · Computer Science 2020-07-03 Raul Castro Fernandez , Pranav Subramaniam , Michael J. Franklin

The common pipeline of training deep neural networks consists of several building blocks such as data augmentation and network architecture selection. AutoML is a research field that aims at automatically designing those parts, but most…

Machine Learning · Computer Science 2021-01-13 Taiga Kashima , Yoshihiro Yamada , Shunta Saito

Federated learning (FL) is increasingly recognized for its efficacy in training models using locally distributed data. However, the proper valuation of shared data in this collaborative process remains insufficiently addressed. In this…

Machine Learning · Computer Science 2024-02-06 Yue Cui , Liuyi Yao , Yaliang Li , Ziqian Chen , Bolin Ding , Xiaofang Zhou
‹ Prev 1 2 3 10 Next ›