中文
相关论文

相关论文: Data Quality Over Quantity: Pitfalls and Guideline…

200 篇论文

Industrial processes generate vast amounts of time series data, yet extracting meaningful relationships and insights remains challenging. This paper introduces a framework for automated knowledge graph learning from time series data,…

机器学习 · 计算机科学 2024-07-03 Lolitta Ammann , Jorge Martinez-Gil , Michael Mayr , Georgios C. Chasparis

The recent advances in information and communication technology (ICT) have promoted the evolution of conventional computer-aided manufacturing industry to smart data-driven manufacturing. Data analytics in massive manufacturing data can…

计算机与社会 · 计算机科学 2019-09-04 Hong-Ning Dai , Hao Wang , Guangquan Xu , Jiafu Wan , Muhammad Imran

Industry 4.0 factories are complex and data-driven. Data is yielded from many sources, including sensors, PLCs, and other devices, but also from IT, like ERP or CRM systems. We ask how to collect and process this data in a way, such that it…

信息检索 · 计算机科学 2026-03-24 Eduard Hirsch , Simon Hoher , Stefan Huber

In computational materials science, mechanical properties are typically extracted from simulations by means of analysis routines that seek to mimic their experimental counterparts. However, simulated data often exhibit uncertainties that…

数据分析、统计与概率 · 物理学 2017-12-07 Paul N. Patrone , Anthony J. Kearsley , Andrew M. Dienstfrey

Fairness-aware machine learning has recently attracted various communities to mitigate discrimination against certain societal groups in data-driven tasks. For fair supervised learning, particularly in pre-processing, there have been two…

机器学习 · 计算机科学 2026-01-21 Jinwon Sohn , Guang Lin , Qifan Song

Data warehousing is continuously gaining importance as organizations are realizing the benefits of decision oriented data bases. However, the stumbling block to this rapid development is data quality issues at various stages of data…

数据库 · 计算机科学 2013-10-09 Vinay Kumar , Reema Thareja

Data collection and labeling are critical bottlenecks in the deployment of machine learning applications. With the increasing complexity and diversity of applications, the need for efficient and scalable data collection and labeling…

数据库 · 计算机科学 2024-07-19 Qianyu Huang , Tongfang Zhao

We have analyzed manufacturing data from several different semiconductor manufacturing plants, using decision tree induction software called Q-YIELD. The software generates rules for predicting when a given product should be rejected. The…

机器学习 · 计算机科学 2007-05-23 Peter D. Turney

When working with real-world insurance data, practitioners often encounter challenges during the data preparation stage that can undermine the statistical validity and reliability of downstream modeling. This study illustrates that…

机器学习 · 统计学 2026-03-20 Jiayi Guo , Panyi Dong , Zhiyu Quan

Traditional statistical and measurements are unable to solve all industrial data in the right way and appropriate time. Open markets mean the customers are increased, and production must increase to provide all customer requirements.…

综合经济学 · 经济学 2020-11-26 Hamza Saad

Smart manufacturing systems are being deployed at a growing rate because of their ability to interpret a wide variety of sensed information and act on the knowledge gleaned from system observations. In many cases, the principal goal of the…

The data processing inequality is an information-theoretic principle stating that the information content of a signal cannot be increased by processing the observations. In particular, it suggests that there is no benefit in enhancing the…

机器学习 · 计算机科学 2025-12-25 Roy Turgeman , Tom Tirer

Process control and optimization have been widely used to solve decision-making problems in chemical engineering applications. However, identifying and tuning the best solution algorithm is challenging and time-consuming. Machine learning…

系统与控制 · 电气工程与系统科学 2024-12-25 Ilias Mitrai , Prodromos Daoutidis

As machine learning (ML) systems get adopted in more critical areas, it has become increasingly crucial to address the bias that could occur in these systems. Several fairness pre-processing algorithms are available to alleviate implicit…

In the emerging era of big data, larger available clinical datasets and computational advances have sparked a massive interest in machine learning-based approaches. The number of manuscripts related to machine learning or artificial…

机器学习 · 统计学 2020-06-29 Julius M. Kernbach , Victor E. Staartjes

In this paper we address the application of pre-processing techniques to multi-channel time series data with varying lengths, which we refer to as the alignment problem, for downstream machine learning. The misalignment of multi-channel…

Strategic planning in a corporate environment is often based on experience and intuition, although internal data is usually available and can be a valuable source of information. Predicting merger & acquisition (M&A) events is at the heart…

应用统计 · 统计学 2022-04-26 Kainat Khowaja , Danial Saef , Sergej Sizov , Wolfgang Karl Härdle

This paper presents the results of an industry expert survey about event log generation in process mining. It takes academic assumptions as a starting point and elicits practitioner's assessments of statements about process execution,…

软件工程 · 计算机科学 2022-04-15 Timotheus Kampik , Mathias Weske

Data is a cornerstone of empirical software engineering (ESE) research and practice. Data underpin numerous process and project management activities, including the estimation of development effort and the prediction of the likely location…

软件工程 · 计算机科学 2020-12-22 Michael F. Bosu , Stephen G. MacDonell

The detection and localization of quality-related problems in industrially mass-produced products has historically relied on manual inspection, which is costly and error-prone. Machine learning has the potential to replace manual handling.…

机器学习 · 计算机科学 2025-06-23 Sebastian Hönel , Jonas Nordqvist
‹ 上一页 1 8 9 10 下一页 ›