中文
相关论文

相关论文: Protecting Time Series Data with Minimal Forecast …

200 篇论文

Enforcing data protection and privacy rules within large data processing applications is becoming increasingly important, especially in the light of GDPR and similar regulatory frameworks. Most modern data processing happens on top of a…

密码学与安全 · 计算机科学 2020-08-13 Zsolt Istvan , Soujanya Ponnapalli , Vijay Chidambaram

Many businesses and industries nowadays rely on large quantities of time series data making time series forecasting an important research area. Global forecasting models that are trained across sets of time series have shown a huge…

机器学习 · 计算机科学 2021-10-25 Rakshitha Godahewa , Christoph Bergmeir , Geoffrey I. Webb , Rob J. Hyndman , Pablo Montero-Manso

We propose a data-driven framework for optimizing privacy-preserving data release mechanisms to attain the information-theoretically optimal tradeoff between minimizing distortion of useful data and concealing specific sensitive…

信息论 · 计算机科学 2019-06-13 Ardhendu Tripathy , Ye Wang , Prakash Ishwar

The $k$-Nearest Neighbor Search ($k$-NNS) is the backbone of several cloud-based services such as recommender systems, face recognition, and database search on text and images. In these services, the client sends the query to the cloud…

数据结构与算法 · 计算机科学 2020-03-10 Hao Chen , Ilaria Chillotti , Yihe Dong , Oxana Poburinnaya , Ilya Razenshteyn , M. Sadegh Riazi

The open data ecosystem is susceptible to vulnerabilities due to disclosure risks. Though the datasets are anonymized during release, the prevalence of the release-and-forget model makes the data defenders blind to privacy issues arising…

密码学与安全 · 计算机科学 2023-04-25 Kaustav Bhattacharjee , Aritra Dasgupta

Differential privacy is a leading protection setting, focused by design on individual privacy. Many applications, in medical / pharmaceutical domains or social networks, rather posit privacy at a group level, a setting we call integral…

机器学习 · 统计学 2019-07-04 Hisham Husain , Zac Cranko , Richard Nock

Access to diverse, high-quality datasets is crucial for machine learning model performance, yet data sharing remains limited by privacy concerns and competitive interests, particularly in regulated domains like healthcare. This dynamic…

机器学习 · 计算机科学 2025-10-20 Keren Fuentes , Mimee Xu , Irene Chen

Deep time series forecasting has emerged as a rapidly growing field in recent years. Despite the exponential growth of community interests, progress on standard benchmarks is often limited to marginal improvements. A common consensus of the…

机器学习 · 计算机科学 2026-05-05 Yuxuan Wang , Haixu Wu , Yuezhou Ma , Yuchen Fang , Ziyi Zhang , Yong Liu , Shiyu Wang , Zhou Ye , Yang Xiang , Jianmin Wang , Mingsheng Long

Advances in sensing and communication capabilities as well as power industry deregulation are driving the need for distributed state estimation in the smart grid at the level of the regional transmission organizations (RTOs). This leads to…

信息论 · 计算机科学 2016-11-18 Lalitha Sankar , Soummya Kar , Ravi Tandon , H. Vincent Poor

This paper considers the single-server Private Linear Transformation (PLT) problem with individual privacy guarantees. In this problem, there is a user that wishes to obtain $L$ independent linear combinations of a $D$-subset of messages…

信息论 · 计算机科学 2021-06-11 Anoosheh Heidarzadeh , Nahid Esmati , Alex Sprintson

Clustering of time series data exhibits a number of challenges not present in other settings, notably the problem of registration (alignment) of observed signals. Typical approaches include pre-registration to a user-specified template or…

机器学习 · 统计学 2021-11-03 Michael Weylandt , George Michailidis

We study a setting where a data holder wishes to share data with a receiver, without revealing certain summary statistics of the data distribution (e.g., mean, standard deviation). It achieves this by passing the data through a…

密码学与安全 · 计算机科学 2023-10-31 Zinan Lin , Shuaiqi Wang , Vyas Sekar , Giulia Fanti

Sensitive datasets are often underutilized in research and industry due to privacy concerns, limiting the potential of valuable data-driven insights. Synthetic data generation presents a promising solution to address this challenge by…

统计计算 · 统计学 2026-01-27 Ali Furkan Kalay

Confidence intervals are a standard technique for analyzing data. When applied to time series, confidence intervals are computed for each time point separately. Alternatively, we can compute confidence bands, where we are required to find…

机器学习 · 计算机科学 2021-12-14 Nikolaj Tatti

There has been a large number of contributions on privacy-preserving smart metering with Differential Privacy, addressing questions from actual enforcement at the smart meter to billing at the energy provider. However, exploitation is…

密码学与安全 · 计算机科学 2018-07-09 Günther Eibl , Kaibin Bao , Philip-William Grassal , Daniel Bernau , Hartmut Schmeck

In this study we show that standard well-known file compression programs (zlib, bzip2, etc.) are able to forecast real-world time series data well. The strength of our approach is its ability to use a set of data compression algorithms and…

信息论 · 计算机科学 2019-04-09 K. S. Chirikhin , B. Ya. Ryabko

Forecasting models that are trained across sets of many time series, known as Global Forecasting Models (GFM), have shown recently promising results in forecasting competitions and real-world applications, outperforming many…

机器学习 · 计算机科学 2020-08-07 Kasun Bandara , Hansika Hewamalage , Yuan-Hao Liu , Yanfei Kang , Christoph Bergmeir

Privacy-preserving analytics is designed to protect valuable assets. A common service provision involves the input data from the client and the model on the analyst's side. The importance of the privacy preservation is fuelled by legal…

密码学与安全 · 计算机科学 2024-04-16 Martin Kodys , Zhongmin Dai , Vrizlynn L. L. Thing

In recent years, machine learning techniques are widely used in numerous applications, such as weather forecast, financial data analysis, spam filtering, and medical prediction. In the meantime, massive data generated from multiple sources…

密码学与安全 · 计算机科学 2018-10-08 Wei Du , Ang Li , Qinghua Li

K-means is one of the most widely used clustering models in practice. Due to the problem of data isolation and the requirement for high model performance, how to jointly build practical and secure K-means for multiple parties has become an…

机器学习 · 计算机科学 2022-08-15 Yingting Liu , Chaochao Chen , Jamie Cui , Li Wang , Lei Wang