English

Synthetic Data and Simulators for Recommendation Systems: Current State and Future Directions

Information Retrieval 2021-12-22 v1 Machine Learning

Abstract

Synthetic data and simulators have the potential to markedly improve the performance and robustness of recommendation systems. These approaches have already had a beneficial impact in other machine-learning driven fields. We identify and discuss a key trade-off between data fidelity and privacy in the past work on synthetic data and simulators for recommendation systems. For the important use case of predicting algorithm rankings on real data from synthetic data, we provide motivation and current successes versus limitations. Finally we outline a number of exciting future directions for recommendation systems that we believe deserve further attention and work, including mixing real and synthetic data, feedback in dataset generation, robust simulations, and privacy-preserving methods.

Keywords

Cite

@article{arxiv.2112.11022,
  title  = {Synthetic Data and Simulators for Recommendation Systems: Current State and Future Directions},
  author = {Adam Lesnikowski and Gabriel de Souza Pereira Moreira and Sara Rabhi and Karl Byleen-Higley},
  journal= {arXiv preprint arXiv:2112.11022},
  year   = {2021}
}

Comments

7 pages, included in SimuRec 2021: Workshop on Simulation Methods for Recommender Systems at ACM RecSys 2021, October 2nd, 2021, Amsterdam, NL and online

R2 v1 2026-06-24T08:25:45.636Z