English

You Do Not Need a Bigger Boat: Recommendations at Reasonable Scale in a (Mostly) Serverless and Open Stack

Machine Learning 2021-07-16 v1

Abstract

We argue that immature data pipelines are preventing a large portion of industry practitioners from leveraging the latest research on recommender systems. We propose our template data stack for machine learning at "reasonable scale", and show how many challenges are solved by embracing a serverless paradigm. Leveraging our experience, we detail how modern open source can provide a pipeline processing terabytes of data with limited infrastructure work.

Keywords

Cite

@article{arxiv.2107.07346,
  title  = {You Do Not Need a Bigger Boat: Recommendations at Reasonable Scale in a (Mostly) Serverless and Open Stack},
  author = {Jacopo Tagliabue},
  journal= {arXiv preprint arXiv:2107.07346},
  year   = {2021}
}

Comments

Manuscript version of a work accepted at RecSys 2021 (camera-ready forthcoming)

R2 v1 2026-06-24T04:13:50.971Z