English

Flexible and Effective Mixing of Large Language Models into a Mixture of Domain Experts

Artificial Intelligence 2024-09-12 v2 Computation and Language

Abstract

We present a toolkit for creating low-cost Mixture-of-Domain-Experts (MOE) from trained models. The toolkit can be used for creating a mixture from models or from adapters. We perform extensive tests and offer guidance on defining the architecture of the resulting MOE using the toolkit. A public repository is available.

Keywords

Cite

@article{arxiv.2408.17280,
  title  = {Flexible and Effective Mixing of Large Language Models into a Mixture of Domain Experts},
  author = {Rhui Dih Lee and Laura Wynter and Raghu Kiran Ganti},
  journal= {arXiv preprint arXiv:2408.17280},
  year   = {2024}
}