English

Better bitmap performance with Roaring bitmaps

Databases 2016-04-12 v9

Abstract

Bitmap indexes are commonly used in databases and search engines. By exploiting bit-level parallelism, they can significantly accelerate queries. However, they can use much memory, and thus we might prefer compressed bitmap indexes. Following Oracle's lead, bitmaps are often compressed using run-length encoding (RLE). Building on prior work, we introduce the Roaring compressed bitmap format: it uses packed arrays for compression instead of RLE. We compare it to two high-performance RLE-based bitmap encoding techniques: WAH (Word Aligned Hybrid compression scheme) and Concise (Compressed `n' Composable Integer Set). On synthetic and real data, we find that Roaring bitmaps (1) often compress significantly better (e.g., 2 times) and (2) are faster than the compressed alternatives (up to 900 times faster for intersections). Our results challenge the view that RLE-based bitmap compression is best.

Keywords

Cite

@article{arxiv.1402.6407,
  title  = {Better bitmap performance with Roaring bitmaps},
  author = {Samy Chambi and Daniel Lemire and Owen Kaser and Robert Godin},
  journal= {arXiv preprint arXiv:1402.6407},
  year   = {2016}
}
R2 v1 2026-06-22T03:15:56.248Z