September 1, 2018

Everything you always wanted to know about compiled and vectorized queries but were afraid to ask

Key Points

Key points are not available for this paper at this time.

Abstract

The query engines of most modern database systems are either based on vectorization or data-centric code generation. These two state-of-the-art query processing paradigms are fundamentally different in terms of system structure and query execution code. Both paradigms were used to build fast systems. However, until today it is not clear which paradigm yields faster query execution, as many implementation-specific choices obstruct a direct comparison of architectures. In this paper, we experimentally compare the two models by implementing both within the same test system. This allows us to use for both models the same query processing algorithms, the same data structures, and the same parallelization framework to ultimately create an apples-to-apples comparison. We find that both are efficient, but have different strengths and weaknesses. Vectorization is better at hiding cache miss latency, whereas data-centric compilation requires fewer CPU instructions, which benefits cache-resident workloads. Besides raw, single-threaded performance, we also investigate SIMD as well as multi-core parallelization and different hardware architectures. Finally, we analyze qualitative differences as a guide for system architects.

Connected Papers

Building similarity graph...

Analyzing shared references across papers

Discussion

Cite this study

Kersten et al. (Sat,) studied this question.

synapsesocial.com/papers/6a21ffe26be264275d4e92ed — DOI: https://doi.org/10.14778/3275366.3284966

Authors

Timo Kersten

Technical University of Munich

Viktor Leis

Technical University of Darmstadt

Alfons Kemper

Karlsruhe Institute of Technology

Journals

Proceedings of the VLDB Endowment

Actions

Institutions

Carnegie Mellon University

Technical University of Munich

Centrum Wiskunde & Informatica

References and Citations

Connected Papers

Building similarity graph...

Analyzing shared references across papers

Everything you always wanted to know about compiled and vectorized queries but were afraid to ask

Key Points

Abstract

Citation Network

Connected Papers

Discussion

Cite this study

Authors

Journals

Actions

Institutions

References and Citations

Citation Network

Connected Papers

Discussion