Data Blocks: Hybrid OLTP and OLAP on Compressed Storage using both Vectorization and Compilation

Data Blocks: Hybrid OLTP and OLAP on Compressed Storage using both Vectorization and Compilation
复制标题

DOI:
10.1145/2882903.2882925
复制
发表时间:
2016-06
期刊:
Proceedings of the 2016 International Conference on Management of Data
影响因子:
--
通讯作者:
Harald Lang;Tobias Mühlbauer;Florian Funke;P. Boncz;Thomas Neumann;A. Kemper
Harald Lang;Tobias Mühlbauer;Florian Funke;P. Boncz;Thomas Neumann;A. Kemper
中科院分区:
其他
文献类型:
--
作者:
Harald Lang;Tobias Mühlbauer;Florian Funke;P. Boncz;Thomas Neumann;A. Kemper

文献摘要

被引文献

相似文献

这项工作旨在减少高性能混合OLTP和OLAP数据库中的主要内存足迹,同时保留高查询性能和交易吞吐量。为此,引入了用于冷数据的创新压缩柱状存储格式,称为数据块。数据块进一步合并了一种称为位置SMA的新的轻重量索引结构,即使整个块不能排除在数据块中,缩小扫描范围也是如此。为了获得最高的OLTP性能,数据块的压缩方案非常轻巧,因此OLTP交易仍然可以快速访问单个元组。这将我们的存储方案与专门分析数据库中使用的存储方案不同,在该数据库中通常必须稍微添加数据。到目前为止,高性能分析系统使用矢量性查询执行或即时(JIT)查询汇编。数据块的细粒度适应性需要通过解释的矢量化扫描子系统将每种方法的最佳特征集成,该功能将其馈送到JIT编译的查询管道中。我们成熟的混合OLTP和OLAP数据库系统Hyper的实验评估表明,数据阻止了各种查询工作负载上的性能,同时保留了高交易吞吐量。
This work aims at reducing the main-memory footprint in high performance hybrid OLTP & OLAP databases, while retaining high query performance and transactional throughput. For this purpose, an innovative compressed columnar storage format for cold data, called Data Blocks is introduced. Data Blocks further incorporate a new light-weight index structure called Positional SMA that narrows scan ranges within Data Blocks even if the entire block cannot be ruled out. To achieve highest OLTP performance, the compression schemes of Data Blocks are very light-weight, such that OLTP transactions can still quickly access individual tuples. This sets our storage scheme apart from those used in specialized analytical databases where data must usually be bit-unpacked. Up to now, high-performance analytical systems use either vectorized query execution or just-in-time (JIT) query compilation. The fine-grained adaptivity of Data Blocks necessitates the integration of the best features of each approach by an interpreted vectorized scan subsystem feeding into JIT-compiled query pipelines. Experimental evaluation of HyPer, our full-fledged hybrid OLTP & OLAP database system, shows that Data Blocks accelerate performance on a variety of query workloads while retaining high transaction throughput.