Larger-than-memory data management on modern storage hardware for in-memory OLTP database systems

Larger-than-memory data management on modern storage hardware for in-memory OLTP database systems
复制标题

现代存储硬件上内存 OLTP 数据库系统的大于内存数据管理

DOI:
--
复制
发表时间:
2016
期刊:
International Workshop on Data Management on New Hardware
影响因子:
--
通讯作者:
S. Zdonik
S. Zdonik
中科院分区:
--
文献类型:
--
作者:
Lin Ma;Joy Arulraj;Sam Zhao;Andrew Pavlo;Subramanya R. Dulloor;Michael J. Giardino;Jeff Parkhurst;J. L. Gardner;K. Doshi;S. Zdonik

文献摘要

被引文献

相似文献

内存数据库管理系统(DBMS)在联机事务处理(OLTP)工作负载方面优于面向磁盘的系统。但是,只有当数据库小于系统中可用的物理内存量时,才能实现这种改进的性能。为了克服这个限制,一些内存中的DBMS可以将冷数据从易失性DRAM移动到辅助存储器。这样的数据看起来好像与数据库的其余部分一起驻留在内存中,即使它没有。 虽然已经提出了几种针对这种类型的冷数据存储的实现,但是在实现这种技术时,还没有对设计决策进行彻底的评估,例如何时驱逐元组以及如何在需要时将它们带回的策略。这些选择由于不同存储设备(包括未来的非易失性存储器技术)的不同性能特性而进一步复杂化。我们在本文中探讨这些问题,并讨论了几种方法来解决它们。我们在内存DBMS中实现了所有这些方法,并使用五种不同的存储技术对其进行了评估。我们的研究结果表明,选择最佳的策略的基础上的硬件提高了92-340%的吞吐量超过一个通用的配置。
In-memory database management systems (DBMSs) outperform disk-oriented systems for on-line transaction processing (OLTP) workloads. But this improved performance is only achievable when the database is smaller than the amount of physical memory available in the system. To overcome this limitation, some in-memory DBMSs can move cold data out of volatile DRAM to secondary storage. Such data appears as if it resides in memory with the rest of the database even though it does not. Although there have been several implementations proposed for this type of cold data storage, there has not been a thorough evaluation of the design decisions in implementing this technique, such as policies for when to evict tuples and how to bring them back when they are needed. These choices are further complicated by the varying performance characteristics of different storage devices, including future non-volatile memory technologies. We explore these issues in this paper and discuss several approaches to solve them. We implemented all of these approaches in an in-memory DBMS and evaluated them using five different storage technologies. Our results show that choosing the best strategy based on the hardware improves throughput by 92-340% over a generic configuration.