A Deep Dive into Common Open Formats for Analytical DBMSs

Chunwei Liu,Anna Pavlenko,Matteo Interlandi,Brandon Haynes

doi:10.14778/3611479.3611507

A Deep Dive into Common Open Formats for Analytical DBMSs

Chunwei Liu, Anna Pavlenko + Show 2 more

https://doi.org/10.14778/3611479.3611507

Copy DOI

Journal: Proceedings of the VLDB Endowment	Publication Date: Jul 1, 2023
Citations: 6

Affiliation: Microsoft Research (United Kingdom)

#Deep Dive #Common Formats + Show 5 more

Abstract
Full-Text PDF
Similar Papers

Abstract

This paper evaluates the suitability of Apache Arrow, Parquet, and ORC as formats for subsumption in an analytical DBMS. We systematically identify and explore the high-level features that are important to support efficient querying in modern OLAP DBMSs and evaluate the ability of each format to support these features. We find that each format has trade-offs that make it more or less suitable for use as a format in a DBMS and identify opportunities to more holistically co-design a unified in-memory and on-disk data representation. Our hope is that this study can be used as a guide for system developers designing and using these formats, as well as provide the community with directions to pursue for improving these common open formats.

Full Text

Paper version not known

Open DOI Link

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

Similar Papers

Paper Title

Journal

Date

Author

View more papers

More From: Proceedings of the VLDB Endowment

Paper Title

Journal

Date

Author

View more papers

Disclaimer: All third-party content on this website/platform is and will remain the property of their respective owners and is provided on "as is" basis without any warranties, express or implied. Use of third-party content does not indicate any affiliation, sponsorship with or endorsement by them. Any references to third-party content is to identify the corresponding services and shall be considered fair use under The CopyrightLaw.