6 ms·
Another Postgres-based project in this vein that makes use of Apache Arrow: https://heterodb.github.io/pg-strom/ https://heterodb.github.io/pg-strom/ > PG-Stro
by refset 3y ago
Another Postgres-based project in this vein that makes use of Apache Arrow: https://heterodb.github.io/pg-strom/ https://heterodb.github.io/pg-strom/
> PG-Strom is an extension module of PostgreSQL designed for version 11 or later. By utilization of GPU (Graphic Processor Unit) device which has thousands cores per chip, it enables to accelerate SQL workloads for data analytics or batch processing to big data set.
> PG-Strom has two storage options. The first one is the heap storage system of PostgreSQL. It is not always optimal for aggregation / analysis workloads because of its row data format, on the other hands, it has an advantage to run aggregation workloads without data transfer from the transactional database. The other one is Apache Arrow files, that have structured columnar format. Even though it is not suitable for update per row basis, it enables to import large amount of data efficiently, and efficiently search / aggregate the data through foreign data wrapper (FDW).