8 ms·
distributed in-memory computing for massive datasets to big to fit into vertically scaled memory. generic tabular files, not tables. delta lake.
by RocketSyntax 7y ago
distributed in-memory computing for massive datasets to big to fit into vertically scaled memory. generic tabular files, not tables. delta lake.
- truth_seeker 7y agoYes those features help and all of the distributed SQL databases have data and query cache.
- RocketSyntax 7y agoThen once the subset of data is in distributed memory... you hittem w pyspark and all compatible libraries.