3 ms·
Thanks. This looks cool. However, my issue is the need to introduce one more tool. I feel that without a single tool to read and write to Iceberg, I would not
by whinvik 2y ago
Thanks. This looks cool.
However, my issue is the need to introduce one more tool. I feel that without a single tool to read and write to Iceberg, I would not want to introduce it to our team.
Spark is cool and all but it requires quite a bit of effort to properly work. And Spark seems to be the only thing right now that can read and write to Iceberg natively with a SQL like interface.
- jaychia 2y agoCheck out Daft (www.getdaft.io) - we've been working really hard on our Iceberg support. Supports full reads/writes (including partitioned writes) and our SQL support is also coming along quite well! Also no cluster, no JVM. Just `pip install daft` and go. Runs locally (as fast as DuckDB for a lot of workloads; faster, if you have S3 cloud data access) and also runs distributed if you have a Ray cluster you can point it at (Disclaimer: I work on it)
- jamesblonde 2y agoDaft is making great progress with Iceberg - faster than PyIceberg in many ways.