3 ms·
Cool. So in a way its Immutable read only HBase (with guaranteed data locality no memstore overhead and compactions overhead). Cool. Nice solution. I wonder thi
by ameyamk 11y ago
Cool. So in a way its Immutable read only HBase (with guaranteed data locality no memstore overhead and compactions overhead). Cool. Nice solution. I wonder this can be patched back to HBase - as a "read only mode" ?
- varunsharma 11y agoSuspect it would be difficult. There are differences like new data is completely independent of previous data and the source of truth for region distribution is HDFS for the block locations. There are multiple replicas per shard while HBase has one region for each shard etc. Across bulk loads, the number of shards (regions) can be changed for the same fileset or table - not possible with HBase. The data sharding is not range based like in HBase but mod based, as output by Hadoop HashPartitioner. There are so many differences that its hard to accommodate it into the HBase code.