3 ms·
For one, Lucene[1] is quite amazing. With a slight stretch of imagination, you can think of it as a document oriented database, with all fields indexed by defau
by mtrn 10y ago
For one, Lucene[1] is quite amazing. With a slight stretch of imagination, you can think of it as a document oriented database, with all fields indexed by default. Imagine loading 100M rows with 100 columns into an RDBMS and then index each column[2]. You don't necessary do that. With Lucene - and SOLR and elasticsearch for that matter - it's not even a thought. Additionally, you get to tweak your indexing with dozens of field-tested analyzers.
[1] http://lucene.apache.org/core/ http://lucene.apache.org/core/
[2] we do this (over 300G of complex data) with a tool called solrbulk and on a single server in less than four hours, https://github.com/miku/solrbulk https://github.com/miku/solrbulk
- jaytaylor 10y agoI agree that Lucene is a mature solution with many excellent use-cases. For this portion of the stack why would it be preferable to start with Lucene rather than Elasticsearch? (for all of the distributed scaling benefits)
- mtrn 10y agoI would probably start with elasticsearch and establish a business case first. Only once that's proven and some need for lower level customizations arises, I would descend the stack.
- swsieber 10y agoPerhaps it's because Elastic Search shouldn't ever be a primary data store - it will eventually eat your data. That said, looking over Chronix I can't tell if it was supposed to be a primary data store.