Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
ivanprado
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
4 ms
·
1.
▲
by
ivanprado
8y ago
That's a great idea.
2.
▲
by
ivanprado
13y ago
You are right that it should be used the proper tool for each particular problem. And Hadoop world is harder than single machine systems (like pandas). So, you shouldn't user Hadoop if you can do the job with simpler systems. But I hav
3.
▲
A Big Data love story: Hive + Splout SQL for a social media reporting webapp
(datasalt.com)
1 points
by
ivanprado
14y ago
|
0 comments
4.
▲
Cascading + Splout SQL for log analysis and serving: A Big Data love story
(datasalt.com)
1 points
by
ivanprado
14y ago
|
0 comments
5.
▲
An example real-time “lambda architecture” using Trident, Hadoop and Splout SQL
(datasalt.com)
2 points
by
ivanprado
14y ago
|
0 comments
6.
▲
by
ivanprado
14y ago
Three things differentiate Splout SQL from using Sqoop for exporting to an existing SQL database: 1) Scalability: Relational databases rarely scales, or are too expensive for big volumes of data. They don't work well with Hadoop. 2) Update
7.
▲
SQL + Hadoop + low-latency = Yes, it is possible
(datasalt.com)
6 points
by
ivanprado
14y ago
|
2 comments
8.
▲
Slides about Tuple MapReduce at IEEE ICDM 2012
(datasalt.com)
1 points
by
ivanprado
14y ago
|
0 comments
9.
▲
MapReduce design patterns and real use cases
(datasalt.com)
2 points
by
ivanprado
14y ago
|
0 comments
10.
▲
Paper: Tuple MapReduce: beyond classic MapReduce
(pangool.net)
2 points
by
ivanprado
14y ago
|
0 comments
11.
▲
A preview of a new SQL low-latency database for Hadoop
(datasalt.com)
3 points
by
ivanprado
14y ago
|
0 comments
12.
▲
Hadoop's Game of Life
(datasalt.com)
28 points
by
ivanprado
14y ago
|
4 comments
13.
▲
Building a scalable text classifier in Hadoop with Pangool
(datasalt.com)
1 points
by
ivanprado
14y ago
|
0 comments
14.
▲
by
ivanprado
15y ago
Just for clarify, split() java function is using regexp for the split as well. The code of String.split() is: return Pattern.compile(regex).split(this, limit); The benchmark seems fair to me.
15.
▲
by
ivanprado
15y ago
Pangool is not an alternative for Cascading. For example, at this point, Pangool does not help you managing workflows. If you are starting a MapReduce application, it is probably the best option to start using higher level abstractions: Cas
16.
▲
by
ivanprado
15y ago
Hi avibryant, According to our initial benchmark ( http://pangool.net/benchmark.html ), secondary sorting in Cascading is slow ( http://bit.ly/wTKOxo ), showing a 243% performance overhead compared to an efficient implementation in MapReduc
17.
▲
by
ivanprado
15y ago
Hi, I'm one of the developers of Pangool. The idea of Pangool is not to be yet another higher level API on top of Hadoop but rather to pose a replacement for the low-level Hadoop Java MapReduce API. Pangool has the same performance and fle
18.
▲
Key/value is dead. Long live tuples: Pangool for Hadoop
(datasalt.com)
79 points
by
ivanprado
15y ago
|
28 comments
19.
▲
MapReduce & Hadoop API revised
(datasalt.com)
3 points
by
ivanprado
15y ago
|
0 comments
20.
▲
Scalable vertical search engine with Hadoop
(datasalt.com)
4 points
by
ivanprado
15y ago
|
0 comments