4 ms·
Mining of Massive Datasets
- leksak 10y agoDeadlines are at the 29th of November. Best of luck
- PaulHoule 10y agoWhat amazed me is how much of this is 1990s stuff.
- cantagi 10y agoHow do you mean? Do you know of any up-to-date books on large scale data mining?
- mrcactu5 10y agomassive is an understatement. I have only dealt with puny GB sized data sets. They deal with vectors which cannot fit into main memory.
- infinitone 10y agoYes, in general what they refer to are things like the IRS Tax records (250 TB), Yahoo Ad data (900 TB). You just can't use a single machine to work with such data.
- joaorico 10y agoThis coursera course was taken down, but it is now up and running at lagunita.stanford.edu [0], which uses edx's open source platform [1]. The same happened to other stanford courses previously on coursera, you can find them here [2], including Compilers, Automata Theory, and Convex Optimization. [0] https://lagunita.stanford.edu/courses/course-v1:ComputerScience+MMDS+Fall2016/about https://lagunita.stanford.edu/courses/course-v1:ComputerScie... [1] https://open.edx.org/ https://open.edx.org/ https://github.com/edx/edx-platform https://github.com/edx/edx-platform [2] https://lagunita.stanford.edu/courses https://lagunita.stanford.edu/courses
- nthcolumn 10y agotldr; parallel map reduce. ;)