As data processing requirements grew beyond the capacities of even large computers, distributed computing systems were developed to spread the load to multiple computers. Hadoop is a distributed computing system with two key features: (1) it is open source, and (2) it can use low-cost commodity computers in its clusters. Hadoop is highly scalable; it is used by data-centric companies like Facebook.
Statistics and data science, defined
Hadoop
Where this gets used
We teach data science and statistics online, one subject at a time, on fixed start dates with an instructor who marks your work.