Skip to content
LearnStatisticspowered by UpThink

Statistics and data science, defined

Hadoop

As data processing requirements grew beyond the capacities of even large computers, distributed computing systems were developed to spread the load to multiple computers. Hadoop is a distributed computing system with two key features: (1) it is open source, and (2) it can use low-cost commodity computers in its clusters. Hadoop is highly scalable; it is used by data-centric companies like Facebook.

Where this gets used

We teach data science and statistics online, one subject at a time, on fixed start dates with an instructor who marks your work.