
edXHadoop and Spark solved the engineering problem of processing petabytes on commodity hardware. But mastery requires understanding both the theory behind distributed computing and the practical coding that deploys it.
This adapted version of a signature MSc program compresses the essentials: 20+ hours of lecture, 100+ quiz questions for conceptual grounding, and 20 coding challenges to force real implementation. You'll work with Spark in cloud environments where these tools are most valuable, bridging what's theoretically sound to what works in production systems handling genuinely massive data.
Big data systems such as Hadoop and Spark emerge as enabling technologies in managing massive amounts of data across hundreds or even thousands of computing nodes. Meanwhile, cloud computing platforms have made these technologies easily accessible to individuals as well as large enterprises. This course is an online adaptation of the signature course MSBD 5003 Big Data Computing offered to our popular MSc Program in Big Data Technology. In addition to 20+ hours of lecture videos, the course contains 100+ multiple-choice questions and 20 coding questions, aimed at equipping learners with both the theory and practical skills of big data systems, using Spark as the exemplary platform.
Price
No active coupon right now
View Course →Advertisement
No coupon right now. We'll tell you when there is one.
More in Software Engineering




Tracking since 1 Aug— not enough history yet to tell you whether today's price is any good. Watch the course and we'll tell you when it drops.
This is what we recorded in US pricing — not every price this course has ever had, and prices differ by country.