For data engineers looking to move beyond basic workflows, this intermediate-level course provides a deep dive into the architecture of robust batch processing. Instead of just running scripts, you will learn how to manage the entire lifecycle of a data pipeline within the Google Cloud ecosystem.
This is an essential resource for professionals aiming to build scalable, production-ready data architectures that can handle evolving data structures and complex processing requirements.
This course teaches you to design, build, and operate batch data pipelines on Google Cloud. Topics include large-scale data transformations with Dataflow and Serverless Spark, batch data validation and cleansing, schema evolution, error handling, and pipeline orchestration with Cloud Composer.
Access
SubscriptionIncluded with a DataCamp subscription
This course is included with a Subscription subscription.
View on DataCamp →Advertisement
No coupon right now. We'll tell you when there is one.