AI, Data Science & Business Certificates from Google, IBM & Microsoft
Free courses from frontend to fullstack and AI
Overview
Google, IBM & Meta Certificates — All 10,000+ Courses at 40% Off
One annual plan covers every course and certificate on Coursera. 40% off for a limited time.
Get Full Access
Learn to set up a Spark master-worker architecture in a Docker container on Azure and perform end-to-end data processing and visualization of Japan visa numbers using PySpark and Plotly. Set up cloud clusters, read and clean data with PySpark, apply data transformation techniques, and create interactive visualizations with Plotly Express. Gain practical experience in data engineering, from system architecture setup to exporting visualizations and cleaned data. Follow along with step-by-step instructions, including timestamps for each section, to master cloud-based data processing and analysis techniques.
Syllabus
Introduction
Setting up the system architecture
Setting up cloud clusters
Coding
Results
Taught by
CodeWithYu