Cloud-Native HR Analytics Pipeline
Data Engineering / Cloud DataBase / Machine Learning
•
Designed a distributed data pipeline on PySpark and Spark MLlib to process and
model HR data at scale.
•
Handled ingestion, cleaning, and feature engineering stages, then trained
predictive models to surface attrition and workforce trends.