Developed a data engineering pipeline to transform raw datasets into clean, structured, and analytics-ready data. Implemented data extraction, cleaning, validation, transformation, and loading (ETL) processes using Python and SQL to improve data quality and consistency. Automated preprocessing workflows, handled missing and duplicate records, standardized data formats, and generated optimized datasets for reporting, business intelligence, and machine learning applications.