Analysts and data engineers lose hours cleaning raw CSV files by hand: null values, inconsistent formats, one-off scripts that break silently.
The project
I built a containerized microservice in Python + FastAPI that automates the whole pipeline. You POST a raw, messy CSV file to a single endpoint and get back a structured, ready-to-use JSON response.
Under the hood, the endpoint processes the file with Pandas: cleaning null values and standardizing formats before returning the clean JSON. The entire microservice runs 100% inside a Docker container, so it behaves the same in any environment.
The result
Manual data-cleaning work becomes one API call. The code is fully functional and open on GitHub.