Designed and built a configurable Python engine for generating large, realistic relational datasets with preserved primary-key and foreign-key relationships. The system supports reusable business templates, weighted distributions, multiple ID strategies, and scalable generation across related tables. Performance work increased throughput to more than 300,000 rows per second in large benchmark runs. The project is being extended with dirty-data injection, validation reports, and additional output formats.