Data Engineers collect and format data. Data wrangling is one of the most tedious and time consuming tasks in the ML stack and eliminating it at this layer solves a lot of problems down stream. Data versioning and database partitioning ensures that old data can be discarded when it becomes irrelevant or non-compliant with privacy laws.