Key Responsibilities
- Lead, coach, and mentor 6+ data engineers in modern, AI-enabled practices.
- Own and optimise Databricks (Delta Lake, Unity Catalog, MLflow, Databricks SQL).
- Redesign data pipelines for maximum scalability, performance, and cost efficiency.
- Champion Python engineering best practices (CI/CD, testing, code reviews).
- Drive DataOps and MLOps to ensure strict observability, governance, and seamless lifecycles.
What We’re Looking For
- 5+ years in data engineering with 2+ years scaling Databricks environments.
- Deep expertise in Apache Spark, PySpark, Unity Catalog, and MLflow.
- Strong production Python background with robust CI/CD experience.
- Proven track record of optimising Databricks environments for cost and performance.
- Experience leading engineering teams and establishing modern best practices.