Overview
We're looking for a Data Engineer to design and implement data extraction, transformation and aggregation pipelines to support Product reporting.
The role will primarily work with AWS data sources, including DynamoDB and CloudWatch, and deliver curated datasets to Power BI and potentially Microsoft Fabric.
Key Responsibilities
- Extract data from DynamoDB, CloudWatch and other AWS data sources
- Design and implement data transformation, cleansing and aggregation processes
- Develop efficient and reusable data pipelines using appropriate AWS services
- Create curated datasets and data models suitable for Power BI and Microsoft Fabric
- Support the development and maintenance of Power BI reports and dashboards
- Work with stakeholders to understand reporting requirements and translate them into appropriate data models and measures
Key Requirements
- Strong experience in data engineering and ETL development
- Hands-on experience with AWS, particularly DynamoDB and CloudWatch
- Experience working with Power BI
- Experience building and maintaining data pipelines
- Understanding of data quality, governance and security
Desirable
- Experience with Microsoft Fabric and/or Azure
- Experience with AWS services such as S3, Glue, Athena, Lambda and Firehose
- Experience working with operational/observability data from CloudWatch
- Understanding of GDPR, NHS information governance and data minimisation principles