All Jobs Vacancy

Aws Sre

Posted 5 days ago by Talent

About The Role

We are looking for an experienced AWS Site Reliability Engineer (SRE) to join our team on an initial 3-month contract.

You will be embedded within a high-impact team dedicated to ensuring the reliability, scalability, and performance of our AWS-hosted Data Platform.

If you live and breathe observability, love tearing down operational toil through automation, and know how to keep complex cloud ecosystems running smoothly, we want to hear from you.

Key Responsibilities

  • Define and Operationalise Reliability: Establish, refine, and operationalise SLIs, SLOs, and error budgets for critical data services, mapping them to the four golden signals (latency, errors, traffic, and saturation).
  • Observability Frameworks: Build and maintain comprehensive SLO dashboards and end-to-end monitoring (metrics, logs, and traces) utilizing Dynatrace and Prometheus.
  • Cloud and Container Management: Navigate the AWS ecosystem confidently, managing and optimizing containerized workloads deployed on Amazon EKS (Kubernetes).
  • Toil Reduction and Automation: Drive aggressive automation initiatives to eliminate repetitive operational tasks and streamline system efficiency.
  • Collaboration and Resilience: Partner closely with developers and architects to improve architecture reliability, contribute to continuous improvement backlogs, and lead root cause analysis (RCA) via blameless post-mortems.

Technical Skills And Experience Required Expertise

  • Strong background as an SRE or DevOps Engineer within an AWS environment.
  • Hands-on experience managing and scaling workloads on Amazon EKS (Kubernetes).
  • Proven track record with observability stacks, specifically Dynatrace and Prometheus.
  • Deep understanding of SRE principles, including error budgets, alerting thresholds, and full-stack tracing.
  • Excellent scripting/automation skills (e.g., Python, Bash, or Go).
  • Data Platform Experience: Prior exposure to data platforms, batch/streaming data pipelines (e.g., Kafka, Spark), and the unique challenges of data observability and workload reliability.
  • Active or recent SC Clearance.
Rate:
£375/day
Location:
London
IR35 Status:
Inside
Remote Status:
Remote
Industry:
IT
Seniority Level:
Not Specified

Take-Home Pay

£5,600 per month

Visit calculators for additional details

Create a free account to view the take-home pay for this contract

Share job