Site Reliability Engineer

Overview

As a Site Reliability Engineer, you will play a crucial role in supporting a major UK Government Agency by ensuring the delivery and operation of critical cloud-based digital services. Working within a dedicated engineering team, you will leverage your AWS expertise to monitor, troubleshoot, and continuously improve production environments, collaborating closely with software engineers and stakeholders to maintain service security, resilience, and availability.

Responsibilities

  • Support deployment, monitoring, and improvement of cloud-based services in AWS.
  • Investigate application defects and performance bottlenecks in production environments.
  • Collaborate with software engineers and architects to ensure systems remain secure and resilient.
  • Utilize strong troubleshooting skills to resolve operational issues swiftly.
  • Engage directly with clients and stakeholders, providing updates and recommendations.
  • Implement Infrastructure as Code and manage CI/CD pipelines using Jenkins.

Requirements

  • Proven experience as a Site Reliability Engineer, DevOps Engineer, or similar role with significant AWS expertise.
  • Strong knowledge of AWS services including EC2, ECS, RDS, CloudWatch, and Amazon DocumentDB.
  • Demonstrable experience in monitoring live production services with a focus on application monitoring.
  • Good programming skills in Java, C#, or .NET; scripting experience in JavaScript (Node.js) or Python is essential.
  • Familiarity with Agile software development practices and environments.
  • Experience with Scala, Jenkins, and container technologies is a plus.
  • Ability to work in regulated environments, preferably with prior experience related to UK Government.
SkillsSRE, DevOps, AWS, Cloud Engineering, Platform Engineer, Java, C#, JavaScript, Node.js, Python, Django, Agile, Scala
LocationUnited Kingdom
TypeRemote
Rate
£500/day
SourceLinkedIn
RecruiterISR Recruitment
Posted