Overview
We are seeking an experienced Site Reliability Engineer with DV clearance to join our team on a contract basis, initially for six months with the possibility of extension. This role requires on-site presence four days per week and focuses on maintaining and improving system reliability through advanced engineering practices and tools.
Responsibilities
- Implement and manage modern configuration management tools such as Ansible or Chef.
- Utilize Terraform for infrastructure as code deployments.
- Manage Docker containers and relevant orchestration tools like Kubernetes or OpenShift.
- Maintain continuous integration and continuous deployment (CI/CD) pipelines using tools like Jenkins.
- Monitor systems and performance using tools such as InfluxDB, Prometheus, or Grafana.
- Integrate event-driven systems using message queuing technologies like RabbitMQ.
- Administer Linux environments and develop shell scripts for automation.
- Develop and maintain cloud hosting services, ideally with AWS components like EC2, RDS, S3, and Lambda.
Requirements
- Must hold live UKIC DV clearance.
- Experience with modern configuration management tools (e.g. Ansible, Chef).
- Proficiency with Terraform for infrastructure deployment.
- Experience with Docker and container orchestration technologies (e.g. Kubernetes, OpenShift).
- Familiarity with CI/CD tools, particularly Jenkins.
- Knowledge of monitoring tools such as InfluxDB, Prometheus, or Grafana.
- Understanding of relational databases and SQL.
- Experience with cloud services, especially AWS EC2, RDS, S3, and Lambda.