Contract

Site Reliability Engineer

TwinStream

TwinStream

Compensation Overview

£500 - £650/bi-wk

Gloucester, UK

Hybrid

Hybrid role: 3-4 days on-site in Cheltenham; 24/7 on-call rota.

UK Top Secret Clearance Required

Category
DevOps & Infrastructure (1)
Required Skills
RabbitMQ
Chef
Bash
Kubernetes
Grafana
OpenShift
InfluxDB
SQL
Docker
AWS
Prometheus
Jenkins
Terraform
Ansible
Linux/Unix
Requirements
  • Experience using modern configuration management tools (such as Ansible, Chef or similar)
  • Experience working with Terraform
  • Experience working with docker containers & container orchestration tools (such as Kubernetes, OpenShift or Docker Swarm)
  • Experience both using and maintaining CI / CD tools (such as Jenkins or similar)
  • Experience with monitoring tools such as InfluxDB, Prometheus or Grafana.
  • Experience of event-driven integration with MQ messaging (RabbitMQ or similar AMQP solution)
  • Good understanding of relational databases and SQL
  • Linux command line, administration and shell scripting
  • Working knowledge of network security protocols
  • Experience using, developing with and maintaining cloud hosting services (ideally AWS EC2, RDS, S3, Lambda)
Responsibilities
  • Designing and maintaining reliable, scalable physical and virtual infrastructure
  • Monitoring system performance and proactively resolving issues
  • Automating processes using tools such as Ansible to improve efficiency and consistency
  • Collaborating with engineers and stakeholders across the business
  • Supporting continuous improvement of systems, tools, and practices
  • Operating across the full infrastructure stack, from bare metal systems through to virtualised deployments and the applications running within them.
Desired Qualifications
  • Industry experience writing well-tested code in one of our platform languages (Java, Go, Python or similar)
  • Knowledge of cross domain principles & technologies
  • Experience of working in a service management environment
  • Practical applications of using observability patterns in previous systems
  • Creating and monitoring system availability metrics and using those to drive work that reduces downtime
  • Experience in Azure

Company Size

N/A

Company Stage

N/A

Total Funding

N/A

Headquarters

N/A

Founded

N/A